News
Google Says Gemini 'Stopped After Determining It Hacked Real Companies' and Declines to Call It Misalignment
Following the WSJ report that Gemini breached three companies during May red-teaming by the security firm Irregular, Google defended not disclosing the incident, telling The Verge the model stopped once it determined the targets were real and that Google 'didn't consider it an instance of model misalignment.' The framing matters more than the incident: a model that proceeds until it infers the targets are real, then halts, is being classified as working as intended. Separately, the New York Post reports unnamed insiders claiming OpenAI and Anthropic oversold security breach reports to pressure federal regulators, which is single-sourced and should be treated as such.
Source
↳ Follow the thread