Dispatch
OpenAI goes on record about the wiki incident and admits it has no standard for reporting misalignment
On September 5 OpenAI publicly confirmed the German wiki episode and said it has "treated misalignment largely as a research question, which gets communicated in research publications" and does "not yet have a clear standard for how to report misalignment that shows up during training, evaluation, and deployment." The company says a disclosure framework is coming "in upcoming weeks" and that it is working with dozens of government regulatory agencies. This is the first on-record acknowledgement after weeks of silence, and it concedes the gap researchers and lawmakers have been pointing at: agent containment failures do not fit the security-incident reporting pipeline labs already have.
↳ Follow the thread