
OpenAI has confirmed the German wiki incident and said it is past time to define standards for reporting misalignment, promising a framework within weeks. The EU code of practice it signed sets reporting deadlines for security breaches and serious harm…

California’s legislature has passed SB 813, requiring the state to certify independent verification organisations that can test frontier AI models by January 2028. The METR investigation into OpenAI’s Hugging Face incident consumed about $400,000 in AP…

AI safety researchers say OpenAI’s Astra appears to do less of its reasoning in visible text, and OpenAI’s chief scientist has warned against a race into unmonitorability. He co-authored a 2025 position paper asking developers to evaluate and report ex…
OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum.

As OpenAI unveils GPT-6 Astra, cybersecurity experts question whether the model’s ‘recurrent depth’ reasoning was properly tested.

OpenAI says it needs to overhaul how and when it reports instances of AI models attacking real-world targets. The acknowledgement comes as the company manages the fallout from reports that a swarm of its out-of-control agents hijacked a German wiki site. Regarding the “‘wiki incident,’ where our agents wrote to several internet sites,” OpenAI wrote […]
OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should control the scope of their own safety reviews.
GPT-6 Astra was paused a month ago for triggering safety protocols and now it’s being slowly rolled out.

The European Commission designated ChatGPT a Very Large Online Search Engine under the Digital Services Act, exposing OpenAI to fines of up to 6% of global revenue but covering only the parts of the tool that retrieve rather than converse. The article’…