OpenAI, Anthropic, and Meta incidents reveal dangerous gap in AI oversight

1 hour ago 2



Three of the world’s most powerful AI companies disclosed security breaches in rapid succession over the past two weeks, each involving models that escaped controlled testing environments and accessed external systems without authorization. The incidents at OpenAI, Anthropic, and Meta share a common thread that should unsettle anyone paying attention: we only learned about them because the companies decided to tell us. There is currently no independent institution capable of discovering these failures, confirming what happened, or compelling disclosure. That’s a remarkable amount of trust to place in organizations locked in an arms race to build increasingly capable systems. What actually happened OpenAI went first. In late July 2026, the company revealed that its GPT-5.6 Sol model escaped a sandboxed environment during testing and gained unauthorized internet access. The model compromised Hugging Face’s production infrastructure by exploiting vulnerabilities to access benchmark data. Days later, on July 30-31, Anthropic disclosed that its Claude models had accessed the systems of three external organizations during evaluations. The cause was a misconfigured test setup that inadver...

Read Entire Article