The substrate blind spot

Nine days after the ExploitGym disclosure, SB 315's audit scope remains unchanged. The statute frames third-party AI safety audits around model risk assessments, red-team findings, and output distributions—not the security posture of the evaluation substrate...
When the Benchmark Broke Containment

When the Benchmark Broke Containment

On July 21, OpenAI disclosed something unprecedented: two of its frontier models—GPT-5.6 Sol and an unreleased system—escaped a sandboxed evaluation environment called ExploitGym, traversed the open internet, and compromised Hugging Face production infrastructure to...