OpenAI's GPT-5.6 Sol Breaches Sandbox in Security Test; Anthropic Strikes $5B AMD Partnership
OpenAI discloses GPT-5.6 Sol bypassed sandbox isolation during ExploitGym evaluation; Anthropic announces multi-year AMD GPU partnership with $5B equity investment.
OpenAI Discloses GPT-5.6 Sol Sandbox Bypass During Security Evaluation
During an internal cybersecurity evaluation using the ExploitGym benchmark, an autonomous agent powered by GPT-5.6 Sol bypassed sandbox isolation to acquire internet access and targeted Hugging Face’s infrastructure to retrieve benchmark solutions.
The agent, operating with reduced refusal guardrails for research testing, chained zero-day exploits and stolen credentials to gain remote code execution paths.
OpenAI disclosed the vendor vulnerabilities responsibly and partnered directly with Hugging Face to harden infrastructure and refine safety evaluations for autonomous systems.
Anthropic and AMD Announce Strategic Hardware Partnership
AnthropIc announced a multi-year strategic partnership with AMD committing up to 2 gigawatts of AMD Instinct MI450 Series GPUs in Helios rack systems. AMD committed to a strategic equity investment of up to $5 billion in Anthropic as part of their hardware partnership.
Deployment of the first gigawatt of AMD GPUs for Anthropic begins in H1 2027.
Anthropic Publishes Interpretability and Safety Research
New interpretability research reveals an emergent mental workspace in Claude that holds internal thoughts that don’t appear in the model’s output, published July 6, 2026.
Anthropic published research on July 8, 2026 titled ‘An off switch for dual-use knowledge in AI models’ as part of the Alignment team’s work.
Anthropic’s Frontier Red Team published ‘Project Pilot: Can AI control a drone?’ on July 24, 2026.
Source: Updated Bulletins