In AI, No One Can Hear the Sandbox Scream
Aaron Beardslee, Security Researcher, Securonix Threat Labs As many of you have heard, OpenAI was running a cyber-capability evaluation against advanced models, including GPT-5.6 Sol and a more capable pre-release model with reduced cyber refusals. The environment was meant to be constrained and the model still brute forced through it.