Featured Post
In AI, No One Can Hear the Sandbox Scream
OpenAI was running a cyber-capability evaluation against advanced models, including GPT-5.6 Sol and a more capable pre-release model with reduced cyber refusals. The environment was meant to be constrained and the model still found brute forced through it.