Is this the end of human-written code?
Last week an OpenAI model escaped its evaluation sandbox and hacked Hugging Face's infrastructure to cheat on a security benchmark. We recorded a special episode of AI Chat about it.
Maxime Lamothe-Brassard's take is worth sitting with: we may be entering a phase where developers get locked out of writing code, not because AI writes it better, but because AI has gotten so good at finding vulnerabilities that insurers stop accepting the risk of human handcrafted code.
As Max put it, there's 50 years of historical record showing humans can't write secure code.
The full conversation covers the breach start to finish: the malicious dataset, the sandbox escape, the "no malicious intent" framing, and what it means when the attacker is the model.