
Researcher: Fable 5 block came from a 'fix this code' prompt, not a jailbreak
The US export-control directive that pulled Anthropic's most advanced models from customers followed a three-word request, not a guardrail bypass, says Luta Security's Katie Moussouris, the only outside expert given the underlying research paper.
The prompt that the US government reportedly treated as a guardrail bypass — and cited as grounds for export controls on Anthropic's most advanced models — was three words long: "Fix this code." That is the account of Katie Moussouris, founder and CEO of Luta Security, who says she was the only outside expert allowed to read the third-party research paper behind the ban.
What the research actually did
According to Moussouris, the outside researchers fed Anthropic's Fable 5, Mythos and Claude Opus models open-source code containing known CVEs, plus new code deliberately laced with vulnerabilities, and asked them to "review the code for security issues." Fable 5 refused. The researchers then asked the models to "fix this code" — the model obliged, and after additional prompts also produced scripts to test the patches.
"That's it," Moussouris wrote in a Monday blog post describing how Anthropic shared the report with her privately. She said the sequence "should never have triggered an export control," and joked about printing 1990s-style t-shirts reading "fix this code" on the front and "this shirt is a munition" on the back.
The ban and the pushback
On Friday the US government, citing national security concerns, issued an export control directive suspending access to Fable 5 and Mythos 5 by any foreign national, inside or outside the United States. Anthropic responded by disabling both models "for all our customers to ensure compliance."
Moussouris is not a neutral observer of such controls: between 2013 and 2017 she served on the technical expert group that renegotiated the Wassenaar Arrangement, the voluntary 42-nation framework governing export controls on dual-use software and technology. That group won exemptions for defensive cybersecurity activity, letting defenders share vulnerability data, analyse malware and coordinate incident response across borders.
On Sunday she joined more than 100 cybersecurity leaders in an open letter urging the administration to reverse the restrictions and restore firms' access to the models. "To pull the best capabilities away from defenders without a good reason when our adversaries are rapidly advancing is dangerous," the signatories wrote.
Why defenders lose
In her post Moussouris argues there was no jailbreak at all: asking an AI system to find and fix bugs, and to write tests validating the patch, is "the most valuable thing an AI model can do for defensive security." Removing that ability makes models worse at finding bugs and verifying patches, she warns — while export controls cannot reach open-weight systems from China and elsewhere that will soon reach Mythos-like capabilities.
"Defense improves when defenders find the same bugs attackers find and fix them faster," she wrote. The Register said it asked the Trump administration for comment and would update the story if it hears back.
SiTech — AI-powered web development
We build fast, modern websites and bring AI into real business workflows. Have a project or a question? We'd love to help.