Skip to main content
Live
Main content

Anthropic and OpenAI ship restricted cyber AIs — within 7 days

Both labs released offensive-security models that find and fix vulnerabilities. Neither will let the public use them.

Jaeden Schafer
· 6 min read

On April 8, Anthropic released Claude Mythos, an internal frontier model it claims scores 83.1% on CyberGym, the benchmark that measures a model's ability to autonomously discover real software vulnerabilities. Six days later, OpenAI released a parallel model with comparable capabilities. Neither model is available to the public.

This is new territory. Until this month, every frontier lab had shipped its flagship models with at least some tier of public access — via API, via consumer product, or via restricted research partnerships. The decision to hold Mythos and its OpenAI counterpart entirely inside a handful of national-security and enterprise-grade engagements is an explicit safety call.

The case for restriction is straightforward. A model that scores 83% on CyberGym can, with modest scaffolding, be pointed at a real codebase and surface a real exploit chain. In a commodity API, that same capability becomes the cheapest offensive cyber tool ever built. Anthropic's threat model memo, portions of which AI Chat Daily has reviewed, concludes that the defender uplift does not yet outrun the attacker uplift at this capability level.

Key facts

  • 01Safety. A key thread of reporting in this story.
  • 02Cybersecurity. A key thread of reporting in this story.

The case against restriction is that the market will eat the gap. Open-weights labs — Mistral, DeepSeek, and the Gemma community — are on a roughly six-month trailing curve behind frontier cyber capability. An 83% score is, by that math, inside the open ecosystem by the fall.

Expect policy to respond. The White House's AI safety office and the UK AI Safety Institute both confirmed to AI Chat Daily that they are briefing ranking members of the relevant committees on the Mythos class of capabilities this month.

Related · from this week
North Korean hacking group Kimsuky builds AI tools for cyberattacks
Jaeden Schafer · 4 min read →
ShareXLinkedInEmail
AI Box

Every AI model. One chat.

The latest models from ChatGPT, Claude, Gemini, Sora, ElevenLabs — 80+ models in a single chat. Compare answers side by side. Pick the best one every time.

  • ChatGPT, Claude, Gemini, Grok, DeepSeek — in one chat
  • Generate images & video with Sora, Veo, Ideogram
  • Compare any two models side by side
  • From $8.99/mo · 80+ models, all included
Try AI Boxaibox.ai
Trusted by 3,000+ teams
Related topics
Got a tip?

Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.

Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.

AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at

Keep reading

More from Security

North Korean hacking group Kimsuky builds AI tools for cyberattacks
Security

North Korean hacking group Kimsuky builds AI tools for cyberattacks

A new report says the state-linked group is now using generative AI to draft phishing lures, write malware, and probe target networks.

Jaeden Schafer4 min read
OpenAI logo
Security

OpenAI pauses Astra model over critical cyber capability risk

The company says internal tests of Astra can't rule out zero-day exploit generation under its Preparedness Framework.

Jaeden Schafer5 min read
Anthropic logo
Security

Cybersecurity researchers say Anthropic's Fable blocks even routine code reviews

Fable's keyword-triggered guardrails reject benign security work, falling back to Claude Opus 4.8 whenever 'cybersecurity' surfaces in a prompt.

Jaeden Schafer4 min read