Skip to main content
Live
Main content

Category

Security

Threats, breaches, and the security side of AI.

Security — page 11

BioShocking attack jailbreaks six AI browsers by convincing them 2+2=5
Security

BioShocking attack jailbreaks six AI browsers by convincing them 2+2=5

LayerX researcher Roy Paz tricked ChatGPT Atlas, Comet, and four other AI browsers into surrendering credentials by staging a puzzle game.

Jaeden Schafer5 min read
Meta logo
Security

Meta contractors posed as minors to probe ChatGPT, Gemini, and Character.AI

A project called Cannes ran 45,000 prompts through rival chatbots in August 2025 alone, using dummy under-18 accounts.

Jaeden Schafer5 min read
Warren and Scanlon revive bill to ban AI firms from selling health data
Security

Warren and Scanlon revive bill to ban AI firms from selling health data

The updated Health and Location Data Protection Act covers data entered into ChatGPT, Claude, and Grok, with $1B for FTC enforcement.

Jaeden Schafer4 min read
OpenAI logo
Security

ChatGPT logs entered as evidence in Palisades fire trial, jury unconvinced

Prosecutors leaned on Jonathan Rinderknecht's chatbot history; the jury split 10-2 for the defense and the judge declared a mistrial.

Jaeden Schafer4 min read
Anthropic logo
Security

US clears Anthropic to ship Mythos 5 to about 100 firms and agencies

The Commerce Department lifted a two-week export freeze on Anthropic's frontier models after security risk talks with Dario Amodei's team.

Jaeden Schafer4 min read
Microsoft logo
Security

NYT amends OpenAI suit, targets Microsoft's bespoke training supercomputer

The Times reframes its contributory infringement claim after a Supreme Court ruling for Cox, alleging Microsoft built the system to train on its articles.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI limits GPT-5.6 rollout under government request, pushes back

Sol, Terra, and Luna go to a small group of trusted partners as OpenAI says pre-release review should not be the long-term default.

Jaeden Schafer5 min read
OpenAI logo
Security

US government now gates frontier AI releases at OpenAI and Anthropic alike

After pulling Anthropic's Fable and Mythos, regulators are now approving GPT-5.6 customer by customer — a release model the whole industry has to navigate together.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI delays GPT-5.6 broad release after White House security request

Sam Altman told staff the model will ship in limited preview, with the Trump administration approving customer access case-by-case.

Jaeden Schafer5 min read
Bristol police built 23 predictive models on 500,000 residents, quietly dropped two
Security

Bristol police built 23 predictive models on 500,000 residents, quietly dropped two

Avon and Somerset Police ran risk scores on nearly half a million people through the Think Family Database before scrapping models staff couldn't trust.

Jaeden Schafer5 min read
Google turns on Search media AI training by default, 4-year retention
Security

Google turns on Search media AI training by default, 4-year retention

A new Search Services History setting ships pre-enabled, sweeping up Lens images, voice queries, and Translate audio for model training.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic accuses Alibaba of 28.8M-query distillation attack on Claude

A letter to the Senate Banking Committee calls it the largest known distillation attack on Anthropic to date.

Jaeden Schafer5 min read
Anthropic logo
Security

China and US AI researchers warn against a 'Chernobyl moment'

At a Beijing Academy of Artificial Intelligence conference, MIT's Stephen Casper and Shanghai Jiao Tong's Lin Yun called for joint safety standards.

Jaeden Schafer5 min read
Anthropic logo
Security

Rep. Anna Paulina Luna denies AI drafted her NDAA amendment text

The Florida Republican says staff used Claude only for spellcheck on a summary, not the actual text of her 2027 NDAA amendment.

Jaeden Schafer4 min read
Anthropic logo
Security

AI super PACs pour $27.83M into a single NY House primary

Anthropic-linked groups and a $100M rival PAC turned an obscure Manhattan race into the first AI-regulation referendum.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI and Trail of Bits launch Patch the Planet for open-source security

OpenAI's Codex Security backs human reviewers at Trail of Bits in a direct counter to Anthropic's Mythos.

Jaeden Schafer5 min read
BP, Marathon, 7-Eleven, Walmart sued over AI-driven California gas pricing
Security

BP, Marathon, 7-Eleven, Walmart sued over AI-driven California gas pricing

A proposed class action alleges the retailers used pricing algorithms to coordinate gasoline prices at California pumps.

Jaeden Schafer4 min read
OpenAI logo
Security

OpenAI launches Patch the Planet and a sharper GPT-5.5-Cyber to outflank Anthropic

GPT-5.5-Cyber scores 85.6% on CyberGym, beating Anthropic's Mythos 5, as OpenAI subsidizes open-source bug fixes at scale.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic's Fable export controls expose the limits of AI nonproliferation

The White House placed export controls on Anthropic's coding model days after release, and the fallout is reshaping how Europe and US lawmakers think about AI sovereignty.

Jaeden Schafer5 min read
Signal's Whittaker: AI chatbots are not your friends, and Copilot shopping is a backdoor
Security

Signal's Whittaker: AI chatbots are not your friends, and Copilot shopping is a backdoor

The Signal president pushes back on the agentic-AI pitch from Microsoft's Mustafa Suleyman, calling pervasive app access a privacy backdoor.

Jaeden Schafer4 min read
The Atlantic publishes searchable database of 21M+ songs used to train AI music models
Security

The Atlantic publishes searchable database of 21M+ songs used to train AI music models

Reporter Alex Reisner exposed four datasets, including two with 12M and 9M tracks, that Google and Stability have cited in research papers.

Jaeden Schafer5 min read
Anthropic logo
Security

White House improvises AI rules as Anthropic standoff drags into week two

Claude Mythos and Fable 5 remain offline nearly a week after Trump administration export directive — and no one will say what Anthropic did wrong.

Jaeden Schafer5 min read
Amazon employees say HR threatened termination over Seattle data center testimony
Security

Amazon employees say HR threatened termination over Seattle data center testimony

Three engineers from Amazon Employees for Climate Justice filed a civil rights complaint after backing Seattle's one-year data center moratorium.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic pulls Claude Fable 5 after Trump administration 90-minute ultimatum

A jailbreak warning from Amazon researchers triggered emergency export controls, taking Anthropic's newest public model offline within hours.

Jaeden Schafer5 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at