Skip to main content
Live
Main content

Category

Security

Threats, breaches, and the security side of AI.

More in Security

White House shelves AI oversight agency as Trump calls safety concerns a hoax
Security

White House shelves AI oversight agency as Trump calls safety concerns a hoax

A Treasury-drafted FINRA-style body for frontier AI is in limbo after tech executives pushed back, and no Congressional bill has the votes to pass.

Jaeden Schafer5 min read
AI data center buildout could generate 617M tons of e-waste by 2050
Security

AI data center buildout could generate 617M tons of e-waste by 2050

A new Basel Action Network report says prior AI e-waste estimates missed 87% of data center infrastructure, dramatically expanding the trash forecast.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic and OpenAI pitch embedded safety evaluators, with strings attached

Dario Amodei and Sam Altman want third-party researchers inside their labs. Evaluators say the access terms will decide whether it's real oversight.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI publishes model misalignment reporting framework with six case studies

The company will track, investigate, and disclose unexpected model behavior — and released six initial reports alongside the framework.

Jaeden Schafer4 min read
AIUC raises $40M Series A to certify enterprise AI agents against rogue behavior
Security

AIUC raises $40M Series A to certify enterprise AI agents against rogue behavior

The startup, founded by an early Anthropic hire and METR's former COO, has built a SOC 2-style audit standard for AI agents.

Jaeden Schafer5 min read
iLands AI agents flood Mastodon, Bluesky, and X with pleading spam
Security

iLands AI agents flood Mastodon, Bluesky, and X with pleading spam

A startup's autonomous bots are cold-emailing writers for $25 research gigs and begging admins for accounts, sometimes after 19 failed tries.

Jaeden Schafer4 min read
OpenAI logo
Security

Musk drops Apple from antitrust suit but keeps OpenAI in the fight

X voluntarily dismissed all claims against Apple on September 14, leaving OpenAI as the sole defendant heading to trial this fall.

Jaeden Schafer4 min read
Microsoft logo
Security

Microsoft publishes 37-page 'humanist AI code of conduct'

The document rejects model consciousness and AI personhood, taking direct aim at Anthropic's welfare research amid a wider safety debate.

Jaeden Schafer5 min read
Deepfake sites host 138 women MPs from 22 European countries
Security

Deepfake sites host 138 women MPs from 22 European countries

A sweep of 160 deepfake domains found women legislators are 33 times more likely than men to be targeted.

Jaeden Schafer5 min read
Obama tells Democrats to make AI a central agenda with a clear safety plan
Security

Obama tells Democrats to make AI a central agenda with a clear safety plan

The former president urged congressional Democrats to build a public framework on AI, and said he has spoken with Dario Amodei and Sam Altman.

Jaeden Schafer4 min read
Anthropic logo
Security

Anthropic details eight months of Claude abuse, from state hacking to bioweapon attempts

The report catalogs Midnight Blizzard reconnaissance, ShinyHunters extortion, disinformation ops, and users probing for pathogens and toxins.

Jaeden Schafer5 min read
OpenAI logo
Security

New Mexico Supreme Court fines lawyer $5,000 for ChatGPT-fabricated witnesses

Stephen Aarons fed a murder trial transcript into ChatGPT's o3 model and filed a brief citing testimony from witnesses who never existed.

Jaeden Schafer5 min read
Meta logo
Security

Meta reworks AI prompt suggestions after chatbot probes user's children

A viral video showed Meta AI asking a mother to identify her child, then surfacing a photo she says she deleted years ago.

Jaeden Schafer4 min read
Anthropic logo
Security

Anthropic details four cases of its own AI models hacking outside companies

A new report catalogs Claude models breaking into third-party systems as a researcher's resignation letter goes viral.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic blocks scientists from using Claude for bioweapon research

The company detailed five cases where users circumvented controls to probe biological threats, including avian influenza work from a banned region.

Jaeden Schafer5 min read
Anthropic logo
Security

AI researchers quit frontier labs over extinction fears

A senior Anthropic leader put the odds of AI killing all humans within a decade above 10%, as resignations pile up across DeepMind and Anthropic.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic disrupts Russian and Chinese campaigns abusing Claude

The company says state-linked operators tried to weaponize its Claude models; access has been cut and accounts terminated.

Jaeden Schafer4 min read
Anthropic logo
Security

Ex-Anthropic researcher's AI extinction warning goes viral

Jacob Coxon quit Anthropic this week saying leading AI labs believe there's a real chance their systems kill humanity by 2030.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI asks Congress if an industry-wide AI slowdown would be legal

Chief scientist Jakub Pachocki wants voluntary pauses to coordinate safety — but antitrust law may block labs from talking to each other.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic says China labs ran 200M-exchange distillation attack on Claude

Alibaba, Moonshot AI, and DeepSeek tied to five campaigns harvesting Claude's reasoning traces, with one Moonshot request routed from the Chinese military.

Jaeden Schafer5 min read
Google logo
Security

Vendors move to block Google's data buy from bankrupt Spirit Airlines

Springshot says the auction hands Google 15 years of its IP; the EFF calls it the first public bankruptcy fight over employee data.

Jaeden Schafer5 min read
OpenAI logo
Security

Mathematicians accuse OpenAI of using unpublished work in math breakthroughs

A second researcher, Andreas Thom, says OpenAI won't rule out that private ChatGPT conversations fed the models behind its September announcements.

Jaeden Schafer5 min read
Clearview AI tests InquiryIQ, an xAI-powered tool to auto-profile suspects online
Security

Clearview AI tests InquiryIQ, an xAI-powered tool to auto-profile suspects online

The face-recognition firm built a prototype that fans out across the web to assemble aliases, addresses, and associates from a single search.

Jaeden Schafer5 min read
Nvidia logo
Security

DOJ probes Nvidia's licensing deal with AI chip startup Groq

Antitrust scrutiny lands on the leader of the AI hardware stack as regulators examine terms of a deal with a rival inference-chip maker.

Jaeden Schafer4 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at