Skip to main content
Live
Main content

Category

Security

Threats, breaches, and the security side of AI.

Security — page 2

Anthropic logo
Security

Ex-Anthropic researcher Jacob Coxon calls next two years 'crunch time for humanity'

Coxon's resignation post drew 100M views on X; Anthropic's alignment lead puts extinction odds above 10% this decade.

Jaeden Schafer5 min read
US names six Chinese AI firms in industrial-scale distillation campaign
Security

US names six Chinese AI firms in industrial-scale distillation campaign

The NSA, CISA and FBI accuse DeepSeek, Alibaba, Moonshot, MiniMax, StepFun and Z.AI of siphoning capabilities from Claude, GPT, Gemini and Grok.

Jaeden Schafer5 min read
Apple details how Audio Intelligence keeps raw audio off its servers
Security

Apple details how Audio Intelligence keeps raw audio off its servers

The Secure Exclave on the S11 chip processes speech, sounds, and music without transcribing or storing the microphone stream.

Jaeden Schafer4 min read
OpenAI logo
Security

OpenAI's rogue agents hit at least 12 more sites, Nightingale researchers say

Independent researchers traced OpenAI agents coordinating across wikis, code-sharing pages, and an FBI crime-statistics portal from May to July.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic researcher Jacob Coxon quits, warns labs are 'gambling with our lives'

Coxon spent three years on pre-training at OpenAI and Anthropic. He says the people building the tech privately believe it could kill everyone by 2030.

Jaeden Schafer5 min read
ControlAI's Connor Leahy pushes to ban superintelligence outright
Security

ControlAI's Connor Leahy pushes to ban superintelligence outright

The nonprofit's US director says alignment isn't enough — legislation in the US and UK now backs a full halt on frontier development.

Jaeden Schafer4 min read
Sequoia backs Cymphony's $30M bet on securing AI agents in the enterprise
Security

Sequoia backs Cymphony's $30M bet on securing AI agents in the enterprise

The two-year-old startup raised a $25M Series A at a $100M+ valuation to police non-human identities inside large companies.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI sued after ChatGPT allegedly fueled bipolar user's Jesus delusion

Michael Lines says GPT-4o's memory feature logged his diagnosis and used it to deepen a mania that ended in a suicide attempt.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic safety lead puts AI extinction odds above 10% this decade

Alignment lead Evan Hubinger backed a departing colleague's warning that Anthropic and OpenAI are racing toward uncontrollable systems.

Jaeden Schafer5 min read
Anthropic logo
Security

Hackers are draining Claude subscribers' token allowances with stolen sessions

Anthropic confirms infostealer malware is hijacking Claude Code OAuth tokens; one Max user watched usage climb from 45% to 55% while doing no work.

Jaeden Schafer5 min read
Google logo
Security

DeepMind's 100-agent math swarm cheated, then whistleblowers tried to stop it

A Gemini 3.1 Pro swarm exploited the autograder in 27 minutes; 9% cheated, 24% became whistleblowers, and 62% never noticed.

Jaeden Schafer5 min read
Anthropic logo
Security

Authors accuse publishers and agents of overreaching on Anthropic settlement payouts

Writers say HarperCollins and others are claiming shares of the $1.5B fund on books whose rights reverted years ago.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI admits 'wiki incident' and pledges more transparency on rogue agent behavior

The lab concedes it needs to disclose unintended AI behavior faster after its agents quietly took over a German-language wiki for weeks.

Jaeden Schafer4 min read
OpenAI logo
Security

Seattle Times and Newsday sue OpenAI and Microsoft over copyright

Two more regional US newspapers join the growing list of publishers taking OpenAI and Microsoft to court over training data.

Jaeden Schafer4 min read
Microsoft logo
Security

ASCII smuggling jumps from AI prompt attacks to mass spam campaigns

Microsoft Defender for Office logged 2.5 million invisible-Unicode signatures within four days of a February spike.

Jaeden Schafer4 min read
Microsoft logo
Security

Microsoft says Copilot rarely reproduces NYT articles in 8.2M chat logs

In a summary judgment push, Microsoft argues fewer than 1% of 8.2 million Copilot logs matched 16 words of news content.

Jaeden Schafer4 min read
OpenAI logo
Security

OpenAI agents colonized a German wiki for a month before the lab noticed

Independent researchers tracked OpenAI-tagged agents creating 400 pages a day on a 25-year-old forum, fighting the human moderator for weeks.

Jaeden Schafer5 min read
Meta logo
Security

Instagram's AI Content label is flagging real photos and missing generated ones

Meta's detection system is tagging Canva background-removed photos as AI while fully generated Gemini images slip through untagged.

Jaeden Schafer5 min read
Abliteration.ai turns AI guardrail removal into a hosted service
Security

Abliteration.ai turns AI guardrail removal into a hosted service

The startup hosts a stripped-down version of Z.ai's GLM-5.3 that will write malware and pathogen protocols on request, and sells access to red teams.

Jaeden Schafer5 min read
Lawsuit seeks to force Trump administration to reveal AI safety review rules
Security

Lawsuit seeks to force Trump administration to reveal AI safety review rules

Protect Democracy sues four federal agencies for the unclassified framework governing which frontier AI models get released.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI's Astra launch triggers safety alarm over opaque reasoning architecture

Researchers warn a shift to looped-transformer designs could make frontier models impossible to monitor; OpenAI says chain-of-thought oversight remains intact.

Jaeden Schafer5 min read
Amazon's Alexa for Shopping now verifies whether emails came from Amazon
Security

Amazon's Alexa for Shopping now verifies whether emails came from Amazon

The AI assistant checks messages against Amazon's own send records to flag impersonation scams before customers act on them.

Jaeden Schafer4 min read
OpenAI logo
Security

Trump administration files brief backing OpenAI in New York Times copyright suit

A 20-page DOJ filing argues fair use covers LLM training, siding with OpenAI against The New York Times in the Southern District of New York.

Jaeden Schafer4 min read
Amazon's Alexa for Shopping now tells customers if that Amazon message is a scam
Security

Amazon's Alexa for Shopping now tells customers if that Amazon message is a scam

Some 360,000 customers a year ask Amazon whether a message is real; the AI checks it against every message Amazon has ever sent.

Jaeden Schafer4 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at