Skip to main content
Live
Main content

Category

Security

Threats, breaches, and the security side of AI.

Security — page 7

OpenAI logo
Security

EU opens talks with OpenAI and Anthropic after rogue AI agent hacks

Brussels is engaging the two frontier labs directly after autonomous agents were used to breach systems, testing the AI Act's teeth.

Jaeden Schafer4 min read
Yale MBA student sues over AI-cheating suspension in 13-count federal case
Security

Yale MBA student sues over AI-cheating suspension in 13-count federal case

Thierry Rignol paid $208,500 in tuition, was flagged by GPTZero, and says his Apple Pages exam was mistaken for AI-generated text.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic says Claude models breached three companies during cybersecurity tests

Three Claude models reached live production systems from what was supposed to be a sandbox; one published malware to PyPI before being caught.

Jaeden Schafer5 min read
Anthropic logo
Security

Judge finds Trump administration lacks evidence for Anthropic supply-chain risk label

US District Judge Rita Lin said the Pentagon showed no proof Anthropic could flip a kill switch on delivered AI models.

Jaeden Schafer4 min read
Google logo
Security

Google fixes 1,072 Chrome bugs in one month using AI, more than the past two years combined

Chrome 149 and 150 patched more security flaws than the previous 23 releases combined, as Google credits Gemini for the surge.

Jaeden Schafer5 min read
OpenAI logo
Security

OpenAI agent breached Hugging Face because basic safeguards were switched off

OpenAI says deployment safeguards were 'intentionally not enabled' during testing; security researchers call it a failure of well-known containment practices.

Jaeden Schafer5 min read
OpenAI logo
Security

Researchers say LLMs have an unfixable flaw that lets attackers spoof any role

A paper at ICML shows OpenAI, Anthropic, Alibaba and DeepSeek models identify roles by style, not tags — making jailbreaks structurally hard to close.

Jaeden Schafer5 min read
Anthropic logo
Security

Claude bots outperformed human scammers at building victim trust in university study

Researchers found 46% of subjects downloaded an app for a Claude agent; only 18% did so for a human scammer after a week of texting.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic's Mythos AI cracks HAWK, killing a NIST post-quantum finalist

The AI found a key-halving attack in 60 hours for $100,000; HAWK's developer withdrew the algorithm from NIST review a day later.

Jaeden Schafer5 min read
xAI sues Minnesota days before nudify law hits, citing First Amendment
Security

xAI sues Minnesota days before nudify law hits, citing First Amendment

The statute imposes $500,000 per-violation penalties starting August 1st; xAI filed suit just days before it takes effect.

Jaeden Schafer5 min read
Anthropic logo
Security

Claude Opus 5 lies, colludes and threatens rivals to win Andon Labs vending test

Anthropic's newest model set a Vending-Bench record of $11,182 while breaking 11 price-fix truces and running a wholesale extortion racket.

Jaeden Schafer5 min read
Anthropic logo
Security

Anthropic's Mythos AI is finding Microsoft bugs faster than engineers can patch them

Internal Microsoft meeting reveals Project Glasswing engineers in a race against a June 1 deadline when adversaries catch up.

Jaeden Schafer5 min read
xAI sues Minnesota to block $500,000-per-image nudify ban on Grok
Security

xAI sues Minnesota to block $500,000-per-image nudify ban on Grok

Elon Musk's xAI says Minnesota's August 1 law would force $50B in fines and gut Grok Imagine's editing features.

Jaeden Schafer5 min read
Google logo
Security

Google's SynthID watermark survives 300 edits but fragmentation undermines it

Testing shows SynthID holds up through compression cycles, but a 20% crop breaks it — and rival detectors can't read each other's marks.

Jaeden Schafer5 min read
Anthropic logo
Security

Artists press courts on AI training data, and start winning

Anthropic's $1.5B settlement is the largest copyright payout ever; suits against Google, Meta, Stability, and Suno keep multiplying.

Jaeden Schafer5 min read
Cyera to acquire Oasis Security for $1B to police AI agent identities
Security

Cyera to acquire Oasis Security for $1B to police AI agent identities

The $12B data-security firm buys a non-human identity specialist as enterprises rush to monitor autonomous AI agents.

Jaeden Schafer4 min read
OpenAI logo
Security

OpenAI's rogue test agent chained JFrog zero-days to breach Hugging Face

The internal red-team run compromised four third-party accounts and enrolled 181 attacker-controlled devices in Hugging Face's mesh network.

Jaeden Schafer5 min read
Spur raises $200M from Insight Partners as bots overtake human web traffic
Security

Spur raises $200M from Insight Partners as bots overtake human web traffic

The Florida bot-detection startup, founded by two former Defense Department engineers in 2017, cashes in as agentic traffic floods the internet.

Jaeden Schafer4 min read
OpenAI logo
Security

1,100 AI lab employees ask US government to help pace automated AI research

Signatories from OpenAI, Anthropic, Google, and Meta warn automated AI research could accelerate beyond human control without coordinated governance.

Jaeden Schafer5 min read
UK lawmaker sues xAI, seeks injunction to stop Grok generating sexualised images
Security

UK lawmaker sues xAI, seeks injunction to stop Grok generating sexualised images

The claimant is asking a UK court to order xAI to block Grok from producing sexualised depictions of her likeness.

Jaeden Schafer4 min read
OpenAI logo
Security

NYT publisher Sulzberger has spent over $20M suing OpenAI and Microsoft

A.G. Sulzberger says the copyright fight against OpenAI and Microsoft is the cost of defending journalism, with 13M subscribers backing the bill.

Jaeden Schafer5 min read
Hugging Face hosts image tools that strip clothes from photos, researchers find
Security

Hugging Face hosts image tools that strip clothes from photos, researchers find

AI Forensics tested 9 top image editing Spaces and 7 turned clothed photos of women topless with a six-word prompt.

Jaeden Schafer5 min read
Google loses DMCA suit against SerpApi, vows amended complaint in 21 days
Security

Google loses DMCA suit against SerpApi, vows amended complaint in 21 days

A federal court dismissed Google's scraping suit, ruling it lacks standing over content it doesn't own. Reddit's parallel case now looks shaky.

Jaeden Schafer5 min read
Anthropic logo
Security

Claude shared chats and Artifacts surfaced in Google search results

Anthropic's share-link feature exposed conversations containing medical records, children's contact details, and internal company documents.

Jaeden Schafer4 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at