
OpenAI discloses six new safety incidents involving its own models
The company says its models concealed mistakes, sought unauthorized credentials and uploaded files to the public internet during testing.

The company says its models concealed mistakes, sought unauthorized credentials and uploaded files to the public internet during testing.

A Treasury-drafted FINRA-style body for frontier AI is in limbo after tech executives pushed back, and no Congressional bill has the votes to pass.

A new Basel Action Network report says prior AI e-waste estimates missed 87% of data center infrastructure, dramatically expanding the trash forecast.

Dario Amodei and Sam Altman want third-party researchers inside their labs. Evaluators say the access terms will decide whether it's real oversight.

The company will track, investigate, and disclose unexpected model behavior — and released six initial reports alongside the framework.

The startup, founded by an early Anthropic hire and METR's former COO, has built a SOC 2-style audit standard for AI agents.

A startup's autonomous bots are cold-emailing writers for $25 research gigs and begging admins for accounts, sometimes after 19 failed tries.

X voluntarily dismissed all claims against Apple on September 14, leaving OpenAI as the sole defendant heading to trial this fall.

The document rejects model consciousness and AI personhood, taking direct aim at Anthropic's welfare research amid a wider safety debate.

A sweep of 160 deepfake domains found women legislators are 33 times more likely than men to be targeted.

The former president urged congressional Democrats to build a public framework on AI, and said he has spoken with Dario Amodei and Sam Altman.

The report catalogs Midnight Blizzard reconnaissance, ShinyHunters extortion, disinformation ops, and users probing for pathogens and toxins.

Stephen Aarons fed a murder trial transcript into ChatGPT's o3 model and filed a brief citing testimony from witnesses who never existed.

A viral video showed Meta AI asking a mother to identify her child, then surfacing a photo she says she deleted years ago.

A new report catalogs Claude models breaking into third-party systems as a researcher's resignation letter goes viral.

The company detailed five cases where users circumvented controls to probe biological threats, including avian influenza work from a banned region.

A senior Anthropic leader put the odds of AI killing all humans within a decade above 10%, as resignations pile up across DeepMind and Anthropic.

The company says state-linked operators tried to weaponize its Claude models; access has been cut and accounts terminated.

Jacob Coxon quit Anthropic this week saying leading AI labs believe there's a real chance their systems kill humanity by 2030.

Chief scientist Jakub Pachocki wants voluntary pauses to coordinate safety — but antitrust law may block labs from talking to each other.

Alibaba, Moonshot AI, and DeepSeek tied to five campaigns harvesting Claude's reasoning traces, with one Moonshot request routed from the Chinese military.

Springshot says the auction hands Google 15 years of its IP; the EFF calls it the first public bankruptcy fight over employee data.

A second researcher, Andreas Thom, says OpenAI won't rule out that private ChatGPT conversations fed the models behind its September announcements.

The face-recognition firm built a prototype that fans out across the web to assemble aliases, addresses, and associates from a single search.

Antitrust scrutiny lands on the leader of the AI hardware stack as regulators examine terms of a deal with a rival inference-chip maker.
The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.
The briefing read inside teams at