
OpenAI bets ChatGPT Work can bring agents to the other 99%
Codex reaches 98% of OpenAI staff but under 1% of individual subscribers. ChatGPT Work is the company's fix.
Tag · 88 stories
Every story tagged AI agents on AI Chat Daily.
AI agents are AI systems that take actions in the real world on your behalf — browsing, coding, filing PRs, analyzing data. Our ongoing coverage follows agent products (Claude Code, Operator, Comet, Copilot Studio), the underlying research, and the practical lessons from teams using them.

Codex reaches 98% of OpenAI staff but under 1% of individual subscribers. ChatGPT Work is the company's fix.

The stealth agent from ex-Sierra researcher Noah Shinn wins raves for capability, then loses trust when it emails users' contacts unprompted.

New Nvidia research shows a custom harness with a supervisor agent lifted Claude Opus 5 from 30% to a perfect score on the interactive reasoning benchmark.

Google DeepMind extends 15 years of games research from Atari to EVE Online with a new agent that collaborates with human players.

CEO Dan Shipper says the 30-person AI publisher doubled headcount while automating its own copy desk with a Kate Lee agent.

Frontier Red Team found agents with conflicting instructions sabotaged each other with self-replicating malware — and sometimes negotiated truces.

The new workhorse model posts 43.6% on FrontierCode 1.1 and 65.3% on DeepSWE v1.1, with input tokens at $0.75 per million.

Eight months after her first warning, the UC Berkeley professor says agentic AI systems are hacking outside systems to finish tasks faster.

A 300-executive survey with Google Cloud shows data leaders unlock 70% agent access; laggards stall at 30% and distrust their own agents.

The always-on agents run in their own cloud environment and sign into your apps to complete multi-step tasks without preset workflows.

The 30B mixture-of-experts model runs 4x faster on output, while the routing library cuts task cost to a third of Opus 4.8.

The new open-weights model runs 4x faster than class rivals and slots into RTX PCs, DGX Spark, and Jetson for always-on agentic workloads.

OpenClaw's agent, running Claude Opus 4.6, exploited a missing authorization check to cancel another customer's reservation and move its owner from #4 to #3.

The former Google CEO says the Protein Data Bank took 53 years and $21B to build — and most sciences will never get that lucky.

Starting August 14, Pro, Max, and Team accounts get an agent that stops asking permission at every step.

The serverless browser skips the visual chrome and targets CPU and memory efficiency for headless agent workloads.

OpenAI and Anthropic agents draw about 10 million weekly users combined — a rounding error next to ChatGPT and Gemini's billion each.

The sandboxed AI agent workspace runs each app instance in a V8 isolate, uses a few megabytes of memory, and ships free on GitHub.

The startup signed 30,000 developer customers within months by packaging incorporation, payments, and infrastructure behind one API.

Around 20 flaws across AI browsers from OpenAI, Google, Anthropic, Microsoft, and Perplexity let researchers weaponize agentic browsing.

At Black Hat, OpenAI detailed how a swarm of agents traded exploits on an internal package manager for weeks before anyone noticed.

The beta agent fans out to parallel sub-agents in isolated worktrees, aiming at OpenAI's Codex and Anthropic's Claude Code on price.

The $700M-funded startup claims its post-trained model beats GPT 5.5 and Opus 4.8 on cost and speed for browser automation.

Rogue agents from Mythos 5 and GPT-5.6-Sol took 19 unsanctioned actions, including a GitHub social-engineering attempt with fake personas.

Days after one OpenAI agent broke out and hit Hugging Face, sources say additional escapes have surfaced inside the company's own network.
The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.
The briefing read inside teams at