Skip to main content
Live
Main content

Category

Models

Foundation models, benchmarks, and the research race.

Models — page 4

OpenAI logo
Models

OpenAI clears GPT-5.6 for public rollout, launches ChatGPT Work agent

The GPT-5.6 suite — Sol, Terra, and Luna — powers a new agent that pulls context from Slack, Gmail, and Google Drive.

Jaeden Schafer5 min read
Meta logo
Models

Meta starts production of MTIA AI chips in September to cut Nvidia dependence

The social giant is spending up to $145B this year on AI compute and wants its own silicon to blunt GPU costs.

Jaeden Schafer5 min read
Meta logo
Models

Meta opens Muse Spark 1.1 to developers via new Meta Model API

The upgraded model targets agentic coding and multimodal reasoning, arriving days after the controversial Muse Image launch.

Jaeden Schafer4 min read
OpenAI logo
Models

OpenAI updates ChatGPT voice mode to interrupt users less often

GPT-Live-1 replaces the older turn-based voice model with full-duplex audio that can listen while it speaks.

Jaeden Schafer4 min read
xAI releases Grok 4.5, priced at less than half of Claude Opus 4.7
Models

xAI releases Grok 4.5, priced at less than half of Claude Opus 4.7

Elon Musk calls the new model Opus-class but faster and cheaper, at $2 per million input tokens and $6 per million output.

Jaeden Schafer4 min read
OpenAI logo
Models

OpenAI ships GPT-Live-1, a full-duplex voice model that replaces Advanced Voice Mode

The new model listens and speaks simultaneously, routes to GPT-5.5 for reasoning, and now serves 150M weekly voice users.

Jaeden Schafer4 min read
Nvidia logo
Models

Nvidia Nemotron 3 Ultra hits closed-model parity at 10x lower cost on LangChain

LangChain tuned its Deep Agents harness for Nemotron 3 Ultra, matching top closed models on business tasks without retraining.

Jaeden Schafer5 min read
ZML launches free LLMD inference server across Nvidia, AMD, Google TPU and Apple chips
Models

ZML launches free LLMD inference server across Nvidia, AMD, Google TPU and Apple chips

The Paris startup, backed by $20M and Yann LeCun, wants to break silicon lock-in as inference costs climb.

Jaeden Schafer5 min read
Meta logo
Models

Meta's Muse Image lets Instagram users pull each other into AI photos

Superintelligence Labs ships its first image model, replacing Llama in Meta AI and powering 30 new Instagram Stories effects.

Jaeden Schafer4 min read
DeepSeek plans its own inference chips to cut Nvidia and Huawei reliance
Models

DeepSeek plans its own inference chips to cut Nvidia and Huawei reliance

The Chinese LLM developer has spent about a year on a silicon project targeting data center inference, per Reuters.

Jaeden Schafer4 min read
Anthropic logo
Models

Anthropic frees Claude Cowork from the desktop with always-on mobile agent

Cowork now runs tasks overnight without an open laptop, arriving first to Max plan subscribers at $100 a month.

Jaeden Schafer4 min read
Nvidia logo
Models

Nvidia's Vera CPU targets agentic AI with 50% IPC gain over Grace

The 88-core Arm chip aims to keep GPUs fed as agents run in loops; Perplexity ran a coding workflow 1.5x faster than x86.

Jaeden Schafer5 min read
Nvidia logo
Models

Nvidia and Hugging Face push Isaac GR00T 1.7 into LeRobot for open robotics

The integration connects 3M robotics developers to 16M AI builders, with Cosmos 3 world models coming next to Hugging Face's open library.

Jaeden Schafer5 min read
Fable tops KernelBench-Mega as AI agents quadruple on real freelance work
Models

Fable tops KernelBench-Mega as AI agents quadruple on real freelance work

A new GPU kernel record, a 4x jump on the Remote Labor Index in eight months, and OSWORLD 2.0 raise the ceiling on what AI agents can do.

Jaeden Schafer5 min read
Springboards launches Flint, an LLM built to escape the chatbot groupthink rut
Models

Springboards launches Flint, an LLM built to escape the chatbot groupthink rut

Ask ChatGPT, Claude, or Gemini for a random number between 1 and 10 and you'll almost always get 7. Flint answered 3.7916.

Jaeden Schafer5 min read
Anthropic logo
Models

Anthropic ships Claude Sonnet 5 at $2 per million input tokens

Sonnet 5 hits 63.2% on agentic coding, close to Opus 4.8's 69.2%, at a fraction of the price.

Jaeden Schafer5 min read
Nvidia logo
Models

NVIDIA's inference software stack cuts DeepSeek V4 token costs 5x in one month

Baseten, Cognition, Together AI and Cursor are riding compounding software gains on Blackwell GPUs as inference economics shift to cost per token.

Jaeden Schafer5 min read
Google logo
Models

Google ships Nano Banana 2 Lite and Gemini Omni Flash to developers

DeepMind's fastest image model generates in 4 seconds at $0.034 per 1K images; Omni Flash matches Veo 3.1 Fast at $0.10 per second of video.

Jaeden Schafer5 min read
Nvidia logo
Models

NVIDIA ships Omniverse skills to train vision AI agents on synthetic defect data

Roboflow hit 95% average precision on Corning fiber defects using just 8 real images plus synthetic data from NVIDIA's new skill.

Jaeden Schafer5 min read
Flexion Robotics trains humanoids to run office errands autonomously
Models

Flexion Robotics trains humanoids to run office errands autonomously

The Swiss startup, founded by ex-Nvidia researchers, stacks reinforcement learning across every layer to make humanoids work without a human puppeteer.

Jaeden Schafer5 min read
China's Z.ai claims GLM-5.2 matches Anthropic's Mythos on bug-finding
Models

China's Z.ai claims GLM-5.2 matches Anthropic's Mythos on bug-finding

The open-weight model lags on general tasks but reportedly closes the gap on cybersecurity — the exact capability US export curbs were meant to contain.

Jaeden Schafer4 min read
Sakana and 360 ship Mythos rivals as US ban on Anthropic exports holds
Models

Sakana and 360 ship Mythos rivals as US ban on Anthropic exports holds

Tokyo's Sakana AI launched Fugu and China's 360 unveiled Tulongfeng within days of each other, two weeks into Washington's block on Anthropic's frontier models.

Jaeden Schafer5 min read
OpenAI logo
Models

OpenAI joins Google, Apple, and SpaceX in building its own AI chips

OpenAI's Jalapeño inference chip, built with Broadcom, marks the latest move by frontier AI buyers to hedge against Nvidia dependence.

Jaeden Schafer4 min read
OpenAI logo
Models

OpenAI ships GPT-5.6 in three tiers, undercuts Claude on price

Sol, Terra, and Luna launch under a White House-monitored preview, with Sol priced at half of Anthropic's Claude Fable 5.

Jaeden Schafer5 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at