Models — page 6

Google DeepMind launches Gemma 4 12B, an encoder-free multimodal model for laptops
The mid-sized open model runs on 16GB of VRAM and nears 26B benchmark performance, with native audio and vision baked into the LLM backbone.

Google DeepMind ships Gemini 3.5 Live Translate across 70+ languages
The new audio model streams speech-to-speech translation continuously, lagging the speaker by just a few seconds across Meet, Translate, and the Live API.

Anthropic logs 8x code merge jump as researchers benchmark AI gaming society's rules
Anthropic sees early signs of recursive self-improvement, a new benchmark tests AI loophole-hunting, and RL drones beat a human champion at 22 m/s.

ECMWF's AIFS runs forecasts on 1,000x less energy than physics-based model
Machine learning is rewriting weather forecasting at a fraction of the compute cost, but climate modeling won't surrender to neural nets.

Estonia's new benchmark ranks Claude Opus 4.7 best at resisting Russian propaganda
The Estonian Language Institute scored dozens of LLMs across 14 propaganda categories; Anthropic took six of the top 10 spots.

Google releases Gemma 4 12B, sized to run locally on a 16GB laptop
The new mid-weight Gemma slots between mobile and workstation variants, with model weights just under 18GB available on Hugging Face and Kaggle.

Nvidia and Unitree pair Thor chip with H2 Plus humanoid in US-China robot blueprint
Jensen Huang's blueprint puts a Thor T5000 brain inside a 6-foot, 150-pound Chinese-built body — and undercuts rival humanoids on price.

Nvidia unveils GraspGen-X, LCDrive and NitroGen foundation models at CVPR
Three new physical AI models target zero-shot robot grasping, faster autonomous-vehicle reasoning, and gameplay-trained embodied agents.

Nvidia ships physical AI agent skills at CVPR, anchored by 32B Alpamayo 2 Super
New agent skills built on Cosmos 3 automate scene reconstruction, simulation and policy training for AVs, robots and vision AI.

Microsoft to unveil MAI-Thinking-1 reasoning model and Copilot super app at Build
Build 2026 lands in San Francisco June 2 with new in-house models, a Windows 11 developer mode, and an RTX Spark push.

Windborne's WeatherMesh 6 out-forecasts ECMWF with 400 balloons feeding the model
The Stanford-founded startup says its sixth model is as accurate five days out as a traditional forecast is the day before.

Intel's Crescent Island AI chip ships this year, undercutting Nvidia on cost and cooling
The air-cooled inference GPU uses LPDDR5 memory instead of HBM, betting that cheaper silicon wins the inference market Nvidia hasn't locked down.

OpenAI model disproves Erdős unit distance conjecture after 80 years
An internal OpenAI system produced the first proof resolving a major open conjecture in discrete geometry, drawing praise from Fields Medalist Tim Gowers.

Nvidia unveils Cosmos 3, a foundation model for physical AI reasoning
Nvidia Research positions Cosmos 3 as the bridge from scripted robotics demos to embodied autonomy in open-world environments.

LLMs believe false claims even when training data labels them as lies
A new preprint finds Qwen, Kimi, and GPT-4.1 absorb fabricated facts at an 88.6% belief rate even after explicit negation warnings.

Nvidia shows eight sim-to-real robotics papers at ICRA, led by COMPASS and PEEK
The research push targets multi-arm scheduling, cross-body navigation, and grasping, with one pipeline delivering a 41x real-world accuracy gain.

Anthropic ships Claude Opus 4.8 with 84% on Online-Mind2Web and cheaper fast mode
The new Opus holds Opus 4.7 pricing at $5/$25 per million tokens, while fast mode drops to a third of its prior cost at 2.5x speed.

Huawei's HiSilicon claims chip breakthrough to close China's 5-year AI lag
Tingbo He says Tau's Scaling Law will deliver 1.4-nanometer equivalent performance by 2031, three years behind TSMC.

NVIDIA Vera CPU delivers 1.5x performance edge over 128-core x86 in Phoronix tests
Custom Olympus cores and 1.2TB/s memory bandwidth push Vera past Intel and AMD in agentic AI workloads.

Claude Code and OpenClaw ignite AI agent revolution among developers
Anthropic's Opus 4.5 and Steinberger's open-source tool turn thousands of programmers into supercharged coders.

Grok barely appears in federal AI records as agencies favor OpenAI and Anthropic
Reuters reviewed 400+ government AI deployments and found xAI's chatbot in just three entries.

Google shifts AI science strategy toward agentic research systems
AlphaFold co-creator John Jumper moved to coding as Google pivots from specialized tools to general-purpose AI scientists.

OpenAI reasoning model disproves 80-year-old Erdős geometry conjecture
The proof marks the first autonomous AI solution to a prominent open math problem, verified by field experts.

Stability AI releases Stable Audio 3.0 with 6-minute song generation
The new large model creates full compositions, more than doubling the length of last year's release.
Stay ahead of everyone in AI.
The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.
The briefing read inside teams at