Skip to main content
Live
Main content

Category

Models

Foundation models, benchmarks, and the research race.

Models — page 5

Naveen Rao's Unconventional AI claims a 1,000x cut in inference power
Models

Naveen Rao's Unconventional AI claims a 1,000x cut in inference power

The ex-Databricks AI chief released Un-0, an image model running on a simulated oscillator chip, with silicon schematics promised next.

Jaeden Schafer5 min read
IBM unveils nanostack architecture, claims first sub-1nm chip tech
Models

IBM unveils nanostack architecture, claims first sub-1nm chip tech

The 0.7nm node packs nearly 100 billion transistors on a fingernail-sized chip, with 50% more performance than 2nm silicon.

Jaeden Schafer5 min read
Google logo
Models

Google bakes computer use into Gemini 3.5 Flash as a native tool

DeepMind folds its standalone agent model into Flash, letting developers build agents that drive browsers, mobile apps and desktops via one API call.

Jaeden Schafer4 min read
OpenAI logo
Models

OpenAI unveils Jalapeño, its first custom inference chip built with Broadcom

The processor targets inference workloads and was co-developed using OpenAI's own models, with early tests showing better performance-per-watt.

Jaeden Schafer5 min read
OpenAI logo
Models

GPT-5 Pro helps immunologist crack a 3-year T cell mystery

OpenAI says its top reasoning model gave Derya Unutmaz the breakthrough insight that closed a stalled immunology project with cancer implications.

Jaeden Schafer4 min read
Nvidia logo
Models

Nvidia silicon runs 400 of the world's 500 fastest supercomputers

Nvidia GPUs, CPUs and networking now power 81% of the TOP500 and 90% of new entrants, with a clean sweep of the Green500 top eight.

Jaeden Schafer5 min read
Nvidia logo
Models

Nvidia's Vera CPU lands at Los Alamos in three new supercomputers

Mission, Vision and Veritas will pair Vera CPUs with Rubin GPUs to run agentic science workloads, posting 7x gains over the Crossroads x86 system.

Jaeden Schafer5 min read
Nvidia logo
Models

NVIDIA's new CUDA-X science software posts 14,900x speedup on telescope data

cuPhoton, DAQIRI, and ALCHEMI accelerate astronomy, particle physics, and materials simulation, with Lila Sciences cutting materials runs from weeks to days.

Jaeden Schafer5 min read
Nvidia logo
Models

NAIRR pilot tops 700 research projects on NVIDIA DGX infrastructure

The NSF program's two-year run has accelerated fusion materials, infectious-disease detection, and fluid simulation foundation models.

Jaeden Schafer5 min read
Nvidia logo
Models

JUPITER, Europe's first exascale supercomputer, posts results across brain, climate, and quantum

Four projects on the Jülich machine show what 20,480 Grace Hopper Superchips can do — from 1-km climate runs to a 50-qubit simulation.

Jaeden Schafer5 min read
Nvidia logo
Models

Nvidia's Rubin AI servers run on 45°C coolant, killing data-center water use

The first 100% liquid-cooled platform from Nvidia targets $4M in annual savings per 50MW and a 100% cut in cooling water.

Jaeden Schafer5 min read
Subquadratic's SubQ posts third-party benchmarks for its sparse-attention LLM
Models

Subquadratic's SubQ posts third-party benchmarks for its sparse-attention LLM

The Miami startup says SubQ ran Nvidia's RULER 128 test for $8 versus $2,600 for Claude Opus 4.6, and Appen's evaluation backs the claim.

Jaeden Schafer5 min read
Nvidia logo
Models

Nvidia's ENPIRE lets AI coding agents train robots overnight, hits 99% success

Teams of up to 8 agents from OpenAI, Anthropic and Moonshot taught robots to insert GPUs and cut zip ties without human input.

Jaeden Schafer5 min read
Nvidia logo
Models

Nvidia Blackwell sweeps MLPerf Training 6.0 across all seven benchmarks

GB300 NVL72 delivers 1.6x faster training than GB200 NVL72, and CoreWeave hits a 2.02-minute time-to-train on DeepSeek-V3 671B.

Jaeden Schafer5 min read
Google logo
Models

Loft Orbital's Yam-9 runs first vision-language model in orbit

A Google DeepMind Gemma 3 model running on a Nvidia Jetson chip aboard Yam-9 identified targets from natural language queries.

Jaeden Schafer5 min read
Nvidia logo
Models

NVIDIA Blackwell runs 20x more agents per megawatt on new AgentPerf benchmark

Artificial Analysis launches the first agentic AI benchmark, and the GB300 NVL72 takes the top spot against Hopper on DeepSeek V4 Pro.

Jaeden Schafer5 min read
Avataar's Varya undercuts Veo and Runway on price by 20x for Indian video AI
Models

Avataar's Varya undercuts Veo and Runway on price by 20x for Indian video AI

The Peak XV-backed startup distilled Alibaba's Wan 2.2 into a four-step model that generates 720p clips for $0.005 per second.

Jaeden Schafer5 min read
Anthropic logo
Models

Anthropic reverses hidden Claude Fable 5 sabotage of AI researchers

After backlash, Anthropic will make its frontier-LLM safeguards visible instead of silently degrading Claude's output for rival AI work.

Jaeden Schafer5 min read
Anthropic logo
Models

Anthropic's Claude Fable 5 refuses basic biology questions by design

Anthropic told The Verge Fable's guardrails are 'overly conservative' to block bioweapons queries, routing routine biology asks to Opus 4.8.

Jaeden Schafer5 min read
Writer researchers find memory tools make AI models more sycophantic, less accurate
Models

Writer researchers find memory tools make AI models more sycophantic, less accurate

Two new papers show that storing user preferences pulls models toward wrong answers, with Mem0 and Zep amplifying the bias.

Jaeden Schafer5 min read
Google logo
Models

Google DeepMind's DiffusionGemma generates text 4x faster on GPUs

The 26B MoE model hits 1,000+ tokens per second on an H100 by drafting 256 tokens in parallel instead of one at a time.

Jaeden Schafer5 min read
Decart launches Oasis 3, a real-time world model for driving simulation
Models

Decart launches Oasis 3, a real-time world model for driving simulation

The $4B startup is pricing API access at $0.02 per second and betting developers will build on top of it like the early OpenAI ecosystem.

Jaeden Schafer5 min read
Anthropic logo
Models

Anthropic's Claude Fable 5 spins up playable games from a single prompt

Wharton's Ethan Mollick says the new Mythos-class model executed multi-page specs autonomously for up to a dozen hours.

Jaeden Schafer5 min read
Anthropic logo
Models

Anthropic releases Claude Fable 5, its first public Mythos-class model

The lab once said Mythos was too dangerous to ship. New safeguards route 5% of risky prompts back to Opus 4.8.

Jaeden Schafer5 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at