
ModelsLead story
Nvidia ships Groq 3 LPX to accelerate Vera Rubin agentic inference
The new accelerator hits 3,400 output tokens per second on 100,000-token contexts, 4x the nearest platform, with Nebius as first cloud adopter.
Jaeden SchaferEditor in Chief5 min read

