Skip to main content
Live
Main content

Tag · 4 stories

AI Inference

Every story tagged AI Inference on AI Chat Daily.

More tagged AI Inference

Kog claims 30x faster LLM inference on existing Nvidia and AMD GPUs
Business

Kog claims 30x faster LLM inference on existing Nvidia and AMD GPUs

The 11-person French startup hit 3,000 tokens per second on a 2B model and now needs to prove the same trick works on real LLMs by September.

Jaeden Schafer5 min read
ZML launches free LLMD inference server across Nvidia, AMD, Google TPU and Apple chips
Models

ZML launches free LLMD inference server across Nvidia, AMD, Google TPU and Apple chips

The Paris startup, backed by $20M and Yann LeCun, wants to break silicon lock-in as inference costs climb.

Jaeden Schafer5 min read
Baseten nears $1.5B round at $13B valuation, up 160% in five months
Business

Baseten nears $1.5B round at $13B valuation, up 160% in five months

The AI inference startup is closing a split-priced round five months after a $300M Series E, riding the inference gold rush.

Jaeden Schafer4 min read
AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at