Skip to main content
Live
Main content

OpenAI ships GPT-5.4 Thinking to all Plus users

The o-series-style reasoning model lands in ChatGPT with adaptive budgets — and a new opt-in 'extended thinking' toggle.

Jaeden Schafer
Editor in Chief · · 4 min read

GPT-5.4 Thinking rolled out to all ChatGPT Plus users today, bringing reasoning-model capabilities to the mainstream consumer tier. OpenAI's adaptive thinking budget automatically selects compute based on perceived problem difficulty, allowing the model to spend 5 seconds on straightforward queries and up to 60 seconds on research or proof-writing tasks.

The model scores 93% on GPQA Diamond, Google's science benchmark, and shows strong performance on AIME competition math problems. This closes the gap with Gemini 3.1 Pro's 94.3% GPQA score, narrowing what was a two-point lead a month ago. The Extended Thinking toggle is opt-in; users can enable it to force extended reasoning on hard problems, trading latency for accuracy on questions where quick answers fail.

Adaptive budgeting is the differentiator here. Rather than a fixed token budget for thinking, OpenAI's system routes queries to a routing classifier that estimates difficulty and allocates compute accordingly. A question like "What is 2+2?" bypasses thinking entirely; a research synthesis or multi-step proof gets the full budget. This reduces user-facing latency while preserving accuracy on complex queries.

Key facts

  • 01OpenAI. A key thread of reporting in this story.
  • 02GPT-5. A key thread of reporting in this story.
  • 03Reasoning. A key thread of reporting in this story.

Pricing remains unchanged at Plus tier rates. OpenAI has not announced GPT-5.4 Thinking availability for Pro subscribers, though internal documents suggest an enterprise pilot is underway. The model is available via the ChatGPT web interface and iOS/Android apps, with API access rolling out in the coming weeks.

GPT-5.4 Thinking reaches 93% on GPQA Diamond, closing the gap with Gemini 3.1 Pro while keeping thinking latency under 30 seconds for most queries.
Jaeden Schafer

The release underscores a shift in OpenAI's strategy toward runtime compute over model scaling. Rather than a new o1 variant, GPT-5.4 Thinking layers adaptive reasoning atop a smaller base model, reducing inference costs while matching or exceeding last year's o1-Pro performance on benchmarks.

Related · from this week
GPT-5 Pro helps immunologist crack a 3-year T cell mystery
Jaeden Schafer · 4 min read →
ShareXLinkedInEmail
AI Box

Every AI model. One chat.

The latest models from ChatGPT, Claude, Gemini, Sora, ElevenLabs — 80+ models in a single chat. Compare answers side by side. Pick the best one every time.

  • ChatGPT, Claude, Gemini, Grok, DeepSeek — in one chat
  • Generate images & video with Sora, Veo, Ideogram
  • Compare any two models side by side
  • From $8.99/mo · 80+ models, all included
Try AI Boxaibox.ai
Trusted by 3,000+ teams
Related topics
Got a tip?

Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.

Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.

AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at

Keep reading

More from Models

OpenAI logo
Models

GPT-5 Pro helps immunologist crack a 3-year T cell mystery

OpenAI says its top reasoning model gave Derya Unutmaz the breakthrough insight that closed a stalled immunology project with cancer implications.

Jaeden Schafer4 min read
OpenAI logo
Models

OpenAI reasoning model disproves 80-year-old Erdős geometry conjecture

The proof marks the first autonomous AI solution to a prominent open math problem, verified by field experts.

Jaeden Schafer5 min read
OpenAI logo
Models

OpenAI adds GPT-Realtime-2, Translate and Whisper to its Realtime API

The new voice stack handles 70 input languages and 13 output languages, pushing the API beyond simple call-and-response.

Jaeden Schafer4 min read