Close Menu
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
What's Hot

China and Europe agree to cut China’s hybrid vehicle exports by more than half

October 9, 2026

US to livestream Fort Hood mass shooting execution: Is it legal? death penalty news

October 9, 2026

Masked robber steals more than $5 million worth of wine from Italy’s top vineyard

October 9, 2026
Facebook X (Twitter) Instagram
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Facebook X (Twitter) Instagram
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Home » OpenAI launches GPT-5.4 with Pro and Thinking versions
AI

OpenAI launches GPT-5.4 with Pro and Thinking versions

Editor-In-ChiefBy Editor-In-ChiefMarch 5, 2026No Comments3 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
Follow Us
Google News Flipboard
Share
Facebook Twitter LinkedIn Pinterest Email


On Thursday, OpenAI released GPT-5.4. This is a new foundation model that is being touted as “the most capable and efficient frontier model for professional work.” In addition to the standard version, GPT-5.4 is also available as an inference model (GPT-5.4 Thinking) or a model optimized for high performance (GPT-5.4 Pro).

The API version of this model will be available in a context window of 1 million tokens, the largest context window ever available in OpenAI.

OpenAI also stated that GPT-5.4 was able to solve the same problem with significantly fewer tokens than its predecessor, highlighting improved token efficiency.

The new model offers significantly improved benchmark results, including record scores for computer usage benchmarks OSWorld-Verified and WebArena Verified. The new model also achieved a record score of 83% on OpenAI’s GDPval test for knowledge work tasks.

GPT-5.4 also led in Mercor’s APEX-Agents benchmark, designed to test professional skills in law and finance, according to a statement from Mercor CEO Brendan Foody.

“(GPT-5.4) excels at creating long-term deliverables such as slide decks, financial models, and legal analysis, delivering the best performance while running faster and at lower cost than competing Frontier models,” Foody said in a statement.

GPT-5.4 continues the company’s efforts to limit hallucinations and factual errors. OpenAI said the new model was 33% less likely to have an error on individual claims and 18% less likely to have an error in the overall response compared to GPT 5.2.

tech crunch event

San Francisco, California
|
October 13-15, 2026

As part of the release, OpenAI overhauled the way the GPT-5.4 API version manages tool calls and introduced a new system called Tool Search. Previously, the system prompt would lay out the definitions of all available tools when calling a model, but this process could consume large amounts of tokens as the number of available tools increased. The new system allows models to search for tool definitions on demand, making requests faster and cheaper in systems with many available tools.

OpenAI also includes new safety assessments to test the thought chain of your models. This is a running commentary provided by the model to demonstrate the thought process through a multi-step task. AI safety researchers have long worried that inference models can misrepresent an AI’s chain of thought, and tests have shown that under the right circumstances, this can happen.

A new OpenAI assessment shows that deception is less likely to occur in the Thinking version of GPT-5.4, “suggesting that the model lacks the ability to hide inferences and that CoT monitoring remains an effective safety tool.”



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Editor-In-Chief
  • Website

Related Posts

A16z’s Olivia Moore talks about the current state of consumer AI

October 9, 2026

Disrupt 2026 begins in 4 days, prices rise

October 9, 2026

OpenAI’s math solutions do not yet meet standards in this field

October 8, 2026
Add A Comment

Comments are closed.

News

US to livestream Fort Hood mass shooting execution: Is it legal? death penalty news

By Editor-In-ChiefOctober 9, 2026

The Pentagon announced it will livestream the execution of convicted mass shooter Nidal Malik Hasan.…

President Trump’s planned D.C. Arch faces new lawsuit | Politics News

October 9, 2026

US sanctions the International Criminal Court itself over recent escalation | ICC News

October 9, 2026
Top Trending

A16z’s Olivia Moore talks about the current state of consumer AI

By Editor-In-ChiefOctober 9, 2026

I wrote this week about the murky economics of consumer AI, which…

Disrupt 2026 begins in 4 days, prices rise

By Editor-In-ChiefOctober 9, 2026

4 days. That’s all you have until TechCrunch Disrupt 2026, when 10,000…

OpenAI’s math solutions do not yet meet standards in this field

By Editor-In-ChiefOctober 8, 2026

As OpenAI announced hundreds of purported solutions to some of the world’s…

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Welcome to WhistleBuzz.com (“we,” “our,” or “us”). Your privacy is important to us. This Privacy Policy explains how we collect, use, disclose, and safeguard your information when you visit our website https://whistlebuzz.com/ (the “Site”). Please read this policy carefully to understand our views and practices regarding your personal data and how we will treat it.

Facebook X (Twitter) Instagram Pinterest YouTube

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Facebook X (Twitter) Instagram Pinterest
  • Home
  • Advertise With Us
  • Contact US
  • DMCA Policy
  • Privacy Policy
  • Terms & Conditions
  • About US
© 2026 whistlebuzz. Designed by whistlebuzz.

Type above and press Enter to search. Press Esc to cancel.