Skip to content
Chatt & Assistenter· UpdateBeta

OpenAI Unveils New Ultrafast Mode for Faster API Responses

OpenAI has introduced a new performance mode called 'Ultrafast' via its API. The new mode is designed to enable processing speeds up to 14 times faster, reaching up to 750 tokens per second.

By the Aheadline editorial team·15 aug. 2026·2 min read·Source: Entity-watch: OpenAIVerifierad signalAI-generated
OpenAI Unveils New Ultrafast Mode for Faster API Responses
OpenAI Unveils New Ultrafast Mode for Faster API Responses
OpenAI Unveils New Ultrafast Mode for Faster API Responses
By · Policy- & EU-reporter
Last updated

What happened?

OpenAI has launched a new performance mode named 'Ultrafast' for its language model via its API. According to the company, the new mode can increase processing speeds by up to 14 times compared to standard processing and deliver up to 750 tokens per second. The feature is available in an initial preview for developers.

Key facts

Maximal hastighet750 token per sekund
PrestandaökningUpp till 14 gånger snabbare
TillgänglighetAPI-förhandsvisning (preview)

Until now, getting real-time speed typically meant choosing a smaller or more specialized model.

OpenAI, Företag · NewsBytes

Why it matters

High-speed token generation is critical for building more responsive AI services. By offering higher speeds directly within its larger models, OpenAI aims to enable companies to run advanced AI without needing to switch to smaller, more limited models.

Who is affected?

The update is primarily aimed at system developers and companies building large-scale enterprise applications. Users requiring fast, real-time responses—such as for interactive assistants or the analysis of large datasets—will benefit most from this increase in capacity.

Impact on the EU

The feature is currently being rolled out in preview via the OpenAI API. As this is an API-based update, it also affects developers and companies within the EU building applications on top of the OpenAI platform, provided they have access to the beta testing.

What else you should know

In its announcement, OpenAI noted that developers previously had to choose smaller or more specialized models to achieve real-time performance. The new solution aims to eliminate that trade-off for more demanding applications.

Frequently asked questions

Quick answers about this story

Vad har hänt?
OpenAI har introducerat ett nytt prestandaläge kallat 'Ultrafast' i sitt API som genererar upp till 750 token per sekund.
När hände det?
Nyheten rapporterades den 14 augusti 2026 i samband med en förhandsvisning av uppdateringen för API-utvecklare.
Varför spelar det roll?
Det gör det möjligt för företag att köra mer avancerade språkmodeller i realtid utan att behöva byta till mindre modeller för bättre hastighet.
Vem har tillgång till det nya läget?
Prestandaläget finns tillgängligt via API:et i en första förhandsvisning för registrerade utvecklare.
Original source
Entity-watch: OpenAI·newsbytesapp.com

The link opens in a new window and leads to the publisher's own site.

Verifierad signal

Källan har spårats automatiskt från utgivaren via Aheadlines signalkedja.

AI-verktyg i artikeln

Topics

#GPT-5.6 Sol#Large Language Models (LLMs)#OpenAI GPT-5.6#API
[ STAY UP TO DATE ]

Get similar news straight to your inbox

No affiliate linksCancel anytimeGDPR-friendly
[ Frequency ]
[ What do you want to read about? ]

You'll receive updates on 2 topics.

The reader's room

Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.

Sign in to submit a comment or question.

Loading comments…
How this affects you

Read the article through your role

  • Decide whether this affects strategy over 6–12 months or is just noise.
  • Discuss with leadership: do we own the right question or does ownership need to move?
  • Ask: what risk are we taking by NOT acting on this this quarter?

Generated angle — not editorial analysis of "OpenAI Unveils New Ultrafast Mode for Faster API Responses"