Procurement

Post Product

  • Post Supply
  • Manage Supplies

OpenAI Previews Ultrafast: GPT-5.6 Sol at Up to 14x the Speed on Cerebras

   2026-08-13 OpenAI10

OpenAI previewed an Ultrafast API tier running GPT-5.6 Sol at up to 14x speed, up to 750 tokens per second on Cerebras.

OpenAI shared an early look at Ultrafast, a new API service tier that runs GPT-5.6 Sol up to 14 times faster than standard processing, generating up to 750 output tokens per second. The tier is powered by Cerebras hardware and is in limited preview with a selected group of customers in coding, commerce, financial research, and support. OpenAI frames the point as removing the usual trade-off in which real-time latency forced users to drop to a smaller model, citing internal use in incident response — reading logs, analyzing traces, and preparing fixes while an outage is still unfolding — and in research, where overnight batches compress into same-day loops.
 
ReportCollect0Reward 0Comment 0
Disclaimer
• 
This article is an original work by {author}. Reproduction is welcome, but please indicate the original source: {linkurl}. The views expressed in this article are those of the author alone, and the website has not verified the content. Readers are advised to use it for reference only. If the article involves content that violates public morality or laws, it will be deleted immediately upon discovery, and the author shall bear the corresponding responsibilities. In case of copyright or other issues, please contact us in a timely manner.
 
More>Similar News
Recommended images and text
Recommend News
Click to rank

Sign Up

For enterprise

Tel:132-7094-5888

Scan with Phone

Support

Tel:132-7094-5888

Program

Mini Program

Procurement Advisory

Scan with WeChat

WeChat

Business Opportunities

Scan with WeChat

Top