Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the API

OpenAI is previewing Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing.

Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, bringing frontier intelligence to products and workflows where every second matters.

GPT‑5.6 Sol on Ultrafast mode is launching first in the OpenAI API. It is currently available as a limited preview to a select group of customers, with access expanding as capacity grows.

Sign up for Ultrafast access updates

Learn more in the official announcement.

5 Likes