OpenAI launches Ultrafast inference mode for GPT-5.6 Sol
OpenAI launched Ultrafast, an inference mode that runs GPT-5.6 Sol at up to 750 output tokens per second using Cerebras hardware. It joins Standard and Fast to form a three-tier pricing structure.
First seen 14 Aug, 14:21 UTC on The Decoder3 sourcesLast update 34d ago
OpenAILaunch · 14 Aug
GPT-5.6 Sol
New flagship model
What the sources say
Linked, never rewritten. Official means the lab itself.
How it unfolded
Every source in the order it appeared. Times in UTC.