OpenAI has previewed Ultrafast, an API tier that runs GPT-5.6 Sol at up to 14x speed on Cerebras hardware, betting that latency, not just intelligence, is what makes AI agents usable.

OpenAI is previewing a new way to run its most capable GPT-5.6 model at dramatically higher speeds. The company says...

Google's Gemini 3.7 Flash model is live and built for cheap agents; OpenAI's GPT-5.6 Sol Ultrafast is quicker but locked behind a waitlist.