OpenAI launches Ultrafast mode for GPT-5.6 Sol, delivering 750 tokens per second via Cerebras hardware for enterprise real-time AI applications.

OpenAI is previewing a new way to run its most capable GPT-5.6 model at dramatically higher speeds. The company says...

Google's Gemini 3.7 Flash model is live and built for cheap agents; OpenAI's GPT-5.6 Sol Ultrafast is quicker but locked behind a waitlist.