Agents & InferenceTechCrunch

OpenAI introduces ‘Ultrafast,’ a new mode that makes GPT-5.6 Sol work at 14x the speed

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

GPT-5.6 Sol now runs at up to 750 output tokens/sec in an "Ultrafast" mode—roughly 14x standard speed—powered by Cerebras hardware, meaning you no longer have to drop to a smaller model to hit real-time latency for your top-tier model. This unlocks latency-sensitive agentic and interactive workflows (incident response, live support, market analysis) at frontier quality, but it's preview-only to a limited customer set gated by Cerebras capacity, so don't architect production dependencies on it yet.