750 tokens per second. That's not a typo.
OpenAI just previewed "Ultrafast mode" for GPT-5.6 Sol — up to 14x faster than standard, powered by a new partnership with chip company Cerebras.
This isn't about smarter answers, it's about speed unlocking entirely new use cases: real-time incident response, live customer support, split-second financial analysis — workflows where waiting on AI was never an option before.
It's rolling out first in the API to a small group of customers, with wider access coming as capacity grows.
Would your workflow actually use AI at 14x speed, or is intelligence still the bottleneck? Let me know below.
Save this for when Ultrafast opens up, and follow @latinaailab for weekly AI news, tools & tips for creators.
.
.
.
























































































