AI BriefingOpenAIFeature Updates17:01
AI summarized from verified sources
High-speed inference makes real-time work smoother
Use frontier models instantly in time-sensitive workflows.
SOURCE CHECK
1 sources
Sources
Key Points
- 1750 tokens/sec via Cerebras
- 2Ready for voice, support, finance
- 3Rolling out as capacity expands
OpenAI announced Ultrafast in partnership with Cerebras. It generates up to 750 tokens per second, enabling frontier models for real-time voice, support, coding and more. Expanding from initial business customers.
What happened
OpenAI released Ultrafast using Cerebras hardware for much faster generation. Initially for selected businesses.
Impact
Accelerates frontier model adoption where real-time performance matters, directly boosting efficiency.