Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
OpenAI previewed an Ultrafast API tier that runs the GPT-5.6 Sol model up to 14 times faster, delivering up to 750 output tokens per second. The tier is powered by Cerebras hardware.

- Ultrafast tier provides up to 14× faster token generation for GPT-5.6 Sol.
- Capable of producing about 750 tokens per second using Cerebras hardware.
- Targets developers needing low‑latency AI responses, expanding OpenAI's API portfolio.
- Represents a strategic partnership between OpenAI and Cerebras to enhance inference speed.
OpenAI announced a preview of a new service tier called Ultrafast, designed to accelerate the GPT-5.6 Sol model.
The tier claims up to a 14‑fold increase in generation speed, reaching roughly 750 tokens per second, thanks to integration with Cerebras' specialized AI processors.
This performance boost aims to reduce latency for applications that require near‑real‑time text generation, such as chatbots, coding assistants, and content creation tools.
The Ultrafast tier sits alongside OpenAI's existing API offerings, giving developers the option to trade higher cost for substantially faster output.
Reduces API latency, enabling faster real‑time applications.
Allows products to deliver quicker AI responses, improving user experience.
Signals OpenAI's focus on performance and a partnership with hardware leader Cerebras.
AI services become noticeably faster for end users.
- Cerebras
- A company that builds large‑scale AI processors optimized for high‑throughput inference.
AI ToolsRecord, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets
AI ToolsSuno is trying to look more like a real music production tool
Amazon Quick for Microsoft 365: Agentic AI where you work - Amazon Web Services (AWS)
AI ToolsWe Open Sourced R-CLI, the Coding Harness Above Every Published Terminal Bench 2.1 Result
Chester County’s AI-Powered Self-Service Kiosk Project Wins National Award - Chester County
BusinessOpenAI hires new CRO as executive shake-up continues
OpenAI has hired Dali Rajic as its new Chief Revenue Officer to lead sales operations during an ongoing executive restructuring phase.
AI ResearchIntroducing Gemini 3.7 Flash
DeepMind has released Gemini 3.7 Flash, an updated version of its Gemini large language model. The new version boasts improved performance and efficiency.
Election Briefing: Artificial Intelligence: What Campaigns Are Already Doing, and What They Should Be Ready for Between Now and Election Day - Campaigns & Elections
A briefing on artificial intelligence's role in US election campaigns, highlighting current uses and future preparations.
Researchers Explore What It Means To Say AI 'Thinks' - Carnegie Mellon University
Researchers at Carnegie Mellon University are exploring the concept of AI 'thinking' and what it means for artificial intelligence to be conscious.
NWACC, MIT program collaborate on artificial intelligence curriculum - talkbusiness.net
NWACC and MIT are collaborating on an artificial intelligence curriculum. This partnership aims to provide students with a comprehensive education in AI.
BusinessMicrosoft kills off unsuccessful AI features while merging its separate Copilot apps
Microsoft is merging its consumer and business Copilot apps into a single platform while discontinuing several AI features, including AI-generated podcasts and the Mico virtual assistant.