OpenAI’s new Ultrafast mode runs GPT-5.6 Sol 14 times faster, on Cerebras chips

OpenAI is launching an "Ultrafast" API tier for its GPT-5.6 Sol model, significantly boosting performance. This new tier, powered by Cerebras hardware, achieves speeds up to 14 times faster than current offerings, processing approximately 750 output tokens per second.
Key takeaways
- GPT-5.6 Sol now offers a 14x speed boost.
- New API tier utilizes specialized Cerebras hardware.
- Achieves 750 output tokens per second.
- Promises faster AI-powered application performance.
Why it matters
For professionals leveraging AI tools, this speed increase means more responsive applications and quicker data processing. Tasks like real-time analysis, content generation, and complex query responses will become significantly more efficient, enhancing productivity in daily workflows.
Try this on SynaBot
Related AI assistants, prompts, and tools from the SynaBot catalog.
- GodmodeGodmode is a user interface for creating and managing autonomous AI agents powered by OpenAI's language models. It simplifies the process of giving an AI a goal and letting it break it down into subtasks to achieve it.
- OpenAI CodexOpenAI Codex is a large language model fine-tuned for programming, capable of translating natural language into code across multiple programming languages. It powers tools like GitHub Copilot.
- Role Model AIRole Model AI — Virtual assistant with 3D avatars, phone integration, and Fortnite connectivity. It sits in the 3d category and is built to generate or edit 3D assets, prototypes, and visualizations for product, game, or architectural workflows.

