GPT-5.6-Sol-Ultrafast Financial Terminal Clone
We gave OpenAI GPT-5.6 Sol the same prompt on Ultrafast and Standard: build a financial terminal-style dashboard for analysts.
Ultrafast: 1 min 50 seconds
Standard: 12 min 20 seconds
Same result, nearly 7x faster.
Learn more about GPT-5.6 Sol Ultrafast in OpenAI’s blog post.
Getting the most out of GPT-5.6: Sol, Terra, and Luna
Your Codex subscription now comes with 3 main models, Sol, Terra, and Luna, each independently trained and served, with reasoning dials. Together, they let you choose a balance of speed, cost, and intelligence for each task.
Build ultra-fast LLM applications with Cerebras and DeepLearning.AI
Fast inference makes a new class of real-time LLM applications possible.
In our new short course, Fast LLM Inference with Cerebras, taught by Zhenwei Gao, Seb Duerr, and Sarah Chieng, you'll build them on the Cerebras Wafer-Scale Engine (WSE), where a model's weights sit on-chip and tokens come out several times faster than a typical GPU setup.
First Look @ Gemma 4 on Cerebras
We spent a week building multimodal use cases with Gemma 4, running on Cerebras at ~2,300 tok/s. Here are some of the things we built and our learnings.
Cognition SWE 1.6 Demo
Cognition SWE 1.6 Fast, powered by Cerebras, in side-by-side demo to Claude Sonnet 4.6, showcasing fast inference advantage.
Build Your Own Content Fact Checker
In collaboration with OpenAI and Khushi Shelat from Parallel Web Systems.
Full coding tutorial available on OpenAI Developers website: ‘Build your own content fact checker with gpt-oss-120B, Cerebras, and Parallel’.
Video workshop available on Cerebras Youtube.
Build a Real-time AI Sales Agent
Co-presented with Sarah Chieng, Cerebras Head of DevX, at AI Engineer Code Summit, New York City.
Video available here.