Cutting Agent "Wait Time" by 14x: OpenAI Runs GPT-5.6 Sol on Cerebras at 750 Tokens per Second
OpenAI is previewing "Ultrafast," an API-only tier that runs full-size GPT-5.6 Sol on Cerebras hardware at up to 750 tokens per second (14x standard). Intelligence stays the same; only speed is carved out into a separately priced tier. It helps agents' back-and-forth loops, but this developer-focuse