Product LaunchNew

OpenAI previews Ultrafast mode for GPT-5.6 Sol

OpenAI has announced a preview of Ultrafast mode, a new API service tier designed for its GPT-5.6 Sol model. According to OpenAI, the tier delivers generation speeds up to 14 times faster than standard execution, reaching up to 750 output tokens per second. Claims are as reported; this summary makes no determination about accuracy or significance.

First detected
Aug 24, 2026
Last updated
Aug 27, 2026

Moderate confidence

Reported by the organization responsible for the announcement.

Limited corroboration

No independent reporting recorded yet.

What does this mean?

Corroboration measures how many genuinely independent sources support the event. Confidence measures how reliable the available evidence appears.

Stable

No recent reporting has materially changed the known facts.

Follow this development to see meaningful updates as new evidence emerges.

Save keeps this for later. Follow tracks meaningful changes as new evidence emerges — it shapes your Following Feed, alerts, and digest eligibility, and doesn't promise an instant notification.

What we know

Why it matters

Inference speeds reaching 750 tokens per second materially improve the viability of real-time voice systems, complex multi-step reasoning, and low-latency software development pipelines. The partnership also marks a notable infrastructure diversification for OpenAI, demonstrating commercial deployment of Cerebras hardware alongside traditional GPU clusters in production enterprise environments.

Coverage

How this developed

  1. Aug 24, 2026

    1. Development detected

  2. Aug 13, 2026

    1. New reporting added