OpenAI DevDay Leaks | Ultrafast around 750 tokens/s, GPT-6 Sol on chat
Summary
Leaks from OpenAI DevDay indicate that GPT-6 Sol is ultrafast, achieving around 750 tokens per second in chat interactions.
Similar Articles
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
OpenAI previews Ultrafast, a new service tier for GPT-5.6 Sol that runs up to 14× faster via Cerebras, generating up to 750 tokens per second in the OpenAI API.
@OpenAI: Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of c…
OpenAI is previewing Ultrafast mode for GPT-5.6 Sol, offering up to 14x speed, initially in the OpenAI API to select customers with expanded access over time.
GPT-5.6 Sol can run now at an incredible rate of ~750 tokens per second
GPT-5.6 Sol now runs at an impressive inference speed of about 750 tokens per second.
@danshipper: Yesterday, GPT-5.6 Sol Ultra became generally available. Today, I'm sharing that I accidentally spent 2b tokens overnig…
Dan Shipper announces that GPT-5.6 Sol Ultra is generally available, and shares a tweet from Ethan Knight claiming the model produced a proof of the Cycle Double Cover Conjecture using 64 subagents in under an hour.
@OpenAI: After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. …
OpenAI deployed GPT-5.6 Sol, achieving 20% lower serving costs and 15%+ better token-generation efficiency through improved GPU kernels and speculative decoding.