GPT-6.1 Sol feels unlimited, because it runs at 20 tokens per second
Summary
A user claims that OpenAI's GPT-6.1 Sol generates text about 2.5x slower than GPT-6 Sol and 2.3x slower than GPT-5.6 Sol, suggesting OpenAI is throttling compute for paying customers to reallocate resources to internal experiments.
Similar Articles
GPT-5.6 Sol can run now at an incredible rate of ~750 tokens per second
GPT-5.6 Sol now runs at an impressive inference speed of about 750 tokens per second.
GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence
OpenAI's GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, scoring near GPT-6 Astra on the Intelligence Index while costing less than a quarter per task and pushing the cost-efficiency Pareto frontier.
GPT-5.6 Sol Uses Twice the Tokens of GPT-5.5 (2 minute read)
GPT-5.6 Sol uses more than twice the tokens per session compared to GPT-5.5 in Codex workflows, leading to higher costs and faster depletion of subscription quotas.
GPT-6.1 Sol is cheap. We made it 77% cheaper by never letting it write code
A practical test demonstrates that using GPT-6.1 Sol as an orchestrator without write access and Qwen 3.8 27B as workers reduces costs by 77% but increases execution time for small tasks in AI agent setups.
GPT-6 Sol vs Sonnet 5.5 at the same cost per task: Sol is more efficient, Sonnet 5.5 has the higher ceiling
GPT-6 Sol is more cost-efficient than Sonnet 5.5 at similar budgets, but Sonnet 5.5 achieves higher performance at increased costs, according to Artificial Analysis data.