Intelligence density went up a lot this year and my bill didn't move. The $/M number is not where the money goes.
Summary
The author reflects on how AI model pricing per token has dropped dramatically, but real-world costs remain flat because cheaper models get re-run more often. They argue that cost per completed step is the metric that matters, not cost per million tokens.
Similar Articles
Intelligence Per Dollar (2 minute read)
Microsoft introduces 'average token usage' as a new metric on model release cards to measure intelligence per dollar, shifting AI competition toward efficiency and cost-effectiveness. This metric benchmarks models on both performance and the cost of achieving that intelligence.
Why Your AI Bill Went Up Even Though Token Prices Are Falling (5 minute read)
Token prices have fallen dramatically but enterprise AI bills are rising due to increased token consumption from agents and background inference, illustrating Jevons paradox.
Price per 1M tokens is meaningless
This article argues that comparing AI models by price per million tokens is misleading due to differences in tokenizers and token efficiency. It provides a benchmark cost analysis showing that models with higher per-token prices can be cheaper per completed task, with DeepSeek V4 Pro being a strong cost-efficiency outlier.
Every AI prompt costs money — and that changes everything
The article argues that the real challenge in AI isn't just building smarter models but making them cost-efficient at scale, highlighting the importance of reducing token usage, improving speed, and optimizing infrastructure.
How I'm charged for AI usage feels broken.
The author argues that current AI usage pricing models are broken because users are charged for hidden 'thinking' tokens that are not visible to them, creating a trust-me billing system. They propose that labs should either adjust output token pricing or bill explicitly for compute.