@FinanceYF5: 1/ The cost of intelligence is collapsing. The price of LLM intelligence has dropped 100-fold in 18 months; this is a fact. But David says pessimists stop there—they miss the second half: as inputs get cheaper, demand expands outward.
Summary
The article notes that the price of LLM intelligence has dropped 100-fold in 18 months, and argues that this cost reduction will drive demand to expand outward, countering purely pessimistic views.
View Cached Full Text
Cached at: 05/10/26, 04:27 PM
1/⚡ The cost of intelligence is collapsing
The price of LLM intelligence has dropped 100-fold in 18 months. That’s a fact.
But as David points out, pessimists stop at this observation—missing the second half of the story: as inference costs plummet, demand expands outward. https://t.co/Iv6F6zaCfV
Similar Articles
Why current LLM costs are not sustainable
The article argues that current high LLM pricing is unsustainable due to diminishing performance gains, the rise of open-weight models, specialized AI chips reducing inference costs, and zero switching costs, predicting significant price drops as competition intensifies.
I track LLM prices every 3 hours. GLM-5.2 quietly went from ~$0.57/$1.80 to $0.90/$3.08 per 1M this week, with no announcement.
A price tracker found that GLM-5.2 and Tencent's Hy3 quietly changed prices multiple times in a week, highlighting volatility in Chinese LLM pricing.
What happens when they stop subsidizing LLM subscriptions?
A commentary on the unsustainable subsidization of LLM subscriptions, predicting price hikes and ecosystem shifts as VC funding tightens, with concerns about open-source model availability and hardware costs.
@rohanpaul_ai: Yann LeCun says LLMs aren’t a bubble in value or investment—they’ll drive many real-world applications and justify curr…
Yann LeCun argues that LLMs are not a bubble in value or investment, as they will drive many real-world applications and justify current infrastructure spending; the actual bubble is in assuming LLMs can achieve human-level thinking.
@h100envy: Ex-vLLM core contributor explained how to make LLM inference 10x cheaper in 34 minutes - better than $3000 inference op…
An ex-vLLM core contributor explains how to reduce LLM inference cost by 10x using LMCache with KV cache offloading to CPU/SSD/remote storage, a technique used by production stacks like Bloomberg.