Tag
A Twitter thread speculates that Anthropic's extension of Claude Fable 5 access is driven by competitive pressure from OpenAI's upcoming GPT-5.6, not generosity, and predicts Fable will remain on subscription until a cheaper replacement is available.
A price tracker found that GLM-5.2 and Tencent's Hy3 quietly changed prices multiple times in a week, highlighting volatility in Chinese LLM pricing.
Deepseek has raised prices for its Flash model, which was previously the only genuinely cheap and capable option. This price change raises concerns about the upcoming V4 model.
The article argues that current high LLM pricing is unsustainable due to diminishing performance gains, the rise of open-weight models, specialized AI chips reducing inference costs, and zero switching costs, predicting significant price drops as competition intensifies.
An analysis questioning whether OpenRouter's API pricing for open models like GLM-5.2 implies more aggressive quantization than assumed, given the economics of running large models on expensive hardware like 8xH200.
Chinese AI models like DeepSeek and Qwen deliver competitive performance at 5x–20x lower cost than Western counterparts, reshaping the economics of AI and driving multi-model deployment strategies.
Google's Gemini 3.5 Flash is priced three times higher than its predecessor and 30 times more expensive than Gemini 1.5 Flash, raising concerns about cost scaling for future models.