Tag
AI spending per employee slumped at top firms in August, raising concerns about whether this is a seasonal slowdown or a warning sign for AI revenue growth amid falling token costs and slower adoption.
The TeamoRouter platform provides free access to AI models like Ox Alpha, along with intelligent routing and discounts up to 90%, making it convenient for developers to access multiple models.
Mistral is now hosting competitor Z.ai's GLM-5.2, pricing it cheaper than its own flagship Mistral Medium 3.5, sparking speculation about a strategic pivot toward compute sales and smaller specialized models.
The author gives a lukewarm take on the 3.7 flash model, noting the price is still high for a flash model but no longer completely unreasonable.
A Twitter thread speculates that Anthropic's extension of Claude Fable 5 access is driven by competitive pressure from OpenAI's upcoming GPT-5.6, not generosity, and predicts Fable will remain on subscription until a cheaper replacement is available.
A price tracker found that GLM-5.2 and Tencent's Hy3 quietly changed prices multiple times in a week, highlighting volatility in Chinese LLM pricing.
Deepseek has raised prices for its Flash model, which was previously the only genuinely cheap and capable option. This price change raises concerns about the upcoming V4 model.
The article argues that current high LLM pricing is unsustainable due to diminishing performance gains, the rise of open-weight models, specialized AI chips reducing inference costs, and zero switching costs, predicting significant price drops as competition intensifies.
An analysis questioning whether OpenRouter's API pricing for open models like GLM-5.2 implies more aggressive quantization than assumed, given the economics of running large models on expensive hardware like 8xH200.
Chinese AI models like DeepSeek and Qwen deliver competitive performance at 5x–20x lower cost than Western counterparts, reshaping the economics of AI and driving multi-model deployment strategies.
Google's Gemini 3.5 Flash is priced three times higher than its predecessor and 30 times more expensive than Gemini 1.5 Flash, raising concerns about cost scaling for future models.