@GergelyOrosz: This is very interesting. Coinbase seems to have lowered their token spend ($$) to about half, by 1) routing to cheap i…
Summary
Coinbase reportedly reduced AI token spend by half through smart routing to cheaper models like GLM 5.2 and Kimi 2.7 and implementing caching, highlighting a trend in AI cost optimization.
View Cached Full Text
Cached at: 06/27/26, 06:00 PM
This is very interesting. Coinbase seems to have lowered their token spend ($$) to about half, by
-
routing to cheap inference like GLM 5.2 and Kimi 2.7 that are still pretty performant
-
Smart routing + caching
They still use the same tokens as before. Start of a trend?
Brian Armstrong (@brian_armstrong): How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and caching.
Better Defaults (not Usage Caps) – Engineers can choose any model they want, but defaults matter. We’re experimenting with defaulting
Similar Articles
Coinbase and DoorDash shift more workloads to Chinese AI models
Coinbase and DoorDash are shifting workloads to Chinese AI models like Moonshot's Kimi and Z.ai's GLM, citing significantly lower costs compared to US alternatives, highlighting a broader trend of cost-driven AI adoption.
@rohanpaul_ai: Coinbase CEO Brian Armstrong said Coinbase is experimenting with defaulting to Chinese open-weight models such as GLM 5…
Coinbase CEO Brian Armstrong announced the company is experimenting with using Chinese open-weight AI models like GLM 5.2 and Kimi 2.7 for its LLM gateway, routing prompts by difficulty, suggesting that frontier models may be overkill for execution tasks.
@DeRonin_: https://x.com/DeRonin_/status/2054235707791778034
A practical guide on reducing AI coding expenses by 80% through smarter token management, including multi-model routing, prompt caching, and context discipline, rather than simply switching to cheaper models.
@GergelyOrosz: Talked with a company where, a year ago they had unlimited AI budgets + CEO is technical and very bullish in AI “We now…
A company with a strong AI focus has implemented daily budget limits for state-of-the-art models, shifting to cheaper models for most tasks, indicating a trend in AI cost management.
@DeRonin_: My entire AI stack is now Chinese 87% cheaper. same revenue swaps by task: 1. reasoning / backend brain Opus 4.8 → Kimi…
A user reports replacing American AI models with Chinese alternatives across reasoning, code generation, agent loops, bulk processing, and image/video generation, achieving 87% cost reduction with only 4% average quality drop and unchanged revenue.