@GergelyOrosz: This is very interesting. Coinbase seems to have lowered their token spend ($$) to about half, by 1) routing to cheap i…

X AI KOLs Following News

Summary

Coinbase reportedly reduced AI token spend by half through smart routing to cheaper models like GLM 5.2 and Kimi 2.7 and implementing caching, highlighting a trend in AI cost optimization.

This is very interesting. Coinbase seems to have lowered their token spend ($$) to about half, by 1) routing to cheap inference like GLM 5.2 and Kimi 2.7 that are still pretty performant 2) Smart routing + caching They still use the same tokens as before. Start of a trend?
Original Article
View Cached Full Text

Cached at: 06/27/26, 06:00 PM

This is very interesting. Coinbase seems to have lowered their token spend ($$) to about half, by

  1. routing to cheap inference like GLM 5.2 and Kimi 2.7 that are still pretty performant

  2. Smart routing + caching

They still use the same tokens as before. Start of a trend?

Brian Armstrong (@brian_armstrong): How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and caching.

Better Defaults (not Usage Caps) – Engineers can choose any model they want, but defaults matter. We’re experimenting with defaulting

Similar Articles

Coinbase and DoorDash shift more workloads to Chinese AI models

Reddit r/ArtificialInteligence

Coinbase and DoorDash are shifting workloads to Chinese AI models like Moonshot's Kimi and Z.ai's GLM, citing significantly lower costs compared to US alternatives, highlighting a broader trend of cost-driven AI adoption.