token-usage

Tag

Cards List
#token-usage

@FinanceYF5: Latest Update: According to Bloomberg, as the token usage of Agent has surged 20-fold, Harvey's costs for renting OpenA…

X AI KOLs Timeline · 2d ago

According to Bloomberg, Harvey's costs for renting OpenAI and Anthropic models have risen sharply due to a 20-fold surge in token usage, causing gross margins to drop from about 50% to -50% by June.

0 favorites 0 likes
#token-usage

@omooretweets: We’re entering the “$5 Uber” era of consumer - inference edition VC $ subsidizing huge token usage on products aiming t…

X AI KOLs Timeline · 4d ago Cached

The article discusses the trend of venture capital subsidizing token usage in consumer AI products, likening it to the '$5 Uber' model for long-term monetization and competitive dominance.

0 favorites 0 likes
#token-usage

@10xmylife: cool 啊

X AI KOLs Following · 5d ago Cached

An AI system named 'jev' is showcased playing Super Smash Bros. against itself, controlling multiple characters and making rapid decisions within fractions of a second, using over 22 million tokens.

0 favorites 0 likes
#token-usage

Question: UkisAI Swift Ternary Bonsai 2 27B?

Reddit r/LocalLLaMA · 6d ago

Jovan from UkisAI discusses improvements in their Swift Qwen3.8 27B model and seeks community feedback on creating a Swifted version of Bonsai 2 to address overthinking loops and high token usage.

0 favorites 0 likes
#token-usage

@VadimStrizheus: i just spent $260,000 in API credits on Codex in the last 24 hours I used up 5.2B tokens with GPT 6 Astra + GPT 5.6 Sol…

X AI KOLs Following · 6d ago Cached

A user reports spending $260,000 in API credits on Codex within 24 hours, utilizing 5.2 billion tokens with GPT 6 Astra and GPT 5.6 Sol models.

0 favorites 0 likes
#token-usage

1.3B token used recently, qwen3.8 27b

Reddit r/LocalLLaMA · 2026-09-16

A Reddit post highlights the strong performance of the Qwen 3.8 27B model, with mention of 1.3B tokens used recently.

0 favorites 0 likes
#token-usage

@FinanceYF5: Last week's two chart-toppers on OpenRouter, completely different: The one that spent the most money was GPT-6 Astra. T…

X AI KOLs Timeline · 2026-09-16

The tweet reports that GPT-6 Astra had the highest spending on OpenRouter last week, while GPT-5.6 Luna used the most tokens, leading by a wide margin.

0 favorites 0 likes
#token-usage

@gdb: Astra and Luna are taking off

X AI KOLs Timeline · 2026-09-15 Cached

The article reports that Astra is the top AI model by spending last week, while Luna leads in token usage, highlighting trends between Anthropic and OpenAI.

0 favorites 0 likes
#token-usage

@Chris_Wozniczek: I'm building a plugin for Devin Desktop and CLI actually it was 85% done and needed testing, so devin was testing this …

X AI KOLs Timeline · 2026-09-12 Cached

A developer is building a plugin for Devin Desktop and CLI, testing it with Devin, and sharing insights from extensive token consumption and lessons learned.

0 favorites 0 likes
#token-usage

@Xiaomi: For AI, we have our own large language model, Xiaomi MiMo. Xiaomi MiMo-V2.5 ranked top globally by monthly token usage …

X AI KOLs Following · 2026-09-03 Cached

Xiaomi's MiMo-V2.5 large language model ranked top globally by monthly token usage on OpenRouter in July, driving AI integration into smart manufacturing and the Human x Car x Home strategy.

0 favorites 0 likes
#token-usage

@FinanceYF5: OpenRouter has once again weathered a "perfectly ordinary" month. The platform's weekly Token usage has surged from 4.6…

X AI KOLs Following · 2026-09-02

OpenRouter's weekly token usage has surged from 4.6 trillion a year ago to 113 trillion, with a doubling in the past month, indicating significant growth in the AI platform's adoption.

0 favorites 0 likes
#token-usage

@tetsuoai: We’re giving all @Grok @Bot users another free reset on token usage

X AI KOLs Following · 2026-09-01 Cached

Elon Musk announced that all users of Grok Bot will receive another free reset on token usage.

0 favorites 0 likes
#token-usage

@gdb: Jevons Paradox is counterintuitive and inspiring

X AI KOLs Timeline · 2026-08-28 Cached

The post highlights how discounting GPT 5.6 models on OpenRouter led to a 13.8x increase in token usage, illustrating the Jevons Paradox where efficiency gains result in higher total resource consumption.

0 favorites 0 likes
#token-usage

GPT 5.6 Discounts & Jevons Paradox (4 minute read)

TLDR AI · 2026-08-28 Cached

OpenAI's discount program on Terra and Luna models led to a sharp increase in token usage, competitive displacement from other labs, and notable user retention after discounts ended.

0 favorites 0 likes
#token-usage

@zainhas: PSA: for everyone using GLM-5.3 Flash just use "high" reasoning_effort unless you're asking it to curing cancer or some…

X AI KOLs Timeline · 2026-08-27 Cached

The tweet provides a public service announcement advising users of the GLM-5.3 Flash AI model to use the 'high' reasoning_effort setting for better efficiency, as it achieves similar accuracy with significantly fewer tokens compared to the 'max' setting.

0 favorites 0 likes
#token-usage

Stop shortening your prompts. Six agents, 97-99% cache hit rate - and why the standard advice is backwards.

Reddit r/AI_Agents · 2026-08-27

The article argues that with prompt caching, longer, stable prompts can be cheaper than frequently changing short ones, sharing insights from running AI agents with high cache hit rates.

0 favorites 0 likes
#token-usage

@jietang: Ox Alpha = GLM-5.3 Flash AA = 57 , 1/100 frontier price, Powered by pure Chinese chips. Delivered nearly 20% weekly tok…

X AI KOLs Timeline · 2026-08-27 Cached

Ox Alpha, based on GLM-5.3 Flash AA, is announced at one-hundredth the price of frontier models, powered by Chinese chips, and has achieved nearly 20% weekly token share on OpenRouter.

0 favorites 0 likes
#token-usage

Most Used OpenRouter Models Over Time

Reddit r/ArtificialInteligence · 2026-08-26 Cached

A video visualizes the top 20 AI models by weekly token usage on OpenRouter from December 2024 to August 2026, showing a shift from US to Chinese labs dominating the leaderboard.

0 favorites 0 likes
#token-usage

@jieyi_ai: 80-100 USD per day for unlimited use, 20 RMB at this unmatched cost-effectiveness. An emergency tool for small companies.

X AI KOLs Timeline · 2026-08-26 Cached

This article introduces a cost-effective tool for accessing AI large models, ideal for emergency use by small companies, at a daily cost of only 80-100 USD or 20 RMB.

0 favorites 0 likes
#token-usage

@cHHillee: I've realized when people are talking about how many tokens they use, they're usually including cached input tokens... …

X AI KOLs Following · 2026-08-24

The author criticizes the common practice of including cached input tokens in discussions about token usage in AI models.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback