token-usage

Tag

Cards List
#token-usage

@rauchg: Most token aggregators have extreme noise from ① token promos from/to providers that train on your data or ② providers …

X AI KOLs Timeline ↗ · 12h ago Cached

Vercel AI Gateway的CEO @rauchg 表示,大多数token聚合器充斥着噪音,包括那些要求你把数据用于训练的token促销,以及声称支持ZDR(零数据保留)但信誉存疑的提供商;而Vercel AI Gateway凭借真实客户用量、40万+付费客户规模和零加价,成为理解全球AI token流向的最可信数据源。

0 favorites 0 likes
#token-usage

@heyshrutimishra: By 2030, AI agents will use 100x more tokens than humans. Right now most AI usage is limited by how fast we can type an…

X AI KOLs Timeline ↗ · 2d ago Cached

The tweet discusses predictions that AI agents will use 100x more tokens than humans by 2030 and introduces Fo, a personal AI product that combines AI with human assistance for real-world tasks.

0 favorites 0 likes
#token-usage

@thermalpastor: The scariest thing about Opus 5.5 is that it goes back to fill in the arm realising the lines were dotted.

X AI KOLs Timeline ↗ · 5d ago Cached

The tweet compares Opus 5.5's performance to Astra on a task, noting it performs better but uses 10x more tokens and draws faster.

0 favorites 0 likes
#token-usage

@FinanceYF5: Latest Update: According to Bloomberg, as the token usage of Agent has surged 20-fold, Harvey's costs for renting OpenA…

X AI KOLs Timeline ↗ · 2026-09-22

According to Bloomberg, Harvey's costs for renting OpenAI and Anthropic models have risen sharply due to a 20-fold surge in token usage, causing gross margins to drop from about 50% to -50% by June.

0 favorites 0 likes
#token-usage

@omooretweets: We’re entering the “$5 Uber” era of consumer - inference edition VC $ subsidizing huge token usage on products aiming t…

X AI KOLs Timeline ↗ · 2026-09-19 Cached

The article discusses the trend of venture capital subsidizing token usage in consumer AI products, likening it to the '$5 Uber' model for long-term monetization and competitive dominance.

0 favorites 0 likes
#token-usage

@10xmylife: cool 啊

X AI KOLs Following ↗ · 2026-09-19 Cached

An AI system named 'jev' is showcased playing Super Smash Bros. against itself, controlling multiple characters and making rapid decisions within fractions of a second, using over 22 million tokens.

0 favorites 0 likes
#token-usage

Question: UkisAI Swift Ternary Bonsai 2 27B?

Reddit r/LocalLLaMA ↗ · 2026-09-18

Jovan from UkisAI discusses improvements in their Swift Qwen3.8 27B model and seeks community feedback on creating a Swifted version of Bonsai 2 to address overthinking loops and high token usage.

0 favorites 0 likes
#token-usage

@VadimStrizheus: i just spent $260,000 in API credits on Codex in the last 24 hours I used up 5.2B tokens with GPT 6 Astra + GPT 5.6 Sol…

X AI KOLs Following ↗ · 2026-09-18 Cached

A user reports spending $260,000 in API credits on Codex within 24 hours, utilizing 5.2 billion tokens with GPT 6 Astra and GPT 5.6 Sol models.

0 favorites 0 likes
#token-usage

1.3B token used recently, qwen3.8 27b

Reddit r/LocalLLaMA ↗ · 2026-09-16

A Reddit post highlights the strong performance of the Qwen 3.8 27B model, with mention of 1.3B tokens used recently.

0 favorites 0 likes
#token-usage

@FinanceYF5: Last week's two chart-toppers on OpenRouter, completely different: The one that spent the most money was GPT-6 Astra. T…

X AI KOLs Timeline ↗ · 2026-09-16

The tweet reports that GPT-6 Astra had the highest spending on OpenRouter last week, while GPT-5.6 Luna used the most tokens, leading by a wide margin.

0 favorites 0 likes
#token-usage

@gdb: Astra and Luna are taking off

X AI KOLs Timeline ↗ · 2026-09-15 Cached

The article reports that Astra is the top AI model by spending last week, while Luna leads in token usage, highlighting trends between Anthropic and OpenAI.

0 favorites 0 likes
#token-usage

@Chris_Wozniczek: I'm building a plugin for Devin Desktop and CLI actually it was 85% done and needed testing, so devin was testing this …

X AI KOLs Timeline ↗ · 2026-09-12 Cached

A developer is building a plugin for Devin Desktop and CLI, testing it with Devin, and sharing insights from extensive token consumption and lessons learned.

0 favorites 0 likes
#token-usage

@Xiaomi: For AI, we have our own large language model, Xiaomi MiMo. Xiaomi MiMo-V2.5 ranked top globally by monthly token usage …

X AI KOLs Following ↗ · 2026-09-03 Cached

Xiaomi's MiMo-V2.5 large language model ranked top globally by monthly token usage on OpenRouter in July, driving AI integration into smart manufacturing and the Human x Car x Home strategy.

0 favorites 0 likes
#token-usage

@FinanceYF5: OpenRouter has once again weathered a "perfectly ordinary" month. The platform's weekly Token usage has surged from 4.6…

X AI KOLs Following ↗ · 2026-09-02

OpenRouter's weekly token usage has surged from 4.6 trillion a year ago to 113 trillion, with a doubling in the past month, indicating significant growth in the AI platform's adoption.

0 favorites 0 likes
#token-usage

@tetsuoai: We’re giving all @Grok @Bot users another free reset on token usage

X AI KOLs Following ↗ · 2026-09-01 Cached

Elon Musk announced that all users of Grok Bot will receive another free reset on token usage.

0 favorites 0 likes
#token-usage

@gdb: Jevons Paradox is counterintuitive and inspiring

X AI KOLs Timeline ↗ · 2026-08-28 Cached

The post highlights how discounting GPT 5.6 models on OpenRouter led to a 13.8x increase in token usage, illustrating the Jevons Paradox where efficiency gains result in higher total resource consumption.

0 favorites 0 likes
#token-usage

GPT 5.6 Discounts & Jevons Paradox (4 minute read)

TLDR AI ↗ · 2026-08-28 Cached

OpenAI's discount program on Terra and Luna models led to a sharp increase in token usage, competitive displacement from other labs, and notable user retention after discounts ended.

0 favorites 0 likes
#token-usage

@zainhas: PSA: for everyone using GLM-5.3 Flash just use "high" reasoning_effort unless you're asking it to curing cancer or some…

X AI KOLs Timeline ↗ · 2026-08-27 Cached

The tweet provides a public service announcement advising users of the GLM-5.3 Flash AI model to use the 'high' reasoning_effort setting for better efficiency, as it achieves similar accuracy with significantly fewer tokens compared to the 'max' setting.

0 favorites 0 likes
#token-usage

Stop shortening your prompts. Six agents, 97-99% cache hit rate - and why the standard advice is backwards.

Reddit r/AI_Agents ↗ · 2026-08-27

The article argues that with prompt caching, longer, stable prompts can be cheaper than frequently changing short ones, sharing insights from running AI agents with high cache hit rates.

0 favorites 0 likes
#token-usage

@jietang: Ox Alpha = GLM-5.3 Flash AA = 57 , 1/100 frontier price, Powered by pure Chinese chips. Delivered nearly 20% weekly tok…

X AI KOLs Timeline ↗ · 2026-08-27 Cached

Ox Alpha, based on GLM-5.3 Flash AA, is announced at one-hundredth the price of frontier models, powered by Chinese chips, and has achieved nearly 20% weekly token share on OpenRouter.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback