tps

Tag

Cards List
#tps

ThinkingCap-Qwen3.6-27B warrants a look

Reddit r/LocalLLaMA ↗ · 2026-07-28

User reports improved tokens per second (tps) with ThinkingCap-Qwen3.6-27B compared to Qwen3.5-27B, with no quality loss, recommending it as a daily driver until the next Qwen release.

0 favorites 0 likes
#tps

@scaling01: DeepSeek just made their inference ~5x cheaper at 50 TPS

X AI KOLs Following ↗ · 2026-06-27 Cached

DeepSeek has reduced inference costs by approximately 5x while maintaining 50 tokens per second throughput.

0 favorites 0 likes
#tps

1000 tps generation on Qwen3.6 27B with V100s

Reddit r/LocalLLaMA ↗ · 2026-05-25

Achieved 1000 tokens per second generation on Qwen3.6 27B using V100 GPUs with 128 concurrent requests, and 80 t/s for single user.

0 favorites 0 likes
← Back to home

Submit Feedback