256k-context

Tag

Cards List
#256k-context

Kimi K3-256k

Hacker News Top · 2026-07-29 Cached

Kimi Code releases Kimi K3-256k, a 256k-context version of its flagship K3 coding model, offering reduced quota consumption while maintaining similar performance for most tasks.

0 favorites 0 likes
#256k-context

@ciruai: Finally 256k context for 16GB cards on smart models at high speeds! Using a 4080 Super 16GB I show you how to get full …

X AI KOLs Timeline · 2026-07-16 Cached

Demonstrates achieving 256k context on a 16GB RTX 4080 Super using Ternary Bonsai 27B Q2_0 model with llama.cpp, achieving up to 141 tok/s generation speed.

0 favorites 0 likes
#256k-context

@AdinaYakup: Keye VL 2.0-30B-A3B New multimodal model from @KwaiKeye 30B/3B active - Apache 2.0 256K context via DeepSeek Sparse Att…

X AI KOLs Following · 2026-06-01 Cached

KwaiKeye releases Keye VL 2.0-30B-A3B, a multimodal model with 30B total / 3B active parameters, 256K context via DeepSeek Sparse Attention, and Apache 2.0 license, claiming it matches Qwen3 VL and Gemini 3 in accuracy.

0 favorites 0 likes
#256k-context

@iotcoi: Qwen3.6-27B-FP8 + Dflash + DDTree, 256k context, 10 agents ~200 tokens/sec max decode 136t/s average on a single tiny G…

X AI KOLs Timeline · 2026-04-22 Cached

Quantized 27B Qwen3.6 model achieves 200 tok/s peak (136 avg) with 256k context and 10 agents on a single 49W GB10 GPU using Dflash+DDTree optimizations.

0 favorites 0 likes
← Back to home

Submit Feedback