kda

Tag

Cards List
#kda

@SemiAnalysis_: Similar to the panic over DeepSeek R1, some uneducated people think Kimi K3’s use of linear attention (KDA) is bad for …

X AI KOLs Following · 4d ago Cached

SemiAnalysis argues that Kimi K3's linear attention (KDA) is not detrimental to NVIDIA, HBM, DRAM, and networking, contrary to uninformed panic, and explains why reduced KV-cache requirements are actually beneficial.

0 favorites 0 likes
#kda

@songhan_mit: We develop an agent-native approach to accelerate genAI, continuing the success of KDA (Kernel Design Agent) at a highe…

X AI KOLs Following · 2026-06-25 Cached

Enze Xie announces Sol Video Inference Engine, an agent-native, training-free full-stack accelerator for video diffusion that auto-tunes cache, sparse attention, token pruning, quantization, and kernel fusion, achieving >2× end-to-end speedup on large models like 64B Cosmos3-Super and 22B LTX-2.3.

0 favorites 0 likes
← Back to home

Submit Feedback