Cached at:
07/28/26, 06:22 PM
# Thread by @Kimi_Moonshot on Thread Reader App
Source: [https://threadreaderapp.com/thread/2081760186235289764.html](https://threadreaderapp.com/thread/2081760186235289764.html)
## More from @Kimi\_Moonshot
[](https://threadreaderapp.com/user/Kimi_Moonshot)
Jul 16
Introducing Kimi K3: Open Frontier Intelligence
πΉ 2\.8 Trillion Parameters, 1 Million Context, Native Multimodal πΉ Kimi Delta Attention enables up to 6\.3x faster decoding in million\-token contexts πΉ Attention Residuals deliver ~25% higher training efficiency at <2% additional cost πΉ Built for long\-horizon agentic coding and self\-evolving workflows
Kimi K3 is now live on on[Kimi\.com](http://kimi.com/), Kimi Work, Kimi Code, and the Kimi API\. Open Weights by July 27, 2026\.
π API:[platform\.kimi\.ai](http://platform.kimi.ai/) π Tech blog:[kimi\.com/blog/kimi\-k3](http://kimi.com/blog/kimi-k3)[](https://pbs.twimg.com/media/HNXu0kobMAAPljb.png) [](https://pbs.twimg.com/media/HNXu2GWaYAAwH4w.png)
K3 is built on Kimi Delta Attention \(KDA\) and Attention Residuals \(AttnRes\), two architectural updates designed to improve how information flows across sequence length and model depth\.
We have also scaled up Mixture of Experts \(MoE\) sparsity, effectively activating 16 out of 896 experts when paired with a Stable LatentMoE framework\.
Together with refined training and data recipes, these structural changes yield an approximate 2\.5Γ improvement in overall scaling efficiency compared to K2, allowing the model to convert compute into intelligence more effectively\.[](https://pbs.twimg.com/media/HNXu7q4akAABQAM.jpg)
Internal knowledge work bench
Beyond public benchmarks, Kimi K3 Max also shows consistent gains on our internal benchmarks, which are built from recurring patterns and challenges in real\-world user\-agent workflows\.
It scores 75\.5 on Online Exp Bench, 73\.5 on DECK\-Bench, and 62\.6 on Finance\-Bench, outperforming Claude Opus 4\.8 \(max\) and GPT\-5\.5 \(xhigh\) across all three\.
These results reflect broad improvements in Kimi K3's agentic knowledge work capabilities, enabling more capable and reliable performance in real\-world use cases\.[](https://pbs.twimg.com/media/HNXvBPrbgAABDsU.jpg)
Read 6 tweets
[](https://threadreaderapp.com/user/Kimi_Moonshot)
May 14
Meet Kimi Web Bridge \- Kimi's browser extension\.
Agent can now interact with websites like a human: search, scroll, click, type and complete tasks\.
Supports Kimi Code CLI, Claude Code, Cursor, Codex, Hermes, and more\.
Available now on and the Chrome Web Store\.[kimi\.com/features/webbrβ¦](http://kimi.com/features/webbridge) 
Search across multiple platforms at scale and auto\-fill results directly into your spreadsheet\. 
With K2\.6's multimodal capability, your agent will open a website, navigate through it, and replicate it\. 
Read 6 tweets
[](https://threadreaderapp.com/user/Kimi_Moonshot)
Apr 20
Meet Kimi K2\.6 agent \- Video hero section, WebGL shaders, real backends\. From one prompt\.
πΉ Video hero sections \- cinematic aesthetic, auto\-composited πΉ WebGL shader animations \- native GLSL / WGSL, liquid metal, caustics, raymarching πΉ Motion design \- GSAP \+ Framer Motion πΉ Backend database: Kimi wires up auth \+ database \+ backend in one pass\. πΉ Website stack \- React 19 \+ TypeScript \+ Vite \+ Tailwind \+ shadcn/ui πΉ 3D w/ physically\-based lighting \- Three\.js \+ React Three Fiber 
Video hero sections, built right in\.
K2\.6 agent calls video generation APIs to create real cinematic footage for your hero, not stock placeholders\. Composited into the page, synced to scroll, with shader overlays\. 
Speaks fluent WebGL shader\.
Writes GLSL / WGSL directly \- fragment shaders, vertex shaders, noise, SDF, raymarching\. Prompt: "a liquid\-metal hero with soft caustics\." 
Read 7 tweets
[](https://threadreaderapp.com/user/Kimi_Moonshot)
Mar 16
Introducing π¨ππππππππ πΉππππ
ππππ: Rethinking depth\-wise aggregation\.
Residual connections have long relied on fixed, uniform accumulation\. Inspired by the duality of time and depth, we introduce Attention Residuals, replacing standard depth\-wise recurrence with learned, input\-dependent attention over preceding layers\.
πΉ Enables networks to selectively retrieve past representations, naturally mitigating dilution and hidden\-state growth\. πΉ Introduces Block AttnRes, partitioning layers into compressed blocks to make cross\-layer attention practical at scale\. πΉ Serves as an efficient drop\-in replacement, demonstrating a 1\.25x compute advantage with negligible \(<2%\) inference latency overhead\. πΉ Validated on the Kimi Linear architecture \(48B total, 3B activated parameters\), delivering consistent downstream performance gains\.
πFull report: [github\.com/MoonshotAI/Attβ¦](https://github.com/MoonshotAI/Attention-Residuals/blob/master/Attention_Residuals.pdf)[](https://pbs.twimg.com/media/HDgCpkHb0AA0a7_.jpg)
Scaling law experiments reveal a consistent 1\.25Γ compute advantage across varying model sizes\.[](https://pbs.twimg.com/media/HDgCt49a8AA14V_.png)
Analysis of training dynamics demonstrates how AttnRes naturally mitigates hidden\-state magnitude growth and yields a more uniform gradient distribution across depth\.[](https://pbs.twimg.com/media/HDgCxm0bUAA7KrD.jpg)
Read 4 tweets
[](https://threadreaderapp.com/user/Kimi_Moonshot)
Nov 28, 2025
Meet Kimi Agentic Slides\! Now with Nano Banana Pro π
π Thanksgiving Gift: 48H FREE & UNLIMITED ACCESS
πΈ Agentic search \(Kimi K2\) πΈ Files β Slides \(PDFs, images, docs\+\) πΈ Fully editable \+ PPTX export πΈ Designer\-level visuals \(infographics, illustrations\)
Try now:[kimi\.com/slides](https://www.kimi.com/slides)
Here's a quick guide\. π
Research paper \-\> Presentation Ready Deck[](https://pbs.twimg.com/media/G61dkYva0AARFcB.jpg)
Read 5 tweets