All oneshots from Kimi-K3, looks better than opus4.8.

Reddit r/LocalLLaMA Models

Summary

A user ran Kimi K3 through 34 oneshot prompts and found it outperformed Opus 4.8 in HTML/screenshot/gif generation evals while being far cheaper.

I've ran Kimi-k3 through 34 oneshot prompts and evaluated the generated htmls, screenshots and gifs using sonnet 4.6. It came out to be better than opus4.8 from the evals. Kimi K3: https://oneshotlm.com/model/moonshotai-kimi-k3/ Opus 4.8: https://oneshotlm.com/model/anthropic-claude-opus-4-8/ Also opus costed $7.16 to go through all 34 prompts whereas $0.44 for kimi k3, so its token efficient as well. More evaluations to come.
Original Article

Similar Articles

Kimi K2.6 is a legit Opus 4.7 replacement

Reddit r/LocalLLaMA

A user reports that Kimi K2.6 is a strong alternative to Claude Opus 4.7, capable of handling ~85% of tasks at comparable quality while offering vision and browser-use capabilities, suggesting frontier models may not always offer unique advantages.

I got Kimi-k3 running.....

Reddit r/LocalLLaMA

User successfully runs the Kimi-k3 model using llama.cpp on high-end hardware, achieving low tokens per second (0.41 prompt eval, 0.23 generation).