@cjzafir: Models that I'm using daily: > Codex 5.5 high (fast) > Deepseek v4 pro via API > Kimi 2.6 via API Models that I am fine…
Summary
User shares a personal list of AI models they use daily (Codex 5.5, Deepseek v4 pro, Kimi 2.6) and for fine-tuning (Qwen 3.5 variants, Gemma4 E4B, GPT-oss 20B), aiming to fine-tune Small Language Models into Expert Language Models.
Similar Articles
@cjzafir: Qwen 3.5 4B model and 8B are too good. I fine-tuned a 4B model today and got 98% accuracy on full precision and Q8 quan…
A developer reports achieving high accuracy with fine-tuned Qwen 3.5 4B and 8B models using Unsloth, suggesting a shift towards specialized Expert Language Models (ELMs) for niche tasks.
@0xSero: DeepSeek-V4-Pro & Kimi-K2.6 running in Codex app. Cheapest way to taste the frontier. Works w local models and they can…
The Codex app now supports DeepSeek-V4-Pro and Kimi-K2.6, offering the cheapest way to use frontier AI models, with local model support and computer-use capabilities.
Less Than a Month: Kimi K3, Qwen3.8, DeepSeek-V4-Pro-0813, GLM-5.3
Multiple leading Chinese AI labs have released new flagship models within the past month, including Kimi K3, Qwen3.8, DeepSeek-V4-Pro, and GLM-5.3, signaling rapid progress in China's AI race.
DeepSeek 0813 "Pro" vs GLM 5.2 & Kimi K3 🐋
A comparison of DeepSeek 0813 'Pro' against GLM 5.2 and Kimi K3, likely covering benchmark performance and capability differences between these AI models.
@jtdavies: Coding on small models... My default model for my 4xDGX Spark cluster is @UnslothAI's Qwen3.6-35B-A3B-NVFP4. I get exce…
A user tests various small AI models for coding tasks, finding Qwen3.6-27B-NVFP4 to be the best balance of speed and accuracy, and notes poor Java performance in these models.