Models are now training models.
Summary
Intology's Locus system post-trained Qwen3 base models beyond the Qwen3 instruct checkpoint, following the PostTrainBench setting.
Similar Articles
@intology: The models are improving the models. Locus, our automated AI research system, is SOTA on PostTrainBench and post-trains…
Locus, an automated AI research system, achieves SOTA on PostTrainBench and post-trains Qwen3 base models that surpass human post-trained models. It also shows strong performance on Kaggle competitions.
@no_stp_on_snek: model testing logpost in the past few days been training more models, working towards something where my TUI can have a…
A detailed logpost sharing lessons learned from training four models across three families, covering invariants in LLM fine-tuning and architecture-specific challenges such as reasoning model evaluation traps, quantization effects, and the waterbed effect of behavioral fine-tuning.
Qwen 3.7 Max
Qwen 3.7 is an impressive new AI model from Chinese labs, with discussion on whether weights will be available for download.
I RL-trained Qwen3.6-35B-A3B to RL-train small task-specific Qwen models. Fully open source! 🤓
The author trained a Qwen3.6-35B-A3B model using reinforcement learning to then RL-train small task-specific Qwen models, and has released everything fully open source.
@no_stp_on_snek: Just a few hours away from Qwen 3.8! Clear your benches! I’ll be working on behavioral tests and comparisons against 3.…
Qwen3.8-27B is a new AI model with enhanced capabilities in coding, agentic tasks, and vision-language understanding, offering flexible thinking control and long context lengths. It is available on Hugging Face and designed for deployment-friendly use.