Union stealth?

Reddit r/LocalLLaMA Models

Summary

A user shares observations about the new 'Union stealth' model on OpenRouter, comparing its performance to Qwen models and speculating it could be a Qwen4 MoE variant with 256k context and a training cutoff in late 2025.

Anyone else trying out the new Union stealth model on openrouter? It’s pretty good for a free model, and of course I had to run tokenizer tests and training cutoff tests which established it’s the 99% the same tokenizer as qwen3/3.5/4 (some multilingual samples ran longer, and some spacing/formatting uses fewer tokens), 256k context, and a training cutoff date of roughly December 2025-January 2026. It’s definitely a step above qwen3.6 35B A3B or that Nex n2.5 mini model based on it, has a newer training cutoff, and easily competes with 3.8 27b in my subjective opinion. Reasoning is hidden, which made me think closed weights, but it’s definitely some sort of Qwen model, and I think it’s not implausible it’s the new small qwen4 MoE. Then again, I got excited about ox stealth and thought it might end up being a poor person model and it turned out to be glm5.3flash, so maybe my perception is off. I’ll have to analyze the agent trace and code it produced more closely to see if it’s qwen3.8 27b level or if it’s clearly a smarter model. Harder to do without CoT visible, but you can usually see model family traits in code. But… if this is qwen4 MoE 30-40b, we’re gonna eat like royalty soon enough. Can wait to throw a brain dead iq2xs quant with 128k q4_0 kv cache on my 16gb M4 and watch it rip at 25 tok/s!
Original Article

Similar Articles

New stealth model: Union Alpha

Reddit r/singularity

A new stealth AI model named Union Alpha has been released for free on OpenRouter and OpenCode, featuring multimodal capabilities with a 256K context and claiming frontier-level general-purpose performance, sparking speculation about its origin.

Are we getting Qwen 3.8 35-A3B?

Reddit r/LocalLLaMA

Speculation about the upcoming Qwen 3.8 release, questioning whether it will be a dense 27B model or a MoE variant like the previous 35B-A3B, with discussion of performance implications for local hardware.