OpenRouter's Mystery Union Model Reverse-Engineered: Likely a Qwen4 MoE
unsane_imagination · reddit · 2026-09-17
- A Reddit user ran tokenizer and cutoff tests on Union, a new free stealth model on OpenRouter: 99% identical tokenizer to Qwen3/3.5/4, 256k context, and a training cutoff around Dec 2025-Jan 2026.
- Subjectively it beats qwen3.6 35B A3B and the Nex n2.5 mini built on it, and competes with qwen3.8 27b; reasoning is hidden, suggesting closed weights. The author suspects it's the new small Qwen4 MoE, while cautioning they over-hyped ox stealth before (which turned out to be glm5.3flash). If it's really a 30-40B Qwen4 MoE, it could rip at 25 tok/s on a 16GB M4 with aggressive quantization.
More from Models
- OpenAI discloses six model misalignment incidents and launches a public disclosure framework — TansuYegen · 2026-09-17
- AGI lab researcher: pretraining counts as RLCD — willcb · 2026-09-17
- Developer uses Muse as orchestrator for industrial app, Scale's Wang endorses it — alexandr_wang · 2026-09-17
- Vitalik: Qwen 3.8 flash runs impressively fast on laptops, local-first AI workflows near — anselm · 2026-09-17
- Abliteration explained: HF isn't banning uncensored models, but the author archives them anyway — anselm · 2026-09-17
- Professor: Anthropic's extreme filtering of bio queries is over the top — anshulkundaje · 2026-09-17