Reflection's Billion-Dollar US Open-Weight Model Underperforms Every Major Chinese Model, Mocked Online
npinto · x · 2026-10-06
Reflection AI's new US-made open-weight model scores worse than every major Chinese model, prompting mockery despite billions raised. Hossein quipped that its "advancing the Western open frontier" claim is "like congratulating yourself for becoming the US national Mahjong champion," while one user jokingly "fixed" the outdated benchmarks to congratulate the lab.
More from Models
- Zuckerberg says Llama 4 failed because it was staffed like Instagram, not a frontier lab — rohanpaul_ai · 2026-10-06
- Controlled Study Finds No Encoding Dominates: Pixels, Bytes and Tokens Each Win on Different Tasks — delliott · 2026-10-06
- Old 'gpt-next' codename resurfaces; researcher speculates an intentional OpenAI easter egg — RileyRalmuto · 2026-10-06
- Claude leaves 50% of plan quota after burning one model's cap; OpenAI goes to zero — xiaohu · 2026-10-06
- Sander Dieleman on why continuous diffusion language models are making a comeback — LucaAmb · 2026-10-06
- Reflection AI launches Beam, a 501B-A23B model scoring between GLM-5.2 and GLM-5.3 — iaziaz · 2026-10-06