User finds Tencent Hunyuan Hy4 cheaper than GLM 5.3 via fewer output tokens, high cache hit rate
mariofilhoml · x · 2026-09-06
A user compared Tencent's Hunyuan Hy4 with GLM 5.3 on OpenRouter. He admits his capability comparison is 'just vibes,' but says Hy4 is significantly cheaper for him: it uses fewer output tokens even at high effort settings and has a very good cache hit rate on OpenRouter, which is why he keeps choosing it. His interlocutor notes Hy4 seems about the same size as GLM 5.3 and questions whether pricing is really a big advantage.
Related event: Tencent Hunyuan Hy4 Preview Tops OpenRouter Weekly Usage Chart(5 posts)→
More from Models
- User: Opus 4.8 nails UI work while OpenAI lacks focus for everyday coding needs — aloncarmel · 2026-09-07
- Netherlands builds 'Dutch AI' by finetuning Qwen 3.5 27B in subsidized datacenter — teortaxesTex · 2026-09-07
- Abacus.AI CEO: DeepSeek Handles 80% of Everyday Tasks at 100x Lower Cost — bindureddy · 2026-09-07
- Users Report Day-One Bans Over 'Distilling' as Opaque Moderation Draws Fire — QuixiAI · 2026-09-07
- Naval's bike analogy for SFT vs RL explains why DeepSeek R1 shocked the industry — McDonaghMatthew · 2026-09-07
- DeepSeek-R1 grew reasoning with pure RL, no SFT — and that's what changed everything — McDonaghMatthew · 2026-09-07