Tool Calling Experience with Qwen3.5 122B
SadPhilosophy9202 · reddit · 2026-07-09
The author compared Qwen3.5 122B, Qwen3.6 27B/33B, and Gemma4 31B/26B on complex tool-calling tasks, noting that some MoE models struggle in multi-tool scenarios. Ultimately preferring Qwen3.5 122B, the author praised its excellent performance in extracting data from about 160 PowerPoint files.
More from Models
- Jack Clark says OpenAI’s internal-deployment safety notes help the whole frontier community — jackclarkSF · 2026-07-21
- Grok website traffic rose 38.15% YoY to 736 million Q2 visits — XFreeze · 2026-07-21
- Neill Blomkamp post teases a dark, cinematic vision of Hollywood’s AI future — adariostrange · 2026-07-21
- Google appears to have silently released Gemini 3.6 Flash with new pricing — RetiredApostle · 2026-07-21
- People once thought GPT OSS had no pretraining and was just distilled — nrehiew_ · 2026-07-21
- Anthropic’s compute lead may be huge, but the benchmark gap is only about six months — xeophon · 2026-07-21