Qwen J-space ablation sharply cuts multi-hop reasoning accuracy in a new bar-chart result
Sauers_ · x · 2026-07-21
A post showing a J-space ablation on Qwen, via Silico/Fable, with a bar chart comparing baseline, the ablation case, and a matched-norm control.
The image reports two tasks:
- Multi-hop reasoning drops sharply under the J-space ablation.
- Single-hop recall also falls, but less dramatically.
Overall, the result suggests the J-space component matters materially for both reasoning and recall performance, especially on harder multi-step tasks.
More from Models
- Grok 4.5 is now free inside Cursor, the popular AI coding IDE — mark_k · 2026-07-21
- GPT often converges on the same near-miss ideas in math problems — yacineMTB · 2026-07-21
- Eno Reyes says model distillation is basically unstoppable — LangChain · 2026-07-21
- Sakana says multiple diffusion models plus MCTS beat test-time scaling on coding and math — SakanaAILabs · 2026-07-21
- OpenAI hackathon project stalls as Codex struggles on voice, while Claude spots the issue — ColleenMBrady · 2026-07-21
- Kimi K3 lands exactly on China’s 2-year AI capability trend line — peterwildeford · 2026-07-21