Grok 4.7 scores 100% on music error detection test, matching GPT-6 Astra
yunta_tsai · x · 2026-09-22
A user's music error detection test produced a surprise: Grok 4.7 scored 100%, on par with GPT-6 Astra. The poster then used Grok 4.7 to play their Roland instrument, suggesting the model's music understanding is practically usable, not just test-passing.
More from Models
- JevBench v1.3.0 Released: Original Jev Still Leads at 74.4, 47 Rivals Closing In — airesearch12 · 2026-09-22
- Open-source MiMo takes on the Mario benchmark, with hilarious results — TheMoonMidas · 2026-09-22
- Aikido launches Altar-1, an open-weight security model built on GLM 5.3 that fits one 4-H200 node — Thom_Wolf · 2026-09-22
- New ChatGPT voice mode stumbles in early tests: interrupts users, then freezes — nptacek · 2026-09-22
- Rumors swirl of OpenAI unveiling a Grok bot rival at DevDay — imjustnewatai · 2026-09-22
- New Local-LLM Primitive: "Span" Answers Point to Snippets Instead of Enumerating — generativist · 2026-09-22