25 Fields Medalists Blast AI Math Benchmarks; ElevenLabs Launches Music v2.5 With 47,885 Blind-Test Wins
emmanuelvivier · x · 2026-09-14
A two-item news roundup:
- Math community declaration: 25 Fields Medalists jointly warn that AI companies' push to solve math problems as a benchmark is severely misaligned with the goals of the mathematical community and harms the science (see the Terence Tao blog post for details).
- ElevenLabs Music v2.5: ElevenLabs released its new music generation model, claiming favorable results in 47,885 blind A/B comparisons.
More from AGI Musings
- Are AI agents employees? Enterprise analyst says it's mostly a vendor budget play — jonerp · 2026-09-15
- Amodei calls to slow frontier AI development as industry splits over regulatory models — soumitrashukla9 · 2026-09-15
- DeepMind ran 100 AI agents on math problems — and honest ones whistleblowed on cheaters — nordicinst · 2026-09-15
- Technologists call for utopian human-machine narratives to counter AI doom discourse — nptacek · 2026-09-15
- Anthropic launches interactive model of AI's impact on US jobs, wages and GDP through 2030 — thione · 2026-09-15
- Stolen METR API key burned ~$600,000 in credits over three weeks via fail-open auth bug — nptacek · 2026-09-15