jxnl's robust music transcription pipeline: Astra + spectrogram error correction
jxnlco · x · 2026-09-24
jxnl shares his music transcription workflow for learning: use Astra to transcribe audio, then have the model review the spectrogram for error correction, and finally check the music theory, producing very robust transcriptions.
In an actual correction example, the model spotted clear issues: some wide vibrato became extra chromatic notes; around 3:22 the pitch tracker briefly jumped to a lower piano note while the clarinet continued above it. It corrected those, simplified isolated sixteenth rests, and noted the quiet ending remains ambiguous.
More from Multimodal
- HeyGen guide: wiring Meta's Muse Agent to MCP for automated avatar videos — HeyGen · 2026-09-24
- Nunchux runs MiniMax-H3 on AMD MI355X with up to 26.7x faster inference — junyanz89 · 2026-09-24
- Claude + Thrixel Build an Interactive 3D Room Designer You Can Walk Through — RanaHanocka · 2026-09-24
- Hands-On With Pexo: An AI Video Agent That Builds Full Promos via Chat — HeyAmit_ · 2026-09-24
- First Test of Typography Lyrics with Minimax H3 Shows Text Errors in 9:16 — Chiduk99 · 2026-09-24
- Opus 5.5 tested on music: agent skill rewrites melodies into virtuosic showpieces — doodlestein · 2026-09-24