Burkov's AGI test: survive 3-5 rounds of interactive image-editing corrections without rage

burkov · x · 2026-10-05

ML author Andriy Burkov argues that the real strength of frontier AI shouldn't be judged by coding — a special case that doesn't generalize — but by interactive image editing with follow-up correction prompts. His bar: if after 3-5 rounds of correction requests you don't want to smash your keyboard, that's AGI. The jab: current multimodal models still fail badly at understanding feedback and making precise edits.

Original post →

More from AGI Musings

AGI Musings channel →