VIGA agent vibe-codes editable 3D scenes in Blender as author warns academic CV research lags 2 years
FinanceYF5 · x · 2026-09-18
VIGA (Vision-as-Inverse-Graphics Agent), from UC Berkeley, CMU, and Max Planck researchers, is a multimodal agent that reconstructs any input image as an editable scene program in Blender via an analysis-by-synthesis loop — using interleaved multimodal reasoning and evolving contextual memory to 'vibe code' the scene, its physics, and interactions, building assets from primitives or tools like Meshy and SAM-3D.
But the bigger issue, the author argues, is that VIGA's impact came mainly from arXiv and social media rather than ECCV, and follow-up inverse-graphics work advanced mostly outside academia. Academia lacks the compute and tokens needed to evaluate top models; model companies, governments, and institutions must share resources. With industry iterating monthly and conferences publishing yearly, academia must rethink how it measures contribution and allocates credit.
Related event: Post-GPT-6 Gloom at ECCV: CV Papers Already Two Years Behind(6 posts)→
More from AGI Musings
- Ex-Anthropic security engineer: humanity may still steer superintelligence — JeffLadish · 2026-09-18
- Reddit poster: AI doomsayers may be right, but they're terrible at telling the story — TevecQ · 2026-09-18
- Anil Seth's 'Conscious AI and Biological Naturalism' collection published in BBS with 50 commentaries — anilkseth · 2026-09-18
- "Humans can't keep control" — the contrarian optimism argument around ASI — flowersslop · 2026-09-18
- AI circle's sharpest quip yet: you can't talk a man off doom if his ego depends on it — basedjensen · 2026-09-18
- "This ship can't sink": mocking ASI containment optimism with the Titanic meme — flowersslop · 2026-09-18