Why the ICCV 2019 'draw via actions + renderer' idea explains visual agents
ziqiao_ma · x · 2026-09-07
Prompted by GPT-6 Astra's near-production Blender work, the author revisits a classic ICCV 2019 paper whose idea still holds: create through sequential actions plus a renderer rather than predicting pixels directly. Back then: parameterized Bézier strokes with a learned neural renderer; today: Blender code controlling geometry, materials, lighting and cameras. Key points: pixels are only final observations and collapse structure; graphics programs keep structure explicit while rendering gives a forward model for evaluation and refinement; the hard problem may be search and control before rasterization, not image synthesis itself.
More from coding & agent
- User burns $6,000 of Claude Code usage in 7 days, mostly cache read tokens — guess_wat · 2026-09-07
- Grok Bot opens template marketplace; Haggle Bot found $100K+ in savings in a week — FinanceYF5 · 2026-09-07
- Developer builds AI agent harness on iMessage that trades and pays via PayBox — kleffew94 · 2026-09-07
- Codex tasked with designing its own sheet-metal part via DFM/quote MCP — PaulYacoubian · 2026-09-07
- Copilot CLI's Astra fixes stuck PowerPoint on Mac via computer-use MCP tools — DanWahlin · 2026-09-07
- Reddit devs ask: are AI coding agents' costs worth the productivity gains at scale? — lowkeyskibidi · 2026-09-07