Collaborative Generative Art Driven by VLM Agents
hardmaru · x · 2026-07-10
The author showcases an AI Picbreeder-style visual system that employs 10 parallel VLM "breeder" agents to collaboratively generate images.
The system operates by:
- Continuously sampling candidates from a shared archive
- Interactively evolving new images via mutation and crossover
- Writing preferred results back to the archive
- Utilizing a VLM "critic" agent to evaluate the growing artistic lineage
This ultimately forms a continuous, agent-driven loop of digital cultural production.
Related event: Sakana AI Replicates Picbreeder with VLM Agents for Open-Ended Creativity(14 posts)→
More from Multimodal
- Reddit user chains Ideogram 4 and Krea2 to mimic bbox-based image positioning — v3lh0t05c0 · 2026-07-22
- Ultimate Face Fix: Open-Source Multi-Face Repair Node for ComfyUI — Merserk13 · 2026-07-22
- Getting Started with AI Video: Solving Consistency and Censorship — cynicalnewenglander · 2026-07-22
- Storyboard-first workflows are making AI dance videos and influencers more consistent — aftahi_ai · 2026-07-22
- Runpod MCP and Claude help spin up image and video generation workflows — 802high · 2026-07-22
- Midjourney prompt turns a bee into a glitching pixel explosion — michaelrabone · 2026-07-22