ASCIITermDraw-Bench tests whether VLMs can draw ASCII diagrams
East-Muffin-6472 · reddit · 2026-07-20
ASCIITermDraw-Bench is a new benchmark for testing whether vision-language models can generate and edit ASCII diagrams. The authors argue that ASCII can be useful for communicating architectures, topologies, and cluster layouts, and that this capability is different from simply describing a diagram. The benchmark contains 80 tasks across four areas: basic box/layout drawing, network topologies, software architecture diagrams, and image-conditioned diagram editing. Each answer is scored with both a structural metric and an LLM-based semantic judge repeated five times, with a 95% confidence interval reported. The current leaderboard is headed by Gemma-4-31B-IT, followed by Qwen3.7-Plus, Kimi-K2.6, MiniMax-M3, Qwen3.5-9B, and Ternary-Bonsai-27B.
Related event: ASCIITermDraw-Bench Tests VLMs on ASCII Drawing(2 posts)→
More from Multimodal
- Reddit user chains Ideogram 4 and Krea2 to mimic bbox-based image positioning — v3lh0t05c0 · 2026-07-22
- Ultimate Face Fix: Open-Source Multi-Face Repair Node for ComfyUI — Merserk13 · 2026-07-22
- Getting Started with AI Video: Solving Consistency and Censorship — cynicalnewenglander · 2026-07-22
- Storyboard-first workflows are making AI dance videos and influencers more consistent — aftahi_ai · 2026-07-22
- Runpod MCP and Claude help spin up image and video generation workflows — 802high · 2026-07-22
- Midjourney prompt turns a bee into a glitching pixel explosion — michaelrabone · 2026-07-22