A New Benchmark for Evaluating VLM ASCII Art

East-Muffin-6472 · reddit · 2026-07-19

This is a new benchmark named ASCIITermDraw-Bench, designed to test whether VLMs can genuinely generate and edit ASCII diagrams as instructed.

The author's motivation: while many models can "describe" what an architecture or topology diagram looks like, accurately formatting that content into ASCII text requires a different, harder capability. The benchmark features 80 tasks across four categories:

The evaluation methodology is comprehensive:

In the current leaderboard, Gemma-4-31B-IT leads with 73.8%, followed by Qwen3.7-Plus (70.2%), Kimi-K2.6 (61.8%), and MiniMax-M3 (59.5%). The project has open-sourced 12 sample tasks and the complete methodology, with reproducible content available on Hugging Face.

Related event: ASCIITermDraw-Bench Tests VLMs on ASCII Drawing(2 posts)→

Original post →

More from Multimodal

Multimodal channel →