3DHarnessBench probes agentic 3D-to-code skills of frontier VLMs

ftm_guney · x · 2026-09-09

Researchers from Institut Polytechnique de Paris introduce 3DHarnessBench, a benchmark evaluating how frontier vision-language models reconstruct 3D geometry as Blender Python code. Four harness settings progressively grant agent access: single-view, multi-view, active visual (arbitrary viewpoints), and full 3D interaction via Blender MCP function calls.

Original post →

More from coding & agent

coding & agent channel →