ExtractBench Reveals VLMs and Coding Agents Fail at Document Grounding

llama_index · x · 2026-08-17

LlamaIndex released ExtractBench, a benchmark strictly evaluating document extraction grounding, requiring correct citations with correct values (IoU 0.5). Key findings:

Related event: LlamaIndex Releases ExtractBench, VLMs Fail at Attribution(2 posts)→

Original post →

More from coding & agent

coding & agent channel →