GPT-6 Astra hits 97.2% on short-doc extraction SOTA but only 31.7% on long docs

llama_index · x · 2026-09-06

LlamaIndex (Jerry Liu) benchmarked GPT-6 Astra on hard document parsing and extraction:

Bottom line: best frontier model for short/medium document extraction, but cost and long-doc performance are real caveats.

Original post →

More from Models

Models channel →