Will Frontier Multimodal Models Replace Dedicated PDF Parsers Like Docling and Marker?
lucasbennett_1 · reddit · 2026-09-23
A Reddit discussion on whether dedicated PDF parsing layers (Docling, Marker, LlamaParse, Surya) will be absorbed by frontier multimodal models that read pages natively and keep getting cheaper. The case for absorption: fewer moving parts and native vision/PDF handling. The case for parsers sticking around: layout, tables, and grounding (pointing back to exact page regions) remain contested benchmark territory.
More from coding & agent
- Dev's 'software factory' burns 5-10B tokens a day with 32 parallel coding agents — BLUECOW009 · 2026-09-23
- Dev shares how he built app moderation with Jev and LLMs — amos_gyamfi · 2026-09-23
- Dev argues the real AI agent problem is unmeasurable requirements, not capability — algo_diver · 2026-09-23
- $/task beats token price: outputs are <5% of tokens in agentic coding — xeophon · 2026-09-23
- Rethinking multiplayer agent UX: shared artifacts may beat a single group-chat box — max__drake · 2026-09-23
- Dev: if your eval costs $20 for 100 runs, the problem set isn't hard enough — pvncher · 2026-09-23