Beyond OCR Scores: Where Document Parsers Fail
October 4, 2026
A 1,250-page OCR benchmark exposes the scrambled reading order, missing tables, and broken formulas that RAG pipelines and AI agents inherit. I compare document parsers, OCR specialists, the open-weights Qwen 3.6, the hosted Claude Fable 5.1 and GPT-6 Astra, and Apple Vision to show where each fails—and what aggregate scores hide.