parsebench

Tag

Cards List
#parsebench

@jerryjliu0: We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol…

X AI KOLs Timeline · 2026-07-09 Cached

LlamaIndex benchmarked GPT-5.6 on document understanding and found no improvement over GPT-5.5; the model performs well on text and tables but struggles with charts and layout.

0 favorites 0 likes
#parsebench

@jerryjliu0: We've provided some updated results on Mistral OCR that make use of the annotation feature for charts. The overall scor…

X AI KOLs Following · 2026-06-24 Cached

Jerry Liu reports updated results for Mistral OCR on ParseBench, showing it outperforms GPT-5.5 and trails only Gemini 3.1 Pro, with strong performance on content faithfulness and semantic formatting.

0 favorites 0 likes
#parsebench

@jerryjliu0: Our team is at CVPR 2026 if you want to come say hi :)

X AI KOLs Following · 2026-06-04 Cached

Jerry Liu's team is presenting ParseBench, a comprehensive document understanding benchmark for VLMs, at CVPR 2026. The benchmark includes 2,000 pages of real-world enterprise documents with evaluation metrics for tables, charts, and visual grounding.

0 favorites 0 likes
← Back to home

Submit Feedback