Tag
This paper presents a multi-stage LLM pipeline for structure-preserving Marathi-to-English translation of government documents, integrating layout-aware OCR and HTML reconstruction to maintain formatting and domain terminology.
HakushoBench is a Japanese chart and table VQA benchmark built from governmental white papers to evaluate vision-language models' understanding of complex visual data, challenging open-weight models with a 58.6% accuracy and a 34.9-point gap to proprietary models.