Tag
This paper compares fine-tuned MahaBERT-based models with large language models (Gemini, LLaMA-3.3-70B, Gemma) for Marathi named entity recognition, finding that the specialized BERT models significantly outperform the LLMs, achieving F1-scores of 0.88–0.91 versus 0.57–0.69.
This paper presents a multi-stage LLM pipeline for structure-preserving Marathi-to-English translation of government documents, integrating layout-aware OCR and HTML reconstruction to maintain formatting and domain terminology.