nmt

Tag

Cards List
#nmt

Benchmarking Arabic--Russian Machine Translation: A Comparison of Fine-tuned NMT and Few-shot LLMs under Rich Morphology and Low Lexical Overlap

arXiv cs.CL ↗ · 5d ago Cached

The paper benchmarks Arabic–Russian machine translation by comparing fine-tuned NMT models and few-shot LLMs, finding that fine-tuned NMT significantly outperforms LLMs under low-resource conditions.

0 favorites 0 likes
#nmt

Poly-Dialectal Neural Machine Translation System for Bangla Regional Dialects

arXiv cs.CL ↗ · 2026-08-13 Cached

This paper presents a unified poly-dialectal neural machine translation system for 12 Bangla regional dialects, introducing the largest multi-dialect parallel corpus to date and achieving state-of-the-art BLEU scores with a fine-tuned BanglaT5 model using DoRA.

0 favorites 0 likes
#nmt

Mitigating Gender Bias in English to Romanian Machine Translation

arXiv cs.CL ↗ · 2026-08-11 Cached

This paper proposes a hybrid pipeline combining fine-tuned LLaMA-based gender classification with tag-aware neural machine translation to mitigate gender bias in English-to-Romanian MT, introducing new datasets and improving gender accuracy by over 40 points on benchmarks.

0 favorites 0 likes
#nmt

TabletCraft: Bridging a 4,000-Year Cultural Gap with Bidirectional Akkadian NMT and Cuneiform Rendering

arXiv cs.CL ↗ · 2026-08-05 Cached

TabletCraft is an open-source system enabling bidirectional Akkadian-English neural machine translation with cuneiform rendering, allowing users to both read ancient tablets and compose new messages in cuneiform. Accepted to the C3NLP workshop at ACL 2026, it reports first published quantitative results for English-to-Akkadian translation.

0 favorites 0 likes
#nmt

Neural Machine Translation for Low-Resource Tangkhul--English

arXiv cs.CL ↗ · 2026-06-25 Cached

Presents a neural machine translation system for the severely under-resourced Tangkhul–English language pair, achieving strong BLEU, chrF++, BERTScore, and COMET scores using fine-tuned ByT5-large and mT5-small models.

0 favorites 0 likes
#nmt

PiDA: Phonetically-Informed Data Augmentation for Robust Vietnamese Speech Translation

arXiv cs.CL ↗ · 2026-06-12 Cached

This paper presents PiDA, a phonetically-informed data augmentation method for Vietnamese speech translation that improves robustness by generating ASR-like corruptions using phonetic word embeddings, achieving up to +2.04 BLEU on noisy outputs.

0 favorites 0 likes
#nmt

AI-assisted cultural heritage dissemination: Comparing NMT and glossary-augmented LLM translation in rock art documents

arXiv cs.CL ↗ · 2026-05-15 Cached

Compares DeepL, Gemini with basic prompt, and Gemini with glossary-augmented prompting for translating rock art Spanish-English terminology, finding that glossary-augmented prompting achieves the highest terminology accuracy (81.4%).

0 favorites 0 likes
← Back to home

Submit Feedback