dialect

Tag

Cards List
#dialect

BavGround: A Benchmark for Regional Cultural Grounding and Dialect Competence in Bavarian

arXiv cs.CL · 2026-08-14 Cached

Introduces BavGround, a benchmark for evaluating LLMs' regional cultural grounding and dialect competence in Bavarian across English, German, and Bavarian, finding that models struggle with dialectal and localized cultural knowledge.

0 favorites 0 likes
#dialect

Poly-Dialectal Neural Machine Translation System for Bangla Regional Dialects

arXiv cs.CL · 2026-08-13 Cached

This paper presents a unified poly-dialectal neural machine translation system for 12 Bangla regional dialects, introducing the largest multi-dialect parallel corpus to date and achieving state-of-the-art BLEU scores with a fine-tuned BanglaT5 model using DoRA.

0 favorites 0 likes
#dialect

How Robust Are LLMs to Vietnamese Dialects?

arXiv cs.CL · 2026-08-12 Cached

This paper introduces VialectBench, a benchmark evaluating LLM robustness to six Vietnamese dialect groups across four tasks, finding average performance drops of 2.82% and no model fully dialect-invariant.

0 favorites 0 likes
#dialect

Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs

arXiv cs.CL · 2026-07-07 Cached

This paper investigates methods to steer Arabic LLMs toward dialect-specific generation by identifying sparse neuron populations and extracting dialect activation directions, enabling dialect control at inference time without fine-tuning.

0 favorites 0 likes
#dialect

could refusal layers be masking dialect-conditioned safety failures in MoE models [d]

Reddit r/MachineLearning · 2026-05-18

Tests on Qwen3.5-35B-A3B show that AAVE-coded prompts cause MoE models to respond differently, with refusal layers masking dialect-conditioned safety failures that become visible when refusal is weakened.

0 favorites 0 likes
← Back to home

Submit Feedback