Tag
Proposes a unified framework, LCF and LCFEdit, that jointly optimizes construction and injection of code-mixing fingerprints for LLMs using low-resource languages to achieve imperceptible and robust ownership verification.
Introduces Indi-RomCoM, a benchmark for evaluating LLMs on Romanized Code-Mixed (RCM) instructions in four Indic languages, finding that LLMs underperform on RCM tasks and performance degrades with higher code-mixing density.
Developer seeks advice on handling English-Hindi code-mixed text classification without heavy LLMs, as sentence transformers fail on Romanized Hindi.