Tag
The paper curates and releases open datasets and a model for Armenian, demonstrating that continued pretraining with news and STEM data improves performance and addresses data scarcity in low-resource NLP.
This paper from Xiaomi introduces reference-free post-training for multilingual machine translation, applying GRPO with quality-estimation rewards to the MiLMMT-46-v0.1 SFT models, producing MiLMMT-46-v1.0 that improves translation across 46 languages and outperforms open and proprietary baselines.
This paper introduces MiLMMT-46-v1.0, a multilingual machine translation model improved via reference-free post-training with GRPO and checkpoint interpolation, surpassing strong open and proprietary baselines across 46 languages.
A tweet announcing a torrent-based site for downloading open LLM model weights, positioning it as a pirate bay alternative for open models.
The poster states that the MoQ GGUFs of the LFM2.5 8B A1B model offer the best accuracy-to-size ratio, advising against using versions with less than 95% accuracy recovery.
This paper measures maximum activation magnitudes across 27 checkpoints from 8 open LLM families, finding significant variance across families, architectures, and training stages, with implications for low-bit quantization and deployment.