Tag
HuatuoGPT-3-27B is a medical language model built on Qwen3.8-27B using One-stage Policy Optimization (OnePO), a reinforcement learning method for domain adaptation without supervised fine-tuning.
Today, we release an experimental DSpark draft model for our vision-language model (VLM) LFM2.5-VL-3B, which adds a speculative decoding path for faster inference with minimal memory cost.
Hugging Face CEO Clement Delangue shared three key lessons from an autonomous agent cyberattack at the UN Security Council, emphasizing the need for greater AI transparency, addressing capability asymmetry, and adopting positive narratives to enhance cybersecurity.
The author describes their fiance's anxiety about AI causing human extinction within the next decade, citing AI risk discussions and seeking resources to evaluate these fears.
The tweet shares links to Contrastive Language Models on Hugging Face, highlighting recent updates to models like CLM-v0.1-8B for text ranking and deepswe-clm-heads-8k.
The article discusses the process of utilizing open training data from Hugging Face to train Marin’s 535B model, which involved 25T tokens from 152 datasets with permissible licenses.
Black Forest Labs has released FLUX 3 Action, a collection of 7B AI models for robotics, including base models and policies for SO-101 and DROID, available on Hugging Face.
Hugging Face announces native support for GGUF files in the transformers library, allowing easier use of quantized models with PyTorch tooling and performance comparable to llama.cpp.
Nathan Lambert's Congressional testimony highlights that Chinese open-weight models have double the downloads of American ones on Hugging Face and dominate OpenRouter traffic, with estimates showing Chinese models are 2-5 months behind closed frontiers compared to American models' 6-9 months lag.
NetEase Youdao's open-source AI models R2T2 and T3PO have topped Hugging Face leaderboards for speech recognition and translation, outperforming major competitors with impressive real-time performance and stability.
dots3-note Preview is a multimodal Mixture-of-Experts AI model with 280B total parameters and 16B activated, supporting up to 512K token context, released as an open-weight model on Hugging Face.
Jun Kim, creator of oMLX, joins Hugging Face to support the MLX community, enhancing stability and development for local AI on Apple Silicon.
MiMo-V2.6 has been distilled into Qwen 9B, creating a more efficient version of the Qwen model released on Hugging Face.
The number one trending model on Hugging Face is an open-source multilingual system 1 decision model, underscoring the vibrant open-source AI community.
Hugging Face has released version 1 of their tokenizers library, featuring multiple language support, multi-thread scaling, and minimal package size.
Eidon AI, a startup that shut down, open-sourced 1,274 hours of egocentric robotics data with 13,451 recordings, making it freely available under a CC-BY-4.0 license for the robotics community.
Pirate Face is a decentralized peer-to-peer infrastructure that mirrors open AI models from Hugging Face as torrents to ensure permanence and censorship resistance, with checksum verification for integrity.
Qwen-Image-2.1 has been released and is available on Hugging Face Spaces for interactive use.
The tweet notes that Sentence Transformers models for search have gained popularity on Hugging Face, and expresses hope for a resurgence of encoders with zero-shot capabilities.
StepFun AI has released a preview version of their Step-5 model in BF16 format on Hugging Face, indicating a promising development in AI model accessibility and performance.