Hugging Face releases The Stack v3 – largest open code dataset yet
Summary
Hugging Face released The Stack v3, the largest open-source code dataset to date, designed to enhance training for code-focused AI models.
Similar Articles
State of Open Source on Hugging Face: Spring 2026
This report analyzes the state of the open source AI ecosystem on Hugging Face in Spring 2026, highlighting significant growth in users, models, and datasets, as well as trends in derivative model creation and specialized sub-communities.
@HuggingPapers: NVIDIA just released a paper review dataset on Hugging Face APRES, Agents4Science, and Sakana v2 subsets covering human…
NVIDIA released a dataset on Hugging Face containing paper reviews for human and AI-authored papers, including subsets APRES, Agents4Science, and Sakana v2.
huggingface/transformers Release 5.8.0
Hugging Face has released version 5.8.0 of the Transformers library, a widely used open-source framework for natural language processing and deep learning.
@RoundtableSpace: HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, itera…
Hugging Face replaced its post-training team with an autonomous agent that reads papers, runs GPU experiments, and improves models, achieving a 22-point benchmark jump in under 10 hours and beating Codex on HealthBench by 60%.
HuggingFace benchmark datasets now let you filter by model size
HuggingFace benchmark datasets now allow filtering by model size, enabling comparisons like 'best model under 32B on swebenchverified'.