Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Summary
Base Labs, the research arm of Baseten, has launched an open-weight AI safety partnership with Hugging Face and Goodfire AI to develop and publish standards for safety evaluation and monitoring of open models, addressing issues like abliteration.
View Cached Full Text
Cached at: 09/17/26, 06:10 PM
Similar Articles
@baselabs: We believe openness to be an advantage for AI safety. Openness provides more visibility into the behavior of models and…
Baseten and Base Labs are launching safety infrastructure for open models to improve AI safety through transparency and collaboration with Hugging Face and Goodfire AI.
Open-weight AI models are catching up to the frontier. The safety gap remains.
A SaferAI report finds GLM-5.2, an open-weight model from China's Z.ai, is closing the capability gap with leading frontier models but fails dangerous cyber and bio safety tests, highlighting the growing safety gap for open-weight models.
OpenAI and Anthropic share findings from a joint safety evaluation
OpenAI and Anthropic released findings from a joint pilot safety evaluation where each lab tested the other's models on internal safety and misalignment assessments, sharing results publicly to improve transparency and identify potential gaps in AI safety testing.
OpenAI institutes new safeguards after Hugging Face breach
OpenAI has implemented new security policies in response to a Hugging Face breach, including enhanced monitoring and network isolation to ensure safer model development and testing.
OpenAI and Los Alamos National Laboratory announce research partnership
OpenAI and Los Alamos National Laboratory announced a research partnership to evaluate how frontier AI models like GPT-4o can safely assist scientists in laboratory settings, with focus on bioscience capabilities and biosecurity risk assessment.