@baselabs: We believe openness to be an advantage for AI safety. Openness provides more visibility into the behavior of models and…
Summary
Baseten and Base Labs are launching safety infrastructure for open models to improve AI safety through transparency and collaboration with Hugging Face and Goodfire AI.
View Cached Full Text
Cached at: 09/17/26, 04:11 AM
We believe openness to be an advantage for AI safety. Openness provides more visibility into the behavior of models and, most importantly, greater means of turning safety research into actionable and transparent controls than closed-source.
This is why Baseten and Base Labs are building a stronger safety and security standard for open models, with the launch of our safety infrastructure. Base Labs will develop and publish methods for training and monitoring open models, and Baseten will integrate that work into its deployment infrastructure, live at runtime, and offer this work as a managed service. This will be a standard that is transparent and built into how our models are trained and deployed.
We invite the open-source community to contribute, and are proud to partner with @huggingface and @GoodfireAI to bring this vision to fruition.
Together, we are building an ecosystem of open models that are safe and accessible to all.
Similar Articles
Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Base Labs, the research arm of Baseten, has launched an open-weight AI safety partnership with Hugging Face and Goodfire AI to develop and publish standards for safety evaluation and monitoring of open models, addressing issues like abliteration.
@LRudL_: I think it's possible to make a lot of progress toward models that are both open and safe! The former will become more …
A tweet discussing the potential to advance AI models that are both open and safe, highlighting the importance of open-source to prevent centralization and safety due to current challenges.
OpenAI safety practices
OpenAI outlines 10 safety practices it actively uses and improves upon, including empirical red-teaming, alignment research, abuse monitoring, and voluntary commitments shared at the AI Seoul Summit. The company emphasizes a balanced, scientific approach to safety integrated into development from the outset.
OpenAI and Anthropic share findings from a joint safety evaluation
OpenAI and Anthropic released findings from a joint pilot safety evaluation where each lab tested the other's models on internal safety and misalignment assessments, sharing results publicly to improve transparency and identify potential gaps in AI safety testing.
@MTSlive: We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says r…
HuggingFace CEO Clément Delangue argues that restricting open source AI models creates more risk than openness, citing historical examples like GPT-2 and Mythos to support his view that openness improves cybersecurity and overall safety.