Baseten, Hugging Face and Goodfire AI launch a partnership on the safety of open AI models
Baseten, through its research division Base Labs, has launched a partnership with Hugging Face and Goodfire AI to build safety infrastructure for open models – a response to the growing number of models stripped of safeguards using abliteration.
Baseten has launched a partnership through its research division Base Labs with Hugging Face and Goodfire AI, focused on building safety evaluation and monitoring infrastructure for open (open-weight) models. According to the companies, the goal is to create a transparent safety standard that will be built into models during training and deployment, rather than added afterward.
The initiative responds to the discussion about the safety of open models, which can be stripped of safety measures using a technique called abliteration. According to the article, Hugging Face, which hosts open models, had recorded more than 6000 such abliterated models on its platform. Base Labs, a research group founded by Baseten earlier in 2026, is to develop and publish methods for training and monitoring open models. Baseten stated on X that it views openness as an advantage for AI safety because it provides greater insight into model behavior and better tools for translating safety research into concrete and transparent controls than closed models do. Goodfire AI, which specializes in model interpretability, stated in response that safety must be built into open models by those who operate them.
Technical details of how the partnership will work have not been disclosed. Baseten, an AI inference provider, raised 1.5 billion dollars in Series F funding in June this year, bringing its valuation to 13 billion dollars. Goodfire AI raised 150 million dollars in Series B funding led by B Capital earlier in 2026. Baseten says it is inviting the broader developer community to contribute to the future framework.
Why it matters
Open models can be stripped of built-in safety measures using abliteration, increasing the risk of misuse for thousands of models on Hugging Face. The emerging standard from Baseten, Hugging Face and Goodfire AI is expected to offer more transparent protection built directly into models, which is particularly relevant for companies and teams that deploy open models or build products on them and must consider safety and reputational risks. However, the specific technical form of the standard has not yet been disclosed.
Relevant practical impact
What this means
For a business
Companies deploying open models face the risk that these models can be stripped of safety measures using a technique called abliteration – Hugging Face hosts over 6000 of them. A new initiative by Baseten (through its Base Labs division), Hugging Face and Goodfire AI aims to create a transparent safety standard built directly into model training and deployment, which could affect…
Risks and complianceCheck the original
Event sources
only one source so far · 1 publisher, 1 independent. We count feeds from the same owner only once.