Skip to content
worth noting Security

Baseten, Hugging Face and Goodfire AI launch a partnership on the safety of open AI models

only one source so far

Baseten, through its research division Base Labs, has launched a partnership with Hugging Face and Goodfire AI to build safety infrastructure for open models – a response to the growing number of models stripped of safeguards using abliteration.

Baseten has launched a partnership through its research division Base Labs with Hugging Face and Goodfire AI, focused on building safety evaluation and monitoring infrastructure for open (open-weight) models. According to the companies, the goal is to create a transparent safety standard that will be built into models during training and deployment, rather than added afterward.

The initiative responds to the discussion about the safety of open models, which can be stripped of safety measures using a technique called abliteration. According to the article, Hugging Face, which hosts open models, had recorded more than 6000 such abliterated models on its platform. Base Labs, a research group founded by Baseten earlier in 2026, is to develop and publish methods for training and monitoring open models. Baseten stated on X that it views openness as an advantage for AI safety because it provides greater insight into model behavior and better tools for translating safety research into concrete and transparent controls than closed models do. Goodfire AI, which specializes in model interpretability, stated in response that safety must be built into open models by those who operate them.

Technical details of how the partnership will work have not been disclosed. Baseten, an AI inference provider, raised 1.5 billion dollars in Series F funding in June this year, bringing its valuation to 13 billion dollars. Goodfire AI raised 150 million dollars in Series B funding led by B Capital earlier in 2026. Baseten says it is inviting the broader developer community to contribute to the future framework.

What changed

Why it matters

Open models can be stripped of built-in safety measures using abliteration, increasing the risk of misuse for thousands of models on Hugging Face. The emerging standard from Baseten, Hugging Face and Goodfire AI is expected to offer more transparent protection built directly into models, which is particularly relevant for companies and teams that deploy open models or build products on them and must consider safety and reputational risks. However, the specific technical form of the standard has not yet been disclosed.

Relevant practical impact

What this means

01

For a business

Companies deploying open models face the risk that these models can be stripped of safety measures using a technique called abliteration – Hugging Face hosts over 6000 of them. A new initiative by Baseten (through its Base Labs division), Hugging Face and Goodfire AI aims to create a transparent safety standard built directly into model training and deployment, which could affect…

Risks and compliance
What to decide Monitor the development of this safety standard and consider incorporating it when selecting or deploying open models, especially if the company works with models downloaded from Hugging Face.
More business impacts →
abliterace Baseten bezpečnost AI Goodfire AI Hugging Face otevřené modely

Check the original

Event sources

only one source so far · 1 publisher, 1 independent. We count feeds from the same owner only once.

1
TechCrunch AI independent context · first detected Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire