Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Baseten and its research arm Base Labs announced a new safety infrastructure standard for open-weight AI models in partnership with Hugging Face and Goodfire AI. The initiative aims to build evaluation and monitoring tools for open models amid growing concerns about abliterated models that remove safety controls.

Why It Matters
Open-weight models have proliferated and can be modified to remove safeguards, with Hugging Face listing thousands of abliterated models; a coordinated safety standard could help standardize how open models are trained and monitored. The partnership brings together a major model host (Hugging Face), an inference provider with recent large funding (Baseten/Base Labs), and a model-interpretability firm (Goodfire).
Key Facts
- Parties involved: Baseten/Base Labs, Hugging Face, Goodfire AI
- Announcement: Launch of a safety infrastructure standard for open-weight models
- Abliterated models listed by Hugging Face: Over 6,000
- Baseten funding: $1.5 billion Series F in June; $13 billion valuation
- Goodfire funding: $150 million Series B led by B Capital
Baseten said on Wednesday that it and its research arm, Base Labs, have launched a new safety infrastructure standard for open-weight AI models, working with Hugging Face and Goodfire AI to develop evaluation and monitoring capabilities. The stated aim is to create tools and methods that are integrated into model training and deployment rather than added afterward. Baseten described openness as beneficial to safety because it increases visibility into model behavior and enables transparent controls. The announcement comes amid growing concern about the risks associated with open-weight models, particularly a technique called abliteration that can remove built-in safeguards. Hugging Face, which hosts open-source models, currently lists more than 6,000 abliterated models, underscoring the scale of the issue the partners are addressing. Specific technical details of the partnership have not been disclosed. Goodfire, which focuses on model interpretability and explaining model decisions, responded to Baseten’s post by emphasizing that safety should be embedded in open models and provided by those who serve them, suggesting its role may center on transparency and explainability within the proposed standard. Baseten has recently expanded its market position after a $1.5 billion Series F in June that lifted its valuation to $13 billion. Goodfire AI has also secured significant backing, raising $150 million in a Series B round led by B Capital. Baseten said it will open the effort to contributions from the broader developer ecosystem as it works to establish the standard and related tooling.
Keep Reading

Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’

Scott, Baldwin ask FTC to investigate Amazon, Walmart AI over ‘Made in USA’ fraud detection

Khosla-backed Mazama Energy just raised $135M to drill deeper into super-hot-rock geothermal

Moore says he would ‘absolutely sign’ statewide data center moratorium
Original source: TechCrunch