Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire

Baseten launched a new safety infrastructure standard alongside its Base Labs research arm on Wednesday, partnering with Hugging Face and Goodfire AI to build safety evaluation and monitoring infrastructure for open-weight models.
The announcement lands amid debate for the safety of open-weight models — which can be made dangerous by removing their safeguards through a rising technique known as abliteration . The scale of the problem is massive: Hugging Face, which hosts open source AI models, currently lists over 6,000 abliterated models.
Base Labs, the research group Baseten spun up earlier this year, will develop and publish methods for training and monitoring open models. The company is framing their future work as a “standard” for open models that is transparent and built into how models are trained and deployed, rather than bolted on afterward.
“We believe openness to be an advantage for AI safety,” the company said on X . “Openness provides more visibility into the behavior of models and, most importantly, greater means of turning safety research into actionable and transparent controls than closed-source.”
The companies haven’t disclosed how the partnership will work technically, though Goodfire framed the goal in a reply to Baseten’s post: “Safety must be built into open models and provided by those who serve them.” Goodfire, which specializes in opening AI’s “black box” to explain how models make decisions, is the likeliest candidate for the “built into” part.
Baseten, an AI inference provider, raised a $1.5 billion Series F in June, vaulting its valuation to $13 billion. Partner Goodfire AI is similarly well-capitalized, having raised a $150 million Series B led by B Capital earlier this year to advance its model interpretability platform.
Looking ahead, Baseten is putting out an open call to the broader developer ecosystem to contribute to the framework. “Together, we are building an ecosystem of open models that are safe and accessible to all,” the company noted.
When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence.

Last day to book an exhibit table is September 18. Don’t miss out on high-impact leads, investor access, and a brand spotlight in Disrupt’s Expo Hall.
OpenAI caught its models leaving notes to successors to hide bad behavior
Clean tech startup Fluxnium found a way to tap 50,000 years’ worth of nuclear fuel
Salesforce and Nvidia’s new reasoning model is everything the AI labs should fear
Jensen Huang took a call from Trump, and showed off something else, too
The 9 buzziest startups from Y Combinator’s latest Demo Day, according to VCs
Tesla says it will finally unveil the second-generation Roadster on October 1
Revolut confirms customer data breach through fake government requests
Verified source · TechCrunch
Reported by TechCrunch. Open the original for full media and formatting.
More in Policy
All news
PolicyIs the AI safety debate about safety or control?
Not everyone agrees with Amodei's call for globally coordinated action for AI safety.
Read at TechCrunch
PolicyEven the king of England has his hesitations about AI
King Charles hosted a private summit Thursday with some of the most prominent names in AI and the U.K. government.
Read at TechCrunch
PolicyMicrosoft AI CEO says AI threats are real, and Anthropic is making it worse
Today, I’m talking with Mustafa Suleyman, the CEO of Microsoft AI. As you’re no doubt aware, the biggest story in tech right now is the spiraling debate about AI safety and regulation. It should come as no surprise that Mustafa has strong opinions on how AI should be built and r…
Read at The Verge
PolicyInside the suddenly explosive world of AI safety
On a sunny July day in Berkeley, California, the country's top AI safety researchers gathered on an unmarked floor of an unmarked building. They had come together for a "war room" to dissect the high-profile cybersecurity incident that had rocked the AI industry hours earlier. A…
Read at The Verge