Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Baseten launched a new safety infrastructure standard alongside its Base Labs research arm on Wednesday, partnering with Hugging Face and Goodfire AI to build safety evaluation and monitoring infrastructure for open-weight models.
The announcement lands amid debate for the safety of open-weight models — which can be made dangerous by removing their safeguards through a rising technique known as abliteration. The scale of the problem is massive: Hugging Face, which hosts open-source AI models, currently lists over 6,000 abliterated models.
Base Labs, the research group Baseten spun up earlier this year, will develop and publish methods for training and monitoring open models. The company is framing their future work as a “standard” for open models that is transparent and built into how models are trained and deployed, rather than bolted on afterward.
“We believe openness to be an advantage for AI safety,” the company said on X. “Openness...
Copyright of this story solely belongs to techcrunch.com. To see the full text click HERE