Topic: model abliteration risks
-
Base Labs, Hugging Face Launch Open-Weight AI Safety Partnership
Baseten, Hugging Face, and Goodfire AI have formed a partnership to establish a new safety infrastructure standard for open-weight AI models, addressing the growing threat of "abliterated" models that lack built-in safeguards. The initiative prioritizes embedding security directly into the core a...
Read More »