← r/LocalLLaMA
▲
466
+5
37👁
r/LocalLLaMA · u/returnity · 21d ago

Is HF starting to move against abliterated models?

Baseten launched a new safety infrastructure standard alongside its Base Labs research arm on Wednesday, partnering with Hugging Face and Goodfire AI to build safety evaluation and monitoring infrastructure for open-weight models.
The announcement lands amid debate for the safety of open-weight models — which can be made dangerous by removing their safeguards through a rising technique known as abliteration. The scale of the problem is massive: Hugging Face, which hosts open source AI models, currently lists over 6,000 abliterated models.

I can't really tell what exactly the implications are of this "partnership" or what it exactly would impact on HF's model-hosting side. However, I do find it concerning that HF is announcing a collaboration on 'infrastructure safety' with publicity that specifically calls out "dangerous" uncensored models. Thoughts?

468 0 466 10/2 17:28 10/9 06:40 UTC
scorecomments37 sightings
first seen 2026-10-02 17:28 UTClast seen 2026-10-09 06:40 UTCscore then 461score now 466gained +5sightings 37
open on reddit ↗ 💬 213 (-1)