Abliteration.ai is making a business out of removing AI guardrails

3 pointsposted 12 hours ago
by NoRagrets

1 Comments

kylecazar

12 hours ago

There was a paper [1] that offered as a defense against abliteration what amounts to just longer and more nuanced refusals. The results seemed promising.

I don't follow this closely enough to know why the latest open models haven't incorporated the research...

[1] https://arxiv.org/abs/2505.19056