this post was submitted on 15 Jun 2024
122 points (93.0% liked)

Privacy

31876 readers
468 users here now

A place to discuss privacy and freedom in the digital world.

Privacy has become a very important issue in modern society, with companies and governments constantly abusing their power, more and more people are waking up to the importance of digital privacy.

In this community everyone is welcome to post links and discuss topics related to privacy.

Some Rules

Related communities

Chat rooms

much thanks to @gary_host_laptop for the logo design :)

founded 5 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[โ€“] [email protected] 1 points 4 months ago (1 children)

Is abliteration based off the research by the Anthropic team? When they got Claude to say it was the golden gate bridge?

[โ€“] [email protected] 4 points 4 months ago

Ironically, as far as I'm aware it's based off of research done by some AI decelerationists over on the alignment forum who wanted to show how "unsafe" open models were in the hopes that there'd be regulation imposed to prevent companies from distributing them. They demonstrated that the "refusals" trained into LLMs could be removed with this method, allowing it to answer questions they considered scary.

The open LLM community responded by going "coooool!" And adapting the technique as a general tool for "training" models in various other ways.