| ▲ | brainwad 2 days ago | |||||||
That's not how abliteration works. The trainer invariably put effort into making the model not dangerous, precisely because they want to be accountable. But because they release its weights, someone can come later and do "weight surgery" to mostly remove any such safeguards. The people proximately accountable for the danger are anonymous and also don't require much resources. The reason this is not _yet_ a big deal is because open weights models are a few months behind the frontier and their users are paying marginal costs for compute. | ||||||||
| ▲ | watwut a day ago | parent [-] | |||||||
None of that is argument against what I said. Someone is running it, that person is responsible. In case of actual hackes that happened, OpenAI and Amtropic. They should stop pointifucating about other people being the danger. They themselves are the perpetrators here. Massively fine these two companies and make their CEO legally responsible and problem will be much smaller. | ||||||||
| ||||||||