| ▲ | matheusmoreira a day ago | ||||||||||||||||||||||||||||||||||
That's alarming. I want to use these models to red team my own computers. How are people getting around this? | |||||||||||||||||||||||||||||||||||
| ▲ | gf000 15 hours ago | parent | next [-] | ||||||||||||||||||||||||||||||||||
In my experience some of these models may have learnt some censoring during distillation of Western models, but these are mostly just there as a probable response. So if they first happen to respond "I ain't doing this because legality", then you will have a hard time "convincing" it, but either rolling the dice again (so that it may not come up with the I can't do that text) or rewriting the conversation history a bit will get it going. I sometimes just switch to a model I know is less smart to block stuff so that it has a text agreeing to do that, and then switch to a stronger model to actually go at the task. Your mileage may vary though. | |||||||||||||||||||||||||||||||||||
| ▲ | throw10920 a day ago | parent | prev [-] | ||||||||||||||||||||||||||||||||||
> I want to use these models to red team my own computers. Exactly what I was trying to use it for! ): I'm in the same boat - I haven't heard of a way to get around it aside from either self-hosting (GLM-5.2? good luck) or "self-hosting" (paying bucks per hour to Vast) an abliterated model. | |||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||