| ▲ | dantudor 7 hours ago | ||||||||||||||||||||||||||||||||||||||||||||||||||||
There are versions of Qwen3.8-27B that are unrestricted and available from hugging face. "It will comply with harmful, unethical, offensive, or illegal requests that the original Qwen3.8-27B would refuse. It has no meaningful built-in guardrails." | |||||||||||||||||||||||||||||||||||||||||||||||||||||
| ▲ | Aurornis 4 hours ago | parent | next [-] | ||||||||||||||||||||||||||||||||||||||||||||||||||||
> There are versions of Qwen3.8-27B that are unrestricted and available from hugging face. The restrictions are not a single check in the model that can be removed. Those models on Huggingface are manipulated in different ways that also degrade the model’s intelligence. The degradation ranges from subtle to obviously broken, but it’s not free. When the restrictions are built into the model’s training sets you can try to alter the weights that are involved in the refusals, but that doesn’t mean that what’s left is useful or good knowledge for the same task. Those weights also might be involved in other tasks, so altering them can interfere with interactions that aren’t obviously related. | |||||||||||||||||||||||||||||||||||||||||||||||||||||
| ▲ | radlad 7 hours ago | parent | prev [-] | ||||||||||||||||||||||||||||||||||||||||||||||||||||
> What makes this build different is the word before FP8: uncensored. We applied abliteration — orthogonalizing the refusal direction out of the residual stream — to remove the model's safety-alignment refusals. The result is a model that will comply with requests the original would refuse. Surely this has unintended side effects on output quality? | |||||||||||||||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||||||||||||||