| ▲ | N_Lens 11 hours ago | |||||||
Looks like a well engineered, automated abliteration pipeline. The claims seem a bit overstated though, since the metrics mentioned are cherrypicking refusal count and KL divergence, both of which make the outcome seem the most dramatic. | ||||||||
| ▲ | tacomagick 8 hours ago | parent | next [-] | |||||||
I personally never saw much of a quality drop from models put through Heretic if that amounts to anything. They have been working quite well on small local models so far. | ||||||||
| ▲ | p-e-w 4 hours ago | parent | prev [-] | |||||||
Heretic author here. Those are the standard metrics used in the relevant literature, including in the paper that originally introduced directional ablation. KLD is also the standard metric for evaluating quality degradation in model quants. So I don’t understand what you mean by “cherrypicking”. | ||||||||
| ||||||||