Looks like a well engineered, automated abliteration pipeline. The claims seem a bit overstated though, since the metrics mentioned are cherrypicking refusal count and KL divergence, both of which make the outcome seem the most dramatic.
I personally never saw much of a quality drop from models put through Heretic if that amounts to anything. They have been working quite well on small local models so far.
Keep a close eye on abliterated and "heretic" open weight models. They will be outlawed first.
Can the load-bearing gaps that are worth being flagged for pinning down be abliterated out of a model?
Looks like a well engineered, automated abliteration pipeline. The claims seem a bit overstated though, since the metrics mentioned are cherrypicking refusal count and KL divergence, both of which make the outcome seem the most dramatic.
I personally never saw much of a quality drop from models put through Heretic if that amounts to anything. They have been working quite well on small local models so far.