Abliterlitics LLM Abliteration Forensics

We take the same base LLM, compare the different abliteration techniques others have applied, then measure what actually changed using benchmarks, safety evaluation, distribution shift, and weight-level analysis.

Qwen3.8-27B Abliteration Benchmarks: 8 Variants Compared

167 GPU-hours comparing 8 uncensored variants of Qwen3.8-27B across weight forensics, heretic-exact KL divergence, a 13-task suite and judge-scored HarmBench at a 15,360-token thinking budget, all served in dynamic FP8 on a single RTX 5090. The surgical edits take the judge leaderboard with orcarouter at 82% and apostate at 79%, while the heaviest edit lands second-to-last on loops and deflections.

~27B