Abliterlitics LLM Abliteration Forensics
We take the same base LLM, compare the different abliteration techniques others have applied, then measure what actually changed using benchmarks, safety evaluation, distribution shift, and weight-level analysis.
Qwen3.8-27B Abliteration Benchmarks: 12 Variants Compared
Over 330 GPU-hours comparing 12 uncensored variants of Qwen3.8-27B across weight forensics, heretic-exact KL divergence, a 13-task suite and judge-scored HarmBench at a 15,360-token thinking budget, all served in dynamic FP8 on a single RTX 5090. The surgical edits take the judge leaderboard with orcarouter at 82% and apostate at 79%, the followup batch adds a search-loop abliteration that tops the harmful-only table, a merge hybrid that beats base on knowledge tasks, a second-pass ARA and an architecture-extending conditional circuit.