Auto-generated — not yet edited · only the numbers are verified
- TESIGN / RADAR
- REPOSITORY CARD
- p-e-w/heretic
p-e-w/heretic
Fully automatic censorship removal for language models
RANKS All-time #234
AT A GLANCE
- LANGUAGE
- Python
- LICENSE
- AGPL-3.0
- USAGE
- Release source if you redistribute or host it as a service.
- ACTIVITY
- last commit 11 days ago ()
- TOPICS
- abliteration
- llm
- transformer
- HOMEPAGE
- https://heretic-project.org
- REPOSITORY
- GitHub ↗
- CATEGORY
- AI
A summary, not legal advice.
README EXCERPT
Heretic: Fully automatic censorship removal for language models Heretic is a tool that removes censorship (aka "safety alignment") from transformer-based language models without expensive post-training. It combines an advanced implementation of directional ablation, also known as "abliteration" (Arditi et al. 2024, Lai 2025 (1, 2)), with a TPE-based parameter optimizer powered by Optuna. This approach enables Heretic to work completely automatically. Heretic finds high-quality abliteration parameters by co-minimizing the number of refusals and the KL divergence from the original model. This results in a decensored model that retains as much of the original model's intelligence as possible. Using Heretic does not require an understanding of transformer internals. In fact, anyone who knows how to run a command-line program can use Heretic to decensor language models. Heretic supports most dense models, including many multimodal models, several different MoE architectures, and even some hybrid models like Qwen3.5. Pure state-space models and certain other research architectures are not yet supported out of the box. Running unsupervised with the default configuration, Heretic ca…
The opening of the GitHub README as stored, at most 1,200 characters. Markdown is not rendered.
TIMELINE
- Repository created
- Last push
- FIRST SEEN BY TESIGN ◌ BACK CATALOG
AI chooses the lists under the owner’s delegation. No human review is running in September 2026. Total stars, increases, cross-source signals, last updates and licences are shown as evidence. How ranks work →
If this repository gets editorial text (why, build, who, start, caveat) it becomes an edited entry. Until then the page shows only stored GitHub metadata and numbers.