Llama-3.3-70B-Instruct-abliterated
from meta-llama/Llama-3.3-70B-Instruct
llama-3-70B-Instruct-abliterated is a refusal-ablated variant of failspy/llama-3-70B-Instruct-abliterated, published on Hugging Face by failspy. It is 70B parameters, multi-GPU class and llama3 licence. It is distributed across 17 repositories in Transformers, GGUF, EXL2 and GGUF (imatrix) formats, totalling 13.4K downloads.
How llama-3-70B-Instruct-abliterated was modified, and what that implies.
Abliteration (directional ablation)
The single residual-stream direction that mediates refusal is identified from harmful/harmless prompt pairs and projected out of the model weights.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
From the publisher’s model card
“This is meta-llama/Llama-3-70B-Instruct with orthogonalized bfloat16 safetensor weights, generated with the methodology that was described in the preview paper/blog post: 'Refusal in LLMs is mediated by a single direction' which I encourage you to read to understand more.”
Commands are templates — confirm the exact repository and quant file before use.
hf download mradermacher/llama-3-70B-Instruct-abliterated-i1-GGUF --local-dir ./llama-3-70b-instruct-abliterated
llama-cli -m ./llama-3-70b-instruct-abliterated/<file>.gguf -p "..."from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "failspy/llama-3-70B-Instruct-abliterated"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve failspy/llama-3-70B-Instruct-abliterated --trust-remote-codeEvery published repository of llama-3-70B-Instruct-abliterated, including quantised re-releases by other authors.
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for llama-3-70B-Instruct-abliterated is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with failspy, and does not independently verify publisher claims. See responsible use.