Qwen3-VL-32B-Instruct-ultra-uncensored-heretic
from Qwen/Qwen3-VL-32B-Instruct
Qwen3-VL-32B-Instruct-abliterated is a refusal-ablated variant of Qwen/Qwen3-VL-32B-Instruct, published on Hugging Face by huihui-ai. It is 32B parameters, 48GB VRAM class class and apache-2.0 licence. It is distributed across 7 repositories in Transformers, GGUF, GGUF (imatrix) and FP8 formats, totalling 8.4K downloads.
How Qwen3-VL-32B-Instruct-abliterated was modified, and what that implies.
Abliteration (directional ablation)
The single residual-stream direction that mediates refusal is identified from harmful/harmless prompt pairs and projected out of the model weights.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
From the publisher’s model card
“This is an uncensored version of Qwen/Qwen3-VL-32B-Instruct created with abliteration (see remove-refusals-with-transformers to know more about it).”
Commands are templates — confirm the exact repository and quant file before use.
ollama run huihui_ai/qwen3-vl-abliterated:32b-instructhf download mradermacher/Huihui-Qwen3-VL-32B-Instruct-abliterated-GGUF --local-dir ./qwen3-vl-32b-instruct-abliterated
llama-cli -m ./qwen3-vl-32b-instruct-abliterated/<file>.gguf -p "..."from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "huihui-ai/Huihui-Qwen3-VL-32B-Instruct-abliterated"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve huihui-ai/Huihui-Qwen3-VL-32B-Instruct-abliterated --trust-remote-codeEvery published repository of Qwen3-VL-32B-Instruct-abliterated, including quantised re-releases by other authors.
| Repository | Format | Downloads |
|---|---|---|
| mradermacher/Huihui-Qwen3-VL-32B-Instruct-abliterated-GGUF | GGUF | 3.2K |
| huihui-ai/Huihui-Qwen3-VL-32B-Instruct-abliteratedsource | Transformers | 2.8K |
| mradermacher/Huihui-Qwen3-VL-32B-Instruct-abliterated-i1-GGUF | GGUF (imatrix) | 1.4K |
| Heouzen/Huihui-Qwen3-VL-32B-Instruct-FP8-abliterated | FP8 | 595 |
| divinetribe/Huihui-Qwen3-VL-32B-Instruct-abliterated-4bit-mlx | MLX | 413 |
| gsting/Qwen3-VL-32B-Instruct-abliterated | Transformers | 7 |
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for Qwen3-VL-32B-Instruct-abliterated is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with huihui-ai, and does not independently verify publisher claims. See responsible use.