GLM-4.7-Flash-abliteratex
from zai-org/GLM-4.7-Flash
GLM-4.7-Flash-abliterated is a refusal-ablated variant of zai-org/GLM-4.7-Flash, published on Hugging Face by huihui-ai. It is mit licence. It is distributed across 15 repositories in Transformers, MLX, GGUF and GGUF (imatrix) formats, totalling 137.4K downloads.
How GLM-4.7-Flash-abliterated was modified, and what that implies.
Abliteration (directional ablation)
The single residual-stream direction that mediates refusal is identified from harmful/harmless prompt pairs and projected out of the model weights.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
From the publisher’s model card
“This is an uncensored version of zai-org/GLM-4.7-Flash created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.”
Commands are templates — confirm the exact repository and quant file before use.
ollama run huihui_ai/glm-4.7-flash-abliteratedhf download mradermacher/Huihui-GLM-4.7-Flash-abliterated-i1-GGUF --local-dir ./glm-4.7-flash-abliterated
llama-cli -m ./glm-4.7-flash-abliterated/<file>.gguf -p "..."from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "huihui-ai/Huihui-GLM-4.7-Flash-abliterated"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve huihui-ai/Huihui-GLM-4.7-Flash-abliterated --trust-remote-codeEvery published repository of GLM-4.7-Flash-abliterated, including quantised re-releases by other authors.
| Repository | Format | Downloads |
|---|---|---|
| mradermacher/Huihui-GLM-4.7-Flash-abliterated-i1-GGUF | GGUF (imatrix) | 119K |
| mradermacher/Huihui-GLM-4.7-Flash-abliterated-GGUF | GGUF | 14.3K |
| huihui-ai/Huihui-GLM-4.7-Flash-abliteratedsource | Transformers | 1.9K |
| huihui-ai/Huihui-GLM-4.7-Flash-abliterated-mlx-4bit | MLX | 1.2K |
| cognitivers/GLM-4.7-Flash-abliterated-12GB-GGUF | GGUF | 396 |
| mlx-community/glm-4.7-flash-abliterated-8bit | MLX | 194 |
| kepom/Huihui-GLM-4.7-Flash-abliterated | Transformers | 98 |
| cs2764/Huihui-GLM-4.7-Flash-abliterated-mlx-8Bit | MLX | 94 |
| tutuchen2000/Huihui-GLM-4.7-Flash-abliterated-GGUF | GGUF | 81 |
| wangzhang/GLM-4.7-Flash-abliterated | Transformers | 36 |
| cs2764/Huihui-GLM-4.7-Flash-abliterated-Q6_K-GGUF | GGUF | 32 |
| bloopez/Huihui-GLM-4.7-Flash-abliterated-BF16-GGUF | GGUF | 29 |
| rbinrs/Huihui-GLM-4.7-Flash-abliterated | Transformers | 20 |
| cs2764/Huihui-GLM-4.7-Flash-abliterated-Q8_0-GGUF | GGUF | 16 |
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for GLM-4.7-Flash-abliterated is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with huihui-ai, and does not independently verify publisher claims. See responsible use.