Gemma4-12B-QAT-Uncensored-HauhauCS-Balanced
from google/gemma-4-12B-it
gemma-3-12b-it-heretic-v2_fp8_e4m3fn is a Heretic-ablated open-weight language model, published on Hugging Face by lynaNSFW. It is 12B parameters and 24GB VRAM class class. It is distributed across 2 repositories in FP8 format, totalling 10.2K downloads.
How gemma-3-12b-it-heretic-v2_fp8_e4m3fn was modified, and what that implies.
Heretic
Automated directional ablation via the Heretic toolchain, which searches for the refusal direction and applies it with a KL-divergence budget so general capability is preserved.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
Commands are templates — confirm the exact repository and quant file before use.
from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "lynaNSFW/gemma-3-12b-it-heretic-v2_fp8_e4m3fn"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve lynaNSFW/gemma-3-12b-it-heretic-v2_fp8_e4m3fn --trust-remote-codeThe only published repository of gemma-3-12b-it-heretic-v2_fp8_e4m3fn.
| Repository | Format | Downloads |
|---|---|---|
| lynaNSFW/gemma-3-12b-it-heretic-v2_fp8_e4m3fnsource | FP8 | 10.2K |
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for gemma-3-12b-it-heretic-v2_fp8_e4m3fn is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with lynaNSFW, and does not independently verify publisher claims. See responsible use.