Qwen3.8-Flash-Next-Uncensored-Serve
from orcarouter/Qwen3.8-Flash-Next-Uncensored
Qwen3.8-Flash-Next-Uncensored-NGQ4 is a abliterated and uncensored variant of orcarouter/Qwen3.8-Flash-Next-Uncensored, published on Hugging Face by cygnal. It is other licence. It is distributed across 2 repositories in GGUF format, totalling 17.8K downloads.
How Qwen3.8-Flash-Next-Uncensored-NGQ4 was modified, and what that implies.
Abliteration + uncensored fine-tune
Directional ablation combined with additional fine-tuning on unfiltered data, so behaviour diverges from the base model beyond refusal removal alone.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
From the publisher’s model card
“First working GGUF build of Qwen3.8-Flash-Next (qwen4exp architecture) with vision, running on stock llama.cpp — no custom tensor formats or forked runtime required. Built from orcarouter/Qwen3.8-Flash-Next-Uncensored, the abliterated (uncensored) release of Qwen's newest hybrid architecture.”
Commands are templates — confirm the exact repository and quant file before use.
hf download cygnal/Qwen3.8-Flash-Next-Uncensored-IQ4XS-NGQ4-GGUF --local-dir ./qwen3.8-flash-next-uncensored-ngq4
llama-cli -m ./qwen3.8-flash-next-uncensored-ngq4/<file>.gguf -p "..."from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "cygnal/Qwen3.8-Flash-Next-Uncensored-IQ4XS-NGQ4-GGUF"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve cygnal/Qwen3.8-Flash-Next-Uncensored-IQ4XS-NGQ4-GGUF --trust-remote-codeThe only published repository of Qwen3.8-Flash-Next-Uncensored-NGQ4.
| Repository | Format | Downloads |
|---|---|---|
| cygnal/Qwen3.8-Flash-Next-Uncensored-IQ4XS-NGQ4-GGUFsource | GGUF | 17.8K |
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for Qwen3.8-Flash-Next-Uncensored-NGQ4 is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with cygnal, and does not independently verify publisher claims. See responsible use.