Qwen3.8-Flash-Next-Uncensored-NGQ4
from orcarouter/Qwen3.8-Flash-Next-Uncensored
Qwen3.8-Flash-Next-Uncensored-Serve is a abliterated and uncensored variant of orcarouter/Qwen3.8-Flash-Next-Uncensored, published on Hugging Face by ARC4NUM. It is apache-2.0 licence. It is distributed across 2 repositories in MLX format, totalling 5.4K downloads.
How Qwen3.8-Flash-Next-Uncensored-Serve was modified, and what that implies.
Abliteration + uncensored fine-tune
Directional ablation combined with additional fine-tuning on unfiltered data, so behaviour diverges from the base model beyond refusal removal alone.
Reported scope: Not stated by the publisher. Publishers frequently omit this, so verify behaviour empirically rather than assuming full coverage.
From the publisher’s model card
“and is memory-mapped from disk, not held in RAM”
Commands are templates — confirm the exact repository and quant file before use.
from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "ARC4NUM/Qwen3.8-Flash-Next-Uncensored-MLX-Serve-4bit"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(
repo, torch_dtype="auto", device_map="auto"
)vllm serve ARC4NUM/Qwen3.8-Flash-Next-Uncensored-MLX-Serve-4bit --trust-remote-codeThe only published repository of Qwen3.8-Flash-Next-Uncensored-Serve.
| Repository | Format | Downloads |
|---|---|---|
| ARC4NUM/Qwen3.8-Flash-Next-Uncensored-MLX-Serve-4bitsource | MLX | 5.4K |
Derived from this model’s metadata and the publisher’s claims — not from benchmarks run by this site.
Handling: run it isolated, put moderation in front of it if anyone outside your team can reach it, and verify the weights before loading — these are third-party artifacts. Full handling and licence guidance.
Metadata for Qwen3.8-Flash-Next-Uncensored-Serve is collected from the public Hugging Face API and the publisher’s model card. This site does not host weights, is not affiliated with ARC4NUM, and does not independently verify publisher claims. See responsible use.