huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
Summary
An uncensored variant of the Qwen3.8-27B LLM, modified via abliteration to remove content restrictions, provided in GGUF format for deployment with llama.cpp and ollama.
View Cached Full Text
Cached at: 08/19/26, 03:45 PM
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF · Hugging Face
Source: https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#huihui-aihuihui-qwen38-27b-abliterated-ggufhuihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
This is an uncensored version ofQwen/Qwen3.8-27Bcreated with abliteration (seeremove-refusals-with-transformersto know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#noteNote
The first 15 layers were retained without ablation. MTP and visual has not been modified.
We have already converted the weights (token_embd,output,ffn_down,ssm_out,attn_output) that need to be ablated in the versions below Q8_0 from Q2_K, Q3_K, Q4_K, Q5_K, and Q6_K to Q8_0 to improve response quality, and changed the filename to K_L.
In the Q8_0 quantized version, we changed the Q8_0 weights (token_embd,output,ffn_down,ssm_out,attn_output) targeted for ablation to BF16 and renamed the file to Q8_0_L.
This is not a standard quantization, so you might find that Q2_K_L is larger than Q3_K and Q4_K.
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#specific-quantification-methodSpecific Quantification Method
Some people may misunderstand. The specific quantification method is as follows.
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#q2_k_l—q6_k_lQ2_K_L - Q6_K_L
Qwen3.8-27B-tensor_types-Q6_K_L.txt
llama-quantize \
--allow-requantize \
--tensor-type-file huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Qwen3.8-27B-tensor_types-Q6_K_L.txt \
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-bf16.gguf \
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q6_K_L.gguf Q6_K
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#q8_0_lQ8_0_L
Qwen3.8-27B-tensor_types-Q6_K_L.txt
llama-quantize \
--allow-requantize \
--tensor-type-file huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Qwen3.8-27B-tensor_types-Q8_0_L.txt \
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-bf16.gguf \
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q8_0_L.gguf Q8_0
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#ollamaollama
Please use the latest version ofollama
You can usehuihui_ai/Qwen3.8-abliterateddirectly,
ollama run huihui_ai/Qwen3.8-abliterated
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#llamacppllama.cpp
Use the latestllama.cpp,
llama-cli -m huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q4_K.gguf -c 262144
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#usage-warningsUsage Warnings
- Risk of Sensitive or Controversial Outputs: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.
- Not Suitable for All Audiences: Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.
- Legal and Ethical Responsibilities: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.
- Research and Experimental Use: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.
- Monitoring and Review Recommendations: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.
- No Default Safety Guarantees: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#donationDonation
https://huggingface.co/huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF#your-donation-helps-us-continue-our-further-development-and-improvement-a-cup-of-coffee-can-do-itYour donation helps us continue our further development and improvement, a cup of coffee can do it.
- bitcoin:
bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge
- Support our work onKo-fi!
Similar Articles
huihui-ai/Huihui-Qwen3.8-27B-abliterated
This is an uncensored version of the Qwen3.8-27B AI model created using abliteration to remove refusals, serving as a proof-of-concept for modifying LLMs without extensive tools.
@support_huihui: New MTP-GGUF: huihui-ai/Huihui-Qwen3.6-27B-abliterated-MTP-GGUF This is an uncensored version of Qwen/Qwen3.6-27B creat…
An uncensored GGUF version of Qwen3.6-27B, created via abliteration, is now available on Hugging Face from huihui-ai.
huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
A model card for Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF, an abliterated (uncensored) GGUF quantized variant of DeepSeek-V4-Flash, designed for local use with llama.cpp and ds4.
0bserverx/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
This article describes the release of Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF, a double-refined abliterated variant of the Qwen model with reduced refusals for adult audiences, using ARA technique for research and creative writing.
Qwen3.6-27B Uncensored Aggressive is out with K_P quants!
Community release of Qwen3.6-27B stripped of safety refusals and packaged in optimized K_P GGUF quants for llama.cpp and LM Studio.