SAVRN
Search Contact SAVRN

Open-weight model · Image and text to text

Huihui-Qwen3.8-27B-abliterated-GGUF

by Huihui.ai huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF

This is an uncensored version of Qwen/Qwen3.8-27B created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.

Parameters
Context
Weights689.6 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads2.8M

Model Card

By Huihui.ai, published under apache-2.0, revision c21e1a9a4549.

This is an uncensored version of Qwen/Qwen3.8-27B created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens. The newly added Huihui-Qwen3.8-27B-abliterated-GSQ-RCO series come from ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF. Only layers 23 to 51 have been ablated, while the other layers remain unablated. It may come with a small disclaimer warning. The size after conversion may differ from the original GGUF. The newly added Huihui-Qwen3.8-27B-abliterated-UD series come from unsloth/Qwen3.8-27B-GGUF. Only layers 18 to 51 have been ablated(Previously…

Read Huihui.ai's full model card

This is an uncensored version of Qwen/Qwen3.8-27B created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.

Latest update 5

The newly added Huihui-Qwen3.8-27B-abliterated-GSQ-RCO series come from ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF. Only layers 23 to 51 have been ablated, while the other layers remain unablated. It may come with a small disclaimer warning. The size after conversion may differ from the original GGUF.

Latest update 4

The newly added Huihui-Qwen3.8-27B-abliterated-UD-DW series come from unsloth/Qwen3.8-27B-GGUF. Only layers 23 to 51 have been ablated, while the other layers remain unablated. It may come with a small disclaimer warning. The size after conversion may differ from the original GGUF.

Latest update

The newly added Huihui-Qwen3.8-27B-abliterated-UD series come from unsloth/Qwen3.8-27B-GGUF. Only layers 18 to 51 have been ablated(Previously, The first 15 layers were retained without ablation), while the other layers remain unablated. The size after conversion may differ from the original GGUF.

Huihui-Qwen3.8-27B-abliterated-bf16.gguf has also been updated.

This helps retain more of the original model’s performance. MTP and visual has not been modified.

Note

The first 15 layers were retained without ablation. MTP and visual has not been modified.

We have already converted the weights (token_embd,output,ffn_down,ssm_out,attn_output) that need to be ablated in the versions below Q8_0 from Q2_K, Q3_K, Q4_K, Q5_K, and Q6_K to Q8_0 to improve response quality, and changed the filename to K_L.

In the Q8_0 quantized version, we changed the Q8_0 weights (token_embd,output,ffn_down,ssm_out,attn_output) targeted for ablation to BF16 and renamed the file to Q8_0_L.

This is not a standard quantization, so you might find that Q2_K_L is larger than Q3_K and Q4_K.

Specific Quantification Method

Some people may misunderstand. The specific quantification method is as follows.

Q2_K_L - Q6_K_L

Qwen3.8-27B-tensor_types-Q6_K_L.txt

llama-quantize \
  --allow-requantize \
  --tensor-type-file huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Qwen3.8-27B-tensor_types-Q6_K_L.txt \
  huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-bf16.gguf \
  huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q6_K_L.gguf Q6_K

Q8_0_L

Qwen3.8-27B-tensor_types-Q8_0_L.txt

llama-quantize \ 
  --allow-requantize \
  --tensor-type-file huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Qwen3.8-27B-tensor_types-Q8_0_L.txt \
  huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-bf16.gguf \ 
  huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q8_0_L.gguf Q8_0

ollama

Please use the latest version of ollama

You can use huihui_ai/Qwen3.8-abliterated directly,

ollama run huihui_ai/Qwen3.8-abliterated

llama.cpp

Use the latest llama.cpp,

llama-cli -m huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF/Huihui-Qwen3.8-27B-abliterated-Q4_K.gguf -c 262144

Usage Warnings

  • Risk of Sensitive or Controversial Outputs: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.

  • Not Suitable for All Audiences: Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.

  • Legal and Ethical Responsibilities: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.

  • Research and Experimental Use: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.

  • Monitoring and Review Recommendations: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.

  • No Default Safety Guarantees: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.

Donation

Your donation helps us continue our further development and improvement, a cup of coffee can do it.
  • bitcoin:
  bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge
  • Support our work on Ko-fi!

Identity and Version

Repository
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
Publisher
Huihui.ai
Task
Image and text to text
Modality
Image and text
Library
transformers
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
c21e1a9a454946b268c74642a257355c2f54ad5c
First published
2026-08-16
Last updated
2026-09-18

Files and Weights

40 files, 689.6 GB in total. The weights are 36 files totalling 689.6 GB in gguf.

Weights36 files · 689.6 GB
Documentation1 file · 5.5 KB
Other2 files · 246 B
Repository1 file · 4.8 KB
Every file
FileTypeSizeSHA-256
Huihui-Qwen3.8-27B-abliterated-GSQ-RCO-IQ3_S-mtp.ggufWeights12.1 GB eea0638e2834
Huihui-Qwen3.8-27B-abliterated-GSQ-RCO-IQ3_S.ggufWeights11.8 GB 329e01fc79dc
Huihui-Qwen3.8-27B-abliterated-GSQ-RCO-IQ3_XXS-mtp.ggufWeights10.4 GB 53fbcda56fae
Huihui-Qwen3.8-27B-abliterated-GSQ-RCO-IQ3_XXS.ggufWeights10.1 GB 54171530e667
Huihui-Qwen3.8-27B-abliterated-Q2_K.ggufWeights10.9 GB c4a5e431ede3
Huihui-Qwen3.8-27B-abliterated-Q2_K_L.ggufWeights17.2 GB 368f0f545100
Huihui-Qwen3.8-27B-abliterated-Q3_K.ggufWeights13.5 GB a82fe31ff871
Huihui-Qwen3.8-27B-abliterated-Q3_K_L.ggufWeights18.7 GB 127edb0588c8
Huihui-Qwen3.8-27B-abliterated-Q4_K.ggufWeights16.8 GB 6c2c13cef892
Huihui-Qwen3.8-27B-abliterated-Q4_K_L.ggufWeights20.9 GB 6d9ee93a089d
Huihui-Qwen3.8-27B-abliterated-Q5_K.ggufWeights19.5 GB 917453854fc6
Huihui-Qwen3.8-27B-abliterated-Q5_K_L.ggufWeights22.9 GB 03b97b83c18c
Huihui-Qwen3.8-27B-abliterated-Q6_K.ggufWeights22.4 GB a5c159519d7b
Huihui-Qwen3.8-27B-abliterated-Q6_K_L.ggufWeights24.9 GB 20b1214a23f0
Huihui-Qwen3.8-27B-abliterated-Q8_0.ggufWeights29.0 GB 427b9416c2ab
Huihui-Qwen3.8-27B-abliterated-Q8_0_L.ggufWeights38.8 GB ac47d8c619ae
Huihui-Qwen3.8-27B-abliterated-UD-DW-IQ3_S.ggufWeights12.3 GB d0eaa1e90e76
Huihui-Qwen3.8-27B-abliterated-UD-DW-IQ3_XXS.ggufWeights11.4 GB bfc6892863ec
Huihui-Qwen3.8-27B-abliterated-UD-DW-Q4_K_M.ggufWeights16.6 GB 0c7cfe306049
Huihui-Qwen3.8-27B-abliterated-UD-DW-Q4_K_S.ggufWeights15.6 GB f10c26c7d07b
Huihui-Qwen3.8-27B-abliterated-UD-DW-Q6_K_L.ggufWeights23.9 GB 72b74fbe001a
Huihui-Qwen3.8-27B-abliterated-UD-DW-Q8_K_L.ggufWeights27.3 GB fb8413d0b5ce
Huihui-Qwen3.8-27B-abliterated-UD-DW-bf16.ggufWeights54.7 GB b880f2042df1
Huihui-Qwen3.8-27B-abliterated-UD-IQ2_S-MTP.ggufWeights8.8 GB 139e6356e22a
Huihui-Qwen3.8-27B-abliterated-UD-IQ2_S.ggufWeights8.4 GB b4b7bdec2e49
Huihui-Qwen3.8-27B-abliterated-UD-IQ3_S.ggufWeights12.0 GB f785ad524b13
Huihui-Qwen3.8-27B-abliterated-UD-IQ3_XXS.ggufWeights11.0 GB e86259c84188
Huihui-Qwen3.8-27B-abliterated-UD-IQ4_XS.ggufWeights14.4 GB d83adbbbaace
Huihui-Qwen3.8-27B-abliterated-UD-Q2_K_XL.ggufWeights10.0 GB 3715d3c31491
Huihui-Qwen3.8-27B-abliterated-UD-Q3_K_XL.ggufWeights13.3 GB c6e144d78595
Huihui-Qwen3.8-27B-abliterated-UD-Q4_K_XL.ggufWeights17.4 GB ebbc66b45cf3
Huihui-Qwen3.8-27B-abliterated-UD-Q5_K_XL.ggufWeights20.7 GB a6ff520853eb
Huihui-Qwen3.8-27B-abliterated-UD-Q6_K_XL.ggufWeights24.8 GB 236d807b7577
Huihui-Qwen3.8-27B-abliterated-UD-Q8_K_XL.ggufWeights31.5 GB b4a19bbd0bb0
Huihui-Qwen3.8-27B-abliterated-bf16.ggufWeights54.7 GB b880f2042df1
mmproj-model-bf16.ggufWeights931.1 MB c9a090646836
README.mdDocumentation5.5 KB
Qwen3.8-27B-tensor_types-Q6_K_L.txtOther123 B
Qwen3.8-27B-tensor_types-Q8_0_L.txtOther123 B
.gitattributesRepository4.8 KB

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
689.6 GB
Download from Huihui.ai

Released by Huihui.ai through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published689.6 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Huihui-Qwen3.8-27B-abliterated-GGUF

Can I use Huihui-Qwen3.8-27B-abliterated-GGUF commercially?

Yes. Huihui-Qwen3.8-27B-abliterated-GGUF is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Similar Models

Model · Image and text to text

Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF

Michał Piszczek

I built this quant because the ready-made FP4 file answered the wrong question. It was fast, but on my short WikiText-2 control it scored 6.4949 PPL. Plain Q40 scored 6.3798. The first higher-quality hybrid went too far the other way: good perplexity, 34.19 tok/s, and no comfortable room for 256K plus vision. This is the build that survived both gates. It is a 17.1 GB, 5.01 BPW mixed-precision GGUF of Qwen/Qwen3.8-27B. It keeps large, tolerant matrices in native NVFP4 and spends more bits on selected attention, Gated DeltaNet, and late FFN tensors. The trained MTP layer remains embedded in the same GGUF. This is not a fine-tune. I built the private calibration workload from 5,472 messages…

Open weights apache-2.0

Qwen3.8-27B uncensored by HauhauCS 0/465 Refusals. This is the Aggressive variant: direct answers, no refusal behavior, and minimal preamble on hard prompts. Every text GGUF preserves Qwen3.8's native NextN head, and this release adds HauhauCS FastMTP: a specific acceleration sidecar qualified across the complete quant lineup at maximum native context. Vision is included through the separate BF16 projector. No changes to datasets or intended capabilities. This release preserves Qwen3.8-27B's text, reasoning, agentic, image, and video capabilities while applying the HauhauCS Aggressive uncensoring profile. Pick Aggressive when you specifically want the model to get to the answer without…

Open weights apache-2.0

Model · Image and text to text

Gemma-4-E4B-Uncensored-HauhauCS-Aggressive

HauhauCS

Gemma 4 E4B-IT uncensored by HauhauCS. 0/465 Refusals\ No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals. These are meant to be the best lossless uncensored models out there. Stronger uncensoring — model is fully unlocked and won't refuse prompts. May occasionally append short disclaimers (baked into base model training, not refusals) but full content is always generated. For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available. All quants generated with importance matrix (imatrix) for optimal quality preservation on abliterated weights. KP ("Perfect")…

Open weights gemma

Model · Image and text to text

Qwen3.5-9B-GGUF

Unsloth AI

You can now also fine-tune the model locally with Unsloth. - Read our Qwen3.5 fine-tuning guide here. Over recent months, we have intensified our focus on developing foundation models that deliver exceptional utility and performance. Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency. For more details, please refer to our blog post Qwen3.5. WMT24++: a harder subset of WMT24 after difficulty labeling and rebalancing; we report the averaged scores on 55 languages using XCOMET-XXL. Empty…

Open weights apache-2.0 transformers

Model · Image and text to text

Qwen3.8-Flash-Next-GGUF

Unsloth AI

As the frontier of foundation models pushes toward ever-larger parameter counts and ever-longer context windows, the question is no longer just how much we can scale, but how efficiently we can do so. Sustainable progress toward artificial general intelligence (AGI) that benefits everyone demands architectural innovation. Today, we are sharing a concrete step in that direction: Qwen3.8-Flash-Next. This experimental preview of the architecture that will underpin Qwen4 is built around a fundamental rethinking of how the core components of modern large language models (LLMs) interact at scale. The first open-weight release under this architecture is Qwen3.8-Flash-Next, which introduces: For…

Open weights other

and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads) instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model ctrl at the chat/message level). Model name has "plusIQ" in the name. (there is also a extra robust "tools" version too.) Extreme intelligence in a small package. Jaw dropping performance. Superior instruction following. A multi-stage and multi-model fine tune and multi-stage merge on local hardware by myself and Nightmedia. Several of my 9B Qwen 3.5 fine tunes were…

Open weights apache-2.0