SAVRN
Search Contact SAVRN

Open-weight model · Image and text to text

Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF

by AJ Gazin ajgazin/Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF

Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF is an open-weight model for image and text to text from AJ Gazin, released under other. Its published files total 243.9 GB. It draws 26.6k downloads a month.

GGUF quants of ajgazin/Swift-Qwen3.8-27B-Uncensored-MTP, an abliterated Swift-Qwen3.8-27B (UkisAI's reasoning-efficient fine-tune of Qwen3.8-27B). For vLLM and version is size, with Unsloth's importance matrix.

Parameters—
Context—
Weights243.9 GB
Licenseother
AccessOpen weights
Monthly Downloads26.6k

Model Card

GGUF quants of ajgazin/Swift-Qwen3.8-27B-Uncensored-MTP, an abliterated Swift-Qwen3.8-27B (UkisAI's reasoning-efficient fine-tune of Qwen3.8-27B). For vLLM and version is size, with Unsloth's importance matrix. - MTP head included in every main GGUF, for self-speculative decoding in llama.cpp. Each quant is one file, Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-.gguf; the BF16 is Swift-Qwen3.8-27B-Uncensored-MTP-BF16.gguf. Each quant compared with the BF16 it was made from, token by token: wikitext-2 test set, 36 × 8192 tokens, f16 KV cache, llama.cpp 94659b076. Lower KLD and higher top-1 are better. - UD-IQ3S and UD-IQ4XS were added after this measurement run, so they have no numbers yet and…

Excerpt from the card by AJ Gazin, licensed other.

Identity and Version

Repository
ajgazin/Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF
Publisher
AJ Gazin
Task
Image and text to text
Modality
Image and text
Library
gguf
Parameters
Not stated by the source
Languages
mtp
Revision
b82911f48602c6aa2be9aa60759db6489c2b8beb
First published
2026-09-15
Last updated
2026-09-25

Files and Weights

18 files, 243.9 GB in total. The weights are 13 files totalling 243.9 GB in gguf.

Weights13 files · 243.9 GB
Documentation1 file · 12.7 KB
Other3 files · 143.9 KB
Repository1 file · 2.7 KB
Every file
FileTypeSizeSHA-256
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-IQ3_S.ggufWeights12.0 GB 23be95d1e4ac
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-IQ3_XXS.ggufWeights10.9 GB 945575f5769c
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-IQ4_XS.ggufWeights14.3 GB 6f40b1eaf6f0
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q2_K_XL.ggufWeights9.8 GB 4495eb0ad263
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q3_K_XL.ggufWeights13.1 GB efe0a2142d91
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q4_K_S.ggufWeights15.4 GB 8fc3a71d2899
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q4_K_XL.ggufWeights17.6 GB e7e063b078de
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q5_K_M.ggufWeights19.8 GB 274e40f4eb3c
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q5_K_S.ggufWeights18.7 GB 2b790ff0195f
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q6_K_XL.ggufWeights25.3 GB 65d5921988a2
Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-UD-Q8_K_XL.ggufWeights31.5 GB 5ae62a04251b
Swift-Qwen3.8-27B-Uncensored-MTP-BF16.ggufWeights54.7 GB 53da4d4c71cc
mmproj-BF16.ggufWeights931.1 MB b343ceeb860c
README.mdDocumentation12.7 KB —
kld_results.csvOther4.3 KB —
kld_vs_size.pngOther96.5 KB —
tensor_types.tsvOther43.2 KB —
.gitattributesRepository2.7 KB —

License and Download

License
other
Access
Open weights, no gate
Download size
243.9 GB
Download from AJ Gazin

Released by AJ Gazin through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published243.9 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF

What license is Swift-Qwen3.8-27B-Uncensored-Dynamic-MTP-GGUF released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image and text to text

Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF

Michał Piszczek

I built this quant because the ready-made FP4 file answered the wrong question. It was fast, but on my short WikiText-2 control it scored 6.4949 PPL. Plain Q40 scored 6.3798. The first higher-quality hybrid went too far the other way: good perplexity, 34.19 tok/s, and no comfortable room for 256K plus vision. This is the build that survived both gates. It is a 17.1 GB, 5.01 BPW mixed-precision GGUF of Qwen/Qwen3.8-27B. It keeps large, tolerant matrices in native NVFP4 and spends more bits on selected attention, Gated DeltaNet, and late FFN tensors. The trained MTP layer remains embedded in the same GGUF. This is not a fine-tune. I built the private calibration workload from 5,472 messages…

Open weights apache-2.0

Model · Image and text to text

Huihui-Qwen3.8-27B-abliterated-GGUF

Huihui.ai

This is an uncensored version of Qwen/Qwen3.8-27B created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens. The newly added Huihui-Qwen3.8-27B-abliterated-Ternary series come from prism-ml/Ternary-Bonsai-2-27B-gguf have been ablated, while the other layers remain unablated. It may come with a small disclaimer warning. The size after conversion may differ from the original GGUF (Some of the weights are converted from PTQ1 to Q2K or Q3K.). This is just a test/validation. The ternary hybrid-attention kernels live in the PrismML-Eng/llama.cpp fork.…

Open weights apache-2.0 transformers

Qwen3.8-27B uncensored by HauhauCS 0/465 Refusals. This is the Aggressive variant: direct answers, no refusal behavior, and minimal preamble on hard prompts. Every text GGUF preserves Qwen3.8's native NextN head, and this release adds HauhauCS FastMTP: a specific acceleration sidecar qualified across the complete quant lineup at maximum native context. Vision is included through the separate BF16 projector. No changes to datasets or intended capabilities. This release preserves Qwen3.8-27B's text, reasoning, agentic, image, and video capabilities while applying the HauhauCS Aggressive uncensoring profile. Pick Aggressive when you specifically want the model to get to the answer without…

Open weights apache-2.0

Model · Image and text to text

Qwen3.8-Flash-Next-GGUF

Unsloth AI

As the frontier of foundation models pushes toward ever-larger parameter counts and ever-longer context windows, the question is no longer just how much we can scale, but how efficiently we can do so. Sustainable progress toward artificial general intelligence (AGI) that benefits everyone demands architectural innovation. Today, we are sharing a concrete step in that direction: Qwen3.8-Flash-Next. This experimental preview of the architecture that will underpin Qwen4 is built around a fundamental rethinking of how the core components of modern large language models (LLMs) interact at scale. The first open-weight release under this architecture is Qwen3.8-Flash-Next, which introduces: For…

Open weights other

and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads) instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model ctrl at the chat/message level). Model name has "plusIQ" in the name. (there is also a extra robust "tools" version too.) A 12+12 (12 reasoning and 12 instruct) model with interactive optimization/help system will be releasing shortly too. Extreme intelligence in a small package. Jaw dropping performance. Superior instruction following. A multi-stage and multi-model…

Open weights apache-2.0

Model · Image and text to text

Gemma-4-E4B-Uncensored-HauhauCS-Aggressive

HauhauCS

Gemma 4 E4B-IT uncensored by HauhauCS. 0/465 Refusals\ No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals. These are meant to be the best lossless uncensored models out there. Stronger uncensoring — model is fully unlocked and won't refuse prompts. May occasionally append short disclaimers (baked into base model training, not refusals) but full content is always generated. For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available. All quants generated with importance matrix (imatrix) for optimal quality preservation on abliterated weights. KP ("Perfect")…

Open weights gemma