SAVRN
Search Contact SAVRN

Open-weight model

Qwen3.8-27B-ULT3-QNext-Stage1.1

by David Belton DavidAU/Qwen3.8-27B-ULT3-QNext-Stage1.1

Qwen3.8-27B-ULT3-QNext-Stage1.1 is a model from David Belton, released under Apache License 2.0 (access requested at publisher). It has 27.8B parameters. At 16-bit it needs about 66.7 GB of GPU memory, which fits on 1x MI300X from $1.85 an hour, at the lowest prices in the SAVRN Index. It draws 2 downloads a month.

This model is part of this project: https://huggingface.co/DavidAU/Qwen3.8-27B-UltimateDetails2-stage1The-Harley-Pelican This is a work in progress, if you want to be notified the FINAL release, go to ABOVE REPO to request access. - Stage 1 built, tested.

Parameters27.8B
Context—
Weights55.6 GB
Licenseapache-2.0
AccessAccess requested at publisher
Monthly Downloads2

Runs On

What it takes to serve Qwen3.8-27B-ULT3-QNext-Stage1.1 (27.8B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 55.6 GB 66.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
8-bit 27.8 GB 33.3 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 13.9 GB 16.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Oct 7, 2026.

Qwen3.8-27B-ULT3-QNext-Stage1.1 on every accelerator the SAVRN Index prices, at every precision

Model Card

By David Belton, published under apache-2.0, revision 536e76a30c7d.

This model is part of this project:

https://huggingface.co/DavidAU/Qwen3.8-27B-UltimateDetails2-stage1__The-Harley-Pelican

This is a work in progress, if you want to be notified the FINAL release, go to ABOVE REPO to request access.

  • Stage 1 built, tested.
  • Stage 1.1 and Stage 1.5 => Built, in testing.

Benches: [see full project notes/benches at above repo]

          arc/c arc/e boolq hswag obkqa piqa  wino

[NEW BRANCH CORE START]

Ultimate 3 [Qwen Next Branch; ground up build, components first, then combine.]
Qwen3.8-27B-ULT3-QNext-Stage1 => Qnext Reasoning Only Training Mild.
mxfp8     0.617,0.809,0.906,...

Stage 1.1 :
Qwen3.8-27B-ULT3-QNext-Stage1.1
mxfp8    0.622,0.801,0.901,...

Stage 1.5
- pending


[QWENS]
[base, non heretic, untuned]

Qwen3.8-27B-Instruct: 
mxfp8     0.591,0.782,0.896,0.746,0.448,0.801,0.711

Qwen3.6-27B-Instruct: 
mxfp8     0.647,0.803,0.910,0.773,0.450,0.806,0.742

Qwen3.5-27B-Instruct: 
mxfp8     0.557,0.711,0.868,0.533,0.452,0.706,0.695

Read the full model card (185 words)

Identity and Version

Repository
DavidAU/Qwen3.8-27B-ULT3-QNext-Stage1.1
Publisher
David Belton
Task
Not stated by the source
Modality
Other
Library
Not stated by the source
Parameters
27.8B parameters
Languages
Not stated by the source
Revision
536e76a30c7d5ae3ae06cf9ba1bbc6d9c8a8e9c3
First published
2026-10-03
Last updated
2026-10-05

Files and Weights

27 files, 55.6 GB in total. The weights are 18 files totalling 55.6 GB in safetensors.

Weights18 files · 55.6 GB
Configuration4 files · 118.1 KB
Tokenizer2 files · 20.0 MB
Documentation1 file · 1.5 KB
Other1 file · 9.0 KB
Repository1 file · 1.6 KB
Every file
FileTypeSizeSHA-256
model-00001-of-00018.safetensorsWeights4.0 GB —
model-00002-of-00018.safetensorsWeights3.0 GB —
model-00003-of-00018.safetensorsWeights2.5 GB —
model-00004-of-00018.safetensorsWeights4.0 GB —
model-00005-of-00018.safetensorsWeights2.1 GB —
model-00006-of-00018.safetensorsWeights4.0 GB —
model-00007-of-00018.safetensorsWeights2.1 GB —
model-00008-of-00018.safetensorsWeights4.0 GB —
model-00009-of-00018.safetensorsWeights2.1 GB —
model-00010-of-00018.safetensorsWeights4.0 GB —
model-00011-of-00018.safetensorsWeights2.1 GB —
model-00012-of-00018.safetensorsWeights4.0 GB —
model-00013-of-00018.safetensorsWeights2.1 GB —
model-00014-of-00018.safetensorsWeights4.0 GB —
model-00015-of-00018.safetensorsWeights2.1 GB —
model-00016-of-00018.safetensorsWeights4.0 GB —
model-00017-of-00018.safetensorsWeights2.1 GB —
model-00018-of-00018.safetensorsWeights3.4 GB —
config.jsonConfiguration4.5 KB —
generation_config.jsonConfiguration213 B —
model.safetensors.index.jsonConfiguration112.2 KB —
processor_config.jsonConfiguration1.2 KB —
README.mdDocumentation1.5 KB —
chat_template.jinjaOther9.0 KB —
.gitattributesRepository1.6 KB —
tokenizer.jsonTokenizer20.0 MB —
tokenizer_config.jsonTokenizer16.4 KB —

License and Download

License
apache-2.0
Access
Access requested at publisher
Download size
55.6 GB
Request access from David Belton

David Belton grants access through its official repository on Hugging Face. Read the license.

Memory Requirements

PrecisionWeights in memory
As published55.6 GB
16-bit55.6 GB
8-bit27.8 GB
4-bit13.9 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Qwen3.8-27B-ULT3-QNext-Stage1.1

How much GPU memory does Qwen3.8-27B-ULT3-QNext-Stage1.1 need?

About 66.7 GB at 16-bit and 16.7 GB at 4-bit: the weights (27.8B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run Qwen3.8-27B-ULT3-QNext-Stage1.1 on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use Qwen3.8-27B-ULT3-QNext-Stage1.1 commercially?

Yes. Qwen3.8-27B-ULT3-QNext-Stage1.1 is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.