SAVRN
Search Contact SAVRN

Open-weight model · Text generation

Qwen3.8-27B-UD-Q8_K_XL-layers

by Maveric Studio studiomaveric/Qwen3.8-27B-UD-Q8_K_XL-layers

GGUF layer package for running Qwen3.8-27B-UD-Q8KXL across a local Mesh LLM cluster. This package is derived from unsloth/Qwen3.8-27B-GGUF and keeps the original GGUF distribution split into per-layer artifacts for distributed inference.

Parameters
Context
Weights33.1 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads

Model Card

By Maveric Studio, published under apache-2.0, revision ee0ffb5bd12f.

GGUF layer package for running Qwen3.8-27B-UD-Q8KXL across a local Mesh LLM cluster. This package is derived from unsloth/Qwen3.8-27B-GGUF and keeps the original GGUF distribution split into per-layer artifacts for distributed inference. - Local and private inference with Mesh LLM. - Multi-machine serving when the full GGUF is too large for one host. - OpenAI-compatible chat/completions workflows through Mesh LLM's local API. For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the source model card: unsloth/Qwen3.8-27B-GGUF. Generated by the Mesh LLM HF Jobs splitter from mesh-llm ref main. Each artifact is checksummed…

Read Maveric Studio's full model card

Qwen3.8-27B-UD-Q8_K_XL

Distributed GGUF inference package for Mesh LLM

GGUF layer package for running Qwen3.8-27B-UD-Q8_K_XL across a local Mesh LLM cluster.

This package is derived from unsloth/Qwen3.8-27B-GGUF and keeps the original GGUF distribution split into per-layer artifacts for distributed inference.

Highlights

Run locally Pool multiple machines OpenAI-compatible Package variant
Private inference on your hardware Split layers across peers Serve /v1/chat/completions locally UD-Q8_K_XL layer package

Model Overview

Property Value
Source model unsloth/Qwen3.8-27B-GGUF
Model id unsloth/Qwen3.8-27B-GGUF:UD-Q8_K_XL
Family Qwen3
Parameter scale 27B
Quantization UD-Q8_K_XL
Layer count 65
Activation width not recorded
Package size 0 B
Source file Qwen3.8-27B-UD-Q8_K_XL.gguf
Package repo studiomaveric/Qwen3.8-27B-UD-Q8_K_XL-layers
License apache-2.0 from unsloth/Qwen3.8-27B-GGUF

Recommended Use

  • Local and private inference with Mesh LLM.
  • Multi-machine serving when the full GGUF is too large for one host.
  • OpenAI-compatible chat/completions workflows through Mesh LLM's local API.

For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the source model card: unsloth/Qwen3.8-27B-GGUF.

Quickstart

# Run this on each machine that should contribute memory/compute.
mesh-llm serve --model "studiomaveric/Qwen3.8-27B-UD-Q8_K_XL-layers" --split
# Check the mesh and discover the OpenAI-compatible model name.
curl -s http://localhost:3131/api/status
curl -s http://localhost:3131/v1/models
# Send an OpenAI-compatible chat request.
curl -s http://localhost:3131/v1/chat/completions \
  -H "Content-Type: application/json" \
  -d '{
    "model": "unsloth/Qwen3.8-27B-GGUF:UD-Q8_K_XL",
    "messages": [{"role": "user", "content": "Write a tiny hello-world function in Rust."}],
    "max_tokens": 128
  }'

Package Variant

Property Value
Format gguf
Canonical source ref unsloth/Qwen3.8-27B-GGUF@4ca720788d1e01f1bff70c033e0d0028fd02e502/Qwen3.8-27B-UD-Q8_K_XL.gguf
Source revision 4ca720788d1e01f1bff70c033e0d0028fd02e502
Source SHA-256 af36ecb6b5db1407953345b746c14ac93f0657dda413910b4348683a2d990377
Skippy ABI not recorded
Package manifest SHA-256 123eafd6c7f280e840d76471fbf0caf4e76eb4037e1247f90921be131672ad4a

What Is Included

Artifact Path Contents SHA-256
Manifest model-package.json Package schema, source identity, checksums 123eafd6c7f280e840d76471fbf0caf4e76eb4037e1247f90921be131672ad4a

Validation

Generated by the Mesh LLM HF Jobs splitter from mesh-llm ref main. Each artifact is checksummed as it is written, uploaded to this repository, and removed from the job workspace before the next artifact is produced.

skippy-model-package write-package "/hf-cache/Qwen3.8-27B-UD-Q8_K_XL.gguf" --out-dir "/tmp/meshllm-layer-job-studiomaveric_Qwen3.8-27B-UD-Q8_K_XL-layers-1/package"

Links

Identity and Version

Repository
studiomaveric/Qwen3.8-27B-UD-Q8_K_XL-layers
Publisher
Maveric Studio
Task
Text generation
Modality
Text
Library
mesh-llm
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
ee0ffb5bd12f2d6c679a054faff69f277cf7aba3
First published
2026-09-18
Last updated
2026-09-18

Files and Weights

72 files, 33.1 GB in total. The weights are 69 files totalling 33.1 GB in gguf.

Weights69 files · 33.1 GB
Configuration1 file · 15.4 KB
Documentation1 file · 5.0 KB
Repository1 file · 5.7 KB
Every file
FileTypeSizeSHA-256
layers/layer-00000.ggufWeights418.3 MB 1c80d4cdf756
layers/layer-00001.ggufWeights418.3 MB 8909f6b23461
layers/layer-00002.ggufWeights418.3 MB e06227a363b1
layers/layer-00003.ggufWeights475.3 MB 70f2a5af85e1
layers/layer-00004.ggufWeights418.3 MB 15688585f8d4
layers/layer-00005.ggufWeights418.3 MB c8dc6b4eec3f
layers/layer-00006.ggufWeights418.3 MB 3afd4ac58a96
layers/layer-00007.ggufWeights475.3 MB db164a4713f0
layers/layer-00008.ggufWeights418.3 MB 006e165d8659
layers/layer-00009.ggufWeights418.3 MB 2cf0a5aa0f38
layers/layer-00010.ggufWeights418.3 MB 83e84f00038e
layers/layer-00011.ggufWeights475.3 MB fa8005ccce9b
layers/layer-00012.ggufWeights418.3 MB 4c3a07c17308
layers/layer-00013.ggufWeights418.3 MB 5f15348dc6a0
layers/layer-00014.ggufWeights418.3 MB 15e4f80f1019
layers/layer-00015.ggufWeights475.3 MB fbaa6cfad44c
layers/layer-00016.ggufWeights418.3 MB 02f8360b88fe
layers/layer-00017.ggufWeights418.3 MB 3a7db863be78
layers/layer-00018.ggufWeights418.3 MB e9ce52c731f4
layers/layer-00019.ggufWeights475.3 MB cb030d834cde
layers/layer-00020.ggufWeights418.3 MB 4ccde1ccb2c6
layers/layer-00021.ggufWeights418.3 MB 7dc2b0f9a60f
layers/layer-00022.ggufWeights418.3 MB d341cc268990
layers/layer-00023.ggufWeights475.3 MB 3fd6029eba36
layers/layer-00024.ggufWeights418.3 MB 1d804741e557
layers/layer-00025.ggufWeights418.3 MB 5218df626664
layers/layer-00026.ggufWeights418.3 MB d9399563582e
layers/layer-00027.ggufWeights475.3 MB 39062a555b6a
layers/layer-00028.ggufWeights418.3 MB e33d968a7b12
layers/layer-00029.ggufWeights418.3 MB 51ceb590ccbe
layers/layer-00030.ggufWeights418.3 MB d5b30ac5f767
layers/layer-00031.ggufWeights475.3 MB c2279862d9d9
layers/layer-00032.ggufWeights418.3 MB 431d34b8438d
layers/layer-00033.ggufWeights418.3 MB 716324510c9b
layers/layer-00034.ggufWeights418.3 MB e1816aa68318
layers/layer-00035.ggufWeights475.3 MB 389fa8cf0919
layers/layer-00036.ggufWeights418.3 MB fe8256d8f9e2
layers/layer-00037.ggufWeights418.3 MB a183127e970b
layers/layer-00038.ggufWeights418.3 MB ab9fbb2e730c
layers/layer-00039.ggufWeights475.3 MB a9bf7af71d4c
layers/layer-00040.ggufWeights418.3 MB e46e2506ed30
layers/layer-00041.ggufWeights418.3 MB 70668cb0a9e9
layers/layer-00042.ggufWeights418.3 MB edb790324615
layers/layer-00043.ggufWeights475.3 MB 8e7ee8398573
layers/layer-00044.ggufWeights418.3 MB 5fa5218d55a7
layers/layer-00045.ggufWeights418.3 MB 494ee3be04c5
layers/layer-00046.ggufWeights418.3 MB 801000c29c73
layers/layer-00047.ggufWeights475.3 MB 9590bc3cd5e2
layers/layer-00048.ggufWeights418.3 MB 56cdaf6aeb27
layers/layer-00049.ggufWeights418.3 MB 8751ed9d090c
layers/layer-00050.ggufWeights418.3 MB c536a8fa93ef
layers/layer-00051.ggufWeights475.3 MB b7a11237b68a
layers/layer-00052.ggufWeights418.3 MB c6da441c10e5
layers/layer-00053.ggufWeights418.3 MB 7a43fc1bb86f
layers/layer-00054.ggufWeights418.3 MB 7de138fda475
layers/layer-00055.ggufWeights475.3 MB f80978dc05a9
layers/layer-00056.ggufWeights418.3 MB c936a0daf3a2
layers/layer-00057.ggufWeights418.3 MB 6ddd8c702c7d
layers/layer-00058.ggufWeights418.3 MB 0a400a4c4992
layers/layer-00059.ggufWeights475.3 MB e6c48beb4ee0
layers/layer-00060.ggufWeights418.3 MB db3b889d5489
layers/layer-00061.ggufWeights418.3 MB 3ba8285e4ddb
layers/layer-00062.ggufWeights418.3 MB 6bac1d87d7ea
layers/layer-00063.ggufWeights475.3 MB 7e2e680010bb
layers/layer-00064.ggufWeights580.2 MB 48a5df89ddf0
projectors/projector-00000.ggufWeights931.1 MB 83ee4f4f205f
shared/embeddings.ggufWeights1.4 GB 343de50fa4b9
shared/metadata.ggufWeights11.0 MB 32d16f1de353
shared/output.ggufWeights2.6 GB 1ebfc808cf90
model-package.jsonConfiguration15.4 KB
README.mdDocumentation5.0 KB
.gitattributesRepository5.7 KB

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
33.1 GB
Download from Maveric Studio

Released by Maveric Studio through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published33.1 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Qwen3.8-27B-UD-Q8_K_XL-layers

Can I use Qwen3.8-27B-UD-Q8_K_XL-layers commercially?

Yes. Qwen3.8-27B-UD-Q8_K_XL-layers is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Similar Models

Fine-tune Qwen3 (14B) for free using our Google Colab notebook! - Read our Blog about Qwen3 support: unsloth.ai/blog/qwen3 - View the rest of our notebooks in our docs here. Qwen3-Coder is available in multiple sizes. Today, we're excited to introduce Qwen3-Coder-30B-A3B-Instruct. This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements: - Significant Performance among open models on Agentic Coding, Agentic Browser-Use, and other foundational coding tasks. - Long-context Capabilities with native support for 256K tokens, extendable up to 1M tokens using Yarn, optimized for repository-scale understanding. - Agentic Coding supporting for…

Open weights apache-2.0 transformers

Model · Text generation

opt-125m

AI at Meta

OPT was first introduced in Open Pre-trained Transformer Language Models and first released in metaseq's repository on May 3rd 2022 by Meta AI. Disclaimer: The team releasing OPT wrote an official model card, which is available in Appendix D of the paper. Content from this model card has been written by the Hugging Face team. To quote the first two paragraphs of the official paper OPT was predominantly pretrained with English text, but a small amount of non-English data is still present within the training corpus via CommonCrawl. The model was pretrained using a causal language modeling (CLM) objective. OPT belongs to the same family of decoder-only models like GPT-3. As such, it was…

Open weights other 2,048 tokens transformers

Model · Text generation

Ornith-1.5-9B-GGUF

Ornith

Chirp Chirp! We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement. Ornith-1.5 extends Ornith-1.0 (which was developed on top of Qwen3.5 and Gemma4 with additional continued pretraining, mid-training, and post-training) by expanding the self-improvement loop from scaffold and rollout optimization to jointly optimizing task generation, scaffold construction, and solution rollouts. Rather than relying on a fixed set of human-curated tasks and manually designed harnesses, Ornith-1.5 continuously generates new training tasks, discovers effective strategies for solving them, and improves the policy through reinforcement learning. For…

Open weights mit transformers

Model · Text generation

Ornith-1.5-35B-A3B-GGUF

Ornith

Chirp Chirp! We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement. Ornith-1.5 extends Ornith-1.0 (which was developed on top of Qwen3.5 and Gemma4 with additional continued pretraining, mid-training, and post-training) by expanding the self-improvement loop from scaffold and rollout optimization to jointly optimizing task generation, scaffold construction, and solution rollouts. Rather than relying on a fixed set of human-curated tasks and manually designed harnesses, Ornith-1.5 continuously generates new training tasks, discovers effective strategies for solving them, and improves the policy through reinforcement learning. For…

Open weights mit transformers

Model · Text generation

Ornith-1.0-9B-GGUF

Ornith

Aloha! Today, we are releasing Ornith-1.0, a self-improving family of open-source models for agentic coding. This model card documents Ornith-1.0-9B, the most lightweight member of the Ornith family, designed for efficient single-GPU deployment. Ornith-1.0-9B is a dense ~9B model (≈19 GB in bf16), so it serves comfortably on a single 80GB GPU. The recipes below stand up an OpenAI-compatible server; add --tensor-parallel-size / --tp if you want to shard across more GPUs. For a quick local test (or to script offline generation), load the model directly with Transformers. Make sure you have a recent release installed — see the Transformers installation guide; Ornith-1.0-9B requires…

Open weights mit transformers

Uncensored Qwen3.8-27B, published as GGUF quantizations with the multi token prediction (MTP) head retained and verified. Refusal behaviour has been substantially reduced, not eliminated. See Measured behaviour for the numbers. Capabilities, training data, and architecture are otherwise unchanged. - Refusal directions removed with Heretic, which co minimizes refusal count against KL divergence from the base model. No handwritten refusal removal code, no finetuning, no additional training data. - Abliteration runs at bf16 (no 4 bit quantization). the resulting LoRA is merged into the bf16 base, so the published weights are not a quantized round trip. - mtp. tensors are copied verbatim from…

Open weights apache-2.0 llama.cpp