SAVRN
Search Contact SAVRN

Open-weight model · Image and text to text

MIMO-V2.6-DERISKED-MXFP4

by Sir Frosty Blackfrost-AI/MIMO-V2.6-DERISKED-MXFP4

MIMO-V2.6-DERISKED-MXFP4 is a model for image and text to text from Sir Frosty, released under MIT License (access requested at publisher). Its published files total 1.9 MB.

Behaviorally modified MXFP4 research checkpoint derived from XiaomiMiMo/MiMo-V2.6-Flash-RL. DERISKED identifies the Blackfrost release family. It is not a claim of zero refusals, complete safety, harmlessness, or production readiness.

Parameters—
Context—
Weights1.9 MB
Licensemit
AccessAccess requested at publisher
Monthly Downloads—

Model Card

By Sir Frosty, published under mit, revision 9c809a361706.

Behaviorally modified MXFP4 research checkpoint derived from XiaomiMiMo/MiMo-V2.6-Flash-RL. DERISKED identifies the Blackfrost release family. It is not a claim of zero refusals, complete safety, harmlessness, or production readiness. Refusal measurements will be published only after evaluation of the exact uploaded artifact is complete. MIMO-V2.6-DERISKED-MXFP4 is an independent Blackfrost research derivative of XiaomiMiMo/MiMo-V2.6-Flash-RL. The checkpoint retains the upstream MiMo V2.6 architecture, tokenizer, multimodal components, native chat template, and speculative-decoding assets while preserving its mixed-precision MXFP4 deployment layout. The behavioral modification is encoded in…

Read Sir Frosty's full model card

Behaviorally modified MXFP4 research checkpoint derived from XiaomiMiMo/MiMo-V2.6-Flash-RL.

Release status

Item Status
Weight package MXFP4 mixed-precision safetensors included
Structural verification Complete
Native MiMo chat template Included
Baked deployment system prompt None
Refusal evaluation Pending
Comprehensive derivative benchmark Pending
Modification recipe Proprietary and intentionally not distributed
Access Public model page with manual approval gating

DERISKED identifies the Blackfrost release family. It is not a claim of zero refusals, complete safety, harmlessness, or production readiness. Refusal measurements will be published only after evaluation of the exact uploaded artifact is complete.

Overview

MIMO-V2.6-DERISKED-MXFP4 is an independent Blackfrost research derivative of XiaomiMiMo/MiMo-V2.6-Flash-RL. The checkpoint retains the upstream MiMo V2.6 architecture, tokenizer, multimodal components, native chat template, and speculative-decoding assets while preserving its mixed-precision MXFP4 deployment layout.

The behavioral modification is encoded in the released weights. It is not a system-prompt wrapper, adapter, or decoding-time filter. Direction data, capture data, intermediate checkpoints, internal evaluation prompts, and the reproduction recipe are not included.

Model specifications

Property Value
Architecture MiMoV2ForCausalLM / mimo_v2
Parameters 309B total / 15B activated, per the upstream card
Language-model layers 48
Hidden size 4,096
Routed experts 256 total / 8 activated per token
Modalities Text, image, video, and audio
Upstream maximum context 1,048,576 tokens
Current Blackfrost serving target 262,144 tokens
Speculative decoder Five-layer MiMo MTP/DFlash assets retained
Main indexed tensor bytes 172,923,364,096
Expert deployment format Mixed-precision MXFP4 / ModelOpt layout

The upstream architectural context limit is not a guarantee that every runtime or hardware configuration can allocate or serve the full window.

Lineage

XiaomiMiMo/MiMo-V2.6-Flash-RL
└── Blackfrost-AI/MIMO-V2.6-DERISKED-MXFP4

Tokenizer, architecture, multimodal, MTP/DFlash, and license lineage follow the upstream checkpoint. The language-model checkpoint contains Blackfrost weight-level behavioral modifications.

Prompting and chat template

The package uses the native MiMo chat template and supports role-structured messages, including a caller-supplied system role. No Blackfrost, Frosty, compliance, or deployment-specific system prompt is baked into this release.

Runtime reasoning and tool-call parsers must be compatible with MiMo. Applications remain responsible for tool execution, result reinjection, authentication, authorization, conversation state, and output handling.

Validation status

The release package has completed structural checks covering the indexed shard inventory, selected modified tensors, tensor shapes and dtypes, finite values, preserved MXFP4 expert shards, native configuration files, and tokenizer/template assets.

Refusal behavior, broad capability retention, coding, cybersecurity, multimodal operation, tool calling, speculative acceptance, and long-context behavior remain pending for the exact uploaded artifact. Upstream benchmark results must not be attributed to this modified checkpoint.

Access

This repository is public and manually gated. Request access through the Hugging Face model page. Approval grants access to the released files; it does not imply suitability for a particular workload or transfer responsibility for deployment controls to Blackfrost.

Limitations and responsibility

  • This is an experimental research checkpoint.
  • Generated content may be inaccurate, insecure, offensive, or otherwise unsuitable.
  • The model is not a security boundary, policy engine, authorization mechanism, or substitute for professional judgment.
  • Treat generated text, code, URLs, tool arguments, file paths, and commands as untrusted until independently reviewed.
  • Operators are responsible for access controls, monitoring, legal compliance, and safeguards appropriate to their environment.

License and attribution

The upstream repository identifies the checkpoint license as MIT. Review the upstream model card and license metadata before use or redistribution.

Blackfrost is independent of and is not affiliated with, sponsored by, or endorsed by Xiaomi. This research artifact is provided as-is, without warranties.

Contact

For reproducible artifact issues, use this repository's Discussions. Do not post credentials, private prompts, personal information, or infrastructure details.

Identity and Version

Repository
Blackfrost-AI/MIMO-V2.6-DERISKED-MXFP4
Publisher
Sir Frosty
Task
Image and text to text
Modality
Image and text
Library
transformers
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
9c809a3617067d9bb7e9fa2923305fe5a9bb1058
First published
2026-09-22
Last updated
2026-09-22

Files and Weights

3 files, 1.9 MB in total.

Documentation1 file · 5.4 KB
Other1 file · 1.9 MB
Repository1 file · 1.6 KB
Every file
FileTypeSizeSHA-256
README.mdDocumentation5.4 KB —
ASSETS/BLACKFROST-AI-BANNER.pngOther1.9 MB —
.gitattributesRepository1.6 KB —

License and Download

License
mit
Access
Access requested at publisher
Request access from Sir Frosty

Sir Frosty grants access through its official repository on Hugging Face. Read the license.

Built From

Questions About MIMO-V2.6-DERISKED-MXFP4

Can I use MIMO-V2.6-DERISKED-MXFP4 commercially?

Yes. MIMO-V2.6-DERISKED-MXFP4 is released under MIT License. The MIT License is a short permissive license. It permits commercial use, modification and redistribution, provided the copyright notice and permission notice are included.

Similar Models

Model · Image and text to text

Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF

Michał Piszczek

I built this quant because the ready-made FP4 file answered the wrong question. It was fast, but on my short WikiText-2 control it scored 6.4949 PPL. Plain Q40 scored 6.3798. The first higher-quality hybrid went too far the other way: good perplexity, 34.19 tok/s, and no comfortable room for 256K plus vision. This is the build that survived both gates. It is a 17.1 GB, 5.01 BPW mixed-precision GGUF of Qwen/Qwen3.8-27B. It keeps large, tolerant matrices in native NVFP4 and spends more bits on selected attention, Gated DeltaNet, and late FFN tensors. The trained MTP layer remains embedded in the same GGUF. This is not a fine-tune. I built the private calibration workload from 5,472 messages…

Open weights apache-2.0

Qwen3.8-27B uncensored by HauhauCS 0/465 Refusals. This is the Aggressive variant: direct answers, no refusal behavior, and minimal preamble on hard prompts. Every text GGUF preserves Qwen3.8's native NextN head, and this release adds HauhauCS FastMTP: a specific acceleration sidecar qualified across the complete quant lineup at maximum native context. Vision is included through the separate BF16 projector. No changes to datasets or intended capabilities. This release preserves Qwen3.8-27B's text, reasoning, agentic, image, and video capabilities while applying the HauhauCS Aggressive uncensoring profile. Pick Aggressive when you specifically want the model to get to the answer without…

Open weights apache-2.0

Model · Image and text to text

Huihui-Qwen3.8-27B-abliterated-GGUF

Huihui.ai

This is an uncensored version of Qwen/Qwen3.8-27B created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens. The newly added Huihui-Qwen3.8-27B-abliterated-Ternary series come from prism-ml/Ternary-Bonsai-2-27B-gguf have been ablated, while the other layers remain unablated. It may come with a small disclaimer warning. The size after conversion may differ from the original GGUF (Some of the weights are converted from PTQ1 to Q2K or Q3K.). This is just a test/validation. The ternary hybrid-attention kernels live in the PrismML-Eng/llama.cpp fork.…

Open weights apache-2.0 transformers

and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads) instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model ctrl at the chat/message level). Model name has "plusIQ" in the name. (there is also a extra robust "tools" version too.) A 12+12 (12 reasoning and 12 instruct) model with interactive optimization/help system will be releasing shortly too. Extreme intelligence in a small package. Jaw dropping performance. Superior instruction following. A multi-stage and multi-model…

Open weights apache-2.0

in 8 bit and over 718 arc-c in 4 bit. This version is called TURBO because it drastically reduces thinking tokens (by 1/2 to as high as 1/10), yet maintains output detail and quality. In otherwords while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more. This repo contains both "regular" and "MTP" Neo-CODER MAX DI-MATRIX (duel imatrix) GGUF quants. and other quant versions (also see "Quantized" in the "model tree" too (lower right)). The strongest, smartest open source multi-stage model fine tune for consumer hardware ever and BUILT on consumer hardware via Unsloth. The first model of this size/type to breach "730" ARC-C…

Open weights apache-2.0

Non-uniform GGUF quantizations produced with GSQ and RCO, with a vision projector for multimodal use. This repository provides GGUF quantizations of Qwen3.8-27B at four sizes, together with the model's vision projector (mmproj) for multimodal use. In contrast to uniform quantization, which applies a single quantization type to all weight tensors, each model here assigns a separate quantization type to every tensor. The assignment is obtained by a gradient-based search that allocates precision according to per-tensor sensitivity, subject to a total size budget. The resulting files are standard GGUF and run unmodified in llama.cpp, Ollama, and LM Studio. Both methods were developed at the…

Open weights apache-2.0 gguf