SAVRN
Search Contact SAVRN

Open-weight model · Any to any

RK182X-MLLM-Qwen2.5-Omni-3B

by RKNNAI RKNNAI/RK182X-MLLM-Qwen2.5-Omni-3B

RK182X-MLLM-Qwen2.5-Omni-3B is an open-weight model for any to any from RKNNAI, released under other. Its published files total 4.3 GB. It draws 1 downloads a month.

本仓库提供由 Qwen/Qwen2.5-Omni-3B 转换得到的 RKNN MLLM 模型。 - Model ID:RKNNAI/RK182X-MLLM-Qwen2.5-Omni-3B - 模型显示名称:RK182X-MLLM-Qwen2.5-Omni-3B - 源模型:Qwen/Qwen2.5-Omni-3B - 模型类型:MLLM ModelScope 完整下载: Hugging Face 完整下载: ModelScope 指定配置下载: Hugging Face 指定配置下载: - 使用配套 RKNN…

Parameters—
Context—
Weights628.3 MB
Licenseother
AccessOpen weights
Monthly Downloads1

Model Card

本仓库提供由 Qwen/Qwen2.5-Omni-3B 转换得到的 RKNN MLLM 模型。 - Model ID:RKNNAI/RK182X-MLLM-Qwen2.5-Omni-3B - 模型显示名称:RK182X-MLLM-Qwen2.5-Omni-3B - 源模型:Qwen/Qwen2.5-Omni-3B - 模型类型:MLLM ModelScope 完整下载: Hugging Face 完整下载: ModelScope 指定配置下载: Hugging Face 指定配置下载: - 使用配套 RKNN Runtime 和驱动;运行前用 rknn-smi -v 检查设备端版本。 源模型许可证见根目录 LICENSE;同时遵守 RKNN Toolkit 和 RKNN Runtime 许可条款。

Excerpt from the card by RKNNAI, licensed other.

Identity and Version

Repository
RKNNAI/RK182X-MLLM-Qwen2.5-Omni-3B
Publisher
RKNNAI
Task
Any to any
Modality
Multimodal
Library
Not stated by the source
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
8fd10c191bb4e5492a050a36a02ae10b9b6d3f0a
First published
2026-09-24
Last updated
2026-09-28

Files and Weights

16 files, 4.3 GB in total. The weights are 2 files totalling 628.3 MB in bin, gguf.

Weights2 files · 628.3 MB
Configuration1 file · 1.7 KB
Documentation3 files · 12.8 KB
Other9 files · 3.7 GB
Repository1 file · 2.3 KB
Every file
FileTypeSizeSHA-256
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-llm.embed.binWeights622.3 MB 1f534a216ded
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-llm.tokenizer.ggufWeights5.9 MB 3614b70e262d
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/config.jsonConfiguration1.7 KB —
LICENSEDocumentation7.4 KB —
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/README.mdDocumentation2.6 KB —
README.mdDocumentation2.8 KB —
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-audio.rknnOther10.9 MB 95f4185829a7
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-audio.weightOther1.3 GB 672c89889c79
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-llm.rknnOther27.4 MB 628b44d759d4
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-llm.weightOther1.9 GB 904bec1ad2cc
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-vision.rknnOther8.2 MB 2b4b6f034995
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/Qwen2.5-Omni-3B-vision.weightOther411.8 MB 7b225b0ea7bf
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/SHA256SUMSOther1.1 KB —
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/llm_model_report.htmlOther1.0 MB —
Qwen2.5-Omni-3B-392x392-600frames-w4a16-8-1024/vision_model_report.htmlOther85.0 KB —
.gitattributesRepository2.3 KB —

License and Download

License
other
Access
Open weights, no gate
Download size
628.3 MB
Download from RKNNAI

Released by RKNNAI through its official repository on Hugging Face.

Built From

Memory Requirements

PrecisionWeights in memory
As published628.3 MB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About RK182X-MLLM-Qwen2.5-Omni-3B

What license is RK182X-MLLM-Qwen2.5-Omni-3B released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Any to any

gemma-4-12B-it-qat-GGUF

Unsloth AI

This model ships a Multi-Token Prediction drafter at the repo root (mtp-gemma-4-12B-it.gguf, a near-lossless smart Q40). A recent llama.cpp auto-discovers it from -hf, so you do not pass --model-draft: The drafter shares the target's KV cache and does not change the output (the target verifies every drafted token). See the MTP/ folder for the other precisions and explicit usage. Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context…

Open weights apache-2.0 transformers

Model · Any to any

gemma-4-E4B-it-GGUF

GGML Org

Run with https://llama.app - https://huggingface.co/google/gemma-4-E4B-it - https://huggingface.co/google/gemma-4-E4B-it-assistant - https://huggingface.co/google/gemma-4-E4B-it-qat-q40-unquantized-assistant - https://huggingface.co/google/gemma-4-E4B-it-qat-q40-unquantized - add info - add dflash

Open weights apache-2.0

Model · Any to any

gemma-4-12B-it-qat-q4_0-gguf

Google

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages. Featuring both Dense and Mixture-of-Experts (MoE) architectures, Gemma 4 is well-suited for tasks like text generation, coding, and reasoning. The models are available in five distinct sizes: E2B, E4B, 12B, 26B A4B, and 31B. Their diverse sizes make them deployable in environments ranging from…

Open weights apache-2.0 transformers

Model · Any to any

gemma-4-E4B-it-qat-q4_0-gguf

Google

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages. Featuring both Dense and Mixture-of-Experts (MoE) architectures, Gemma 4 is well-suited for tasks like text generation, coding, and reasoning. The models are available in five distinct sizes: E2B, E4B, 12B, 26B A4B, and 31B. Their diverse sizes make them deployable in environments ranging from…

Open weights apache-2.0

Model · Any to any

gemma-4-E4B-it-qat-GGUF

Unsloth AI

This model ships a Multi-Token Prediction drafter at the repo root (mtp-gemma-4-E4B-it.gguf, a near-lossless smart Q40). A recent llama.cpp auto-discovers it from -hf, so you do not pass --model-draft: The drafter shares the target's KV cache and does not change the output (the target verifies every drafted token). See the MTP/ folder for the other precisions and explicit usage. Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context…

Open weights apache-2.0 131,072 tokens transformers

Model · Any to any

gemma-4-E2B-it-qat-GGUF

Unsloth AI

This model ships a Multi-Token Prediction drafter at the repo root (mtp-gemma-4-E2B-it.gguf, a near-lossless smart Q40). A recent llama.cpp auto-discovers it from -hf, so you do not pass --model-draft: The drafter shares the target's KV cache and does not change the output (the target verifies every drafted token). See the MTP/ folder for the other precisions and explicit usage. Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context…

Open weights apache-2.0 131,072 tokens transformers