SAVRN
Search Contact SAVRN

SAVRN Model Hub · Models by Task

Any to Any Models

61 open-weight any to any models in the SAVRN Model Hub, with Google, LM Studio Community and Qwen publishing the most.

61Models
15Publishers
79M to 35.3BParameter range
4Licenses

SAVRN's Take

The label on this page covers models that accept more than one kind of input. Mostly that means Google's Gemma 4 line, which takes text and images, adds audio on the E2B, E4B and 12B sizes, and answers in text, with a context of up to 256K tokens and more than 140 languages. Qwen3-Omni-30B-A3B-Instruct goes further, reading text, images, audio and video and streaming back both text and speech. Sixty-one models sit here, 20 with Index pricing.

At the bottom, gemma-4-E4B-it-assistant at 79M parameters wants 0.2 GB at 16-bit. The most downloaded, gemma-4-E4B-it at 8B parameters and 4,473,231 pulls a month, needs 19.2 GB at 16-bit and 4.8 GB at 4-bit. The 12B, with a 262,144 token context where the E-series stops at 131,072, climbs to 28.7 GB. Qwen3-Omni tops the range at 35.3B parameters and 84.6 GB at 16-bit, or 21.2 GB at 4-bit. Every indexed entry lands on a single MI300X, priced on the Index at $1.85 an hour, so the hardware question is how much of one card you leave for the KV cache when the context runs long.

Google publishes 16 of the 61 and LM Studio Community 12; the top GGUF and MLX repacks from GGML Org, LM Studio and Unsloth AI each clear a million downloads a month. Forty-eight carry Apache 2.0, publisher weights and repacks alike. Ten are marked other, Qwen3-Omni among them, so read those terms first. Two state no license; we would not deploy those. Decide on audio in, speech out and context length before you commit.

Most Downloaded

ModelPublisherParametersLicenseMonthly downloadsCheapest GPUs at 16-bit
gemma-4-E4B-it Google 8B apache-2.0 4.5M 1x MI300X, $1.85/hr
gemma-4-E2B-it Google 5.1B apache-2.0 3.5M 1x MI300X, $1.85/hr
gemma-4-12B-it Google 12B apache-2.0 2.7M 1x MI300X, $1.85/hr
gemma-4-E4B-it-GGUF GGML Org apache-2.0 1.5M
gemma-4-E4B-it-MLX-4bit LM Studio Community 8B apache-2.0 1.2M 1x MI300X, $1.85/hr
gemma-4-12B-it-qat-GGUF Unsloth AI apache-2.0 1.2M
gemma-4-E4B-it-MLX-8bit LM Studio Community 8B apache-2.0 1.1M 1x MI300X, $1.85/hr
gemma-4-E4B-it-MLX-5bit LM Studio Community 8B apache-2.0 1.1M 1x MI300X, $1.85/hr
gemma-4-E4B-it-MLX-6bit LM Studio Community 8B apache-2.0 1.1M 1x MI300X, $1.85/hr
gemma-4-12B-it-qat-w4a16-ct Google 13.3B apache-2.0 863.7k 1x MI300X, $1.85/hr

Licenses

LicenseModelsCommercial use
apache-2.048Yes
other10Read the license
not stated2Not stated
mit1Yes

Who Publishes Them

PublisherModels
Google16
LM Studio Community12
Qwen5
Unsloth AI4
OpenBMB4
Bartowski4

All 61 Models, Page 2 of 2

Using llama.cpp release b10068 for quantization. All quants made using imatrix option with dataset from here Run them in your choice of tools: Note: if it's a newly supported model, you may need to wait for an update from the developers. Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings and output weights quantized to Q80 instead of what they would normally default to. First, make sure you have huggingface-cli installed: Then, you can target the specific file you want: If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run: You can either specify a new local-dir…

Open weights apache-2.0

Questions

Which Any to any models are most downloaded?

By monthly downloads reported by the Hugging Face Hub: MirilAI_Miril-Drone-2B-1-GGUF (2.7k).

Other Tasks

See all