SAVRN
Search Contact SAVRN

SAVRN Model Hub · Comparisons

Qwen3-ASR-1.7B vs seamless-m4t-v2-large

Qwen3-ASR-1.7B has 2.3B parameters and seamless-m4t-v2-large has 2.3B parameters; Qwen3-ASR-1.7B is released under Apache License 2.0 and seamless-m4t-v2-large under Creative Commons Attribution-NonCommercial 4.0; at 16-bit, Qwen3-ASR-1.7B needs about 5.6 GB (1x MI300X from $1.85 an hour) and seamless-m4t-v2-large about 5.5 GB (1x MI300X from $1.85 an hour).

Published metadata for 2 models, each read from its own repository.
Field Qwen3-ASR-1.7B
Qwen/Qwen3-ASR-1.7B
seamless-m4t-v2-large
facebook/seamless-m4t-v2-large
Publisher Qwen AI at Meta
Task Speech recognition Speech recognition
Modality Audio Audio
Parameters, as reported 2.3B parameters 2.3B parameters
Architecture Qwen3ASRForConditionalGeneration SeamlessM4Tv2Model
Library Not stated transformers
Context length Not stated 4,096 tokens
Repository size 4.7 GB 29.9 GB
Artifact formats safetensors safetensors
License apache-2.0 cc-by-nc-4.0
Access Open weights, no gate Open weights, no gate
Memory at 16-bit (weights and margin) 5.6 GB 5.5 GB
Cheapest GPUs at 16-bit, per hour 1x MI300X, $1.85 1x MI300X, $1.85
Memory at 4-bit (weights and margin) 1.4 GB 1.4 GB
Cheapest GPUs at 4-bit, per hour 1x MI300X, $1.85 1x MI300X, $1.85
Revision viewed 7278e1e70fe2 5f8cc790b19f
Downloads reported by the hub 2.4M 306.7k
Last observed 2026-09-18 2026-09-18

An evaluation row appears only where at least two of these models report the same benchmark with the same stated configuration, metric, unit and setup. Different evaluators stay named in each cell. Values are shown as reported: no unit conversion, no ranking.

Other Reported Results

These results are listed for each model on its own, because the conditions needed to compare them are not stated or do not match. Two results that leave a condition blank are not assumed to share it.

Qwen3-ASR-1.7B

BenchmarkConditionsResultReported byRevisionDate
hf-audio/open-asr-leaderboard Task ami_werMetric ami_werComparison conditions not established 10.56 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task earnings22_werMetric earnings22_werComparison conditions not established 10.25 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task gigaspeech_werMetric gigaspeech_werComparison conditions not established 8.74 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task librispeech_clean_werMetric librispeech_clean_werComparison conditions not established 1.63 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task librispeech_other_werMetric librispeech_other_werComparison conditions not established 3.4 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task mean_werMetric mean_werComparison conditions not established 5.76 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task rtfxMetric rtfxComparison conditions not established 147.93 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task spgispeech_werMetric spgispeech_werComparison conditions not established 2.84 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task tedlium_werMetric tedlium_werComparison conditions not established 2.28 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28
hf-audio/open-asr-leaderboard Task voxpopuli_werMetric voxpopuli_werComparison conditions not established 6.35 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2026-01-28

SAVRN's Notes on Qwen3-ASR-1.7B

The name says 1.7B, the weight files hold 2.3B parameters, so size by the files. It turns speech into text and identifies which of 52 languages and dialects it hears; a 0.6B sibling shares its Qwen3-Omni foundation, and a separate 0.6B forced aligner handles timestamps. At 16-bit the weights are 4.7 GB and the run needs 5.6 GB, which leaves most of a 192 GB MI300X at $1.85 an hour empty, so scale by instances per GPU, not by GPUs.

Apache 2.0 permits commercial use, modification and redistribution provided the notices stay and changes are stated. Before you commit: no context length is published, no library is named, and the only artifact is safetensors, so 8-bit at 2.8 GB or 4-bit at 1.4 GB is your own quantization work. Released January 28, 2026 and described in arXiv:2601.21337; read the paper before planning around the 52 languages.

Questions

Which is larger, Qwen3-ASR-1.7B or seamless-m4t-v2-large?

Qwen3-ASR-1.7B (2.3B parameters) is larger than seamless-m4t-v2-large (2.3B parameters), by the parameter counts their publishers report.

Which is cheaper to run, Qwen3-ASR-1.7B or seamless-m4t-v2-large?

At 4-bit, Qwen3-ASR-1.7B fits on 1x MI300X from $1.85 an hour and seamless-m4t-v2-large on 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use Qwen3-ASR-1.7B commercially?

Yes. Qwen3-ASR-1.7B is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Can I use seamless-m4t-v2-large commercially?

Not without separate permission. seamless-m4t-v2-large is released under Creative Commons Attribution-NonCommercial 4.0. CC BY-NC 4.0 permits sharing and adapting with credit for non-commercial purposes only. Commercial use needs separate permission from the rights holder.

Related Comparisons