SAVRN
Search Contact SAVRN

SAVRN Model Hub · Comparisons

parakeet-ctc-1.1b vs whisper-large-v3

Parakeet-ctc-1.1b has 1.1B parameters and whisper-large-v3 has 1.5B parameters; parakeet-ctc-1.1b is released under Creative Commons Attribution 4.0 and whisper-large-v3 under Apache License 2.0; at 16-bit, parakeet-ctc-1.1b needs about 2.6 GB (1x MI300X from $1.85 an hour) and whisper-large-v3 about 3.7 GB (1x MI300X from $1.85 an hour).

Published metadata for 2 models, each read from its own repository.
Field parakeet-ctc-1.1b
nvidia/parakeet-ctc-1.1b
whisper-large-v3
openai/whisper-large-v3
Publisher NVIDIA OpenAI
Task Speech recognition Speech recognition
Modality Audio Audio
Parameters, as reported 1.1B parameters 1.5B parameters
Architecture ParakeetForCTC WhisperForConditionalGeneration
Library nemo transformers
Context length Not stated Not stated
Repository size 9.7 GB 24.7 GB
Artifact formats safetensors, gguf, pytorch safetensors, pytorch, jax
License cc-by-4.0 apache-2.0
Access Open weights, no gate Open weights, no gate
Memory at 16-bit (weights and margin) 2.6 GB 3.7 GB
Cheapest GPUs at 16-bit, per hour 1x MI300X, $1.85 1x MI300X, $1.85
Memory at 4-bit (weights and margin) 0.6 GB 0.9 GB
Cheapest GPUs at 4-bit, per hour 1x MI300X, $1.85 1x MI300X, $1.85
Revision viewed 20e63a0fed6a 06f233fe06e7
Downloads reported by the hub 796.2k 4.8M
Last observed 2026-09-18 2026-09-18

An evaluation row appears only where at least two of these models report the same benchmark with the same stated configuration, metric, unit and setup. Different evaluators stay named in each cell. Values are shown as reported: no unit conversion, no ranking.

Other Reported Results

These results are listed for each model on its own, because the conditions needed to compare them are not stated or do not match. Two results that leave a condition blank are not assumed to share it.

parakeet-ctc-1.1b

BenchmarkConditionsResultReported byRevisionDate
AMI (Meetings test) Configuration ihmTask Automatic Speech RecognitionMetric Test WERComparison conditions not established 15.62 nvidia
Publisher reported
Evaluated revision not stated
Earnings-22 Task Automatic Speech RecognitionMetric Test WERComparison conditions not established 13.69 nvidia
Publisher reported
Evaluated revision not stated
GigaSpeech Task Automatic Speech RecognitionMetric Test WERComparison conditions not established 10.27 nvidia
Publisher reported
Evaluated revision not stated
LibriSpeech (clean) Configuration otherTask Automatic Speech RecognitionMetric Test WERComparison conditions not established 1.83 nvidia
Publisher reported
Evaluated revision not stated
LibriSpeech (other) Configuration otherTask Automatic Speech RecognitionMetric Test WERComparison conditions not established 3.54 nvidia
Publisher reported
Evaluated revision not stated
Mozilla Common Voice 9.0 Configuration enTask automatic-speech-recognitionMetric Test WERComparison conditions not established 9.02 nvidia
Publisher reported
Evaluated revision not stated
SPGI Speech Configuration testTask automatic-speech-recognitionMetric Test WERComparison conditions not established 4.2 nvidia
Publisher reported
Evaluated revision not stated
Vox Populi Configuration enTask Automatic Speech RecognitionMetric Test WERComparison conditions not established 6.53 nvidia
Publisher reported
Evaluated revision not stated
hf-audio/open-asr-leaderboard Task ami_werMetric ami_werComparison conditions not established 15.67 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task earnings22_werMetric earnings22_werComparison conditions not established 13.75 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task gigaspeech_werMetric gigaspeech_werComparison conditions not established 10.28 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task librispeech_clean_werMetric librispeech_clean_werComparison conditions not established 1.83 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task librispeech_other_werMetric librispeech_other_werComparison conditions not established 3.51 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task mean_werMetric mean_werComparison conditions not established 7.39875 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task rtfxMetric rtfxComparison conditions not established 2728.52 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task spgispeech_werMetric spgispeech_werComparison conditions not established 4.02 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task tedlium_werMetric tedlium_werComparison conditions not established 3.57 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
hf-audio/open-asr-leaderboard Task voxpopuli_werMetric voxpopuli_werComparison conditions not established 6.56 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-12-28
tedlium-v3 Configuration release1Task automatic-speech-recognitionMetric Test WERComparison conditions not established 3.54 nvidia
Publisher reported
Evaluated revision not stated

whisper-large-v3

BenchmarkConditionsResultReported byRevisionDate
ARTPARK-IISc/Vaani-Benchmark-V1.0 Task Hindi_WERMetric Hindi_WERComparison conditions not established 26.8 Not named
Reported by a third party
Evaluated revision not stated 2026-06-26
hf-audio/open-asr-leaderboard Task ami_werMetric ami_werComparison conditions not established 15.95 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task earnings22_werMetric earnings22_werComparison conditions not established 11.29 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task gigaspeech_werMetric gigaspeech_werComparison conditions not established 10.02 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task librispeech_clean_werMetric librispeech_clean_werComparison conditions not established 2.01 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task librispeech_other_werMetric librispeech_other_werComparison conditions not established 3.91 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task mean_werMetric mean_werComparison conditions not established 7.44 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task rtfxMetric rtfxComparison conditions not established 145.51 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task spgispeech_werMetric spgispeech_werComparison conditions not established 2.94 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task tedlium_werMetric tedlium_werComparison conditions not established 3.86 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07
hf-audio/open-asr-leaderboard Task voxpopuli_werMetric voxpopuli_werComparison conditions not established 9.54 open-asr-leaderboard
Reported by a third party
Evaluated revision not stated 2023-11-07

SAVRN's Notes on parakeet-ctc-1.1b

We would give this one the smallest slice of a card we could carve out. NVIDIA's 1.1B-parameter speech recognizer, built with Suno.ai on FastConformer CTC and served through NeMo, needs 2.6 GB at 16-bit, 1.3 GB at 8-bit and 0.6 GB at 4-bit, and it writes lower-case English. On the cheapest host we price, one MI300X with 192 GB at $1.85 an hour on-demand, that footprint is a rounding error, so transcription belongs beside everything else on the card, not on a second one.

CC BY 4.0 lets you share and adapt it, commercially included, as long as NVIDIA is credited and your changes are indicated. Before committing, look at the training mix, librispeech_asr, fisher_corpus, Switchboard-1, WSJ and vctk among them, and at NVIDIA's own reported word error rates, 1.83 on LibriSpeech clean up to 15.62 on AMI meetings, then run your own audio through it. Released December 28, 2023, last updated August 5, 2026.

Questions

Which is larger, parakeet-ctc-1.1b or whisper-large-v3?

whisper-large-v3 (1.5B parameters) is larger than parakeet-ctc-1.1b (1.1B parameters), by the parameter counts their publishers report.

Which is cheaper to run, parakeet-ctc-1.1b or whisper-large-v3?

At 4-bit, parakeet-ctc-1.1b fits on 1x MI300X from $1.85 an hour and whisper-large-v3 on 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use parakeet-ctc-1.1b commercially?

Yes. parakeet-ctc-1.1b is released under Creative Commons Attribution 4.0. CC BY 4.0 permits sharing and adapting the work, including commercially, provided the creator is credited and changes are indicated.

Can I use whisper-large-v3 commercially?

Yes. whisper-large-v3 is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Related Comparisons