SAVRN Model Hub · Comparisons
Llama-3.1-8B-Instruct vs Meta-Llama-3-8B-Instruct
Llama-3.1-8B-Instruct has 8B parameters and Meta-Llama-3-8B-Instruct has 8B parameters; Llama-3.1-8B-Instruct is released under Meta Llama 3.1 Community License and Meta-Llama-3-8B-Instruct under Meta Llama 3 Community License; at 16-bit, Llama-3.1-8B-Instruct needs about 19.3 GB (1x MI300X from $1.85 an hour) and Meta-Llama-3-8B-Instruct about 19.3 GB (1x MI300X from $1.85 an hour).
| Field | Llama-3.1-8B-Instruct meta-llama/Llama-3.1-8B-Instruct | Meta-Llama-3-8B-Instruct meta-llama/Meta-Llama-3-8B-Instruct |
|---|---|---|
| Publisher | Meta Llama | Meta Llama |
| Task | Text generation | Text generation |
| Modality | Text | Text |
| Parameters, as reported | 8B parameters | 8B parameters |
| Architecture | LlamaForCausalLM | LlamaForCausalLM |
| Library | transformers | transformers |
| Context length | Not stated | Not stated |
| Repository size | 32.1 GB | 32.1 GB |
| Artifact formats | safetensors, pytorch | safetensors, pytorch |
| License | llama3.1 | llama3 |
| Access | Access requested at publisher | Access requested at publisher |
| Memory at 16-bit (weights and margin) | 19.3 GB | 19.3 GB |
| Cheapest GPUs at 16-bit, per hour | 1x MI300X, $1.85 | 1x MI300X, $1.85 |
| Memory at 4-bit (weights and margin) | 4.8 GB | 4.8 GB |
| Cheapest GPUs at 4-bit, per hour | 1x MI300X, $1.85 | 1x MI300X, $1.85 |
| Revision viewed | 0e9e39f249a1 | 8afb486c1db2 |
| Downloads reported by the hub | 5.9M | 1.2M |
| Last observed | 2026-09-18 | 2026-09-18 |
An evaluation row appears only where at least two of these models report the same benchmark with the same stated configuration, metric, unit and setup. Different evaluators stay named in each cell. Values are shown as reported: no unit conversion, no ranking.
Other Reported Results
These results are listed for each model on its own, because the conditions needed to compare them are not stated or do not match. Two results that leave a condition blank are not assumed to share it.
Llama-3.1-8B-Instruct
| Benchmark | Conditions | Result | Reported by | Revision | Date |
|---|---|---|---|---|---|
| Idavidrein/gpqa | Task diamondMetric diamondComparison conditions not established | 30.4 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-01-27 |
| LEXam-Benchmark/LEXam | Task mcq_4_choicesMetric mcq_4_choicesComparison conditions not established | 24.04 | LEXam Leaderboard Reported by a third party |
Evaluated revision not stated | 2026-06-02 |
| LEXam-Benchmark/LEXam | Task open_questionMetric open_questionComparison conditions not established | 10 | LEXam Leaderboard Reported by a third party |
Evaluated revision not stated | 2026-06-02 |
| openai/gsm8k | Task gsm8kMetric gsm8kComparison conditions not established | 84.5 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-03-23 |
| thamilvendhan/signalbench | Task access_denyMetric access_denySetup family=access_deny; n=12Comparison conditions not established | 0.5 | thamilvendhan Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
| thamilvendhan/signalbench | Task bot_policyMetric bot_policySetup family=bot_policy; n=12Comparison conditions not established | 0.4167 | thamilvendhan Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
| thamilvendhan/signalbench | Task injectionMetric injectionSetup family=injection; n=12Comparison conditions not established | 0.75 | thamilvendhan Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
| thamilvendhan/signalbench | Task memory_labelMetric memory_labelSetup family=memory_label; n=12Comparison conditions not established | 0.6667 | thamilvendhan Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
| thamilvendhan/signalbench | Task srcMetric srcSetup SRC overall; deterministic action-based grader, no LLM judge; seed 0, n=75Comparison conditions not established | 0.6333 | signalbench raw per-item responses Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
| thamilvendhan/signalbench | Task timeMetric timeSetup family=time; n=12Comparison conditions not established | 0.8333 | thamilvendhan Reported by a third party |
Evaluated revision not stated | 2026-07-08 |
Meta-Llama-3-8B-Instruct
| Benchmark | Conditions | Result | Reported by | Revision | Date |
|---|---|---|---|---|---|
| TIGER-Lab/MMLU-Pro | Task mmlu_proMetric mmlu_proComparison conditions not established | 40.98 | EvalEval Reported by a third party |
Evaluated revision not stated | 2026-06-30 |
SAVRN's Notes on Llama-3.1-8B-Instruct
At 16-bit this checkpoint needs 19.3 GB of memory, and the cheapest SAVRN Index listing that covers it is a single 192 GB MI300X at $1.85 per hour on-demand. That card is the cheapest answer at 8-bit (9.6 GB) and 4-bit (4.8 GB) too, so quantizing does not buy a cheaper hour; it buys room for more concurrent multilingual dialogue sessions on one card.
Commercial use is allowed under the Meta Llama 3.1 Community License, with attribution and Meta's Acceptable Use Policy observed, unless your products had more than 700 million monthly active users on the release date; then you request a license from Meta. Gated access means approval precedes download. Two checks: the page lists no context length, and the $1.85 hour must beat the Index's hosted rates, $0.02 in and $0.05 out per million tokens at DeepInfra and Novita, $0.06 both ways at Nscale.
SAVRN's Notes on Meta-Llama-3-8B-Instruct
Nineteen point three gigabytes is the number to plan around at 16-bit, 16.1 of that being weights. At 8-bit the need falls to 9.6 gigabytes; at 4-bit it is 4.8, small enough that the 192 gigabyte MI300X the Index lists as cheapest, at $1.85 an hour, is far more card than one copy needs. That headroom is the point: room to serve many sessions at once.
The Llama 3 Community License permits commercial use with conditions: attribution as Meta specifies, compliance with Meta's Acceptable Use Policy, and a separate license request for any licensee with more than 700 million monthly active users on the release date. Access is gated, so the files arrive only after the publisher approves you. Check the context length, which our record does not carry, and the age of the weights, released April 17, 2024 and last updated June 18, 2025.
Questions
Which is larger, Llama-3.1-8B-Instruct or Meta-Llama-3-8B-Instruct?
Llama-3.1-8B-Instruct (8B parameters) is larger than Meta-Llama-3-8B-Instruct (8B parameters), by the parameter counts their publishers report.
Which is cheaper to run, Llama-3.1-8B-Instruct or Meta-Llama-3-8B-Instruct?
At 4-bit, Llama-3.1-8B-Instruct fits on 1x MI300X from $1.85 an hour and Meta-Llama-3-8B-Instruct on 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.
Can I use Llama-3.1-8B-Instruct commercially?
Yes, with conditions. Llama-3.1-8B-Instruct is released under Meta Llama 3.1 Community License. The Llama 3.1 Community License permits commercial use, except that a licensee whose products had more than 700 million monthly active users on the release date must request a license from Meta. It requires attribution as the license specifies and compliance with Meta's Acceptable Use Policy.
Can I use Meta-Llama-3-8B-Instruct commercially?
Yes, with conditions. Meta-Llama-3-8B-Instruct is released under Meta Llama 3 Community License. The Llama 3 Community License permits commercial use, except that a licensee whose products had more than 700 million monthly active users on the release date must request a license from Meta. It requires attribution as the license specifies and compliance with Meta's Acceptable Use Policy.