SAVRN Model Hub · Comparisons
Qwen3.6-27B vs Qwen3.8-27B
Qwen3.6-27B has 27.8B parameters and Qwen3.8-27B has 27.8B parameters; both are released under Apache License 2.0; at 16-bit, Qwen3.6-27B needs about 66.7 GB (1x MI300X from $1.85 an hour) and Qwen3.8-27B about 66.7 GB (1x MI300X from $1.85 an hour).
| Field | Qwen3.6-27B Qwen/Qwen3.6-27B | Qwen3.8-27B Qwen/Qwen3.8-27B |
|---|---|---|
| Publisher | Qwen | Qwen |
| Task | Image and text to text | Image and text to text |
| Modality | Image and text | Image and text |
| Parameters, as reported | 27.8B parameters | 27.8B parameters |
| Architecture | Qwen3_5ForConditionalGeneration | Qwen3_5ForConditionalGeneration |
| Library | transformers | transformers |
| Context length | 262,144 tokens | 262,144 tokens |
| Repository size | 55.6 GB | 55.6 GB |
| Artifact formats | safetensors | safetensors |
| License | apache-2.0 | apache-2.0 |
| Access | Open weights, no gate | Open weights, no gate |
| Memory at 16-bit (weights and margin) | 66.7 GB | 66.7 GB |
| Cheapest GPUs at 16-bit, per hour | 1x MI300X, $1.85 | 1x MI300X, $1.85 |
| Memory at 4-bit (weights and margin) | 16.7 GB | 16.7 GB |
| Cheapest GPUs at 4-bit, per hour | 1x MI300X, $1.85 | 1x MI300X, $1.85 |
| Revision viewed | 6a9e13bd6fc8 | 1d4bf0f2ff60 |
| Downloads reported by the hub | 3.7M | 7.4M |
| Last observed | 2026-09-18 | 2026-09-18 |
An evaluation row appears only where at least two of these models report the same benchmark with the same stated configuration, metric, unit and setup. Different evaluators stay named in each cell. Values are shown as reported: no unit conversion, no ranking.
Other Reported Results
These results are listed for each model on its own, because the conditions needed to compare them are not stated or do not match. Two results that leave a condition blank are not assumed to share it.
Qwen3.6-27B
| Benchmark | Conditions | Result | Reported by | Revision | Date |
|---|---|---|---|---|---|
| Idavidrein/gpqa | Task diamondMetric diamondComparison conditions not established | 87.8 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| MMMU/MMMU_Pro | Task mmmu_pro_visionMetric mmmu_pro_visionComparison conditions not established | 75.8 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-05-15 |
| MathArena/aime_2026 | Task MathArena/aime_2026Metric MathArena/aime_2026Comparison conditions not established | 94.1 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| MathArena/hmmt_feb_2026 | Task MathArena/hmmt_feb_2026Metric MathArena/hmmt_feb_2026Comparison conditions not established | 84.3 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| SWE-bench/SWE-bench_Multilingual | Task swe_bench_multilingual_%_resolvedMetric swe_bench_multilingual_%_resolvedComparison conditions not established | 71.3 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-08-10 |
| SWE-bench/SWE-bench_Verified | Task swe_bench_%_resolvedMetric swe_bench_%_resolvedComparison conditions not established | 77.2 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| ScaleAI/SWE-bench_Pro | Task SWE_Bench_ProMetric SWE_Bench_ProComparison conditions not established | 53.5 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| TIGER-Lab/MMLU-Pro | Task mmlu_proMetric mmlu_proComparison conditions not established | 86.2 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| benchflow/skillsbench | Task skillsbench_v1_1Metric skillsbench_v1_1Setup SkillsBench Avg5. Evaluated via OpenCode on 78 tasks (self-contained subset, excluding API-dependent tasks); avg of 5 runs. Mapped to the Hub's benchflow/skillsbench task skillsbench_v1_1.Comparison conditions not established | 48.2 | Qwen3.6-27B model card — Benchmark Results (Language > Coding Agent) Reported by a third party |
Evaluated revision not stated | 2026-04-21 |
| cais/hle | Task hleMetric hleComparison conditions not established | 24 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-04-22 |
| harborframework/terminal-bench-2.0 | Task terminalbench_2Metric terminalbench_2Setup Harbor/Terminus-2 harness; 3h timeout, 32 CPU/48 GB RAM; temp=1.0, top_p=0.95, top_k=20, max_tokens=80K, 256K ctx; avg of 5 runs.Comparison conditions not established | 59.3 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-06-23 |
| internlm/WildClawBench | Task avg_costMetric avg_costComparison conditions not established | 20.91 | WildClawBench Reported by a third party |
Evaluated revision not stated | 2026-08-11 |
| internlm/WildClawBench | Task avg_timeMetric avg_timeComparison conditions not established | 421 | WildClawBench Reported by a third party |
Evaluated revision not stated | 2026-08-11 |
| internlm/WildClawBench | Task overallMetric overallComparison conditions not established | 43.2 | WildClawBench Reported by a third party |
Evaluated revision not stated | 2026-08-11 |
Qwen3.8-27B
| Benchmark | Conditions | Result | Reported by | Revision | Date |
|---|---|---|---|---|---|
| Idavidrein/gpqa | Task diamondMetric diamondComparison conditions not established | 89.2 | Qwen3.8-27B model card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| ScaleAI/SWE-bench_Pro | Task SWE_Bench_ProMetric SWE_Bench_ProSetup Evaluated with the Claude Code harness, temp=1.0, top_p=0.95, 256K context; baseline models re-evaluated on the same refined task set.Comparison conditions not established | 61.7 | Qwen3.8-27B model card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| cais/hle | Task hleMetric hleSetup Judged by GPT-4o.Comparison conditions not established | 30.8 | Qwen3.8-27B model card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| claw-eval/Claw-Eval | Task multimodalMetric multimodalSetup Reported as ClawEval-MM. Pass@3, the benchmark's own Pass³ methodology; card also reports a secondary 56.9 'Average' metric, not included here.Comparison conditions not established | 57.4 | Qwen3.8-27B model card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| datacurve/deep-swe | Task deep_sweMetric deep_sweComparison conditions not established | 42.2 | Qwen3.8-27B model card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| harborframework/terminal-bench-2.1 | Task terminalbench_2_1Metric terminalbench_2_1Setup Row labeled "(Terminus)" as the harness; no further hyperparameter footnote given for this row.Comparison conditions not established | 73 | Model Card Reported by a third party |
Evaluated revision not stated | 2026-08-14 |
| internlm/WildClawBench | Task avg_timeMetric avg_timeComparison conditions not established | 516 | WildClawBench Reported by a third party |
Evaluated revision not stated | 2026-08-16 |
| internlm/WildClawBench | Task overallMetric overallComparison conditions not established | 48.0152 | WildClawBench Reported by a third party |
Evaluated revision not stated | 2026-08-16 |
| llamaindex/ExtractBench | Task longMetric longSetup Pipeline name: qwen3_8_27b_fp8_vllm_extract_oneshot_structured_output_file; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 38.45 | ExtractBench Reported by a third party |
Evaluated revision not stated | 2026-08-24 |
| llamaindex/ExtractBench | Task meanMetric meanSetup Pipeline name: qwen3_8_27b_fp8_vllm_extract_oneshot_structured_output_file; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 89.75 | ExtractBench Reported by a third party |
Evaluated revision not stated | 2026-08-24 |
| llamaindex/ExtractBench | Task mediumMetric mediumSetup Pipeline name: qwen3_8_27b_fp8_vllm_extract_oneshot_structured_output_file; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 87.54 | ExtractBench Reported by a third party |
Evaluated revision not stated | 2026-08-24 |
| llamaindex/ExtractBench | Task shortMetric shortSetup Pipeline name: qwen3_8_27b_fp8_vllm_extract_oneshot_structured_output_file; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 94.68 | ExtractBench Reported by a third party |
Evaluated revision not stated | 2026-08-24 |
| llamaindex/ParseBench | Task chartMetric chartSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 69.17 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
| llamaindex/ParseBench | Task layoutMetric layoutSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 69.9 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
| llamaindex/ParseBench | Task meanMetric meanSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 70.79 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
| llamaindex/ParseBench | Task tableMetric tableSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 66.82 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
| llamaindex/ParseBench | Task text_contentMetric text_contentSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 88.28 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
| llamaindex/ParseBench | Task text_formattingMetric text_formattingSetup Pipeline name: qwen3_8_27b_thinking_parse_with_layout; served checkpoint: Qwen/Qwen3.8-27B-FP8Comparison conditions not established | 59.77 | ParseBench Reported by a third party |
Evaluated revision not stated | 2026-08-28 |
SAVRN's Notes on Qwen3.6-27B
Two hosts on the SAVRN Index sell this model by the token, so renting and owning can be priced side by side. With 27.8B parameters, images and text in, text out, it needs 66.7 GB at 16-bit, which fits one 192 GB MI300X, the cheapest listed setup at $1.85 an hour on-demand. Eight-bit drops the need to 33.3 GB and 4-bit to 16.7 GB, so one card carries several copies.
Renting: DeepInfra lists $0.32 input and $3.20 output per million tokens, OVHcloud $0.47 and $3.19. Price your monthly token volume both ways before choosing. Apache 2.0 permits commercial use, modification and redistribution, notices kept and changes stated, so a fine-tune is yours to ship. Verify the 262,144-token context, and note the config's architecture is Qwen3_5ForConditionalGeneration, model type qwen3_5, despite the 3.6 name, so confirm your serving stack loads it. No base model or paper is on record.
SAVRN's Notes on Qwen3.8-27B
Plan around 66.7 GB for Qwen3.8-27B at 16-bit. That fits one MI300X with 192 GB, which the SAVRN Index prices at $1.85 an hour on demand, and 8-bit brings it to 33.3 GB, 4-bit to 16.7 GB, all on that one card. It reads images and text and writes text, with a 262,144-token context, so it belongs where one accelerator per instance is the budget and the inputs include pictures or long documents.
Apache 2.0 allows commercial use, modification and redistribution, provided the license and notice files stay with the weights and significant changes are stated. Weigh the card against renting tokens on the Index: Cerebras at $0.99 in and $1.49 out per million, DeepInfra at $0.20 and $2.50, Novita at $0.42 and $3.00, OVHcloud at $0.47 and $3.19. The benchmark figures on the page are reported by the model card and outside evaluators, not measured by SAVRN.
Questions
Which is larger, Qwen3.6-27B or Qwen3.8-27B?
Qwen3.6-27B (27.8B parameters) is larger than Qwen3.8-27B (27.8B parameters), by the parameter counts their publishers report.
Which is cheaper to run, Qwen3.6-27B or Qwen3.8-27B?
At 4-bit, Qwen3.6-27B fits on 1x MI300X from $1.85 an hour and Qwen3.8-27B on 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.
Can I use Qwen3.6-27B commercially?
Yes. Qwen3.6-27B is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.
Can I use Qwen3.8-27B commercially?
Yes. Qwen3.8-27B is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.