Cyber-F1-AWQ-smoke-prod is an open-weight model from Autumn. It has 35.1B parameters and a 262,144-token context. At 16-bit it needs about 84.3 GB of GPU memory, which fits on 1x MI300X from $1.85 an hour, at the lowest prices in the SAVRN Index.
Runs On
What it takes to serve Cyber-F1-AWQ-smoke-prod (35.1B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.
| Precision | Weights | Memory needed | Cheapest setup | Per hour | Also fits |
|---|---|---|---|---|---|
| 16-bit | 70.2 GB | 84.3 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x MI325X $2.00 · 1x MI355X $2.59 |
| 8-bit | 35.1 GB | 42.1 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
| 4-bit | 17.6 GB | 21.1 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 24, 2026.
Cyber-F1-AWQ-smoke-prod on every accelerator the SAVRN Index prices, at every precision
Model Card
The publisher has not written a card for this model.
Configuration
- Architecture
- Qwen3_5MoeForConditionalGeneration
- Context length (tokens)
- 262,144
- Layers
- 40
- Hidden size
- 2,048
- Attention heads
- 16
- Key/value heads
- 2
- Head dimension
- 256
- Vocabulary size
- 248,320
- Experts
- 256
- Experts active per token
- 8
- Model type
- qwen3_5_moe
- Quantization
- compressed-tensors
Identity and Version
- Repository
- autumn10/Cyber-F1-AWQ-smoke-prod
- Publisher
- Autumn
- Task
- Not stated by the source
- Modality
- Other
- Library
- Not stated by the source
- Parameters
- 35.1B parameters
- Languages
- Not stated by the source
- Revision
- cbd71f2fa691dfd4fc81202fd56d9484b1cf7e40
- First published
- 2026-09-21
- Last updated
- 2026-09-21
Files and Weights
8 files, 24.5 GB in total. The weights are 1 file totalling 24.4 GB in safetensors.
Every file
| File | Type | Size | SHA-256 |
|---|---|---|---|
| model.safetensors | Weights | 24.4 GB | d362b18b4d26 |
| config.json | Configuration | 602.9 KB | — |
| generation_config.json | Configuration | 142 B | — |
| recipe.yaml | Configuration | 28.9 KB | — |
| chat_template.jinja | Other | 7.8 KB | — |
| .gitattributes | Repository | 1.6 KB | — |
| tokenizer.json | Tokenizer | 20.0 MB | 06b9509352d2 |
| tokenizer_config.json | Tokenizer | 1.1 KB | — |
License and Download
- License
- Not stated by the source
- Access
- Open weights, no gate
- Download size
- 24.4 GB
Released by Autumn through its official repository on Hugging Face.
Memory Requirements
| Precision | Weights in memory |
|---|---|
| As published | 24.4 GB |
| 16-bit | 70.2 GB |
| 8-bit | 35.1 GB |
| 4-bit | 17.6 GB |
Weights only, from the published parameter count; the key-value cache and runtime add to this.
Questions About Cyber-F1-AWQ-smoke-prod
How much GPU memory does Cyber-F1-AWQ-smoke-prod need?
About 84.3 GB at 16-bit and 21.1 GB at 4-bit: the weights (35.1B parameters) plus a working margin. A long context needs more.
What is the cheapest GPU to run Cyber-F1-AWQ-smoke-prod on?
At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.
What is Cyber-F1-AWQ-smoke-prod's context length?
262,144 tokens, from the maximum position embeddings in its published configuration.