Open-weight model
Shreyansh-STEM-AI-7B-Final
by Shreyansh singh shreyansh12183/Shreyansh-STEM-AI-7B-Final
Shreyansh-STEM-AI-7B-Final is an open-weight model from Shreyansh singh, released under Creative Commons Attribution-NonCommercial 4.0. It has 7.3B parameters and a 4,096-token context. At 16-bit it needs about 17.5 GB of GPU memory, which fits on 1x MI300X from $1.85 an hour, at the lowest prices in the SAVRN Index. It draws 267 downloads a month.
Runs On
What it takes to serve Shreyansh-STEM-AI-7B-Final (7.3B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.
| Precision | Weights | Memory needed | Cheapest setup | Per hour | Also fits |
|---|---|---|---|---|---|
| 16-bit | 14.6 GB | 17.5 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
| 8-bit | 7.3 GB | 8.8 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
| 4-bit | 3.6 GB | 4.4 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Oct 7, 2026.
Shreyansh-STEM-AI-7B-Final on every accelerator the SAVRN Index prices, at every precision
Model Card
The publisher has not written a card for this model.
Configuration
- Architecture
- Olmo2ForCausalLM
- Context length (tokens)
- 4,096
- Layers
- 32
- Hidden size
- 4,096
- Feed-forward size
- 11,008
- Attention heads
- 32
- Key/value heads
- 32
- Vocabulary size
- 100,352
- Model type
- olmo2
Identity and Version
- Repository
- shreyansh12183/Shreyansh-STEM-AI-7B-Final
- Publisher
- Shreyansh singh
- Task
- Not stated by the source
- Modality
- Other
- Library
- Not stated by the source
- Parameters
- 7.3B parameters
- Languages
- Not stated by the source
- Revision
- c9f8ddcd4b2c85149ed8237e09c3ca92278b9876
- First published
- 2026-09-13
- Last updated
- 2026-10-06
Files and Weights
16 files, 14.6 GB in total. The weights are 4 files totalling 14.6 GB in safetensors.
Every file
| File | Type | Size | SHA-256 |
|---|---|---|---|
| model-00001-of-00004.safetensors | Weights | 4.0 GB | 9412b3b43bbe |
| model-00002-of-00004.safetensors | Weights | 3.9 GB | 6a7fc71eeca9 |
| model-00003-of-00004.safetensors | Weights | 4.0 GB | e7d2b334a2a0 |
| model-00004-of-00004.safetensors | Weights | 2.7 GB | 2f5290c16634 |
| active_endpoint.json | Configuration | 156 B | — |
| config.json | Configuration | 674 B | — |
| generation_config.json | Configuration | 246 B | — |
| model.safetensors.index.json | Configuration | 29.6 KB | — |
| README.md | Documentation | 31 B | — |
| Colab_Serve_Shreyansh_7B.ipynb | Other | 4.4 KB | — |
| Kaggle_Serve_Shreyansh_7B.ipynb | Other | 4.4 KB | — |
| Shreyansh_7B_Full_Benchmark_Colab.ipynb | Other | 16.3 KB | — |
| chat_template.jinja | Other | 508 B | — |
| .gitattributes | Repository | 1.5 KB | — |
| tokenizer.json | Tokenizer | 7.1 MB | — |
| tokenizer_config.json | Tokenizer | 556 B | — |
License and Download
- License
- cc-by-nc-4.0
- Access
- Open weights, no gate
- Download size
- 14.6 GB
Released by Shreyansh singh through its official repository on Hugging Face. Read the license.
Memory Requirements
| Precision | Weights in memory |
|---|---|
| As published | 14.6 GB |
| 16-bit | 14.6 GB |
| 8-bit | 7.3 GB |
| 4-bit | 3.6 GB |
Weights only, from the published parameter count; the key-value cache and runtime add to this.
Built on This Model
- Adapter ofVigyan-7B-STEM-Instruct-v1
- Derived fromVigyan-7B-STEM-Instruct-v1
- Derived fromvigyan-7b-moe-gguf
- Derived fromVigyan-7B-STEM-DPO-v1
- Quantized fromShreyansh-STEM-AI-7B-Final-GGUF
- Derived fromShreyansh-STEM-AI-7B-Final-GGUF
Questions About Shreyansh-STEM-AI-7B-Final
How much GPU memory does Shreyansh-STEM-AI-7B-Final need?
About 17.5 GB at 16-bit and 4.4 GB at 4-bit: the weights (7.3B parameters) plus a working margin. A long context needs more.
What is the cheapest GPU to run Shreyansh-STEM-AI-7B-Final on?
At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.
Can I use Shreyansh-STEM-AI-7B-Final commercially?
Not without separate permission. Shreyansh-STEM-AI-7B-Final is released under Creative Commons Attribution-NonCommercial 4.0. CC BY-NC 4.0 permits sharing and adapting with credit for non-commercial purposes only. Commercial use needs separate permission from the rights holder.
What is Shreyansh-STEM-AI-7B-Final's context length?
4,096 tokens, from the maximum position embeddings in its published configuration.