SAVRN
Search Contact SAVRN

Open-weight model

Shreyansh-STEM-AI-7B-Final

by Shreyansh singh shreyansh12183/Shreyansh-STEM-AI-7B-Final

Shreyansh-STEM-AI-7B-Final is an open-weight model from Shreyansh singh, released under Creative Commons Attribution-NonCommercial 4.0. It has 7.3B parameters and a 4,096-token context. At 16-bit it needs about 17.5 GB of GPU memory, which fits on 1x MI300X from $1.85 an hour, at the lowest prices in the SAVRN Index. It draws 267 downloads a month.

Parameters7.3B
Context4,096
Weights14.6 GB
Licensecc-by-nc-4.0
AccessOpen weights
Monthly Downloads267

Runs On

What it takes to serve Shreyansh-STEM-AI-7B-Final (7.3B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 14.6 GB 17.5 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
8-bit 7.3 GB 8.8 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 3.6 GB 4.4 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Oct 7, 2026.

Shreyansh-STEM-AI-7B-Final on every accelerator the SAVRN Index prices, at every precision

Model Card

The publisher has not written a card for this model.

Configuration

Architecture
Olmo2ForCausalLM
Context length (tokens)
4,096
Layers
32
Hidden size
4,096
Feed-forward size
11,008
Attention heads
32
Key/value heads
32
Vocabulary size
100,352
Model type
olmo2

Identity and Version

Repository
shreyansh12183/Shreyansh-STEM-AI-7B-Final
Publisher
Shreyansh singh
Task
Not stated by the source
Modality
Other
Library
Not stated by the source
Parameters
7.3B parameters
Languages
Not stated by the source
Revision
c9f8ddcd4b2c85149ed8237e09c3ca92278b9876
First published
2026-09-13
Last updated
2026-10-06

Files and Weights

16 files, 14.6 GB in total. The weights are 4 files totalling 14.6 GB in safetensors.

Weights4 files · 14.6 GB
Configuration4 files · 30.7 KB
Tokenizer2 files · 7.1 MB
Documentation1 file · 31 B
Other4 files · 25.6 KB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
model-00001-of-00004.safetensorsWeights4.0 GB 9412b3b43bbe
model-00002-of-00004.safetensorsWeights3.9 GB 6a7fc71eeca9
model-00003-of-00004.safetensorsWeights4.0 GB e7d2b334a2a0
model-00004-of-00004.safetensorsWeights2.7 GB 2f5290c16634
active_endpoint.jsonConfiguration156 B —
config.jsonConfiguration674 B —
generation_config.jsonConfiguration246 B —
model.safetensors.index.jsonConfiguration29.6 KB —
README.mdDocumentation31 B —
Colab_Serve_Shreyansh_7B.ipynbOther4.4 KB —
Kaggle_Serve_Shreyansh_7B.ipynbOther4.4 KB —
Shreyansh_7B_Full_Benchmark_Colab.ipynbOther16.3 KB —
chat_template.jinjaOther508 B —
.gitattributesRepository1.5 KB —
tokenizer.jsonTokenizer7.1 MB —
tokenizer_config.jsonTokenizer556 B —

License and Download

License
cc-by-nc-4.0
Access
Open weights, no gate
Download size
14.6 GB
Download from Shreyansh singh

Released by Shreyansh singh through its official repository on Hugging Face. Read the license.

Memory Requirements

PrecisionWeights in memory
As published14.6 GB
16-bit14.6 GB
8-bit7.3 GB
4-bit3.6 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Built on This Model

Questions About Shreyansh-STEM-AI-7B-Final

How much GPU memory does Shreyansh-STEM-AI-7B-Final need?

About 17.5 GB at 16-bit and 4.4 GB at 4-bit: the weights (7.3B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run Shreyansh-STEM-AI-7B-Final on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use Shreyansh-STEM-AI-7B-Final commercially?

Not without separate permission. Shreyansh-STEM-AI-7B-Final is released under Creative Commons Attribution-NonCommercial 4.0. CC BY-NC 4.0 permits sharing and adapting with credit for non-commercial purposes only. Commercial use needs separate permission from the rights holder.

What is Shreyansh-STEM-AI-7B-Final's context length?

4,096 tokens, from the maximum position embeddings in its published configuration.