SAVRN
Search Contact SAVRN

Open-weight model · Robotics

Alpamayo2-Super

by NVIDIA nvidia/Alpamayo2-Super

Alpamayo 2 Super is a 34B-parameter foundation model designed to tackle multiple autonomous vehicle (AV) development tasks. It combines a 32B VLM backbone with a 2B diffusion expert.

Parameters35.8B
Context
Weights71.6 GB
Licenseopenmdw-1.1
AccessOpen weights
Monthly Downloads11.8k

Runs On

What it takes to serve Alpamayo2-Super (35.8B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 71.6 GB 86.0 GB 1x MI300X (192 GB)
Vultr
$1.85 1x MI325X $2.00 · 1x MI355X $2.59
8-bit 35.8 GB 43.0 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 17.9 GB 21.5 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 18, 2026.

Model Card

Alpamayo 2 Super is a 34B-parameter foundation model designed to tackle multiple autonomous vehicle (AV) development tasks. It combines a 32B VLM backbone with a 2B diffusion expert. Alpamayo 2 Super was developed by NVIDIA as a part of the broader Alpamayo Open Platform. Model weights: The model weights are released under the OpenMDW-1.1 license. Source code: Apache License 2.0, as provided in the Alpamayo 2 Super source repository. Global Developers and researchers working on autonomous vehicle systems who need a foundation model for perception, planning, and decision-making tasks. Alpamayo 2 Super supports multiple AV development tasks such as trajectory prediction, visual question…

Excerpt from the card by NVIDIA, licensed openmdw-1.1.

Configuration

Architecture
Alpamayo2Super
Model type
alpamayo2_super

Identity and Version

Repository
nvidia/Alpamayo2-Super
Publisher
NVIDIA
Task
Robotics
Modality
Control
Library
Not stated by the source
Parameters
35.8B parameters
Languages
en
Revision
00554695e729a6ff0b6281fd2c81b18d06e33dbe
First published
2026-05-27
Last updated
2026-08-07

Files and Weights

32 files, 71.6 GB in total. The weights are 15 files totalling 71.6 GB in safetensors.

Weights15 files · 71.6 GB
Configuration7 files · 181.2 KB
Tokenizer4 files · 15.9 MB
Documentation2 files · 12.4 KB
Other3 files · 941.9 KB
Repository1 file · 1.7 KB
Every file
FileTypeSizeSHA-256
model-00001-of-00015.safetensorsWeights4.9 GB 794ffa42440f
model-00002-of-00015.safetensorsWeights4.9 GB faa303461d97
model-00003-of-00015.safetensorsWeights4.9 GB 11ed2a840737
model-00004-of-00015.safetensorsWeights4.9 GB 695a122dc6c1
model-00005-of-00015.safetensorsWeights4.9 GB 67cd73482cb2
model-00006-of-00015.safetensorsWeights4.9 GB d9ef4f035560
model-00007-of-00015.safetensorsWeights4.9 GB a0f639e3a4ab
model-00008-of-00015.safetensorsWeights4.9 GB 73a8e3b12a0e
model-00009-of-00015.safetensorsWeights4.9 GB dbe04e5d540b
model-00010-of-00015.safetensorsWeights4.9 GB 31c6920d91a7
model-00011-of-00015.safetensorsWeights4.9 GB 74a34bb7a2f2
model-00012-of-00015.safetensorsWeights4.9 GB db9afcec2566
model-00013-of-00015.safetensorsWeights4.9 GB 831a7a5cdf85
model-00014-of-00015.safetensorsWeights5.0 GB 86e4f9e3a720
model-00015-of-00015.safetensorsWeights3.2 GB cc9423f8a630
added_tokens.jsonConfiguration707 B
config.jsonConfiguration11.2 KB
generation_config.jsonConfiguration213 B
model.safetensors.index.jsonConfiguration166.9 KB
preprocessor_config.jsonConfiguration782 B
special_tokens_map.jsonConfiguration613 B
video_preprocessor_config.jsonConfiguration817 B
LICENSEDocumentation2.6 KB
README.mdDocumentation9.8 KB
chat_template.jinjaOther5.3 KB
images/Alpamayo-2-Super_Benchmark-Results_A.pngOther345.4 KB a9c164fa731a
images/Alpamayo-2-Super_Benchmark-Results_B.pngOther591.2 KB 0e2737aaf5f1
.gitattributesRepository1.7 KB
merges.txtTokenizer1.7 MB
tokenizer.jsonTokenizer11.4 MB aeb13307a71a
tokenizer_config.jsonTokenizer5.4 KB
vocab.jsonTokenizer2.8 MB

License and Download

License
openmdw-1.1
Access
Open weights, no gate
Download size
71.6 GB
Download from NVIDIA

Released by NVIDIA through its official repository on Hugging Face.

Built From

Memory Requirements

PrecisionWeights in memory
As published71.6 GB
16-bit71.6 GB
8-bit35.8 GB
4-bit17.9 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Alpamayo2-Super

How much GPU memory does Alpamayo2-Super need?

About 86 GB at 16-bit and 21.5 GB at 4-bit: the weights (35.8B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run Alpamayo2-Super on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

What license is Alpamayo2-Super released under?

openmdw-1.1, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Robotics

Alpamayo-1.5-10B

NVIDIA

Alpamayo 1.5 is a significant update to NVIDIA’s open 10B-parameter chain-of-thought reasoning VLA model, designed to be an interactive and steerable reasoning engine for the AV community. Alpamayo 1.5 is built on the Cosmos-Reason2 VLM backbone, is RL post-trained, and introduces support for navigation guidance, flexible camera counts, and user question answering. This model is ready for non-commercial use. Commercial licensing available upon request. Model weights: The model weights are released under the OpenMDW-1.1 license. Source code: Apache License 2.0, as provided in the Alpamayo 1.5 source repository. Global Researchers and autonomous-driving practitioners who are developing and…

Open weights openmdw-1.1 11.1B parameters

Model · Robotics

Alpamayo-R1-10B

NVIDIA

Note: Following the release of NVIDIA Alpamayo at CES 2026, Alpamayo-R1 has been renamed to Alpamayo 1. Alpamayo 1 integrates Chain-of-Causation reasoning with trajectory planning to enhance decision-making in complex autonomous-driving scenarios. Alpamayo 1 (v1.0) was developed by NVIDIA as a vision-language-action (VLA) model that bridges interpretable reasoning with precise vehicle control for autonomous-driving applications. This model is ready for non-commercial use. Commercial licensing available upon request. Model weights: The model weights are released under the OpenMDW-1.1 license. Source code: Apache License 2.0, as provided in the Alpamayo 1 source repository. Global Researchers…

Open weights openmdw-1.1 11.1B parameters transformers

Model · Robotics

Alpamayo-1.5-10B

Z Lab

Flash Vision-Language-Action Inference for Autonomous Driving FlashDrive accelerates Alpamayo 1.5 — one of NVIDIA's 10B-parameter vision-language-action models for autonomous driving — by 4.7× with no loss in accuracy, through streaming inference, DFlash speculative reasoning, ParoQuant W4A8 quantization, adaptive action caching, and torch.compile. This repository mirrors the weights of nvidia/Alpamayo-1.5-10B and is the base checkpoint of the FlashDrive stack. Loading it pulls the derived companions automatically: Install FlashDrive, then load this base checkpoint — the -PARO and -DFlash companions are fetched automatically: The first call per stream only prefills the KV cache and returns…

Open weights other 11.1B parameters

Model · Robotics

GraspMolmo

Ai2

[[Paper]](https://arxiv.org/pdf/2505.13441) [[arXiv]](https://arxiv.org/abs/2505.13441) [[Project Website]](https://abhaybd.github.io/GraspMolmo/) [[Data]](https://huggingface.co/datasets/allenai/PRISM) GraspMolmo is a generalizable open-vocabulary task-oriented grasping (TOG) model for robotic manipulation. Given an image and a task to complete (e.g. "Pour me some tea"), GraspMolmo will point to the most appropriate grasp location, which can then be matched to the closest stable grasp. Running the above code could result in the following output: To predict a grasp point and match it to one of the candidate grasps, refer to the GraspMolmo class. First, install graspmolmo with and then…

Open weights mit 8B parameters 4,096 tokens

Model · Robotics

openvla-7b

OpenVLA Collaboration

OpenVLA 7B (openvla-7b) is an open vision-language-action model trained on 970K robot manipulation episodes from the Open X-Embodiment dataset. The model takes language instructions and camera images as input and generates robot actions. It supports controlling multiple robots out-of-the-box, and can be quickly adapted for new robot domains via (parameter-efficient) fine-tuning. All OpenVLA checkpoints, as well as our training codebase are released under an MIT License. For full details, please read our paper and see our project page. OpenVLA models take a language instruction and a camera image of a robot workspace as input, and predict (normalized) robot actions consisting of 7-DoF…

Open weights mit 7.5B parameters transformers

This repository contains the OpenVLA-OFT checkpoint for LIBERO-Spatial, as described in Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success. OpenVLA-OFT significantly improves upon the base OpenVLA model by incorporating optimized fine-tuning techniques. See here for other OpenVLA-OFT checkpoints: https://huggingface.co/moojink?searchmodels=oft This example demonstrates generating an action chunk using a pretrained OpenVLA-OFT checkpoint. Ensure you have set up the conda environment as described in the GitHub README.

Open weights mit 7.5B parameters transformers