SAVRN
Search Contact SAVRN

Open-weight model · Text to video

Wan2.1-T2V-1.3B-Diffusers

by Wan-AI Wan-AI/Wan2.1-T2V-1.3B-Diffusers

In this repository, we present Wan2.1, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.

Parameters1.4B
Context
Weights28.9 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads285.3k

Runs On

What it takes to serve Wan2.1-T2V-1.3B-Diffusers (1.4B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 2.8 GB 3.4 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
8-bit 1.4 GB 1.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 0.7 GB 0.9 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 18, 2026.

Model Card

By Wan-AI, published under apache-2.0, revision 0fad780a534b.

Wan2.1

Read the full model card (1,925 words)

Identity and Version

Repository
Wan-AI/Wan2.1-T2V-1.3B-Diffusers
Publisher
Wan-AI
Task
Text to video
Modality
Video
Library
diffusers
Parameters
1.4B parameters
Languages
en, zh
Revision
0fad780a534b6463e45facd96134c9f345acfa5b
First published
2025-03-01
Last updated
2025-04-04

Files and Weights

31 files, 28.9 GB in total. The weights are 8 files totalling 28.9 GB in safetensors.

Weights8 files · 28.9 GB
Configuration8 files · 106.0 KB
Tokenizer3 files · 21.4 MB
Documentation1 file · 18.1 KB
Other10 files · 6.7 MB
Repository1 file · 2.2 KB
Every file
FileTypeSizeSHA-256
text_encoder/model-00001-of-00005.safetensorsWeights5.0 GB c0ef3a140898
text_encoder/model-00002-of-00005.safetensorsWeights4.9 GB 481c7b2b3977
text_encoder/model-00003-of-00005.safetensorsWeights5.0 GB f93148bcc040
text_encoder/model-00004-of-00005.safetensorsWeights5.0 GB a451792c739c
text_encoder/model-00005-of-00005.safetensorsWeights2.9 GB 7e76e18d2245
transformer/diffusion_pytorch_model-00001-of-00002.safetensorsWeights5.0 GB 6d011927dbd2
transformer/diffusion_pytorch_model-00002-of-00002.safetensorsWeights677.3 MB b92ec2309b1f
vae/diffusion_pytorch_model.safetensorsWeights507.6 MB d6e524b3fffe
model_index.jsonConfiguration400 B
scheduler/scheduler_config.jsonConfiguration751 B
text_encoder/config.jsonConfiguration854 B
text_encoder/model.safetensors.index.jsonConfiguration22.5 KB
tokenizer/special_tokens_map.jsonConfiguration7.1 KB
transformer/config.jsonConfiguration465 B
transformer/diffusion_pytorch_model.safetensors.index.jsonConfiguration73.3 KB
vae/config.jsonConfiguration724 B
README.mdDocumentation18.1 KB
assets/comp_effic.pngOther1.8 MB b0e225caffb4
assets/data_for_diff_stage.jpgOther528.3 KB 59aec08409f2
assets/i2v_res.pngOther891.7 KB 6823b3206d8d
assets/logo.pngOther56.3 KB 96cddc0f6672
assets/t2v_res.jpgOther301.0 KB 91db57909244
assets/vben_1.3b_vs_sota.pngOther515.8 KB b7705db79f2e
assets/vben_vs_sota.pngOther1.6 MB 9a0e86ca8504
assets/video_dit_arch.jpgOther643.4 KB 195dceec6570
assets/video_vae_res.jpgOther212.6 KB d8f9e7f73538
examples/i2v_input.JPGOther250.6 KB 077e3d965090
.gitattributesRepository2.2 KB
tokenizer/spiece.modelTokenizer4.5 MB e3909a67b780
tokenizer/tokenizer.jsonTokenizer16.8 MB 20a46ac25674
tokenizer/tokenizer_config.jsonTokenizer61.8 KB

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
28.9 GB
Download from Wan-AI

Released by Wan-AI through its official repository on Hugging Face. Read the license.

Memory Requirements

PrecisionWeights in memory
As published28.9 GB
16-bit2.8 GB
8-bit1.4 GB
4-bit0.7 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Wan2.1-T2V-1.3B-Diffusers

How much GPU memory does Wan2.1-T2V-1.3B-Diffusers need?

About 3.4 GB at 16-bit and 0.9 GB at 4-bit: the weights (1.4B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run Wan2.1-T2V-1.3B-Diffusers on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use Wan2.1-T2V-1.3B-Diffusers commercially?

Yes. Wan2.1-T2V-1.3B-Diffusers is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Similar Models

Model · Text to video

Wan2.1-T2V-1.3B

Wan-AI

In this repository, we present Wan2.1, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation. Wan2.1 offers these key features: This repository hosts our T2V-1.3B model, a versatile solution for video generation that is compatible with nearly all consumer-grade GPUs. In this way, we hope that Wan2.1 can serve as an easy-to-use tool for more creative teams in video creation, providing a high-quality foundational model for academic teams with limited computing resources. This will facilitate both the rapid development of the video creation community and the swift advancement of video technology. Your browser does not support the video tag.…

Open weights apache-2.0 1.4B parameters diffusers

convert TurboWan2.1-T2V-1.3B-480P(https://modelscope.cn/models/TurboDiffusion/TurboWan2.1-T2V-1.3B-480P/summary) to TurboWan2.1-T2V-1.3B-Diffusers convert script https://github.com/IPostYellow/TurboWantoDiffusers/blob/main/convertturbowantodiffusers.py To use in sglang

Open weights apache-2.0 1.4B parameters diffusers

Model · Text to video

CogVideoX-2b

Z.ai

Visit QingYing and API Platform to experience commercial video generation models. CogVideoX is an open-source version of the video generation model originating from QingYing. The table below displays the list of video generation models we currently offer, along with their foundational information. Data Explanation + When testing using the diffusers library, all optimizations provided by the diffusers library were enabled. This solution has not been tested for actual VRAM/memory usage on devices other than NVIDIA A100 / H100. Generally, this solution can be adapted to all devices with NVIDIA Ampere architecture and above. If the optimizations are disabled, VRAM usage will increase…

Open weights apache-2.0 1.7B parameters diffusers

Model · Text to video

AnimateLCM

Fu-Yun Wang

AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data by Fu-Yun Wang et al. For more details, please refer to our [paper] | [code] | [proj-page] | [civitai].

Open weights 454M parameters diffusers

Model · Text to video

Wan2.2-T2V-A14B-GGUF

QuantStack

This GGUF file is a direct conversion of Wan-AI/Wan2.2-T2V-A14B Since this is a quantized model, all original licensing terms and usage restrictions remain in effect. Usage The model can be used with the ComfyUI custom node ComfyUI-GGUF by city96 Place model files in ComfyUI/models/unet see the GitHub readme for further installation instructions.

Open weights apache-2.0 gguf

Model · Text to video

MiniMax-H3-Turbo-Lora

Larryvrh

A LoRA for MiniMax-H3 that renders joint video + synchronized stereo audio in as few as 4 sampling steps instead of the usual ~20 — a ~5× sampling speedup — and keeps getting better as you add steps. For most work, use minimaxh3turbov4step600ema.safetensors. It's the markedly better micro-detail (faces, fingers, fine texture), and the over-sharpening / plastic look of the earlier v1 (~850) line is fully resolved. v4 introduced a static-frame enhancement — a big win for static and small-motion content. The one trade-off shows up only at 4 steps with large, fast motion, where v4 can produce motion-smear / trailing ghosting (we're actively fixing this). Two things address it: - Use 6–8 steps.…

Open weights apache-2.0 minimax-h3