SAVRN
Search Contact SAVRN

Open-weight model · Time series forecasting

chronos-t5-large

by Amazon amazon/chronos-t5-large

Update Feb 14, 2025: Chronos-Bolt & original Chronos models are now available on Amazon SageMaker JumpStart! Check out the tutorial notebook to learn how to deploy Chronos endpoints for production use in a few lines of code.

Parameters709M
Context
Weights2.8 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads96.5k

Runs On

What it takes to serve chronos-t5-large (709M parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 1.4 GB 1.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
8-bit 0.7 GB 0.9 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 0.4 GB 0.4 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 18, 2026.

Model Card

By Amazon, published under apache-2.0, revision 0e46c9c7e2e9.

Update Feb 14, 2025: Chronos-Bolt & original Chronos models are now available on Amazon SageMaker JumpStart! Check out the tutorial notebook to learn how to deploy Chronos endpoints for production use in a few lines of code.

Update Nov 27, 2024: We have released Chronos-Bolt models that are more accurate (5% lower error), up to 250 times faster and 20 times more memory-efficient than the original Chronos models of the same size. Check out the new modelshere.

Read the full model card (570 words)

Configuration

Architecture
T5ForConditionalGeneration
Vocabulary size
4,096
Stored precision
float32
Model type
t5

Identity and Version

Repository
amazon/chronos-t5-large
Publisher
Amazon
Task
Time series forecasting
Modality
Time series
Library
chronos-forecasting
Parameters
709M parameters
Languages
Not stated by the source
Revision
0e46c9c7e2e9f74b53db0617fdfcfe42a413e54a
First published
2024-02-21
Last updated
2025-11-21

Files and Weights

6 files, 2.8 GB in total. The weights are 1 file totalling 2.8 GB in safetensors.

Weights1 file · 2.8 GB
Configuration2 files · 1.3 KB
Documentation1 file · 6.1 KB
Other1 file · 232.3 KB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
model.safetensorsWeights2.8 GB dad1e592a5ba
config.jsonConfiguration1.1 KB
generation_config.jsonConfiguration142 B
README.mdDocumentation6.1 KB
figures/main-figure.pngOther232.3 KB
.gitattributesRepository1.5 KB

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
2.8 GB
Download from Amazon

Released by Amazon through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published2.8 GB
16-bit1.4 GB
8-bit0.7 GB
4-bit0.4 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About chronos-t5-large

How much GPU memory does chronos-t5-large need?

About 1.7 GB at 16-bit and 0.4 GB at 4-bit: the weights (709M parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run chronos-t5-large on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use chronos-t5-large commercially?

Yes. chronos-t5-large is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Similar Models

Model · Time series forecasting

timesfm-2.0-500m-pytorch

Google

TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting. This is not an officially supported Google product. timesfm-2.0-500m is the second open model checkpoint: - It performs univariate time series forecasting for context lengths up to 2048 time points and any horizon lengths, with an optional frequency indicator. Note that it can go even beyond 2048 context even though it was trained with that as the maximum context. - It focuses on point forecasts. We experimentally offer 10 quantile heads but they have not been calibrated after pretraining. - It ideally requires the context to be contiguous (i.e. no…

Open weights apache-2.0 499M parameters timesfm

Model · Time series forecasting

moirai-moe-1.0-R-base

Salesforce AI Research

This model has been pushed to the Hub using the PytorchModelHubMixin integration: This release is for research purposes only in support of an academic paper. Our models, datasets, and code are not specifically designed or evaluated for all downstream purposes. We strongly recommend users evaluate and address potential concerns related to accuracy, safety, and fairness before deploying this model. We encourage users to consider the common limitations of AI, comply with applicable laws, and leverage best practices when selecting use cases, particularly for high-risk scenarios where errors or misuse could significantly impact people’s lives, rights, or safety. For further guidance on use…

Open weights cc-by-nc-4.0 935M parameters

Model · Time series forecasting

granite-timeseries-patchtst-fm-r2

IBM Granite

PatchTST-FM-r2, a state-of-the-art zero-shot time series foundation model, represents a continuation of the well-recognized PatchTST model series, building on the original PatchTST and its zero-shot variant PatchTST-FM-r1. PatchTST-FM-r2 brings architectural enhancements as well as an expanded training base on top of its predecessor PatchTST-FM-r1. As of August 31, 2026 Granite-TimeSeries-PatchTST-FM-r2 is the top performing zero-shot model released under a permissive, commercial-friendly open-source license on the GIFT-Eval benchmark. Granite-TimeSeries-PatchTST-FM-r2 ranks #2 when considering all zero-shot, replicable models (see below for more details). The architectural changes in r2…

Open weights openmdw-1.0 385M parameters

Model · Time series forecasting

MOMENT-1-large

Auton Lab

MOMENT is a family of foundation models for general-purpose time-series analysis. The models in this family (1) serve as a building block for diverse time-series analysis tasks (e.g., forecasting, classification, anomaly detection, and imputation, etc.), (2) are effective out-of-the-box, i.e., with no (or few) task-specific exemplars (enabling e.g., zero-shot forecasting, few-shot classification, etc.), and (3) are tunable using in-distribution and task-specific data to improve performance. For details on MOMENT models, training data, and experimental results, please refer to the paper MOMENT: A Family of Open Time-series Foundation Models. Recommended Python Version: Python 3.11 (support…

Open weights mit 346M parameters transformers

Model · Time series forecasting

timesfm-3.0-pytorch

Google

TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting. This repository contains the official PyTorch weights and configurations for TimesFM 3.0. This model is released under the TimesFM Non-Commercial License v1.0. timesfm-3.0 is pretrained using - GiftEvalPretrain excluding the datasets that overlap with fev-bench - Wikipedia Pageviews, cutoff Nov 2023 (see paper for details). - Google Trends top queries, cutoff EoY 2022 (see paper for details). - Synthetic and augmented data. title={A decoder-only foundation model for time-series forecasting}, author={Das, Abhimanyu and Kong, Weihao and Sen, Rajat and…

Open weights other 331M parameters

Model · Time series forecasting

Toto-2.0-313m

Datadog

Toto (Time Series Optimized Transformer for Observability) is a family of time series foundation models for multivariate forecasting developed by Datadog. Toto 2.0 is the current generation, featuring u-μP-scaled transformers ranging from 4m to 2.5B parameters, all trained from a single recipe. Forecast quality improves reliably with parameter count across the family. The family sets a new state of the art on three forecasting benchmarks: BOOM, our observability benchmark; GIFT-Eval, the standard general-purpose benchmark; and the recent contamination-resistant TIME benchmark. Inference code is available on GitHub. For more examples, see the Quick Start notebook and GluonTS integration…

Open weights apache-2.0 313M parameters pytorch