SAVRN
Search Contact SAVRN

Open-weight model · Image segmentation

sam3-onnx

by Waechter Julien greenjava/sam3-onnx

Meta SAM 3 (Segment Anything Model 3), exported to ONNX (FP16) for a C++ ONNX Runtime pipeline: a vision encoder, a text encoder and a decoder with a dynamic prompt-count batch (a single file accepts any number of text prompts at runtime, no re-export…

Parameters
Context
Weights1.7 GB
Licenseother
AccessOpen weights
Monthly Downloads

Model Card

Meta SAM 3 (Segment Anything Model 3), exported to ONNX (FP16) for a C++ ONNX Runtime pipeline: a vision encoder, a text encoder and a decoder with a dynamic prompt-count batch (a single file accepts any number of text prompts at runtime, no re-export needed). The model is derived from Meta's SAM 3 and is provided under the SAM License (see LICENSE in this repository). Use, reproduction and redistribution are subject to that agreement; the license text must be kept with any redistribution.

Excerpt from the card by Waechter Julien, licensed other.

Identity and Version

Repository
greenjava/sam3-onnx
Publisher
Waechter Julien
Task
Image segmentation
Modality
Image
Library
onnx
Parameters
Not stated by the source
Languages
en
Revision
05b1f7103c46391bfc0a7baf912e53fdf9c580b8
First published
2026-09-04
Last updated
2026-09-18

Files and Weights

9 files, 1.7 GB in total. The weights are 4 files totalling 1.7 GB in onnx.

Weights4 files · 1.7 GB
Configuration1 file · 999.7 KB
Documentation2 files · 8.8 KB
Other1 file · 3.2 MB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
fp16/decoder-plugin.onnxWeights57.8 MB e78d242ab949
fp16/decoder.onnxWeights61.6 MB bd60a1a8bf4c
fp16/text-encoder.onnxWeights707.2 MB a54cf145bae4
fp16/vision-encoder.onnxWeights908.1 MB 5a68367d75c8
fp16/clip_vocab.jsonConfiguration999.7 KB
LICENSEDocumentation7.4 KB
README.mdDocumentation1.4 KB
fp16/bpe_simple_vocab_16e6.txtOther3.2 MB
.gitattributesRepository1.5 KB

License and Download

License
other
Access
Open weights, no gate
Download size
1.7 GB
Download from Waechter Julien

Released by Waechter Julien through its official repository on Hugging Face.

Memory Requirements

PrecisionWeights in memory
As published1.7 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About sam3-onnx

What license is sam3-onnx released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image segmentation

segformer-b2-finetuned-ade-512-512

NVIDIA

SegFormer model fine-tuned on ADE20k at resolution 512x512. It was introduced in the paper SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers by Xie et al. and first released in this repository. Disclaimer: The team releasing SegFormer did not write a model card for this model so this model card has been written by the Hugging Face team. SegFormer consists of a hierarchical Transformer encoder and a lightweight all-MLP decode head to achieve great results on semantic segmentation benchmarks such as ADE20K and Cityscapes. The hierarchical Transformer is first pre-trained on ImageNet-1k, after which a decode head is added and fine-tuned altogether on a…

Open weights other transformers

Model · Image segmentation

segformer-b3-finetuned-ade-512-512

NVIDIA

SegFormer model fine-tuned on ADE20k at resolution 512x512. It was introduced in the paper SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers by Xie et al. and first released in this repository. Disclaimer: The team releasing SegFormer did not write a model card for this model so this model card has been written by the Hugging Face team. SegFormer consists of a hierarchical Transformer encoder and a lightweight all-MLP decode head to achieve great results on semantic segmentation benchmarks such as ADE20K and Cityscapes. The hierarchical Transformer is first pre-trained on ImageNet-1k, after which a decode head is added and fine-tuned altogether on a…

Open weights other transformers

Model · Image segmentation

oneformer_ade20k_swin_large

SHI Labs

OneFormer model trained on the ADE20k dataset (large-sized version, Swin backbone). It was introduced in the paper OneFormer: One Transformer to Rule Universal Image Segmentation by Jain et al. and first released in this repository. OneFormer is the first multi-task universal image segmentation framework. It needs to be trained only once with a single universal architecture, a single model, and on a single dataset, to outperform existing specialized models across semantic, instance, and panoptic segmentation tasks. OneFormer uses a task token to condition the model on the task in focus, making the architecture task-guided for training, and task-dynamic for inference, all with a single…

Open weights mit transformers

Model · Image segmentation

segformer-b1-finetuned-ade-512-512

NVIDIA

SegFormer model fine-tuned on ADE20k at resolution 512x512. It was introduced in the paper SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers by Xie et al. and first released in this repository. Disclaimer: The team releasing SegFormer did not write a model card for this model so this model card has been written by the Hugging Face team. SegFormer consists of a hierarchical Transformer encoder and a lightweight all-MLP decode head to achieve great results on semantic segmentation benchmarks such as ADE20K and Cityscapes. The hierarchical Transformer is first pre-trained on ImageNet-1k, after which a decode head is added and fine-tuned altogether on a…

Open weights other transformers

Model · Image segmentation

segformer-b0-finetuned-ade-512-512

Joshua

https://huggingface.co/nvidia/segformer-b0-finetuned-ade-512-512 with ONNX weights to be compatible with Transformers.js. If you haven't already, you can install the Transformers.js JavaScript library from NPM using: Example: Image segmentation with Xenova/segformer-b0-finetuned-ade-512-512. You can visualize the outputs with: Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).

Open weights transformers.js

Model · Image segmentation

DelineateAnything

Mykola Lavreniuk

Delineate Anything v2 extends Delineate Anything into a globally representative, resolution-agnostic foundation model that scales agricultural field boundary detection to a planetary level from any imagery source. Trained on FBIS-73M, a massive 73-million-instance dataset spanning 61 countries with diverse imagery sources ranging from 0.25m to 10m resolution, built through a resolution-specific curation pipeline that solves the parcel-versus-field mismatch, Delineate Anything v2 sets a new state-of-the-art in global zero-shot delineation. It delivers a +103.3% relative gain in [email protected] over Delineate Anything while maintaining extreme efficiency, mapping all of Ukraine (603,000 km²) in 5.4…

Open weights agpl-3.0 ultralytics