SAVRN
Search Contact SAVRN

Organization

OpenGVLab

OpenGVLab

Computer Vision

Models in Library5
Datasets in Library0
Models on Hugging Face286
Followers2k

Models

Model · Video classification

VideoMAEv2-Base

OpenGVLab

VideoMAEv2-Base model pre-trained for 800 epochs in a self-supervised way on UnlabeldHybrid-1M dataset. It was introduced in the paper [[CVPR23]VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking](https://arxiv.org/abs/2203.12602) by Wang et al. and first released in GitHub. You can use the raw model for video feature extraction. Here is how to use this model to extract a video feature

Open weights cc-by-nc-4.0 86M parameters

Model · Video classification

VideoMAEv2-Huge

OpenGVLab

VideoMAEv2-Huge model pre-trained for 1200 epochs in a self-supervised way on UnlabeldHybrid-1M dataset. It was introduced in the paper [[CVPR23]VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking](https://arxiv.org/abs/2203.12602) by Wang et al. and first released in GitHub. You can use the raw model for video feature extraction. Here is how to use this model to extract a video feature

Open weights cc-by-nc-4.0 632M parameters

Model · Video classification

VideoMAEv2-Large

OpenGVLab

VideoMAEv2-Large model pre-trained for 800 epochs in a self-supervised way on UnlabeldHybrid-1M dataset. It was introduced in the paper [[CVPR23]VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking](https://arxiv.org/abs/2203.12602) by Wang et al. and first released in GitHub. You can use the raw model for video feature extraction. Here is how to use this model to extract a video feature

Open weights cc-by-nc-4.0 304M parameters

Model · Video classification

InternVideo2-Stage2_6B

OpenGVLab

This repository contains the 6B model of the paper InternVideo2 in stage 2. Code: https://github.com/OpenGVLab/InternVideo/tree/main/InternVideo2/multimodality Please refer to https://github.com/OpenGVLab/InternVideo/blob/main/InternVideo2/multimodality/INSTALL.md

Open weights mit 6.4B parameters

Model · Video classification

VideoMAEv2-giant

OpenGVLab

VideoMAEv2-giant model pre-trained for 1200 epochs in a self-supervised way on UnlabeldHybrid-1M dataset. It was introduced in the paper [[CVPR23]VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking](https://arxiv.org/abs/2203.12602) by Wang et al. and first released in GitHub. You can use the raw model for video feature extraction. Here is how to use this model to extract a video feature

Open weights cc-by-nc-4.0 1B parameters