SAVRN
Search Contact SAVRN

Dataset

course-assets

by Hugging Face for Computer Vision hf-vision/course-assets

Rows6
Configurations1
Size283.7 MB
Licenseapache-2.0
AccessPublicly accessible
Monthly Downloads124.9k

Dataset Card

The publisher has not written a card for this dataset.

Structure

default 6 rows

SplitRowsSize
train23.3 MB
test42.5 MB
imageImage

Details

Repository
hf-vision/course-assets
Publisher
Hugging Face for Computer Vision
Task category
Not stated by the source
Tags
Not stated by the source
Size category
Not stated by the source
Languages
Not stated by the source
Revision
5e88de973e1bda37a7d8a6b98449c47d309f8927
Last updated
2025-01-24

Files

299 files, 283.7 MB in total.

Documentation1 file · 28 B
Other296 files · 283.7 MB
Repository2 files · 8.6 KB
Every file
FileTypeSizeSHA-256
README.mdDocumentation28 B
3d_stereo_vision_images/annotated_left_img.jpgOther70.8 KB6f5b72ca5333
3d_stereo_vision_images/annotated_right_img.jpgOther72.8 KB063a69d1fdfc
3d_stereo_vision_images/calculated_dim_results.pngOther767.5 KB3db5e8120411
3d_stereo_vision_images/image_formation_simple_stereo.jpgOther15.3 KB8b251a1d7dd1
3d_stereo_vision_images/image_formation_single_camera.pngOther232.2 KB1ac5c99954db
3d_stereo_vision_images/rectified_left_frame.jpgOther61.1 KB773848ae7d21
3d_stereo_vision_images/rectified_overlapping_frames.jpgOther56.8 KB52fdffc900a4
3d_stereo_vision_images/rectified_right_frame.jpgOther63.2 KBf4094eb7b7a2
3d_stereo_vision_images/rectified_stacked_frames.jpgOther129.8 KBef97be99a9d6
3d_stereo_vision_images/unrectified_left_frame.jpgOther73.4 KBd3398b79c513
3d_stereo_vision_images/unrectified_right_frame.jpgOther69.7 KBa08d62ee4a1e
3d_stereo_vision_images/unrectified_stacked_frames.jpgOther148.8 KB9b418b93d359
Attention Comparison.pngOther879.8 KB6269933d799c
BLIP.pngOther427.6 KB270a2f640399
Black_hole_-_Messier_87.jpgOther38.9 KBb8c821fb3312
CNNs/1d_conv.jpgOther95.6 KBd86eeacb2245
CNNs/1d_conv_multiplication.pngOther241.9 KB700a22b616e3
CNNs/1d_conv_result.pngOther187.5 KB8f6048d39dd0
CNNs/2d_conv.pngOther356.7 KB26b66af982f7
CNNs/2d_conv_illustrated.pngOther580.4 KBbe27af80ae36
CNNs/2d_conv_matrix.pngOther330.4 KB5951fc6e0c51
CNNs/2d_feature_map.pngOther523.3 KBe72da3d192b5
CNNs/convolved_illustrated.pngOther382.8 KBccded4ede837
CNNs/kernel_image.pngOther268.7 KB50b5847d1197
CNNs/max_pooling.pngOther489.2 KB09900a81cd95
CNNs/network_illustration.pngOther828.5 KB565b00e046b9
CNNs/network_illustration_2.pngOther654.3 KB95c3a6f940e7
CNNs/network_illustration_3.pngOther682.7 KB3faeb65c808f
CNNs/padding.pngOther497.4 KBc5e5e71f6e6e
CNNs/prewitt_sobel.pngOther150.7 KB1598f2d5f69f
CV_in_defintiion.pngOther160.6 KBc5b3525342b2
CycleGAN2.jpgOther6.3 MBc4d67fd95160
DETR.pngOther53.8 KB2f6a2f211196
DETR_DecoderLayer.pngOther184.8 KB1d43e577febf
DETR_FeatureMaps.pngOther897.5 KB8f8982453944
FPN.pngOther69.8 KBab4fbba54f71
FPN_2.pngOther187.7 KBf7e0b89f3b61
Fast R-CNN.pngOther829.0 KB8608769a795c
Faster_RCNN.pngOther56.1 KB9c5458fd6d3f
KL-Loss.pngOther1.8 KB855bbed75ccb
LLM Challenges.pngOther52.4 KB20d36af73266
MaSA vs Retention.pngOther50.5 KBdf3e03749911
MaSAD.pngOther17.2 KB15bebc3357d2
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/DA2_pipeline.pngOther325.8 KB0013f8da953b
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/dataset.pngOther3.2 MB71193ef20757
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_ambiguity1.gifOther2.9 MB78d3b5cc928f
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_ambiguity2.gifOther7.6 MB61b75357a7a4
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution1.pngOther246.0 KB27815321d9d8
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution2.pngOther627.3 KBcd9644e89867
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution3.pngOther1.3 MB168babd6eac4
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_holes.pngOther326.6 KB007c1f76cc0a
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/different_scales.pngOther1.2 MBc685bd6852b1
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/dpt.pngOther140.6 KBa217a50809cf
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/focal_lenght.pngOther681.1 KBfb5c044ab24e
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/inference.pngOther6.0 MB289d78555aef
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics1.pngOther101.3 KBc497108a8046
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics2.pngOther57.0 KB070afda83246
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics3.pngOther45.0 KBf78b7212a6ba
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics4.pngOther87.3 KB1333e99377b7
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics5.pngOther111.9 KB6a8b043bee0f
Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/stereo_vision_sfm.pngOther308.2 KB0186dc2a0665
MobileVIt - Every Picture Sees Every Block.pngOther188.6 KB3adae6df8101
MobileViT-Architecture.pngOther144.6 KB2c5cb2e6f755
MobileViT-Attention.pngOther177.9 KBb8b81ed3f6a8
MobileViT-CNNPreformance.pngOther199.5 KBce8e3e65b26e
MobileViT-Inference.pngOther177.3 KB39ef53cfc0a9
MobileViT-MobileViTBlock.pngOther85.2 KB962ef8f5326f
MobileVit - Architecture.pngOther154.4 KB29994bc88d5b
MobileVit - DifferingAttensions.pngOther177.9 KBb8b81ed3f6a8
MobileVit - Generalization Graph.pngOther152.4 KB1e9a0ae89c65
MobileVit - MobileVit vs CNNs.pngOther199.5 KBce8e3e65b26e
MobileVit -SelfSeberable.pngOther177.3 KB39ef53cfc0a9
Multimodal_Based_Video_Models/Modality_example.jpgOther342.4 KB7b083f7866e9
Multimodal_Based_Video_Models/Modality_example.pdfOther27.7 KB
Multimodal_Based_Video_Models/Overview_ImageBind.pngOther1.0 MB7a9333265cf1
Multimodal_Based_Video_Models/Overview_VATT.pngOther236.4 KBc05af7d67077
Multimodal_Based_Video_Models/Overview_VideoBERT.pngOther322.0 KB4b19183ac70b
OWL_ViT_example.jpgOther66.8 KB3bc682c31f9e
Object_Detection.pngOther387.0 KB87be2e5a3188
Pale_Blue_Dot.pngOther538.8 KB68d0f037a448
Parallel Representation.pngOther10.7 KB9af76813b50a
Photo_51_x-ray_diffraction_image.jpgOther16.9 KB9af6ab4fae02
Pinhole-camera.pngOther51.8 KBc7364f978e88
Pinhole_transform.pngOther54.3 KBa90d5edffc62
PixelNeRF_input.pngOther2.2 KBe0f2f720fb2c
PixelNeRF_output.gifOther46.9 KB6991cbb9ba7e
PixelNeRF_pipeline.pngOther167.5 KB0bca19fefdbd
RCNN.pngOther148.3 KB80f7a92e6efc
Recurrent Representation.pngOther21.4 KBbe8dee7b50da
ResnetBlock.pngOther37.8 KBa29e210b63b8
Retention_imgs/Chunkwise Representation.jpgOther47.1 KBf0bce516aca5
Retention_imgs/Chunkwise Representation.pngOther119.4 KB6e1becfc2f1d
Retention_imgs/Decay Matrix.pngOther61.0 KB3d5690cc6ddf
Retention_imgs/Multi Scale Retention.jpgOther102.8 KBb963ec821cee
Retention_imgs/Parallel Representation.pngOther10.7 KB9af76813b50a
Retention_imgs/Positional Embedding.jpgOther17.3 KB29a6755ed9eb
Retention_imgs/Positional Embedding.pngOther45.1 KB53593888c339
Retention_imgs/Recurrent Representation.pngOther21.4 KBbe8dee7b50da
Retention_imgs/RetNet Architecture.jpgOther58.6 KB3435ff46437a
Retention_imgs/Retention Block.jpgOther51.2 KBf6b0a00e0100
Screenshot from 2024-12-27 14-25-49.pngOther123.2 KB81cd51c440e2
ViR.pngOther178.9 KB1a4b63537dc2
Zero123.pngOther2.9 MB082094a2e04b
accuracy_vs_latency.pngOther21.0 KBc6b89e7aa412
ai_policy.pngOther193.7 KB307712a3e62b
apple.jpgOther1.5 MB9d9f312c4d75
application_02.pngOther425.4 KB06f9ee79b415
asilomar-ai.pngOther142.4 KB3e16b95f2e7e
axes_handedness.pngOther678.2 KBab8c88994891
balls.pngOther847.1 KB80e39cb93d76
block_comparison.pngOther99.1 KBbe8d7f3eba95
camera_axes.pngOther653.1 KB308c4eb59d12
cas-9-image.mp4Other17.2 MB82d1647e2461
cas_9.gifOther630.6 KBae97335d6c36
cat.mp4Other108.1 KB
cat_kiss.gifOther1.6 MB42a72e5e65d3
clip_paper.pngOther85.6 KBed5d8e87a0e8
combined.pngOther82.0 KB769bd18bbcce
contrastive_learning.pngOther2.3 MB93e3f017a962
cropped_motivation.pngOther1.3 MB2c7596e01d0e
cropped_vision.pngOther1.3 MBdcb2419fba7c
cv-thumbnail.pngOther219.2 KB923bc7467b21
cvt_architecture.pngOther399.2 KB1c1e3deefaf6
cvt_conv_proj.pngOther360.5 KB95a1f84956a9
cycleGAN1.jpgOther6.4 MBc162202e8614
dcgan_training_animation.gifOther1.3 MBef4fa09483a9
depthwise_moveup.pngOther29.7 KB72360d471264
dice-coefficient.pngOther16.8 KB1cef71a2355c
dinat_images/dina_comparison.pngOther1.1 MB0f022a5dfd24
dinat_images/dinat_architecture.pngOther219.4 KB99e6b5e42753
dinat_images/text_gen.svgOther241.9 KB
ethical_considerations.pngOther275.8 KB13dd97abde49
ethics_bias/IR Tweet.jpgOther69.7 KBdb153fd9bf09
ethics_bias/balancing.pngOther5.8 MB9ef27292c288
ethics_bias/crime offendor.jpgOther100.6 KBe28f337684d9
ethics_bias/crime suspect.jpgOther59.4 KBdb27b09f0cd8
ethics_bias/ethnic slur.jpgOther121.0 KB79e2dc5145ca
ethics_bias/goodperson.jpgOther80.3 KB8c1870b2d3d3
ethics_bias/imagenet.pngOther181.0 KBb02b531c663c
fairness.pngOther291.4 KBf884a0abdd02
feature-extraction-feature-matching/FLANN.pngOther219.4 KB992e012eeee9
feature-extraction-feature-matching/Flow-Chart-for-SURF-Feature-Detection.pngOther21.9 KBaf86ada1ab52
feature-extraction-feature-matching/LoFTR.pngOther2.0 MB4fe3c890a9f2
feature-extraction-feature-matching/Original-SIFT-algorithm-flow.pngOther58.6 KB7bdaaa1d2bc2
feature-extraction-feature-matching/SIFT.pngOther194.2 KBf8866ecd7327
feature-extraction-feature-matching/image1.jpgOther4.0 MB9805a824b5bf
feature-extraction-feature-matching/image2.jpgOther3.2 MB8e07f4f85439
fish-embryo.mp4Other46.0 MBdbf928abd4e1
fish.gifOther827.4 KBab874928705c
generative_models/GAN.pngOther157.4 KB5e0ad1c21329
generative_models/autoencoder.pngOther137.3 KB07726792bd2f
generative_models/bedroom.pngOther132.5 KB0a8eab5e6aa8
generative_models/comparison.pngOther95.0 KB5b7721710711
generative_models/vae.pngOther274.9 KB275addf0050d
good_images.pngOther439.5 KB79ec4ee030b6
googlenet_architecture.pngOther74.9 KB68641e3323e9
googlenet_auxiliary_classifier.jpgOther57.2 KBf26182548e81
hiera_images/CNN_architecture.webpOther332.4 KB8bda0b6bee1c
hiera_images/hiera_architecture.pngOther407.6 KB1541de80af09
hiera_images/hiera_changes.pngOther136.2 KBd211be505650
hiera_images/mae.pngOther390.3 KB468cb0ed6bcd
hiera_images/mvitv2.pngOther34.0 KBef3bf21e58cd
horsegif_0.gifOther195.9 KB37061e52c2a9
human_spectrum.jpgOther177.9 KBb963459fd247
i-jepa-1.pngOther51.4 KBcdb69c770d60
i-jepa-2.pngOther101.1 KB3a8e6b432710
inception-resnet.pngOther25.5 KBc1dd9ca9fafa
inception_naive.pngOther30.3 KB062c43ae2ccd
inception_reduced.pngOther68.6 KB0b29dd91a042
inverted_bottleneck.pngOther29.9 KB679725e78655
iou.pngOther13.7 KB1eeee42991b7
knowledge_distillation.pngOther166.6 KB2f3f5f874920
label_dataset_owlv2/0.jpegOther97.4 KB94839acf3183
label_dataset_owlv2/1.jpegOther82.6 KBf178b3940e9a
label_dataset_owlv2/10.jpegOther128.9 KB00669ff99859
label_dataset_owlv2/11.jpegOther176.4 KB81dc37e2871f
label_dataset_owlv2/2.jpegOther122.4 KB0f6fcd02a8ae
label_dataset_owlv2/3.jpegOther167.8 KBee86e3a739d3
label_dataset_owlv2/4.jpegOther156.5 KBa62b60693df7
label_dataset_owlv2/5.jpegOther195.8 KB47d8fc7c03cd
label_dataset_owlv2/6.jpegOther146.7 KB527536dd8ad9
label_dataset_owlv2/7.jpegOther175.3 KBe6016ee91733
label_dataset_owlv2/8.jpegOther173.0 KB8e0334859b33
label_dataset_owlv2/9.jpegOther165.4 KB138f17e49cd9
model_size_vs_accuracy.pngOther18.9 KBa7eb0933cf9a
motivation_01.pngOther3.5 MB47ec0b9b50c1
motivation_02.pngOther3.4 MB67f95e32130e
motivation_03.gifOther585.3 KB720d76e6b6df
motivation_04.pngOther580.7 KB162c1089cd2c
multimodal_fusion_text_vision/Multimodal.jpgOther40.0 KB632298d5f7b2
multimodal_fusion_text_vision/bigbang.jpgOther54.1 KB37ae9115e298
multimodal_fusion_text_vision/cats_remote.jpgOther173.1 KBdea9e7ef9738
multimodal_fusion_text_vision/doc_vqa.jpgOther157.3 KB3ddb9fdbf888
multimodal_fusion_text_vision/image_captioning.pngOther820.0 KB305e6c22de6c
multimodal_fusion_text_vision/image_text_retrieval.pngOther785.2 KB1c17fc620575
multimodal_fusion_text_vision/multimodal_elephant.pngOther167.8 KBd541378a986b
multimodal_fusion_text_vision/text_image_generation.pngOther1.1 MBa88cf1cb477d
multimodal_fusion_text_vision/visual_grounding.jpgOther120.2 KBeb33d3ccef4a
multimodal_fusion_text_vision/vlp_framework.pngOther423.4 KB2a1b66a62690
multimodal_fusion_text_vision/vqa_visual_reasoning.pngOther801.5 KB23fd3a4136ce
multimodal_trasnsfer_learning_images/transfer_learning_dark.svgOther8.4 KB
multimodal_trasnsfer_learning_images/transfer_learning_light.pngOther119.9 KB73e08723921f
multiple-choice-question.pngOther11.5 KBd9cec06b2445
nerf_encodings.pngOther92.9 KB70bc94b3cf37
nerf_pipeline.pngOther597.6 KB1eb40a1723da
nerf_ray_visualisation.pngOther199.0 KB54aad310a171
object-detection-gif.gifOther594.7 KB3539a8ffdfba
object_detection_train_image_with_annotation_plots.pngOther2.0 MBd6b9e4363f0b
object_detection_wiki.pngOther826.2 KBd3ce58639268
oneformer/oneformer.svgOther612.8 KB
oneformer/oneformer_panotic.pngOther293.1 KB69579eaa26e9
oneformer/oneformer_semantic.pngOther339.6 KBc6fce3db6ad6
oneformer/plots.svgOther471.7 KB
oneformer/teaser.svgOther1.7 MB
oneformer/text_gen.svgOther241.9 KB
outlook_hyena_images/hyena-order2-schema.pngOther211.1 KB8d04c1adf770
outlook_hyena_images/hyena_mechanism.pngOther41.5 KBa63046a5004e
outlook_hyena_images/hyena_recurence.pngOther192.0 KB5e43894aa08d
outlook_hyena_images/hyena_vision_benchmarks.pngOther95.1 KBbab9c934e4a2
outlook_hyena_images/nd_hyena.pngOther151.4 KB12eb61ce5538
outlook_hyena_images/self-attention-schema.pngOther97.5 KB8831615f7ac5
outlook_hyena_images/transformer2hyena.pngOther200.9 KB445bfd6cbea2
outlook_hyena_images/vit_vs_hyenavit.pngOther17.9 KB81f975818e7f
paired_images.pngOther300.4 KBc26609e86ce8
pic_1.pngOther1.9 MBd617d49474e6
pic_2.pngOther476.7 KBa3db68ba1845
pic_3.pngOther188.9 KBad4d923586fb
pic_4.pngOther686.6 KB55b11ce4fe44
pic_5.pngOther188.7 KB609d424a6cb0
pic_6.pngOther2.0 MB243fc8199de5
pic_7.pngOther6.7 KBb4c4a617356d
pic_8.jpgOther629.8 KB2054e1ae15e0
pic_9.pngOther744.9 KB27fe0cd9a99f
pixel-accuracy.pngOther147.3 KBc01f2c2682a0
point_cloud_example.jpegOther10.4 KBb83773c7a390
previous sota models/SOTA Models 3D vs (2+1)D convolution..pngOther175.1 KBde7a4264f38c
previous sota models/SOTA Models Residual block. Shortcut connections bypass a signal from the top of the block to the tail. Signals are summed at the tail..pngOther345.9 KB75f6fe570502
previous sota models/SOTA Models Two-Stream architecture for video classification.pngOther600.3 KBa515431b8736
pruning.pngOther113.8 KB292f795e59ef
resnext_same_topology.pngOther15.1 KBf49deb5dee3b
rnn_video_models/attention_model.pngOther214.7 KB6dd2b67d4ea0
rnn_video_models/convlstm.pngOther82.9 KBb3b5bed99ce0
rnn_video_models/lrcn.pngOther346.4 KB2eab8db075cf
rnn_video_models/rnn.pngOther83.6 KBc518c34988f9
rotation.pngOther84.6 KB7677f3207af9
sag-event-image.jpgOther70.1 KBbd2277338383
sample_od_test_set_inference_output.pngOther1.7 MBb284bc332a54
scaling.pngOther88.8 KB04617e990b01
segmentation-example.pngOther412.3 KB7ef67f4fdd91
segmentation-types.pngOther504.3 KBeaf5e9f41790
self_driving_gif.gifOther77.4 MB273161425520
swin_transformer_architecture.pngOther310.2 KBa21e8a6fcd4b
synthetic-dat-creation-PBR:apple.jpgOther1.8 MBf651fd61ee3f
synthetic-data-creation-PBR/colors.pngOther45.3 KB14e51fc52905
synthetic-data-creation-PBR/depth.pngOther13.6 KBd3197a8b8038
synthetic-data-creation-PBR/normals.pngOther24.4 KB0bf3c8a2b2e0
synthetic-data-creation-PBR/rendered_elephant.pngOther508.5 KB434db5ecd465
synthetic-data-creation-diffusion-models/custom_diffusion.pngOther686.4 KB82f71f1ff57c
synthetic-data-creation-diffusion-models/denoising.jpgOther566.8 KB887181309541
synthetic-data-creation-diffusion-models/dreambooth.pngOther576.4 KBf08fac6ba227
synthetic-data-creation-diffusion-models/general-workflow.pngOther255.2 KB58d5aceffadf
synthetic-data-creation-diffusion-models/noising.jpgOther138.1 KBc796ace753f1
synthetic-data-creation-diffusion-models/stable-diffusion-workflow.pngOther254.5 KB768a5045d12d
synthetic-data-creation-diffusion-models/textual_inversion.pngOther331.3 KB335c6440c859
synthetic-data-creation-overfit/good-variety.jpgOther201.5 KB80eef3df25b1
synthetic-data-creation-overfit/overfit-background.jpgOther169.8 KBf2c9ddf120b8
synthetic-data-creation-overfit/overfit-color.jpgOther213.3 KBebda9404971b
synthetic-data-creation-overfit/overfit-location.jpgOther174.6 KB28ce0e951129
synthetic-data-creation-overfit/overfit-size.jpgOther189.5 KB14f1655b8354
tasks_comics.pngOther29.6 KBdb72921f1a38
teaser_static.jpgOther617.3 KB66dbab88a6cf
test-helmet-object-detection.jpgOther75.3 KB3fbfac3e9b3f
test_input_for_od.pngOther395.3 KBb764134a3b66
test_output_for_od.pngOther391.4 KB222a7c427629
transfer_learning.pngOther1.3 MB37405b89785f
transferlearning_vgg19_plot.pngOther46.4 KB8a724fc3f8d3
transformer_based_video_model/unit7_10_timesformer.JPGOther50.6 KB
transformer_based_video_model/unit7_1_vit_architecture.pngOther42.4 KB
transformer_based_video_model/unit7_2_vit_performance.JPGOther34.9 KB
transformer_based_video_model/unit7_3_vivit_architecture.pngOther515.4 KB
transformer_based_video_model/unit7_4_uniform_sampling_1JPG.JPGOther11.5 KB
transformer_based_video_model/unit7_5_tubelet_embedding.JPGOther14.5 KB
transformer_based_video_model/unit7_6_vivit_model2.JPGOther24.8 KB
transformer_based_video_model/unit7_7_vivit_model3.JPGOther18.3 KB
transformer_based_video_model/unit7_8_vivit_model4.JPGOther20.6 KB
transformer_based_video_model/unit7_9_vivit_performance.JPGOther20.3 KB
translation.pngOther85.1 KBd05272656f07
unit3-chapter-3-segmentation-maskformer.pngOther179.8 KBa33f0210d251
unit7 CNN based model/Efficient Video Models X3D (Expanded 3D Networks).pngOther89.4 KBc523bd86de77
unit7 CNN based model/Real-time Video Processing ST-GCN (Spatial-Temporal Graph Convolutional Networks).pngOther244.4 KB355f8aa36bd4
unit7 CNN based model/Self-Supervised Learning_MoCo.pngOther25.8 KBa7de4b8a1182
unpaired_images.pngOther174.2 KB2bdaacada943
vit_architecture.jpgOther177.4 KB142f1b9c0744
winogrand_paper.pngOther181.3 KBad97f0a1c006
yolo_evolution.pngOther94.4 KB04236a3aa25e
yolov1_arch.pngOther65.5 KB986c8989eae3
.gitattributesRepository2.4 KB
Multimodal_Based_Video_Models/.DS_StoreRepository6.1 KB

License and Download

License
apache-2.0
Access
No access gate
Download from Hugging Face for Computer Vision

Released by Hugging Face for Computer Vision through its official repository on Hugging Face. Read the license.