Rows6
Configurations1
Size283.7 MB
Licenseapache-2.0
AccessPublicly accessible
Monthly Downloads124.9k
Dataset Card
The publisher has not written a card for this dataset.
Structure
default 6 rows
| Split | Rows | Size |
|---|---|---|
| train | 2 | 3.3 MB |
| test | 4 | 2.5 MB |
imageImage
Details
- Repository
- hf-vision/course-assets
- Publisher
- Hugging Face for Computer Vision
- Task category
- Not stated by the source
- Tags
- Not stated by the source
- Size category
- Not stated by the source
- Languages
- Not stated by the source
- Revision
- 5e88de973e1bda37a7d8a6b98449c47d309f8927
- Last updated
- 2025-01-24
Files
299 files, 283.7 MB in total.
Documentation1 file · 28 B
Other296 files · 283.7 MB
Repository2 files · 8.6 KB
Every file
| File | Type | Size | SHA-256 |
|---|---|---|---|
| README.md | Documentation | 28 B | — |
| 3d_stereo_vision_images/annotated_left_img.jpg | Other | 70.8 KB | 6f5b72ca5333 |
| 3d_stereo_vision_images/annotated_right_img.jpg | Other | 72.8 KB | 063a69d1fdfc |
| 3d_stereo_vision_images/calculated_dim_results.png | Other | 767.5 KB | 3db5e8120411 |
| 3d_stereo_vision_images/image_formation_simple_stereo.jpg | Other | 15.3 KB | 8b251a1d7dd1 |
| 3d_stereo_vision_images/image_formation_single_camera.png | Other | 232.2 KB | 1ac5c99954db |
| 3d_stereo_vision_images/rectified_left_frame.jpg | Other | 61.1 KB | 773848ae7d21 |
| 3d_stereo_vision_images/rectified_overlapping_frames.jpg | Other | 56.8 KB | 52fdffc900a4 |
| 3d_stereo_vision_images/rectified_right_frame.jpg | Other | 63.2 KB | f4094eb7b7a2 |
| 3d_stereo_vision_images/rectified_stacked_frames.jpg | Other | 129.8 KB | ef97be99a9d6 |
| 3d_stereo_vision_images/unrectified_left_frame.jpg | Other | 73.4 KB | d3398b79c513 |
| 3d_stereo_vision_images/unrectified_right_frame.jpg | Other | 69.7 KB | a08d62ee4a1e |
| 3d_stereo_vision_images/unrectified_stacked_frames.jpg | Other | 148.8 KB | 9b418b93d359 |
| Attention Comparison.png | Other | 879.8 KB | 6269933d799c |
| BLIP.png | Other | 427.6 KB | 270a2f640399 |
| Black_hole_-_Messier_87.jpg | Other | 38.9 KB | b8c821fb3312 |
| CNNs/1d_conv.jpg | Other | 95.6 KB | d86eeacb2245 |
| CNNs/1d_conv_multiplication.png | Other | 241.9 KB | 700a22b616e3 |
| CNNs/1d_conv_result.png | Other | 187.5 KB | 8f6048d39dd0 |
| CNNs/2d_conv.png | Other | 356.7 KB | 26b66af982f7 |
| CNNs/2d_conv_illustrated.png | Other | 580.4 KB | be27af80ae36 |
| CNNs/2d_conv_matrix.png | Other | 330.4 KB | 5951fc6e0c51 |
| CNNs/2d_feature_map.png | Other | 523.3 KB | e72da3d192b5 |
| CNNs/convolved_illustrated.png | Other | 382.8 KB | ccded4ede837 |
| CNNs/kernel_image.png | Other | 268.7 KB | 50b5847d1197 |
| CNNs/max_pooling.png | Other | 489.2 KB | 09900a81cd95 |
| CNNs/network_illustration.png | Other | 828.5 KB | 565b00e046b9 |
| CNNs/network_illustration_2.png | Other | 654.3 KB | 95c3a6f940e7 |
| CNNs/network_illustration_3.png | Other | 682.7 KB | 3faeb65c808f |
| CNNs/padding.png | Other | 497.4 KB | c5e5e71f6e6e |
| CNNs/prewitt_sobel.png | Other | 150.7 KB | 1598f2d5f69f |
| CV_in_defintiion.png | Other | 160.6 KB | c5b3525342b2 |
| CycleGAN2.jpg | Other | 6.3 MB | c4d67fd95160 |
| DETR.png | Other | 53.8 KB | 2f6a2f211196 |
| DETR_DecoderLayer.png | Other | 184.8 KB | 1d43e577febf |
| DETR_FeatureMaps.png | Other | 897.5 KB | 8f8982453944 |
| FPN.png | Other | 69.8 KB | ab4fbba54f71 |
| FPN_2.png | Other | 187.7 KB | f7e0b89f3b61 |
| Fast R-CNN.png | Other | 829.0 KB | 8608769a795c |
| Faster_RCNN.png | Other | 56.1 KB | 9c5458fd6d3f |
| KL-Loss.png | Other | 1.8 KB | 855bbed75ccb |
| LLM Challenges.png | Other | 52.4 KB | 20d36af73266 |
| MaSA vs Retention.png | Other | 50.5 KB | df3e03749911 |
| MaSAD.png | Other | 17.2 KB | 15bebc3357d2 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/DA2_pipeline.png | Other | 325.8 KB | 0013f8da953b |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/dataset.png | Other | 3.2 MB | 71193ef20757 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_ambiguity1.gif | Other | 2.9 MB | 78d3b5cc928f |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_ambiguity2.gif | Other | 7.6 MB | 61b75357a7a4 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution1.png | Other | 246.0 KB | 27815321d9d8 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution2.png | Other | 627.3 KB | cd9644e89867 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_estimation_evolution3.png | Other | 1.3 MB | 168babd6eac4 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/depth_holes.png | Other | 326.6 KB | 007c1f76cc0a |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/different_scales.png | Other | 1.2 MB | c685bd6852b1 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/dpt.png | Other | 140.6 KB | a217a50809cf |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/focal_lenght.png | Other | 681.1 KB | fb5c044ab24e |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/inference.png | Other | 6.0 MB | 289d78555aef |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics1.png | Other | 101.3 KB | c497108a8046 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics2.png | Other | 57.0 KB | 070afda83246 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics3.png | Other | 45.0 KB | f78b7212a6ba |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics4.png | Other | 87.3 KB | 1333e99377b7 |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/metrics5.png | Other | 111.9 KB | 6a8b043bee0f |
| Metric and Relative Monocular Depth Estimation An Overview. Fine-Tuning Depth Anything V2/stereo_vision_sfm.png | Other | 308.2 KB | 0186dc2a0665 |
| MobileVIt - Every Picture Sees Every Block.png | Other | 188.6 KB | 3adae6df8101 |
| MobileViT-Architecture.png | Other | 144.6 KB | 2c5cb2e6f755 |
| MobileViT-Attention.png | Other | 177.9 KB | b8b81ed3f6a8 |
| MobileViT-CNNPreformance.png | Other | 199.5 KB | ce8e3e65b26e |
| MobileViT-Inference.png | Other | 177.3 KB | 39ef53cfc0a9 |
| MobileViT-MobileViTBlock.png | Other | 85.2 KB | 962ef8f5326f |
| MobileVit - Architecture.png | Other | 154.4 KB | 29994bc88d5b |
| MobileVit - DifferingAttensions.png | Other | 177.9 KB | b8b81ed3f6a8 |
| MobileVit - Generalization Graph.png | Other | 152.4 KB | 1e9a0ae89c65 |
| MobileVit - MobileVit vs CNNs.png | Other | 199.5 KB | ce8e3e65b26e |
| MobileVit -SelfSeberable.png | Other | 177.3 KB | 39ef53cfc0a9 |
| Multimodal_Based_Video_Models/Modality_example.jpg | Other | 342.4 KB | 7b083f7866e9 |
| Multimodal_Based_Video_Models/Modality_example.pdf | Other | 27.7 KB | — |
| Multimodal_Based_Video_Models/Overview_ImageBind.png | Other | 1.0 MB | 7a9333265cf1 |
| Multimodal_Based_Video_Models/Overview_VATT.png | Other | 236.4 KB | c05af7d67077 |
| Multimodal_Based_Video_Models/Overview_VideoBERT.png | Other | 322.0 KB | 4b19183ac70b |
| OWL_ViT_example.jpg | Other | 66.8 KB | 3bc682c31f9e |
| Object_Detection.png | Other | 387.0 KB | 87be2e5a3188 |
| Pale_Blue_Dot.png | Other | 538.8 KB | 68d0f037a448 |
| Parallel Representation.png | Other | 10.7 KB | 9af76813b50a |
| Photo_51_x-ray_diffraction_image.jpg | Other | 16.9 KB | 9af6ab4fae02 |
| Pinhole-camera.png | Other | 51.8 KB | c7364f978e88 |
| Pinhole_transform.png | Other | 54.3 KB | a90d5edffc62 |
| PixelNeRF_input.png | Other | 2.2 KB | e0f2f720fb2c |
| PixelNeRF_output.gif | Other | 46.9 KB | 6991cbb9ba7e |
| PixelNeRF_pipeline.png | Other | 167.5 KB | 0bca19fefdbd |
| RCNN.png | Other | 148.3 KB | 80f7a92e6efc |
| Recurrent Representation.png | Other | 21.4 KB | be8dee7b50da |
| ResnetBlock.png | Other | 37.8 KB | a29e210b63b8 |
| Retention_imgs/Chunkwise Representation.jpg | Other | 47.1 KB | f0bce516aca5 |
| Retention_imgs/Chunkwise Representation.png | Other | 119.4 KB | 6e1becfc2f1d |
| Retention_imgs/Decay Matrix.png | Other | 61.0 KB | 3d5690cc6ddf |
| Retention_imgs/Multi Scale Retention.jpg | Other | 102.8 KB | b963ec821cee |
| Retention_imgs/Parallel Representation.png | Other | 10.7 KB | 9af76813b50a |
| Retention_imgs/Positional Embedding.jpg | Other | 17.3 KB | 29a6755ed9eb |
| Retention_imgs/Positional Embedding.png | Other | 45.1 KB | 53593888c339 |
| Retention_imgs/Recurrent Representation.png | Other | 21.4 KB | be8dee7b50da |
| Retention_imgs/RetNet Architecture.jpg | Other | 58.6 KB | 3435ff46437a |
| Retention_imgs/Retention Block.jpg | Other | 51.2 KB | f6b0a00e0100 |
| Screenshot from 2024-12-27 14-25-49.png | Other | 123.2 KB | 81cd51c440e2 |
| ViR.png | Other | 178.9 KB | 1a4b63537dc2 |
| Zero123.png | Other | 2.9 MB | 082094a2e04b |
| accuracy_vs_latency.png | Other | 21.0 KB | c6b89e7aa412 |
| ai_policy.png | Other | 193.7 KB | 307712a3e62b |
| apple.jpg | Other | 1.5 MB | 9d9f312c4d75 |
| application_02.png | Other | 425.4 KB | 06f9ee79b415 |
| asilomar-ai.png | Other | 142.4 KB | 3e16b95f2e7e |
| axes_handedness.png | Other | 678.2 KB | ab8c88994891 |
| balls.png | Other | 847.1 KB | 80e39cb93d76 |
| block_comparison.png | Other | 99.1 KB | be8d7f3eba95 |
| camera_axes.png | Other | 653.1 KB | 308c4eb59d12 |
| cas-9-image.mp4 | Other | 17.2 MB | 82d1647e2461 |
| cas_9.gif | Other | 630.6 KB | ae97335d6c36 |
| cat.mp4 | Other | 108.1 KB | — |
| cat_kiss.gif | Other | 1.6 MB | 42a72e5e65d3 |
| clip_paper.png | Other | 85.6 KB | ed5d8e87a0e8 |
| combined.png | Other | 82.0 KB | 769bd18bbcce |
| contrastive_learning.png | Other | 2.3 MB | 93e3f017a962 |
| cropped_motivation.png | Other | 1.3 MB | 2c7596e01d0e |
| cropped_vision.png | Other | 1.3 MB | dcb2419fba7c |
| cv-thumbnail.png | Other | 219.2 KB | 923bc7467b21 |
| cvt_architecture.png | Other | 399.2 KB | 1c1e3deefaf6 |
| cvt_conv_proj.png | Other | 360.5 KB | 95a1f84956a9 |
| cycleGAN1.jpg | Other | 6.4 MB | c162202e8614 |
| dcgan_training_animation.gif | Other | 1.3 MB | ef4fa09483a9 |
| depthwise_moveup.png | Other | 29.7 KB | 72360d471264 |
| dice-coefficient.png | Other | 16.8 KB | 1cef71a2355c |
| dinat_images/dina_comparison.png | Other | 1.1 MB | 0f022a5dfd24 |
| dinat_images/dinat_architecture.png | Other | 219.4 KB | 99e6b5e42753 |
| dinat_images/text_gen.svg | Other | 241.9 KB | — |
| ethical_considerations.png | Other | 275.8 KB | 13dd97abde49 |
| ethics_bias/IR Tweet.jpg | Other | 69.7 KB | db153fd9bf09 |
| ethics_bias/balancing.png | Other | 5.8 MB | 9ef27292c288 |
| ethics_bias/crime offendor.jpg | Other | 100.6 KB | e28f337684d9 |
| ethics_bias/crime suspect.jpg | Other | 59.4 KB | db27b09f0cd8 |
| ethics_bias/ethnic slur.jpg | Other | 121.0 KB | 79e2dc5145ca |
| ethics_bias/goodperson.jpg | Other | 80.3 KB | 8c1870b2d3d3 |
| ethics_bias/imagenet.png | Other | 181.0 KB | b02b531c663c |
| fairness.png | Other | 291.4 KB | f884a0abdd02 |
| feature-extraction-feature-matching/FLANN.png | Other | 219.4 KB | 992e012eeee9 |
| feature-extraction-feature-matching/Flow-Chart-for-SURF-Feature-Detection.png | Other | 21.9 KB | af86ada1ab52 |
| feature-extraction-feature-matching/LoFTR.png | Other | 2.0 MB | 4fe3c890a9f2 |
| feature-extraction-feature-matching/Original-SIFT-algorithm-flow.png | Other | 58.6 KB | 7bdaaa1d2bc2 |
| feature-extraction-feature-matching/SIFT.png | Other | 194.2 KB | f8866ecd7327 |
| feature-extraction-feature-matching/image1.jpg | Other | 4.0 MB | 9805a824b5bf |
| feature-extraction-feature-matching/image2.jpg | Other | 3.2 MB | 8e07f4f85439 |
| fish-embryo.mp4 | Other | 46.0 MB | dbf928abd4e1 |
| fish.gif | Other | 827.4 KB | ab874928705c |
| generative_models/GAN.png | Other | 157.4 KB | 5e0ad1c21329 |
| generative_models/autoencoder.png | Other | 137.3 KB | 07726792bd2f |
| generative_models/bedroom.png | Other | 132.5 KB | 0a8eab5e6aa8 |
| generative_models/comparison.png | Other | 95.0 KB | 5b7721710711 |
| generative_models/vae.png | Other | 274.9 KB | 275addf0050d |
| good_images.png | Other | 439.5 KB | 79ec4ee030b6 |
| googlenet_architecture.png | Other | 74.9 KB | 68641e3323e9 |
| googlenet_auxiliary_classifier.jpg | Other | 57.2 KB | f26182548e81 |
| hiera_images/CNN_architecture.webp | Other | 332.4 KB | 8bda0b6bee1c |
| hiera_images/hiera_architecture.png | Other | 407.6 KB | 1541de80af09 |
| hiera_images/hiera_changes.png | Other | 136.2 KB | d211be505650 |
| hiera_images/mae.png | Other | 390.3 KB | 468cb0ed6bcd |
| hiera_images/mvitv2.png | Other | 34.0 KB | ef3bf21e58cd |
| horsegif_0.gif | Other | 195.9 KB | 37061e52c2a9 |
| human_spectrum.jpg | Other | 177.9 KB | b963459fd247 |
| i-jepa-1.png | Other | 51.4 KB | cdb69c770d60 |
| i-jepa-2.png | Other | 101.1 KB | 3a8e6b432710 |
| inception-resnet.png | Other | 25.5 KB | c1dd9ca9fafa |
| inception_naive.png | Other | 30.3 KB | 062c43ae2ccd |
| inception_reduced.png | Other | 68.6 KB | 0b29dd91a042 |
| inverted_bottleneck.png | Other | 29.9 KB | 679725e78655 |
| iou.png | Other | 13.7 KB | 1eeee42991b7 |
| knowledge_distillation.png | Other | 166.6 KB | 2f3f5f874920 |
| label_dataset_owlv2/0.jpeg | Other | 97.4 KB | 94839acf3183 |
| label_dataset_owlv2/1.jpeg | Other | 82.6 KB | f178b3940e9a |
| label_dataset_owlv2/10.jpeg | Other | 128.9 KB | 00669ff99859 |
| label_dataset_owlv2/11.jpeg | Other | 176.4 KB | 81dc37e2871f |
| label_dataset_owlv2/2.jpeg | Other | 122.4 KB | 0f6fcd02a8ae |
| label_dataset_owlv2/3.jpeg | Other | 167.8 KB | ee86e3a739d3 |
| label_dataset_owlv2/4.jpeg | Other | 156.5 KB | a62b60693df7 |
| label_dataset_owlv2/5.jpeg | Other | 195.8 KB | 47d8fc7c03cd |
| label_dataset_owlv2/6.jpeg | Other | 146.7 KB | 527536dd8ad9 |
| label_dataset_owlv2/7.jpeg | Other | 175.3 KB | e6016ee91733 |
| label_dataset_owlv2/8.jpeg | Other | 173.0 KB | 8e0334859b33 |
| label_dataset_owlv2/9.jpeg | Other | 165.4 KB | 138f17e49cd9 |
| model_size_vs_accuracy.png | Other | 18.9 KB | a7eb0933cf9a |
| motivation_01.png | Other | 3.5 MB | 47ec0b9b50c1 |
| motivation_02.png | Other | 3.4 MB | 67f95e32130e |
| motivation_03.gif | Other | 585.3 KB | 720d76e6b6df |
| motivation_04.png | Other | 580.7 KB | 162c1089cd2c |
| multimodal_fusion_text_vision/Multimodal.jpg | Other | 40.0 KB | 632298d5f7b2 |
| multimodal_fusion_text_vision/bigbang.jpg | Other | 54.1 KB | 37ae9115e298 |
| multimodal_fusion_text_vision/cats_remote.jpg | Other | 173.1 KB | dea9e7ef9738 |
| multimodal_fusion_text_vision/doc_vqa.jpg | Other | 157.3 KB | 3ddb9fdbf888 |
| multimodal_fusion_text_vision/image_captioning.png | Other | 820.0 KB | 305e6c22de6c |
| multimodal_fusion_text_vision/image_text_retrieval.png | Other | 785.2 KB | 1c17fc620575 |
| multimodal_fusion_text_vision/multimodal_elephant.png | Other | 167.8 KB | d541378a986b |
| multimodal_fusion_text_vision/text_image_generation.png | Other | 1.1 MB | a88cf1cb477d |
| multimodal_fusion_text_vision/visual_grounding.jpg | Other | 120.2 KB | eb33d3ccef4a |
| multimodal_fusion_text_vision/vlp_framework.png | Other | 423.4 KB | 2a1b66a62690 |
| multimodal_fusion_text_vision/vqa_visual_reasoning.png | Other | 801.5 KB | 23fd3a4136ce |
| multimodal_trasnsfer_learning_images/transfer_learning_dark.svg | Other | 8.4 KB | — |
| multimodal_trasnsfer_learning_images/transfer_learning_light.png | Other | 119.9 KB | 73e08723921f |
| multiple-choice-question.png | Other | 11.5 KB | d9cec06b2445 |
| nerf_encodings.png | Other | 92.9 KB | 70bc94b3cf37 |
| nerf_pipeline.png | Other | 597.6 KB | 1eb40a1723da |
| nerf_ray_visualisation.png | Other | 199.0 KB | 54aad310a171 |
| object-detection-gif.gif | Other | 594.7 KB | 3539a8ffdfba |
| object_detection_train_image_with_annotation_plots.png | Other | 2.0 MB | d6b9e4363f0b |
| object_detection_wiki.png | Other | 826.2 KB | d3ce58639268 |
| oneformer/oneformer.svg | Other | 612.8 KB | — |
| oneformer/oneformer_panotic.png | Other | 293.1 KB | 69579eaa26e9 |
| oneformer/oneformer_semantic.png | Other | 339.6 KB | c6fce3db6ad6 |
| oneformer/plots.svg | Other | 471.7 KB | — |
| oneformer/teaser.svg | Other | 1.7 MB | — |
| oneformer/text_gen.svg | Other | 241.9 KB | — |
| outlook_hyena_images/hyena-order2-schema.png | Other | 211.1 KB | 8d04c1adf770 |
| outlook_hyena_images/hyena_mechanism.png | Other | 41.5 KB | a63046a5004e |
| outlook_hyena_images/hyena_recurence.png | Other | 192.0 KB | 5e43894aa08d |
| outlook_hyena_images/hyena_vision_benchmarks.png | Other | 95.1 KB | bab9c934e4a2 |
| outlook_hyena_images/nd_hyena.png | Other | 151.4 KB | 12eb61ce5538 |
| outlook_hyena_images/self-attention-schema.png | Other | 97.5 KB | 8831615f7ac5 |
| outlook_hyena_images/transformer2hyena.png | Other | 200.9 KB | 445bfd6cbea2 |
| outlook_hyena_images/vit_vs_hyenavit.png | Other | 17.9 KB | 81f975818e7f |
| paired_images.png | Other | 300.4 KB | c26609e86ce8 |
| pic_1.png | Other | 1.9 MB | d617d49474e6 |
| pic_2.png | Other | 476.7 KB | a3db68ba1845 |
| pic_3.png | Other | 188.9 KB | ad4d923586fb |
| pic_4.png | Other | 686.6 KB | 55b11ce4fe44 |
| pic_5.png | Other | 188.7 KB | 609d424a6cb0 |
| pic_6.png | Other | 2.0 MB | 243fc8199de5 |
| pic_7.png | Other | 6.7 KB | b4c4a617356d |
| pic_8.jpg | Other | 629.8 KB | 2054e1ae15e0 |
| pic_9.png | Other | 744.9 KB | 27fe0cd9a99f |
| pixel-accuracy.png | Other | 147.3 KB | c01f2c2682a0 |
| point_cloud_example.jpeg | Other | 10.4 KB | b83773c7a390 |
| previous sota models/SOTA Models 3D vs (2+1)D convolution..png | Other | 175.1 KB | de7a4264f38c |
| previous sota models/SOTA Models Residual block. Shortcut connections bypass a signal from the top of the block to the tail. Signals are summed at the tail..png | Other | 345.9 KB | 75f6fe570502 |
| previous sota models/SOTA Models Two-Stream architecture for video classification.png | Other | 600.3 KB | a515431b8736 |
| pruning.png | Other | 113.8 KB | 292f795e59ef |
| resnext_same_topology.png | Other | 15.1 KB | f49deb5dee3b |
| rnn_video_models/attention_model.png | Other | 214.7 KB | 6dd2b67d4ea0 |
| rnn_video_models/convlstm.png | Other | 82.9 KB | b3b5bed99ce0 |
| rnn_video_models/lrcn.png | Other | 346.4 KB | 2eab8db075cf |
| rnn_video_models/rnn.png | Other | 83.6 KB | c518c34988f9 |
| rotation.png | Other | 84.6 KB | 7677f3207af9 |
| sag-event-image.jpg | Other | 70.1 KB | bd2277338383 |
| sample_od_test_set_inference_output.png | Other | 1.7 MB | b284bc332a54 |
| scaling.png | Other | 88.8 KB | 04617e990b01 |
| segmentation-example.png | Other | 412.3 KB | 7ef67f4fdd91 |
| segmentation-types.png | Other | 504.3 KB | eaf5e9f41790 |
| self_driving_gif.gif | Other | 77.4 MB | 273161425520 |
| swin_transformer_architecture.png | Other | 310.2 KB | a21e8a6fcd4b |
| synthetic-dat-creation-PBR:apple.jpg | Other | 1.8 MB | f651fd61ee3f |
| synthetic-data-creation-PBR/colors.png | Other | 45.3 KB | 14e51fc52905 |
| synthetic-data-creation-PBR/depth.png | Other | 13.6 KB | d3197a8b8038 |
| synthetic-data-creation-PBR/normals.png | Other | 24.4 KB | 0bf3c8a2b2e0 |
| synthetic-data-creation-PBR/rendered_elephant.png | Other | 508.5 KB | 434db5ecd465 |
| synthetic-data-creation-diffusion-models/custom_diffusion.png | Other | 686.4 KB | 82f71f1ff57c |
| synthetic-data-creation-diffusion-models/denoising.jpg | Other | 566.8 KB | 887181309541 |
| synthetic-data-creation-diffusion-models/dreambooth.png | Other | 576.4 KB | f08fac6ba227 |
| synthetic-data-creation-diffusion-models/general-workflow.png | Other | 255.2 KB | 58d5aceffadf |
| synthetic-data-creation-diffusion-models/noising.jpg | Other | 138.1 KB | c796ace753f1 |
| synthetic-data-creation-diffusion-models/stable-diffusion-workflow.png | Other | 254.5 KB | 768a5045d12d |
| synthetic-data-creation-diffusion-models/textual_inversion.png | Other | 331.3 KB | 335c6440c859 |
| synthetic-data-creation-overfit/good-variety.jpg | Other | 201.5 KB | 80eef3df25b1 |
| synthetic-data-creation-overfit/overfit-background.jpg | Other | 169.8 KB | f2c9ddf120b8 |
| synthetic-data-creation-overfit/overfit-color.jpg | Other | 213.3 KB | ebda9404971b |
| synthetic-data-creation-overfit/overfit-location.jpg | Other | 174.6 KB | 28ce0e951129 |
| synthetic-data-creation-overfit/overfit-size.jpg | Other | 189.5 KB | 14f1655b8354 |
| tasks_comics.png | Other | 29.6 KB | db72921f1a38 |
| teaser_static.jpg | Other | 617.3 KB | 66dbab88a6cf |
| test-helmet-object-detection.jpg | Other | 75.3 KB | 3fbfac3e9b3f |
| test_input_for_od.png | Other | 395.3 KB | b764134a3b66 |
| test_output_for_od.png | Other | 391.4 KB | 222a7c427629 |
| transfer_learning.png | Other | 1.3 MB | 37405b89785f |
| transferlearning_vgg19_plot.png | Other | 46.4 KB | 8a724fc3f8d3 |
| transformer_based_video_model/unit7_10_timesformer.JPG | Other | 50.6 KB | — |
| transformer_based_video_model/unit7_1_vit_architecture.png | Other | 42.4 KB | — |
| transformer_based_video_model/unit7_2_vit_performance.JPG | Other | 34.9 KB | — |
| transformer_based_video_model/unit7_3_vivit_architecture.png | Other | 515.4 KB | — |
| transformer_based_video_model/unit7_4_uniform_sampling_1JPG.JPG | Other | 11.5 KB | — |
| transformer_based_video_model/unit7_5_tubelet_embedding.JPG | Other | 14.5 KB | — |
| transformer_based_video_model/unit7_6_vivit_model2.JPG | Other | 24.8 KB | — |
| transformer_based_video_model/unit7_7_vivit_model3.JPG | Other | 18.3 KB | — |
| transformer_based_video_model/unit7_8_vivit_model4.JPG | Other | 20.6 KB | — |
| transformer_based_video_model/unit7_9_vivit_performance.JPG | Other | 20.3 KB | — |
| translation.png | Other | 85.1 KB | d05272656f07 |
| unit3-chapter-3-segmentation-maskformer.png | Other | 179.8 KB | a33f0210d251 |
| unit7 CNN based model/Efficient Video Models X3D (Expanded 3D Networks).png | Other | 89.4 KB | c523bd86de77 |
| unit7 CNN based model/Real-time Video Processing ST-GCN (Spatial-Temporal Graph Convolutional Networks).png | Other | 244.4 KB | 355f8aa36bd4 |
| unit7 CNN based model/Self-Supervised Learning_MoCo.png | Other | 25.8 KB | a7de4b8a1182 |
| unpaired_images.png | Other | 174.2 KB | 2bdaacada943 |
| vit_architecture.jpg | Other | 177.4 KB | 142f1b9c0744 |
| winogrand_paper.png | Other | 181.3 KB | ad97f0a1c006 |
| yolo_evolution.png | Other | 94.4 KB | 04236a3aa25e |
| yolov1_arch.png | Other | 65.5 KB | 986c8989eae3 |
| .gitattributes | Repository | 2.4 KB | — |
| Multimodal_Based_Video_Models/.DS_Store | Repository | 6.1 KB | — |
License and Download
- License
- apache-2.0
- Access
- No access gate
Download from Hugging Face for Computer Vision
Released by Hugging Face for Computer Vision through its official repository on Hugging Face. Read the license.