YouTube-8M: A Large and Diverse Labeled Video Dataset for Video Understanding Research
8M Dataset Explore Download Workshop About Games Car Football Food Performance art Trailer String instrument Fashion Recipe Piano Video game Animation Animal Guitar Outdoor recreation Nature Cooking Drummer Action-adventure game Orchestra Road Drums Pet Dish Hair Basketball Transport Lighting iPhone Ball Fishing Choir Aircraft Highlight film Cuisine Train Bollywood Driving Truck Wedding Horse Snow Machine Radio-controlled model Soldier Electric guitar Pianist Skateboard Ice skating Pokémon Airplane Plant Sports game Dashcam Festival Building Hairstyle Album Engine Grand Theft Auto V Winter Tractor Cymbal Bird Gardening LEGO Amusement park Painting Track Agriculture Weight training The Walt Disney Company Gymnastics Call of Duty: Black Ops Call of Duty: Black Ops II Hockey Snare drum Violin Wheel Tire Radio-controlled aircraft Beach Basketball moves Motocross PlayStation 4 FIFA 15 Vegetable Airline Battlefield Eye Medicine Cat Home improvement Farm Samsung Galaxy Puppy Pokémon Bride
YouTube-8M: A Large and Diverse Labeled Video Dataset for Video Understanding Research 8M Dataset Explore Download Workshop 2019 2018 2017 About News Sep 4th, 2019: Released the MediaPipe YouTube-8M feature extractor which extracts both visual and audio features. Jun 27th, 2019: Released the YouTube-8M Segments dataset . May 14th, 2018: Released an update to the dataset, with improved quality machine-generated labels, and reduced size / higher-quality video dataset. ( YouTube-8M 2018 ). YouTube-8M Segments Dataset The YouTube-8M Segments dataset is an extension of the YouTube-8M dataset with h
Explore this link on the map →saved by
related reading
- Dataset list - A list of the biggest machine learning datasetsdatasetlist.com
- The First Fully General Computer Action Model | blogsi.inc
- Reintroducing Sievesieve.ai
- [2206.11795] Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videosarxiv.org
- Video models are zero-shot learners and reasonersarxiv.org
- upload videos to youtube so we can farm jmoon for high quality training data and sell it by anishthite · Pull Request #2 · thecooltechguy/smash-leaderboard-ai · GitHubgithub.com
- [2506.09985] V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planningarxiv.org
- [2206.11795] Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videosarxiv.org
- YouTubestudio.youtube.com
- Deep Temporal Reasoning in Video Language Models: A Cross-Linguistic Evaluation of Action Duration and Completion through Perfect Times - ACL Anthologyaclanthology.org
- how we accidentally solved robotics by watching 1 million hours of YouTube – atharva's blogksagar.bearblog.dev
- Update: Expanding access to Meta Segment Anything 2.1 on Amazon SageMaker JumpStartai.meta.com