YouTube-8M: A Large-Scale Video Classification Benchmark
Sami Abu-El-Haija Nisarg Kothari Joonseok Lee Paul Natsev George Toderici Balakrishnan Varadarajan Sudheendra Vijayanarasimhan Google Research
Abstract
Many recent advancements in Computer Vision are attributed to large datasets. Open-source software packages for Machine Learning and inexpensive commodity hardware have reduced the barrier of entry for exploring novel approaches at scale. It is possible to train models over millions of examples within a few days. Although large-scale datasets exist for image understanding, such as ImageNet, there are no comparable size video classification datasets.
原文 arXiv:1609.08675;中英对照 + 大白话阅读 https://aha.fim.ai/paper/1609.08675v1