AViD Dataset: Anonymized Videos from Diverse Countries
A. J. Piergiovanni, Michael S. Ryoo
Abstract
We introduce a new public video dataset for action recognition: Anonymized Videos from Diverse countries (AViD). Unlike existing public video datasets, AViD is a collection of action videos from many different countries. The motivation is to create a public dataset that would benefit training and pretraining of action recognition models for everybody, rather than making it useful for limited countries. Further, all the face identities in the AViD videos are properly anonymized to protect their privacy. It also is a static dataset where each video is licensed with the creative commons license. We confirm that most of the existing video datasets are statistically biased to only capture action videos from a limited number of countries. We experimentally illustrate that models trained with such biased datasets do not transfer perfectly to action videos from the other countries, and show that AViD addresses such problem. We also confirm that the new AViD dataset could serve as a good dataset for pretraining the models, performing comparably or better than prior datasets 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c29c9d43-c877-4553-a897-c418c6841d17Cited by top-tier papers3
- TokenLearner: Adaptive Space-Time Tokenization for VideosMichael S. Ryoo, A. J. Piergiovanni, Anurag Arnab, Mostafa Dehghani et al.NeurIPS 2021 · 274 citations
- A Study of Face Obfuscation in ImageNetKaiyu Yang, Jacqueline H. Yau, Li Fei-Fei, Jia Deng et al.ICML 2022 · 163 citations
- DartBlur: Privacy Preservation with Detection Artifact SuppressionBaowei Jiang, Bing Bai, Haozhe Lin, Yu Wang et al.CVPR 2023
Builds on4
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video ClipsAntoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi et al.ICCV 2019 · 1,437 citations
- HACS: Human Action Clips and Segments Dataset for Recognition and Temporal LocalizationHang Zhao, Antonio Torralba, Lorenzo Torresani, Zhicheng YanICCV 2019 · 298 citations
- A Multigrid Method for Efficiently Training Video ModelsChao-Yuan Wu, Ross B. Girshick, Kaiming He, Christoph Feichtenhofer et al.CVPR 2020
Related papers
- Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled LearningZhi-Wei Xia, Kun-Yu Lin, Yuan-Ming Li, Wei-Jin Huang et al.ICCV 2025 · 2 citations
- Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-TrainingArun V. Reddy, William Paul, Corban Rivera, Ketul Shah et al.CVPR 2024 · 3 citations
- AutoLabel: CLIP-based framework for Open-Set Video Domain AdaptationGiacomo Zara, Subhankar Roy, Paolo Rota, Elisa RicciCVPR 2023
- Further Understanding Videos through Adverbs: A New Video TaskBo Pang, Kaiwen Zha, Yifan Zhang, Cewu LuAAAI 2020 · 18 citations
- CIAGAN: Conditional Identity Anonymization Generative Adversarial NetworksMaxim Maximov, Ismail Elezi, Laura Leal-TaixéCVPR 2020
