Adaptive Interaction Modeling via Graph Operations Search
Haoxin Li, Wei-Shi Zheng, Yu Tao, Haifeng Hu, Jian-Huang Lai
摘要
Interaction modeling is important for video action analysis. Recently, several works design specific structures to model interactions in videos. However, their structures are manually designed and non-adaptive, which require structures design efforts and more importantly could not model interactions adaptively. In this paper, we automate the process of structures design to learn adaptive structures for interaction modeling. We propose to search the network structures with differentiable architecture search mechanism, which learns to construct adaptive structures for different videos to facilitate adaptive interaction modeling. To this end, we first design the search space with several basic graph operations that explicitly capture different relations in videos. We experimentally demonstrate that our architecture search framework learns to construct adaptive interaction modeling structures, which provides more understanding about the relations between the structures and some interaction characteristics, and also releases the requirement of structures design efforts. Additionally, we show that the designed basic graph operations in the search space are able to model different interactions in videos. The experiments on two interaction datasets show that our method achieves competitive performance with state-of-the-arts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Diversifying Spatial-Temporal Perception for Video Domain GeneralizationKun-Yu Lin, Jia-Run Du, Yipeng Gao, Jiaming Zhou 等NeurIPS 2023 · 被引用 27 次
- Optimization Planning for 3D ConvNetsZhaofan Qiu, Ting Yao, Chong-Wah Ngo, Tao MeiICML 2021 · 被引用 9 次
- Representing Videos As Discriminative Sub-Graphs for Action RecognitionDong Li, Zhaofan Qiu, Yingwei Pan, Ting Yao 等CVPR 2021
它引用的顶会 Paper7
- TSM: Temporal Shift Module for Efficient Video UnderstandingJi Lin, Chuang Gan, Song HanICCV 2019 · 被引用 2,049 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- Video Classification With Channel-Separated Convolutional NetworksDu Tran, Heng Wang, Matt Feiszli, Lorenzo TorresaniICCV 2019 · 被引用 647 次
- STM: SpatioTemporal and Motion Encoding for Action RecognitionBoyuan Jiang, Mengmeng Wang, Weihao Gan, Wei Wu 等ICCV 2019 · 被引用 442 次
- Grouped Spatial-Temporal Aggregation for Efficient Action RecognitionChenxu Luo, Alan L. YuilleICCV 2019 · 被引用 170 次
相关 Paper
- Graph Differentiable Architecture Search with Structure LearningYijian Qin, Xin Wang, Zeyang Zhang, Wenwu ZhuNeurIPS 2021 · 被引用 52 次
- Dynamic Heterogeneous Graph Attention Neural Architecture SearchZeyang Zhang, Ziwei Zhang, Xin Wang, Yijian Qin 等AAAI 2023 · 被引用 44 次
- AutoAC: Towards Automated Attribute Completion for Heterogeneous Graph Neural NetworkGuanghui Zhu, Zhennan Zhu, Wenjie Wang, Zhuoer Xu 等ICDE 2023 · 被引用 14 次
- AutoGSR: Neural Architecture Search for Graph-based Session RecommendationJingfan Chen, Guanghui Zhu, Haojun Hou, Chunfeng Yuan 等SIGIR 2022 · 被引用 25 次
- NAS-CTR: Efficient Neural Architecture Search for Click-Through Rate PredictionGuanghui Zhu, Feng Cheng, Defu Lian, Chunfeng Yuan 等SIGIR 2022 · 被引用 9 次
