Frequency Decoupling for Motion Magnification Via Multi-Level Isomorphic Architecture
Fei Wang, Dan Guo, Kun Li, Zhun Zhong, Meng Wang
摘要
Video Motion Magnification (VMM) aims to reveal subtle and imperceptible motion information of objects in the macroscopic world. Prior methods directly model the motion field from the Eulerian perspective by Representation Learning that separates shape and texture or Multi-domain Learning from phase fluctuations. Inspired by the frequency spectrum, we observe that the low-frequency components with stable energy always possess spatial structure and less noise, making them suitable for modeling the subtle motion field. To this end, we present FD4MM, a new paradigm of Frequency Decoupling for Motion Magnification with a Multi-level Isomorphic Architecture to capture multi-level high-frequency details and a stable low-frequency structure (motion field) in video space. Since high-frequency details and subtle motions are susceptible to information degradation due to their inherent subtlety and unavoidable external interference from noise, we carefully design Sparse High/Low-pass Filters to enhance the integrity of details and motion structures, and a Sparse Frequency Mixer to promote seamless recoupling. Besides, we innovatively design a contrastive regularization for this task to strengthen the model's ability to discriminate irrelevant features, reducing undesired motion magnification. Extensive experiments on both Real-world and Synthetic Datasets show that our FD4MM outperforms SOTA methods. Meanwhile, FD4MM reduces FLOPs by 1.63× and boosts inference speed by 1.68× than the latest method. Our code is available at https://github.com/Jiafei127/FD4MM .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action RecognitionJihao Gu, Kun Li, Fei Wang, Yanyan Wei 等ACM MM 2025 · 被引用 23 次
- MMAD: Multi-Label Micro-Action Detection in VideosKun Li, Pengyu Liu, Dan Guo, Fei Wang 等ICCV 2025 · 被引用 21 次
- Patch-level Sounding Object Tracking for Audio-Visual Question AnsweringZhangbin Li, Jinxing Zhou, Jing Zhang, Shengeng Tang 等AAAI 2025 · 被引用 20 次
- PhysDiff: Physiology-based Dynamicity Disentangled Diffusion Model for Remote Physiological MeasurementWei Qian, Gaoji Su, Dan Guo, Jinxing Zhou 等AAAI 2025 · 被引用 14 次
- OLMD: Orientation-aware Long-term Motion Decoupling for Continuous Sign Language RecognitionYiheng Yu, Sheng Liu, Yuan Feng, Min Xu 等AAAI 2025 · 被引用 5 次
它引用的顶会 Paper24
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks With Octave ConvolutionYunpeng Chen, Haoqi Fan, Bing Xu, Zhicheng Yan 等ICCV 2019 · 被引用 665 次
- Fast Vision Transformers with HiLo AttentionZizheng Pan, Jianfei Cai, Bohan ZhuangNeurIPS 2022 · 被引用 321 次
相关 Paper
- EulerMormer: Robust Eulerian Motion Magnification via Dynamic Filtering within TransformerFei Wang, Dan Guo, Kun Li, Meng WangAAAI 2024 · 被引用 49 次
- 3D Motion Magnification: Visualizing Subtle Motions with Time-Varying Radiance FieldsBrandon Y. Feng, Hadi Alzayer, Michael Rubinstein, William T. Freeman 等ICCV 2023 · 被引用 8 次
- Frequency-aware Dynamic Gaussian SplattingQiaowei Miao, JinSheng Quan, Kehan Li, Yichao Xu 等ICLR 2026
- DeMatch: Deep Decomposition of Motion Field for Two-View Correspondence LearningShihua Zhang, Zizhuo Li, Yuan Gao, Jiayi MaCVPR 2024 · 被引用 5 次
- Mem4D: Decoupling Static and Dynamic Memory for Dynamic Scene ReconstructionXudong Cai, Shuo Wang, Peng Wang, Yongcai Wang 等AAAI 2026 · 被引用 4 次
