Frequency Decoupling for Motion Magnification Via Multi-Level Isomorphic Architecture
Fei Wang, Dan Guo, Kun Li, Zhun Zhong, Meng Wang
Abstract
Video Motion Magnification (VMM) aims to reveal subtle and imperceptible motion information of objects in the macroscopic world. Prior methods directly model the motion field from the Eulerian perspective by Representation Learning that separates shape and texture or Multi-domain Learning from phase fluctuations. Inspired by the frequency spectrum, we observe that the low-frequency components with stable energy always possess spatial structure and less noise, making them suitable for modeling the subtle motion field. To this end, we present FD4MM, a new paradigm of Frequency Decoupling for Motion Magnification with a Multi-level Isomorphic Architecture to capture multi-level high-frequency details and a stable low-frequency structure (motion field) in video space. Since high-frequency details and subtle motions are susceptible to information degradation due to their inherent subtlety and unavoidable external interference from noise, we carefully design Sparse High/Low-pass Filters to enhance the integrity of details and motion structures, and a Sparse Frequency Mixer to promote seamless recoupling. Besides, we innovatively design a contrastive regularization for this task to strengthen the model's ability to discriminate irrelevant features, reducing undesired motion magnification. Extensive experiments on both Real-world and Synthetic Datasets show that our FD4MM outperforms SOTA methods. Meanwhile, FD4MM reduces FLOPs by 1.63× and boosts inference speed by 1.68× than the latest method. Our code is available at https://github.com/Jiafei127/FD4MM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5fea5a5-db66-4b60-81bc-dfceab173969Cited by top-tier papers6
- Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action RecognitionJihao Gu, Kun Li, Fei Wang, Yanyan Wei et al.ACM MM 2025 · 23 citations
- MMAD: Multi-Label Micro-Action Detection in VideosKun Li, Pengyu Liu, Dan Guo, Fei Wang et al.ICCV 2025 · 21 citations
- Patch-level Sounding Object Tracking for Audio-Visual Question AnsweringZhangbin Li, Jinxing Zhou, Jing Zhang, Shengeng Tang et al.AAAI 2025 · 20 citations
- PhysDiff: Physiology-based Dynamicity Disentangled Diffusion Model for Remote Physiological MeasurementWei Qian, Gaoji Su, Dan Guo, Jinxing Zhou et al.AAAI 2025 · 14 citations
- OLMD: Orientation-aware Long-term Motion Decoupling for Continuous Sign Language RecognitionYiheng Yu, Sheng Liu, Yuan Feng, Min Xu et al.AAAI 2025 · 5 citations
Builds on24
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks With Octave ConvolutionYunpeng Chen, Haoqi Fan, Bing Xu, Zhicheng Yan et al.ICCV 2019 · 665 citations
- Fast Vision Transformers with HiLo AttentionZizheng Pan, Jianfei Cai, Bohan ZhuangNeurIPS 2022 · 321 citations
Related papers
- EulerMormer: Robust Eulerian Motion Magnification via Dynamic Filtering within TransformerFei Wang, Dan Guo, Kun Li, Meng WangAAAI 2024 · 49 citations
- 3D Motion Magnification: Visualizing Subtle Motions with Time-Varying Radiance FieldsBrandon Y. Feng, Hadi Alzayer, Michael Rubinstein, William T. Freeman et al.ICCV 2023 · 8 citations
- Frequency-aware Dynamic Gaussian SplattingQiaowei Miao, JinSheng Quan, Kehan Li, Yichao Xu et al.ICLR 2026
- DeMatch: Deep Decomposition of Motion Field for Two-View Correspondence LearningShihua Zhang, Zizhuo Li, Yuan Gao, Jiayi MaCVPR 2024 · 5 citations
- Mem4D: Decoupling Static and Dynamic Memory for Dynamic Scene ReconstructionXudong Cai, Shuo Wang, Peng Wang, Yongcai Wang et al.AAAI 2026 · 4 citations
