EulerMormer: Robust Eulerian Motion Magnification via Dynamic Filtering within Transformer
Fei Wang, Dan Guo, Kun Li, Meng Wang
摘要
Video Motion Magnification (VMM) aims to break the resolution limit of human visual perception capability and reveal the imperceptible minor motion that contains valuable information in the macroscopic domain. However, challenges arise in this task due to photon noise inevitably introduced by photographic devices and spatial inconsistency in amplification, leading to flickering artifacts in static fields and motion blur and distortion in dynamic fields in the video. Existing methods focus on explicit motion modeling without emphasizing prioritized denoising during the motion magnification process. This paper proposes a novel dynamic filtering strategy to achieve static-dynamic field adaptive denoising. Specifically, based on Eulerian theory, we separate texture and shape to extract motion representation through inter-frame shape differences, expecting to leverage these subdivided features to solve this task finely. Then, we introduce a novel dynamic filter that eliminates noise cues and preserves critical features in the motion magnification and amplification generation phases. Overall, our unified framework, EulerMormer, is a pioneering effort to first equip with Transformer in learning-based VMM. The core of the dynamic filter lies in a global dynamic sparse cross-covariance attention mechanism that explicitly removes noise while preserving vital information, coupled with a multi-scale dual-path gating mechanism that selectively regulates the dependence on different frequency features to reduce spatial attenuation and complement motion boundaries. We demonstrate extensive experiments that EulerMormer achieves more robust video motion magnification from the Eulerian perspective, significantly outperforming state-of-the-art methods. The source code is available at https://github.com/VUT-HFUT/EulerMormer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action RecognitionJihao Gu, Kun Li, Fei Wang, Yanyan Wei 等ACM MM 2025 · 被引用 23 次
- MMAD: Multi-Label Micro-Action Detection in VideosKun Li, Pengyu Liu, Dan Guo, Fei Wang 等ICCV 2025 · 被引用 21 次
- Towards Efficient General Feature Prediction in Masked Skeleton ModelingShengkai Sun, Zefan Zhang, Jianfeng Dong, Zhiyong Cheng 等ICCV 2025 · 被引用 3 次
- Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational AdaptationFei Wang, Xinye Zheng, Kun Li, Yanyan Wei 等CVPR 2026 · 被引用 2 次
- Frequency Decoupling for Motion Magnification Via Multi-Level Isomorphic ArchitectureFei Wang, Dan Guo, Kun Li, Zhun Zhong 等CVPR 2024
它引用的顶会 Paper8
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- DeepRhythm: Exposing DeepFakes with Attentional Visual Heartbeat RhythmsHua Qi, Qing Guo, Felix Juefei-Xu, Xiaofei Xie 等ACM MM 2020 · 被引用 224 次
- Proposal-Free Video Grounding with Contextual Pyramid NetworkKun Li, Dan Guo, Meng WangAAAI 2021 · 被引用 138 次
- AnimeSR: Learning Real-World Super-Resolution Models for Animation VideosYanze Wu, Xintao Wang, Gen Li, Ying ShanNeurIPS 2022 · 被引用 46 次
- Gloss Semantic-Enhanced Network with Online Back-Translation for Sign Language ProductionShengeng Tang, Richang Hong, Dan Guo, Meng WangACM MM 2022 · 被引用 44 次
相关 Paper
- Bilateral Video Magnification FilterShoichiro Takeda, Kenta Niwa, Mariko Isogawa, Shinya Shimizu 等CVPR 2022 · 被引用 11 次
- 3D Motion Magnification: Visualizing Subtle Motions with Time-Varying Radiance FieldsBrandon Y. Feng, Hadi Alzayer, Michael Rubinstein, William T. Freeman 等ICCV 2023 · 被引用 8 次
- FMA-Net: Flow-Guided Dynamic Filtering and Iterative Feature Refinement with Multi-Attention for Joint Video Super-Resolution and DeblurringGeunhyuk Youk, Jihyong Oh, Munchurl KimCVPR 2024 · 被引用 16 次
- Video Frame Interpolation with Flow TransformerPan Gao, Haoyue Tian, Jie QinACM MM 2023 · 被引用 4 次
- MSTDiff: Multiscale-Aware Transformer Diffusion Network for Video Object DetectionQiang Qi, Wenqi Shang, Xiao Wang, Yanjie Liang 等AAAI 2026
