FourierHandFlow: Neural 4D Hand Representation Using Fourier Query Flow
Jihyun Lee, Junbong Jang, Donghwan Kim, Minhyuk Sung, Tae-Kyun Kim
Abstract
Recent 4D shape representations model continuous temporal evolution of implicit shapes by (1) learning query flows without leveraging shape and articulation priors or (2) decoding shape occupancies separately for each time value. Thus, they do not effectively capture implicit correspondences between articulated shapes or regularize jittery temporal deformations. In this work, we present FOURIER-HANDFLOW, which is a spatio-temporally continuous representation for human hands that combines a 3D occupancy field with articulation-aware query flows represented as Fourier series. Given an input RGB sequence, we aim to learn a fixed number of Fourier coefficients for each query flow to guarantee smooth and continuous temporal shape dynamics. To effectively model spatio-temporal deformations of articulated hands, we compose our 4D representation based on two types of Fourier query flow: (1) pose flow that models query dynamics influenced by hand articulation changes via implicit linear blend skinning and (2) shape flow that models query-wise displacement flow. In the experiments, our method achieves state-of-the-art results on video-based 4D reconstruction while being computationally more efficient than the existing 3D/4D implicit shape representations. We additionally show our results on motion inter-and extrapolation and texture transfer using the learned correspondences of implicit shapes. To the best of our knowledge, FOURIERHANDFLOW is the first neural 4D continuous hand representation learned from RGB videos. The code will be publicly accessible. Fou rier Que ry Flo w t Figure 1: From monocular RGB sequence inputs, FOURIERHANDFLOW learns 4D hand shapes that are continuous in both space and time. It models temporal shape evolutions with query flows learned as a fixed number of coefficients for Fourier series to guarantee smooth temporal dynamics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2704eb23-38ec-4f37-be00-2b97f57fef66Cited by top-tier papers9
- MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based DynamicsChangmin Lee, Jihyun Lee, Tae-Kyun KimNeurIPS 2025 · 9 citations
- Multi-hypotheses Conditioned Point Cloud Diffusion for 3D Human Reconstruction from Occluded ImagesDonghwan Kim, Tae-Kyun KimNeurIPS 2024 · 8 citations
- SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit MeshesMinje Kim, Tae-Kyun KimNeurIPS 2025 · 2 citations
- Diffusion-Based 3D Hand Motion Recovery with Intuitive PhysicsYufei Zhang, Zijun Cui, Jeffrey O. Kephart, Qiang JiICCV 2025 · 1 citation
- CanFields: Consolidating Diffeomorphic Flows for Non-Rigid 4D Interpolation From Arbitrary-Length SequencesMiaowei Wang, Changjian Li, Amir VaxmanICCV 2025 · 1 citation
Builds on20
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Exploiting Spatial-Temporal Relationships for 3D Pose Estimation via Graph Convolutional NetworksYujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai et al.ICCV 2019 · 504 citations
- Occupancy Flow: 4D Reconstruction by Learning Particle DynamicsMichael Niemeyer, Lars M. Mescheder, Michael Oechsle, Andreas GeigerICCV 2019 · 314 citations
- End-to-End Hand Mesh Recovery From a Monocular RGB ImageXiong Zhang, Qiang Li, Hong Mo, Wenbo Zhang et al.ICCV 2019 · 248 citations
- A-SDF: Learning Disentangled Signed Distance Functions for Articulated Shape RepresentationJiteng Mu, Weichao Qiu, Adam Kortylewski, Alan L. Yuille et al.ICCV 2021 · 138 citations
Related papers
- Learning Parallel Dense Correspondence From Spatio-Temporal Descriptors for Efficient and Robust 4D ReconstructionJiapeng Tang, Dan Xu, Kui Jia, Lei ZhangCVPR 2021
- Neural Radiance Flow for 4D View Synthesis and Video ProcessingYilun Du, Yinan Zhang, Hong-Xing Yu, Joshua B. Tenenbaum et al.ICCV 2021 · 329 citations
- FOF: Learning Fourier Occupancy Field for Monocular Real-time Human ReconstructionQiao Feng, Yebin Liu, Yu-Kun Lai, Jingyu Yang et al.NeurIPS 2022 · 52 citations
- The Structure-Equivalent Prior: Unifying Temporal Dynamics and 3D Evolution in 4D Latent SpaceJingyuan Gao, Tianyu Shen, Ruosen Hao, Te Guo et al.AAAI 2026
- 4RC: 4D Reconstruction via Conditional Querying Anytime and AnywhereYihang Luo, Shangchen Zhou, Yushi Lan, Xingang Pan et al.ICML 2026 · 12 citations
