Advancing Video Synchronization with Fractional Frame Analysis: Introducing a Novel Dataset and Model
Yuxuan Liu, Haizhou Ai, Junliang Xing, Xuri Li, Xiaoyi Wang, Pin Tao
摘要
Multiple views play a vital role in 3D pose estimation tasks. Ideally, multi-view 3D pose estimation tasks should directly utilize naturally collected videos for pose estimation. However, due to the constraints of video synchronization, existing methods often use expensive hardware devices to synchronize the initiation of cameras, which restricts most 3D pose collection scenarios to indoor settings. Some recent works learn deep neural networks to align desynchronized datasets derived from synchronized cameras and can only produce frame-level accuracy. For fractional frame video synchronization, this work proposes an Inter-Frame and Intra-Frame Desynchronized Dataset (IFID), which labels fractional time intervals between two video clips. IFID is the first dataset that annotates inter-frame and intra-frame intervals, with a total of 382, 500 video clips annotated, making it the largest dataset to date. We also develop a novel model based on the Transformer architecture, named InSynFormer, for synchronizing inter-frame and intra-frame. Extensive experimental evaluations demonstrate its promising performance. The dataset and source code of the model are available at https: //github.com/yuxuan-cser/InSynFormer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Multi-View 3D Human Pose Estimation with Weakly Synchronized ImagesLing Li, Ruiwen Gu, Chongyang Wang, Junliang Xing 等AAAI 2025 · 被引用 3 次
- Visual Sync: Multi-Camera Synchronization via Cross-View Object MotionShaowei Liu, David Yifan Yao, Saurabh Gupta, Shenlong WangNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper8
- Optimizing Network Structure for 3D Human Pose EstimationHai Ci, Chunyu Wang, Xiaoxuan Ma, Yizhou WangICCV 2019 · 被引用 267 次
- Direct Multi-view Multi-person 3D Pose EstimationTao Wang, Jianfeng Zhang, Yujun Cai, Shuicheng Yan 等NeurIPS 2021 · 被引用 147 次
- 3D Human Pose Estimation Using Spatio-Temporal Networks with Explicit Occlusion TrainingYu Cheng, Bo Yang, Bo Wang, Robby T. TanAAAI 2020 · 被引用 145 次
- Graph and Temporal Convolutional Networks for 3D Multi-person Pose Estimation in Monocular VideosYu Cheng, Bo Wang, Bo Yang, Robby T. TanAAAI 2021 · 被引用 55 次
- Novel View Synthesis of Human Interactions from Sparse Multi-view VideosQing Shuai, Chen Geng, Qi Fang, Sida Peng 等SIGGRAPH 2022 · 被引用 43 次
相关 Paper
- FreeMan: Towards Benchmarking 3D Human Pose Estimation Under Real-World ConditionsJiong Wang, Fengyu Yang, Bingliang Li, Wenbo Gou 等CVPR 2024 · 被引用 8 次
- PoseSyn: Synthesizing Diverse 3D Pose Data from In-the-Wild 2D DataChangHee Yang, Hyeonseop Song, Seokhun Choi, Seungwoo Lee 等ICCV 2025 · 被引用 1 次
- Self-Supervised Human Pose based Multi-Camera Video SynchronizationLiqiang Yin, Ruize Han, Wei Feng, Song WangACM MM 2022 · 被引用 7 次
- Neural Video Depth StabilizerYiran Wang, Min Shi, Jiaqi Li, Zihao Huang 等ICCV 2023
- 3D Human Pose Estimation with Spatial and Temporal TransformersCe Zheng, Sijie Zhu, Matías Mendieta, Taojiannan Yang 等ICCV 2021 · 被引用 648 次
