RIFTCast: A Template-Free End-to-End Multi-View Live Telepresence Framework and Benchmark
Domenic Zingsheim, Markus Plack, Hannah Dröge, Janelle Pfeifer, Patrick Stotko, Matthias B. Hullin, Reinhard Klein
Abstract
Immersive telepresence aims to authentically reproduce remote physical scenes, enabling the experience of real-world places, objects and people over large geographic distances. This requires the ability to generate realistic novel views of the scene with low latency. Existing methods either depend on depth data from specialized hardware setups or precomputed templates such as human models, which severely restrict their practicality and generalization to diverse scenes. To address these challenges, we introduce RIFTCast, a real-time template-free volumetric reconstruction framework that synthesizes high-fidelity dynamic scenes from a multi-view RGB-only capture setup. The framework is specifically targeted at the efficient reconstruction, transmission and visualization of complex scenes, including extensive human-human and human-object interactions. For this purpose, our method leverages a GPU-accelerated client-server pipeline that computes a visual hull representation to select a suitable subset of images for novel view synthesis, substantially reducing bandwidth and computation demands. This lightweight architecture enables deployment from small-scale configurations to sophisticated multi-camera capture stages, achieving low-latency telepresence even on resource-constrained devices. For evaluation, we provide a comprehensive high-quality multi-view video data benchmark as well as our reconstruction and rendering code, including tools for loading and processing a variety of data input formats, to facilitate future telepresence research.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 96d3a121-a4c8-46ea-aa38-d6f12deafcefRelated papers
- Volumetric Avatar Reconstruction with Spatio-Temporally Offset RGBD CamerasGareth Rendle, Adrian Kreskowski, Bernd FröhlichIEEE VR 2023 · 6 citations
- SceneHub4D: A Dataset and Evaluation Framework for 6-DoF 4D VR ScenesJaehong Kim, Tao Jin, Mallesham Dasari, Srinivasan Seshan et al.IEEE VR 2026 · 1 citation
- MetaStream: Live Volumetric Content Capture, Creation, Delivery, and Rendering in Real TimeYongjie Guan, Xueyu Hou, Nan Wu, Bo Han et al.MobiCom 2023 · 47 citations
- Holoported Characters: Real-Time Free-Viewpoint Rendering of Humans from Sparse RGB CamerasAshwath Shetty, Marc Habermann, Guoxing Sun, Diogo C. Luvizon et al.CVPR 2024 · 9 citations
- FarfetchFusion: Towards Fully Mobile Live 3D Telepresence PlatformKyungjin Lee, Juheon Yi, Youngki LeeMobiCom 2023 · 30 citations
