Flora: Dual-Frequency LOss-Compensated ReAl-Time Monocular 3D Video Reconstruction
Likang Wang, Yue Gong, Qirui Wang, Kaixuan Zhou, Lei Chen
摘要
In this work, we propose a real-time monocular 3D video reconstruction approach named Flora for reconstructing delicate and complete 3D scenes from RGB video sequences in an end-to-end manner. Specifically, we introduce a novel method with two main contributions. Firstly, the proposed feature aggregation module retains both color and reliability in a dual-frequency form. Secondly, the loss compensation module solves missing structure by correcting losses for falsely pruned voxels. The dual-frequency feature aggregation module enhances reconstruction quality in both precision and recall, and the loss compensation module benefits the recall. Notably, both proposed contributions achieve great results with negligible inferencing overhead. Our state-of-the-art experimental results on realworld datasets demonstrate Flora's leading performance in both effectiveness and efficiency. The code is available at https://github.com/NoOneUST/Flora . Taxonomy of Reconstruction Methods Most RGB-based real-time 3D reconstruction methods are based on deep neural networks (
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Hard Sample Aware Network for Contrastive Deep Graph ClusteringYue Liu, Xihong Yang, Sihang Zhou, Xinwang Liu 等AAAI 2023 · 被引用 175 次
- Substructure Aware Graph Neural NetworksDingyi Zeng, Wanlong Liu, Wenyu Chen, Li Zhou 等AAAI 2023 · 被引用 60 次
- Rethinking Alignment and Uniformity in Unsupervised Image Semantic SegmentationDaoan Zhang, Chenming Li, Haoquan Li, Wenjian Huang 等AAAI 2023 · 被引用 21 次
- DisWOT: Student Architecture Search for Distillation WithOut TrainingPeijie Dong, Lujun Li, Zimian WeiCVPR 2023
- Dionysus: Recovering Scene Structures by Dividing into Semantic PiecesLikang Wang, Lei ChenCVPR 2023
它引用的顶会 Paper13
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima 等ICCV 2019 · 被引用 1,411 次
- Hard Sample Aware Network for Contrastive Deep Graph ClusteringYue Liu, Xihong Yang, Sihang Zhou, Xinwang Liu 等AAAI 2023 · 被引用 175 次
- Multi-View Stereo by Temporal Nonparametric FusionYuxin Hou, Juho Kannala, Arno SolinICCV 2019 · 被引用 99 次
- DICNet: Deep Instance-Level Contrastive Network for Double Incomplete Multi-View Multi-Label ClassificationChengliang Liu, Jie Wen, Xiaoling Luo, Chao Huang 等AAAI 2023 · 被引用 68 次
相关 Paper
- SLAM3R: Real-Time Dense Scene Reconstruction from Monocular RGB VideosYuzheng Liu, Siyan Dong, Shuzhe Wang, Yingda Yin 等CVPR 2025
- NeuralRecon: Real-Time Coherent 3D Reconstruction From Monocular VideoJiaming Sun, Yiming Xie, Linghao Chen, Xiaowei Zhou 等CVPR 2021
- DG-Recon: Depth-Guided Neural 3D Scene ReconstructionJihong Ju, Ching Wei Tseng, Oleksandr Bailo, Georgi Dikov 等ICCV 2023 · 被引用 21 次
- Photorealistic Monocular 3D Reconstruction of Humans Wearing ClothingThiemo Alldieck, Mihai Zanfir, Cristian SminchisescuCVPR 2022 · 被引用 136 次
- DeepFaceFlow: In-the-Wild Dense 3D Facial Motion EstimationMohammad Rami Koujan, Anastasios Roussos, Stefanos ZafeiriouCVPR 2020
