Semantic Attention Flow Fields for Monocular Dynamic Scene Decomposition
Yiqing Liang, Eliot Laidlaw, Alexander Meyerowitz, Srinath Sridhar, James Tompkin
摘要
From video, we reconstruct a neural volume that captures time-varying color, density, scene flow, semantics, and attention information. The semantics and attention let us identify salient foreground objects separately from the background across spacetime. To mitigate low resolution semantic and attention features, we compute pyramids that trade detail with whole-image context. After optimization, we perform a saliency-aware clustering to decompose the scene. To evaluate real-world scenes, we annotate object masks in the NVIDIA Dynamic Scene and DyCheck datasets. We demonstrate that this method can decompose dynamic scenes in an unsupervised way with competitive performance to a supervised method, and that it improves foreground/background segmentation over recent static/dynamic split methods. Project webpage: https://visual.cs.brown.edu/saff
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Grid4D: 4D Decomposed Hash Encoding for High-Fidelity Dynamic Gaussian SplattingJiawei Xu, Zexin Fan, Jian Yang, Jin XieNeurIPS 2024 · 被引用 64 次
- When does perceptual alignment benefit vision representations?Shobhita Sundaram, Stephanie Fu, Lukas Muttenthaler, Netanel Tamir 等NeurIPS 2024 · 被引用 24 次
- WorDepth: Variational Language Prior for Monocular Depth EstimationZiyao Zeng, Daniel Wang, Fengyu Yang, Hyoungseob Park 等CVPR 2024 · 被引用 20 次
- DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene RenderingJie Chen, Zhangchi Hu, Peixi Wu, Huyue Zhu 等ICCV 2025 · 被引用 2 次
- DIV-FF: Dynamic Image-Video Feature Fields For Environment Understanding in Egocentric VideosLorenzo Mur-Labadia, Josechu Guerrero, Ruben Martinez-CantinCVPR 2025
它引用的顶会 Paper28
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer 等ICCV 2021 · 被引用 617 次
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 被引用 615 次
相关 Paper
- Semantic Flow: Learning Semantic Fields of Dynamic Scenes from Monocular VideosFengrui Tian, Yueqi Duan, Angtian Wang, Jianfei Guo 等ICLR 2024 · 被引用 7 次
- DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric VoxelizationYanpeng Zhao, Siyu Gao, Yunbo Wang, Xiaokang YangICLR 2024 · 被引用 2 次
- DS-NeRV: Implicit Neural Video Representation with Decomposed Static and Dynamic CodesHao Yan, Zhihui Ke, Xiaobo Zhou, Tie Qiu 等CVPR 2024 · 被引用 18 次
- DeGauss: Dynamic-Static Decomposition with Gaussian Splatting for Distractor-Free 3D ReconstructionRui Wang, Quentin Lohmeyer, Mirko Meboldt, Siyu TangICCV 2025 · 被引用 12 次
- NVFi: Neural Velocity Fields for 3D Physics Learning from Dynamic VideosJinxi Li, Ziyang Song, Bo YangNeurIPS 2023 · 被引用 39 次
