DA2: Depth Anything in Any Direction
Haodong Li, Wangguandong Zheng, Jing He, Yuhao Liu, Xin Lin, Xin Yang, Yingcong Chen, Chunchao Guo
Abstract
Panorama has a full FoV (360180), offering a more complete visual description than perspective images. Thanks to this characteristic, panoramic depth estimation is gaining increasing traction in 3D vision. However, due to the scarcity of panoramic data, previous methods are often restricted to in-domain settings, leading to poor zero-shot generalization. Furthermore, due to the spherical distortions inherent in panoramas, many approaches rely on perspective splitting (e.g., cubemaps), which leads to suboptimal efficiency. To address these challenges, we propose \textbf{DA}$$^{\textbf{2}}: epth nything in ny irection, an accurate, zero-shot generalizable, and fully end-to-end panoramic depth estimator. Specifically, for scaling up panoramic data, we introduce a data curation engine for generating high-quality panoramic depth data from perspective, and create 543K panoramic RGB-depth pairs, bringing the total to 607K. To further mitigate the spherical distortions, we present SphereViT, which explicitly leverages spherical coordinates to enforce the spherical geometric consistency in panoramic image features, yielding improved performance. A comprehensive benchmark on multiple datasets clearly demonstrates DA's SoTA performance, with an average 38% improvement on AbsRel over the strongest zero-shot baseline. Surprisingly, DA even outperforms prior in-domain methods, highlighting its superior zero-shot generalization. Moreover, as an end-to-end solution, DA exhibits much higher efficiency over fusion-based approaches. Both the code and the curated panoramic data have be released. Project page: https://depth-any-in-any-dir.github.io/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0ae6cfc7-6942-41e3-9f63-3c16b75b55eaCited by top-tier papers5
- Depth Any Panoramas: A Foundation Model for Panoramic Depth EstimationXin Lin, Meixi Song, Dizhe Zhang, Wenxuan Lu et al.CVPR 2026 · 27 citations
- More than the Sum: Panorama-Language Models for Adverse Omni-ScenesWeijia Fan, Ruiping Liu, Jiale Wei, Yufan Chen et al.CVPR 2026 · 5 citations
- PVDepth: Panoramic Video Depth Estimation via Geometry-Aware Spatiotemporal AdaptationChuanxin Song, Peixi PengICML 2026
- Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth PriorsChuanqing Zhuang, Xin Lu, Zehui Deng, Zhengda Lu et al.CVPR 2026
- MTPano: Multi-Task Panoramic Scene Understanding via Label-Free Integration of Dense Prediction PriorsJingdong Zhang, Xiaohang Zhan, Lingzhi Zhang, Yizhou Wang et al.SIGGRAPH 2026
Builds on36
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
Related papers
- SO(3)-Equivariant ViT-Adapter for Data-Efficient Zero-Shot Sim-to-Real Indoor Panoramic Depth EstimationZiyan He, Qiudan Zhang, Lin Ma, Xu WangCVPR 2026
- Depth Anywhere: Enhancing 360 Monocular Depth Estimation via Perspective Distillation and Unlabeled Data AugmentationNing-Hsu Wang, Yu-Lun LiuNeurIPS 2024 · 56 citations
- Depth Any Camera: Zero-Shot Metric Depth Estimation from Any CameraYuliang Guo, Sparsh Garg, S. Mahdi H. Miangoleh, Xinyu Huang et al.CVPR 2025
- PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic ImageryYijing Guo, Mengjun Chao, Luo Wang, Tianyang Zhao et al.CVPR 2026 · 11 citations
- VGGT-360: Geometry-Consistent Zero-Shot Panoramic Depth EstimationJiayi Yuan, Haobo Jiang, De Wen Soh, Na ZhaoCVPR 2026 · 6 citations
