Gallant: Voxel Grid-based Humanoid Locomotion and Local-navigation across 3-D Constrained Terrains
Qingwei Ben, Botian Xu, Kailin Li, Feiyu Jia, Wentao Zhang, Jingping Wang, Jingbo Wang, Dahua Lin, Jiangmiao Pang
摘要
Robust humanoid locomotion requires accurate and globally consistent perception of the surrounding 3D environment. However, existing perception modules, mainly based on depth images or elevation maps, offer only partial and locally flattened views of the environment, failing to capture the full 3D structure.This paper presents , a voxel-grid–based framework for humanoid locomotion and local navigation in 3D constrained terrains.It leverages voxelized LiDAR data as a lightweight and structured perceptual representation, and employs a z-grouped 2D CNN to map this representation to the control policy, enabling fully end-to-end optimization. A high-fidelity LiDAR simulation that dynamically generates realistic observations is developed to support scalable, LiDAR-based training and ensure sim-to-real consistency.Experimental results show that Gallant’s broader perceptual coverage facilitates the use of a single policy that goes beyond the limitations of previous methods confined to ground-level obstacles, extending to lateral clutter, overhead constraints, multi-level structures, and narrow passages. Gallant also firstly achieves near-100% success rates in challenging scenarios such as stair climbing and stepping onto elevated platforms through improved end-to-end optimization. This project will be fully open-source.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal TransformersRuihan Yang, Minghao Zhang, Nicklas Hansen, Huazhe Xu 等ICLR 2022 · 被引用 146 次
- OneOcc: Semantic Occupancy Prediction for Legged Robots with a Single Panoramic CameraHao Shi, Ze Wang, Shangwei Guo, Mengfei Duan 等CVPR 2026 · 被引用 11 次
- VoxelNeXt: Fully Sparse VoxelNet for 3D Object Detection and TrackingYukang Chen, Jianhui Liu, Xiangyu Zhang, Xiaojuan Qi 等CVPR 2023
相关 Paper
- Wanderland: Geometrically Grounded Simulation for Open-World Embodied AIXinhao Liu, Jiaqi Li, Youming Deng, Ruxin Chen 等CVPR 2026 · 被引用 5 次
- Volumetric Environment Representation for Vision-Language NavigationRui Liu, Wenguan Wang, Yi YangCVPR 2024 · 被引用 25 次
- LidarMultiNet: Towards a Unified Multi-Task Network for LiDAR PerceptionDongqiangzi Ye, Zixiang Zhou, Weijia Chen, Yufei Xie 等AAAI 2023 · 被引用 108 次
- SGLoc: Scene Geometry Encoding for Outdoor LiDAR LocalizationWen Li, Shangshu Yu, Cheng Wang, Guosheng Hu 等CVPR 2023
- 3D-Aware Object Goal Navigation via Simultaneous Exploration and IdentificationJiazhao Zhang, Liu Dai, Fanpeng Meng, Qingnan Fan 等CVPR 2023
