Language-driven Grasp Detection
Vuong Dinh An, Minh Nhat Vu, Baoru Huang, Nghia Nguyen, Hieu Le, Thieu Vo, Anh Nguyen
摘要
A coffee cup, a grape biscuit, a black oven plate on a white dining table Grasp the coffee cup. Grasp the calculator at its keypad. A black pen, a steel marble and a digital calculator on an office desk A black pot, a green bottle, a bread container arranged on a kitchen counter Grab the neck of the green bottle. Give me the brown-band wristwatch. A brown-band wristwatch, a sleek black pen and an office clip placed on a table A white mug, a spiral notepad, a fountain pen resting on a brown desk Pick the fountain pen at its cap. Figure 1 . We present a new dataset and method for language-driven grasp task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive PromptingSangoh Lee, Sangwoo Mo, Wook-Shin HanICML 2026 · 被引用 5 次
- RAGNet: Large-Scale Reasoning-Based Affordance Segmentation Benchmark Towards General GraspingDongming Wu, Yanping Fu, Saike Huang, Yingfei Liu 等ICCV 2025 · 被引用 2 次
- AffordMatcher: Affordance Learning in 3D Scenes from Visual SignifiersNghia Vu, Tuong Do, Khang Nguyen, Baoru Huang 等CVPR 2026 · 被引用 2 次
- RealVLG-R1: A Large-Scale Real-World Visual-Language Grounding Benchmark for Robotic Perception and ManipulationLinfei Li, Lin Zhang, Ying ShenCVPR 2026
- PhysVLM: Enabling Visual Language Models to Understand Robotic Physical ReachabilityWeijie Zhou, Manli Tao, Chaoyang Zhao, Haiyun Guo 等CVPR 2025
它引用的顶会 Paper33
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- Action-Sketcher: From Reasoning to Action via Visual Sketches for Robotic ManipulationHuajie Tan, Peterson Co, Yijie Xu, Shanyu Rong 等CVPR 2026
- RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to ConcreteYuheng Ji, Huajie Tan, Jiayu Shi, Xiaoshuai Hao 等CVPR 2025
- DexFuncGrasp: A Robotic Dexterous Functional Grasp Dataset Constructed from a Cost-Effective Real-Simulation Annotation SystemJinglue Hang, Xiangbo Lin, Tianqiang Zhu, Xuanheng Li 等AAAI 2024 · 被引用 17 次
- AffordDexGrasp: Open-Set Language-Guided Dexterous Grasp With Generalizable-Instructive AffordanceYi-Lin Wei, Mu Lin, Yuhao Lin, Jian-Jian Jiang 等ICCV 2025 · 被引用 8 次
- Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Video ContentQiuheng Wang, Yukai Shi, Jiarong Ou, Rui Chen 等CVPR 2025
