Towards Unobtrusive Physical AI: Augmenting Everyday Objects with Intelligence and Robotic Movement for Proactive Assistance
Violet Yinuo Han, Jesse T. Gonzalez, Christina Yang, Zhiruo Wang, Scott E. Hudson, Alexandra Ion
2025Year
3Citations
2Top-tier citations
Abstract
Action 1: knife moves away to ensure safety Video streamAction 2: pan trivets move close for baking sheet Action-goal alignment Figure 1: Everyday objects, such as trivets, are brought to life by the Object Agents system.These objects move autonomously, in order to assist and protect users.Our system (1) perceives context using a vision language model backbone, (2) reasons about user goals and object affordances, and (3) generates actions that are delivered to familiar items augmented with robotic motion.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Towards Fluent Interaction with Cyber-Physical ArchitectureJesse T. Gonzalez, Neeta M. Khanuja, Michael Mingxuan Li, Maggie Guo et al.CHI 2026 · 2 citations
- SpeechLess: Micro-utterance with Personalized Spatial Memory-aware Assistant in Everyday Augmented RealityYoonsang Kim, Devshree Jadeja, Divyansh Pradhan, Yalong Yang et al.IEEE VR 2026 · 1 citation
Builds on20
- SWE-agent: Agent-Computer Interfaces Enable Automated Software EngineeringJohn Yang, Carlos E. Jimenez, Alexander Wettig, Kilian Lieret et al.NeurIPS 2024 · 2,059 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
- WebArena: A Realistic Web Environment for Building Autonomous AgentsShuyan Zhou, Frank F. Xu, Hao Zhu, Xuhui Zhou et al.ICLR 2024 · 1,197 citations
- Large Language Models as Commonsense Knowledge for Large-Scale Task PlanningZirui Zhao, Wee Sun Lee, David HsuNeurIPS 2023 · 423 citations
- RoomShift: Room-scale Dynamic Haptics for VR with Furniture-moving Swarm RobotsRyo Suzuki, Hooman Hedayati, Clement Zheng, James L. Bohn et al.CHI 2020 · 122 citations
Related papers
- CookAR: Affordance Augmentations in Wearable AR to Support Kitchen Tool Interactions for People with Low VisionJaewook Lee, Andrew D. Tjahjadi, Jiho Kim, Junpu Yu et al.UIST 2024 · 25 citations
- Factorizing Perception and Policy for Interactive Instruction FollowingKunal Pratap Singh, Suvaansh Bhambri, Byeonghwi Kim, Roozbeh Mottaghi et al.ICCV 2021 · 39 citations
- Shaping embodied agent behavior with activity-context priors from egocentric videoTushar Nagarajan, Kristen GraumanNeurIPS 2021 · 23 citations
- CHEF-VL: Detecting Cognitive Sequencing Errors in Cooking with Vision-language ModelsRuiqi Wang, Peiqi Gao, Patrick Lynch, Tingjun Liu et al.UbiComp 2026 · 1 citation
- VAT-Mart: Learning Visual Action Trajectory Proposals for Manipulating 3D ARTiculated ObjectsRuihai Wu, Yan Zhao, Kaichun Mo, Zizheng Guo et al.ICLR 2022 · 119 citations
