Stargazer: An Interactive Camera Robot for Capturing How-To Videos Based on Subtle Instructor Cues
Jiannan Li, Maurício Sousa, Karthik Mahadevan, Bryan Wang, Paula Akemi Aoyaui, Nicole Yu, Angela Yang, Ravin Balakrishnan, Anthony Tang, Tovi Grossman
摘要
Live and pre-recorded video tutorials are an effective means for teaching physical skills such as cooking or prototyping electronics. A dedicated cameraperson following an instructor’s activities can improve production quality. However, instructors who do not have access to a cameraperson’s help often have to work within the constraints of static cameras. We present Stargazer, a novel approach for assisting with tutorial content creation with a camera robot that autonomously tracks regions of interest based on instructor actions to capture dynamic shots. Instructors can adjust the camera behaviors of Stargazer with subtle cues, including gestures and speech, allowing them to fluidly integrate camera control commands into instructional activities. Our user study with six instructors, each teaching a distinct skill, showed that participants could create dynamic tutorial videos with a diverse range of subjects, camera framing, and camera angle combinations using Stargazer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- AQuA: Automated Question-Answering in Software Tutorial Videos with Visual AnchorsSaelyne Yang, Jo Vermeulen, George W. Fitzmaurice, Justin MatejkaCHI 2024 · 被引用 15 次
- How Do We Research Human-Robot Interaction in the Age of Large Language Models? A Systematic ReviewYufeng Wang, Yuan Xu, Anastasia Nikolova, Yuxuan Wang 等CHI 2026 · 被引用 4 次
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video ClipsAntoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi 等ICCV 2019 · 被引用 1,437 次
- RealitySketch: Embedding Responsive Graphics and Visualizations in AR through Dynamic SketchingRyo Suzuki, Rubaiat Habib Kazi, Li-Yi Wei, Stephen DiVerdi 等UIST 2020 · 被引用 98 次
- RealityTalk: Real-Time Speech-Driven Augmented Presentation for AR Live StorytellingJian Liao, Adnan Karim, Shivesh Singh Jadon, Rubaiat Habib Kazi 等UIST 2022 · 被引用 44 次
- RubySlippers: Supporting Content-based Voice Navigation for How-to VideosMinsuk Chang, Mina Huh, Juho KimCHI 2021 · 被引用 40 次
相关 Paper
- Automatic Instructional Video Creation from a Markdown-Formatted TutorialPeggy Chi, Nathan Frey, Katrina Panovich, Irfan EssaUIST 2021 · 被引用 27 次
- ASTEROIDS: Exploring Swarms of Mini-Telepresence Robots for Physical Skill DemonstrationJiannan Li, Maurício Sousa, Chu Li, Jessie Liu 等CHI 2022 · 被引用 24 次
- AdapTutAR: An Adaptive Tutoring System for Machine Tasks in Augmented RealityGaoping Huang, Xun Qian, Tianyi Wang, Fagun Patel 等CHI 2021 · 被引用 93 次
- ChameleonControl: Teleoperating Real Human Surrogates through Mixed Reality Gestural Guidance for Remote Hands-on ClassroomsMehrad Faridan, Bheesha Kumari, Ryo SuzukiCHI 2023 · 被引用 29 次
- GAZED- Gaze-guided Cinematic Editing of Wide-Angle Monocular Video RecordingsK. L. Bhanu Moorthy, Moneish Kumar, Ramanathan Subramanian, Vineet GandhiCHI 2020 · 被引用 26 次
