ChatDirector: Enhancing Video Conferencing with Space-Aware Scene Rendering and Speech-Driven Layout Transition
Xun Qian, Feitong Tan, Yinda Zhang, Brian Moreno Collins, David Kim, Alex Olwal, Karthik Ramani, Ruofei Du
摘要
Initial full-view scene of ChatDirector (b) Conversations with Alice (c) Conversations between Alice and Bob (d) Full-view scene when Charlie speaks to all
Figure 1: Screenshots of ChatDirector, captured from the local user, Sean's laptop during a remote meeting with Alice (left), Bob (center), and Charlie (right). (a) Using an off-the-shelf laptop or workstation equipped with an RGB camera, ChatDirector depicts remote participants as 3D portrait avatars and renders them in a shared virtual meeting environment. Sean starts his progress update to the team. (b) When Sean inquires about a feature update from Alice, ChatDirector recognizes the speech activity and automatically focuses the camera on Alice, facilitating a more personal one-on-one discussion. (c) Later, Bob steps in and asks Alice further questions, ChatDirector arranges their avatars in a pairwise layout and simulates direct eye contact by orienting their 3D avatars towards each other. (d) When Charlie updates his progress to everyone, the camera is zoomed out with other avatars turning to Charlie, to provide Sean with a visual cue of the speech transition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Since U Been Gone: Augmenting Context-Aware Transcriptions for Re-Engaging in Immersive VR MeetingsGeonsun Lee, Yue Yang, Jennifer Healey, Dinesh ManochaCHI 2025 · 被引用 13 次
- Thing2Reality: Enabling Spontaneous Creation of 3D Objects from 2D Content using Generative AI in XR MeetingsErzhen Hu, Mingyi Li, Jungtaek Hong, Xun Qian 等UIST 2025 · 被引用 13 次
- CoCreatAR: Enhancing Authoring of Outdoor Augmented Reality Experiences Through Asymmetric CollaborationNels Numan, Gabriel J. Brostow, Suhyun Park, Simon Julier 等CHI 2025 · 被引用 12 次
- "I use video calling in all areas of my life": Understanding the Video Calling Experiences of Chronically Ill PeopleHumphrey Curtis, Erin Beneteau, Edward Cutrell, Denae Ford 等CHI 2025 · 被引用 6 次
- DialogLab: Authoring, Simulating, and Testing Dynamic Human-AI Group ConversationsErzhen Hu, Yanhe Chen, Mingyi Li, Vrushank Phadnis 等UIST 2025 · 被引用 3 次
它引用的顶会 Paper13
- GaussianAvatars: Photorealistic Head Avatars with Rigged 3D GaussiansShenhan Qian, Tobias Kirschstein, Liam Schoneveld, Davide Davoli 等CVPR 2024 · 被引用 175 次
- Total relighting: learning to relight portraits for background replacementRohit Pandey, Sergio Orts-Escolano, Chloe LeGendre, Christian Häne 等SIGGRAPH 2021 · 被引用 138 次
- Authentic volumetric avatars from a phone scanChen Cao, Tomas Simon, Jin Kyu Kim, Gabe Schwartz 等SIGGRAPH 2022 · 被引用 123 次
- Partially Blended Realities: Aligning Dissimilar Spaces for Distributed Mixed Reality MeetingsJens Emil Sloth Grønbæk, Ken Pfeuffer, Eduardo Velloso, Morten Astrup 等CHI 2023 · 被引用 75 次
- VirtualCube: An Immersive 3D Video Communication SystemYizhong Zhang, Jiaolong Yang, Zhen Liu, Ruicheng Wang 等IEEE VR 2022 · 被引用 66 次
相关 Paper
- GazeChat: Enhancing Virtual Conferences with Gaze-aware 3D PhotosZhenyi He, Keru Wang, Brandon Yushan Feng, Ruofei Du 等UIST 2021 · 被引用 36 次
- Redirecting Desktop Interface Input to Animate Cross-Reality AvatarsJason W. Woodworth, David Broussard, Christoph W. BorstIEEE VR 2022 · 被引用 12 次
- SealMates: Improving Communication in Video Conferencing using a Collective Behavior-Driven AvatarMark Armstrong, Chi-Lan Yang, Kinga Skiers, Mengzhen Lim 等CSCW 2024 · 被引用 6 次
- VoluMe - Authentic 3D Video Calls from Live Gaussian Splat PredictionMartin de La Gorce, Charlie Hewitt, Tibor Takács, Robert Gerdisch 等ICCV 2025 · 被引用 5 次
- ConeSpeech: Exploring Directional Speech Interaction for Multi-Person Remote Communication in Virtual RealityYukang Yan, Haohua Liu, Yingtian Shi, Jingying Wang 等IEEE VR 2023 · 被引用 13 次
