Context-Aware Head-and-Eye Motion Generation with Diffusion Model
Yuxin Shen, Manjie Xu, Wei Liang
摘要
In humanity’s ongoing quest to craft natural and realistic avatars within virtual environments, the generation of authentic eye gaze behaviors stands paramount. Eye gaze not only serves as a primary non-verbal communication cue, but it also reflects cognitive processes, intent, and attentiveness, making it a crucial element in ensuring immersive interactions. However, automatically generating these intricate gaze behaviors presents significant challenges. Traditional methods can be both time-consuming and lack the precision to align gaze behaviors with the intricate nuances of the environment in which the avatar resides. To overcome these challenges, we introduce a novel two-stage approach to generate context-aware head-and-eye motions across diverse scenes. By harnessing the capabilities of advanced diffusion models, our approach adeptly produces contextually appropriate eye gaze points, further leading to the generation of natural head-and-eye movements. Utilizing Head-Mounted Display (HMD) eye-tracking technology, we also present a comprehensive dataset, which captures human eye gaze behaviors in tandem with associated scene features. We show that our approach consistently delivers intuitive and lifelike head-and-eye motions and demonstrates superior performance in terms of motion fluidity, alignment with contextual cues, and overall user satisfaction.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- TextGaze: Gaze-Controllable Face Generation with Natural LanguageHengfei Wang, Zhongqun Zhang, Yihua Cheng, Hyung Jin ChangACM MM 2024 · 被引用 3 次
- The eyes have it: an integrated eye and face model for photorealistic facial animationGabriel Schwartz, Shih-En Wei, Te-Li Wang, Stephen Lombardi 等SIGGRAPH 2020 · 被引用 54 次
- Talking Together: Synthesizing Co-Located 3D Conversations from AudioMengyi Shan, Shouchieh Chang, Ziqian Bai, Shichen Liu 等CVPR 2026
- EyeNeRF: a hybrid representation for photorealistic synthesis, animation and relighting of human eyesGengyan Li, Abhimitra Meka, Franziska Mueller, Marcel C. Bühler 等SIGGRAPH 2022 · 被引用 39 次
- PrivateEyes: Gaze-Preserving Anonymization for Data SharingSurabhi Gupta, Dinesh Prabhu Muthumariappan, Biplab Ch Das, Anoop Kolar Rajagopal 等CVPR 2026
