ReDirTrans: Latent-to-Latent Translation for Gaze and Head Redirection
Shiwei Jin, Zhen Wang, Lei Wang, Ning Bi, Truong Q. Nguyen
Abstract
Learning-based gaze estimation methods require large amounts of training data with accurate gaze annotations. Facing such demanding requirements of gaze data collection and annotation, several image synthesis methods were proposed, which successfully redirected gaze directions precisely given the assigned conditions. However, these methods focused on changing gaze directions of the images that only include eyes or restricted ranges of faces with low resolution (less than 128 × 128) to largely reduce interference from other attributes such as hairs, which limits application scenarios. To cope with this limitation, we proposed a portable network, called ReDirTrans, achieving latentto-latent translation for redirecting gaze directions and head orientations in an interpretable manner. ReDirTrans projects input latent vectors into aimed-attribute embeddings only and redirects these embeddings with assigned pitch and yaw values. Then both the initial and edited embeddings are projected back (deprojected) to the initial latent space as residuals to modify the input latent vectors by subtraction and addition, representing old status removal and new status addition. The projection of aimed attributes only and subtraction-addition operations for status replacement essentially mitigate impacts on other attributes and the distribution of latent vectors. Thus, by combining ReDirTrans with a pretrained fixed e4e-StyleGAN pair, we created ReDirTrans-GAN, which enables accurately redirecting gaze in full-face images with 1024×1024 resolution while preserving other attributes such as identity, expression, and hairstyle. Furthermore, we presented improvements for the downstream learning-based gaze estimation task, using redirected samples as dataset augmentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single ImageHengfei Wang, Zhongqun Zhang, Yihua Cheng, Hyung Jin ChangAAAI 2026
- Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball RotationYoungChan Choi, HengFei Wang, YiHua Cheng, Boeun Kim et al.ACM MM 2025
- AUEditNet: Dual-Branch Facial Action Unit Intensity Manipulation with Implicit DisentanglementShiwei Jin, Zhen Wang, Lei Wang, Peng Liu et al.CVPR 2024
Builds on17
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or et al.ICCV 2021 · 1,437 citations
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 1,195 citations
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- ReStyle: A Residual-Based StyleGAN Encoder via Iterative RefinementYuval Alaluf, Or Patashnik, Daniel Cohen-OrICCV 2021 · 377 citations
Related papers
- Adaptive Nonlinear Latent Transformation for Conditional Face EditingZhizhong Huang, Siteng Ma, Junping Zhang, Hongming ShanICCV 2023 · 13 citations
- Photo-Realistic Monocular Gaze Redirection Using Generative Adversarial NetworksZhe He, Adrian Spurr, Xucong Zhang, Otmar HilligesICCV 2019 · 54 citations
- Only a matter of style: age transformation using a style-based regression modelYuval Alaluf, Or Patashnik, Daniel Cohen-OrSIGGRAPH 2021 · 143 citations
- Controllable Continuous Gaze RedirectionWeihao Xia, Yujiu Yang, Jing-Hao Xue, Wensen FengACM MM 2020 · 13 citations
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 97 citations
