Lune

ICCV2025顶会

Greg: GEometry-Aware RegIon Refinement for Sign Language Video Generation

Tongkai Shi, Lianyu Hu, Fanhua Shang, Liqing Gao, Wei Feng

2025年份
1被引次数

摘要

Sign Language Video Generation (SLVG) aims to transform sign language sequences into natural and fluent sign language videos. Existing SLVG methods lack geometric modeling of human anatomical structures, leading to anatomically implausible and temporally inconsistent generation. To address these challenges, we propose a novel framework: Geometry-Aware Region Refinement (GReg) for SLVG. GReg uses geometric information (such as normal maps and gradient maps) from the SMPL-X model to ensure anatomical and temporal consistency. To fully leverage the geometric priors, we propose two novel methods: 1) Regional Prior Generation, which uses regional expert networks to generate target-structured regions as generation priors; 2) Gradient-enhanced Refinement, which guides the refinement of detailed structures in key regions using gradient features. Furthermore, we enhance visual realism in key regions through adversarial training on both these regions and their gradient maps. Experimental results demonstrate that GReg achieves state-of-the-art performance with superior structural accuracy and temporal consistency.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper13

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖