Lune

NeurIPS2025顶会

MoCha: Towards Movie-Grade Talking Character Generation

Cong Wei, Bo Sun, Haoyu Ma, Ji Hou, Felix Juefei-Xu, Zecheng He, Xiaoliang Dai, Luxin Zhang, Kunpeng Li, Tingbo Hou, Animesh Sinha, Peter Vajda, Wenhu Chen

2025年份
2被引次数
1顶会引用

摘要

"Two distinct streams of tears trail down her cheeks as she speaks with an angry expression…" "A medium shot of a man interacting warmly with an elephant. the man talks to the camera…" "A tilt up shot of a man standing in a dimly lit room, speaking to the camera…" Action Control Multi-Character Turn-based Talk Talking Character "Close-up shot of a doctor in a white lab coat over blue scrubs, speaking…" Emotion Control Figure 1: MoCha is an end-to-end dialogue-centric video generation model that takes only speech and text as input, without requiring any auxiliary conditions.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper1

问问它们各自怎么用它

它引用的顶会 Paper28

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖