PINO: Person-Interaction Noise Optimization for Long-Duration and Customizable Motion Generation of Arbitrary-Sized Groups
Sakuya Ota, Qing Yu, Kent Fujiwara, Satoshi Ikehata, Ikuro Sato
摘要
Generating realistic group interactions involving multiple characters remains challenging due to increasing complexity as group size expands. While existing conditional diffusion models incrementally generate motions by conditioning on previously generated characters, they rely on single shared prompts, limiting nuanced control and leading to overly simplified interactions. In this paper, we introduce Person-Interaction Noise Optimization (PINO), a novel, training-free framework designed for generating realistic and customizable interactions among groups of arbitrary size. PINO decomposes complex group interactions into semantically relevant pairwise interactions, and leverages pretrained two-person interaction diffusion models to incrementally compose group interactions. To ensure physical plausibility and avoid common artifacts such as overlapping or penetration between characters, PINO employs physics-based penalties during noise optimization. This approach allows precise user control over character orientation, speed, and spatial relationships without additional training. Comprehensive evaluations demonstrate that PINO generates visually realistic, physically coherent, and adaptable multi-person interactions suitable for diverse animation, gaming, and robotics applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Causal Motion Diffusion Models for Autoregressive Motion GenerationQing Yu, Akihisa Watanabe, Kent FujiwaraCVPR 2026 · 被引用 9 次
- ProjFlow: Projection Sampling with Flow Matching for Zero‑Shot Exact Spatial Motion ControlAkihisa Watanabe, Qing Yu, Edgar Simo-Serra, Kent FujiwaraCVPR 2026 · 被引用 6 次
- Stability-Driven Motion Generation for Object-Guided Human-Human Co-ManipulationJiahao Xu, Xiaohan Yuan, Xingchen Wu, Chongyang Xu 等CVPR 2026
- Diffusion Mental AveragesPhonphrm Thawatdamrongkit, Sukit Seripanitkarn, Supasorn SuwajanakornCVPR 2026
它引用的顶会 Paper30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 被引用 672 次
相关 Paper
- InterControl: Zero-shot Human Interaction Generation by Controlling Every JointZhenzhi Wang, Jingbo Wang, Yixuan Li, Dahua Lin 等NeurIPS 2024 · 被引用 27 次
- Multi-Person Interaction Generation from Two-Person Motion PriorsWenning Xu, Shiyu Fan, Paul Henderson, Edmond S. L. HoSIGGRAPH 2025 · 被引用 2 次
- Towards Storytelling Animations: Joint Synthesis of Human and Camera MotionsBoyuan Cheng, Yingjie Xi, Rui He, Jinhe Na 等CVPR 2026
- Optimizing Diffusion Noise Can Serve As Universal Motion PriorsKorrawe Karunratanakul, Konpat Preechakul, Emre Aksan, Thabo Beeler 等CVPR 2024
- MixerMDM: Learnable Composition of Human Motion Diffusion ModelsPablo Ruiz-Ponce, Germán Barquero, Cristina Palmero, Sergio Escalera 等CVPR 2025
