Lune

CVPR2025Top-tier venue

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models

Rick Akkerman, Haiwen Feng, Michael J. Black, Dimitrios Tzionas, Victoria Fernández Abrevaya

2025Year
9Top-tier citations

Abstract

Input image Video generation by InterDyn using only the hand mask sequence as control signal Force propagation Counterfactual dynamics Future #1 Future #2 Denotes a driving object motion | Tracks indicate an object with generated uncontrolled dynamics t=3 t=13 t=0 t=0 Figure 1 . We present InterDyn, a framework for synthesizing realistic interactive dynamics without 3D reconstruction and physics simulation. Our core principle is to rely on the implicit physics knowledge embedded in large-scale video generative models. Given an image and a "driving motion", our model generates the consequential scene dynamics. We investigate the generated interactive dynamics in a simple object collision scenario (bottom) and complex in-the-wild human-object interaction (top).

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext b913bd4b-1837-47ef-977a-a7ffa65d85c3

Cited by top-tier papers9

Ask how each one uses it

Builds on51

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines