Example-driven virtual cinematography by learning camera behaviors
Hongda Jiang, Bin Wang, Xi Wang, Marc Christie, Baoquan Chen
Abstract
Designing a camera motion controller that has the capacity to move a virtual camera automatically in relation with contents of a 3D animation, in a cinematographic and principled way, is a complex and challenging task. Many cinematographic rules exist, yet practice shows there are significant stylistic variations in how these can be applied. In this paper, we propose an example-driven camera controller which can extract camera behaviors from an example film clip and re-apply the extracted behaviors to a 3D animation, through learning from a collection of camera motions. Our first technical contribution is the design of a low-dimensional cinematic feature space that captures the essence of a film's cinematic characteristics (camera angle and distance, screen composition and character configurations) and which is coupled with a neural network to automatically extract these cinematic characteristics from real film clips. Our second technical contribution is the design of a cascaded deep-learning architecture trained to (i) recognize a variety of camera motion behaviors from the extracted cinematic features, and (ii) predict the future motion of a virtual camera given a character 3D animation. We propose to rely on a Mixture of Experts (MoE) gating+prediction mechanism to ensure that distinct camera behaviors can be learned while ensuring generalization. We demonstrate the features of our approach through experiments that highlight (i) the quality of our cinematic feature extractor (ii) the capacity to learn a range of behaviors through the gating mechanism, and (iii) the ability to generate a variety of camera motions by applying different behaviors extracted from film clips. Such an example-driven approach offers a high level of controllability which opens new possibilities toward a deeper understanding of cinematographic style and enhanced possibilities in exploiting real film data in virtual environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- GAIT: Generating Aesthetic Indoor Tours with Deep Reinforcement LearningDesai Xie, Ping Hu, Xin Sun, Sören Pirk et al.ICCV 2023 · 9 citations
- A Reinforcement Learning-Based Automatic Video Editing Method Using Pre-trained Vision-Language ModelPanwen Hu, Nan Xiao, Feifei Li, Yongquan Chen et al.ACM MM 2023 · 8 citations
- Pulp Motion: Framing-aware multimodal camera and human motion generationRobin Courant, Xi WANG, David Loiseaux, Marc Christie et al.ICLR 2026 · 8 citations
- Optimization-based User Support for Cinematographic Quadrotor Camera Target FramingChristoph Gebhardt, Otmar HilligesCHI 2021 · 7 citations
- DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera SynthesisZixuan Wang, Jiayi Li, Xiaoyu Qin, Shikun Sun et al.ACM MM 2024 · 5 citations
Related papers
- Virtual Camera Layout Generation using a Reference VideoJung Eun Yoo, Kwanggyoon Seo, Sanghun Park, Jaedong Kim et al.CHI 2021 · 5 citations
- CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video GenerationKaiyi Huang, Yukun Huang, Yu Li, Jianhong Bai et al.CVPR 2026 · 7 citations
- An Interactive System for Supporting Creative Exploration of Cinematic Composition DesignsRui He, Huaxin Wei, Ying CaoUIST 2024 · 10 citations
- Cinematic Behavior Transfer via NeRF-based Differentiable FilmingXuekun Jiang, Anyi Rao, Jingbo Wang, Dahua Lin et al.CVPR 2024
- Neural animation layering for synthesizing martial arts movementsSebastian Starke, Yiwei Zhao, Fabio Zinno, Taku KomuraSIGGRAPH 2021 · 75 citations
