EditDuet: A Multi-Agent System for Video Non-Linear Editing
Marcelo Sandoval-Castañeda, Bryan C. Russell, Josef Sivic, Gregory Shakhnarovich, Fabian Caba Heilbron
Abstract
Editor 🤖 ✂ : search_collection("kneading …"); add_to_timeline([…]) Draft timeline ! Critic 🤖🔎: Make video_127 last a few more seconds and add a shot of lined-up […] Editor 🤖 ✂ : remove_from_timeline(video_127); add_to_timeline(video_127, start=1.5, end=23.7); search_collection("lined-up dough in tray"); […]
… Critic 🤖🔎: RENDER (The current timeline satisfies the user request) … … Talking head footage Audio track Non-linear Editing (NLE) Environment
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6303ea73-8e99-4a1a-aeb7-cebdf725a7aeCited by top-tier papers4
- Who Does What? Archetypes of Roles Assigned to LLMs During Human-AI Decision-MakingShreya Chappidi, Jatinder Singh, Andra Valentina KrauzeCHI 2026 · 2 citations
- Vidmento: Creating Video Stories through Context-Aware Expansion with Generative VideoCatherine Yeh, Anh Truong, Mira Dontcheva, Bryan WangCHI 2026 · 1 citation
- VlogReward: Learning Multi-Dimensional Evaluation for Vlog EditingYexiang Liu, Wen Zhong, Sijie Zhu, Xin Gu et al.ICML 2026
- Autoregressive Modeling of Film with Applications in Video MontageMarcelo Sandoval-Castañeda, Fabian Caba Heilbron, Shiry Ginosar, Bryan C. Russell et al.SIGGRAPH 2026
Builds on14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan et al.NeurIPS 2023 · 4,972 citations
- SWE-agent: Agent-Computer Interfaces Enable Automated Software EngineeringJohn Yang, Carlos E. Jimenez, Alexander Wettig, Kilian Lieret et al.NeurIPS 2024 · 2,059 citations
- ViperGPT: Visual Inference via Python Execution for ReasoningDÃdac SurÃs, Sachit Menon, Carl VondrickICCV 2023 · 732 citations
- A Survey on In-context LearningQingxiu Dong, Lei Li, Damai Dai, Ce Zheng et al.EMNLP 2024 · 479 citations
Related papers
- Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative TrainingRunze He, Shaofei Huang, Xuecheng Nie, Tianrui Hui et al.CVPR 2024
- Context-Aware Talking-Head Video EditingSonglin Yang, Wei Wang, Jun Ling, Bo Peng et al.ACM MM 2023 · 10 citations
- Interactive Optimization of Scaffolded Procedural PatternsDavide Sforza, Marzia Riso, Filippo Muzzini, Nicola Capodieci et al.SIGGRAPH 2025
- Mind the Time: Temporally-Controlled Multi-Event Video GenerationZiyi Wu, Aliaksandr Siarohin, Willi Menapace, Ivan Skorokhodov et al.CVPR 2025
- Video-Guided Foley Sound Generation with Multimodal ControlsZiyang Chen, Prem Seetharaman, Bryan C. Russell, Oriol Nieto et al.CVPR 2025
