EditDuet: A Multi-Agent System for Video Non-Linear Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sandoval-Castaneda, Marcelo, Russell, Bryan, Sivic, Josef, Shakhnarovich, Gregory, Heilbron, Fabian Caba |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ResidualViT for Efficient Temporally Dense Video Encoding
von: Soldan, Mattia, et al.
Veröffentlicht: (2025)
von: Soldan, Mattia, et al.
Veröffentlicht: (2025)
Discovering Divergent Representations between Text-to-Image Models
von: Dunlap, Lisa, et al.
Veröffentlicht: (2025)
von: Dunlap, Lisa, et al.
Veröffentlicht: (2025)
Improving Personalized Search with Regularized Low-Rank Parameter Updates
von: Ryan, Fiona, et al.
Veröffentlicht: (2025)
von: Ryan, Fiona, et al.
Veröffentlicht: (2025)
Generative Timelines for Instructed Visual Assembly
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
Adapting Dual-encoder Vision-language Models for Paraphrased Retrieval
von: Cheng, Jiacheng, et al.
Veröffentlicht: (2024)
von: Cheng, Jiacheng, et al.
Veröffentlicht: (2024)
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets
von: Dave, Ishan Rajendrakumar, et al.
Veröffentlicht: (2024)
von: Dave, Ishan Rajendrakumar, et al.
Veröffentlicht: (2024)
Grounded Video Caption Generation
von: Kazakos, Evangelos, et al.
Veröffentlicht: (2024)
von: Kazakos, Evangelos, et al.
Veröffentlicht: (2024)
Large-scale Pre-training for Grounded Video Caption Generation
von: Kazakos, Evangelos, et al.
Veröffentlicht: (2025)
von: Kazakos, Evangelos, et al.
Veröffentlicht: (2025)
NewMove: Customizing text-to-video models with novel motions
von: Materzynska, Joanna, et al.
Veröffentlicht: (2023)
von: Materzynska, Joanna, et al.
Veröffentlicht: (2023)
Videogenic: Identifying Highlight Moments in Videos with Professional Photographs as a Prior
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
Scaling Up Video Summarization Pretraining with Large Language Models
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2024)
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2024)
FocalPose++: Focal Length and Object Pose Estimation via Render and Compare
von: Cífka, Martin, et al.
Veröffentlicht: (2023)
von: Cífka, Martin, et al.
Veröffentlicht: (2023)
CineVerse: Consistent Keyframe Synthesis for Cinematic Scene Composition
von: Phung, Quynh, et al.
Veröffentlicht: (2025)
von: Phung, Quynh, et al.
Veröffentlicht: (2025)
GenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos
von: Souček, Tomáš, et al.
Veröffentlicht: (2023)
von: Souček, Tomáš, et al.
Veröffentlicht: (2023)
Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
von: Kwon, Gihyun, et al.
Veröffentlicht: (2024)
von: Kwon, Gihyun, et al.
Veröffentlicht: (2024)
Persistent Robot World Models: Stabilizing Multi-Step Rollouts via Reinforcement Learning
von: Bardhan, Jai, et al.
Veröffentlicht: (2026)
von: Bardhan, Jai, et al.
Veröffentlicht: (2026)
Towards Automated Movie Trailer Generation
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2024)
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2024)
Edit As You Wish: Video Caption Editing with Multi-grained User Control
von: Yao, Linli, et al.
Veröffentlicht: (2023)
von: Yao, Linli, et al.
Veröffentlicht: (2023)
FastVideoEdit: Leveraging Consistency Models for Efficient Text-to-Video Editing
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment
von: Mikeštíková, Anna Šárová, et al.
Veröffentlicht: (2025)
von: Mikeštíková, Anna Šárová, et al.
Veröffentlicht: (2025)
LoopDraw: a Loop-Based Autoregressive Model for Shape Synthesis and Editing
von: Dinh, Nam Anh, et al.
Veröffentlicht: (2022)
von: Dinh, Nam Anh, et al.
Veröffentlicht: (2022)
ParallelEdits: Efficient Multi-object Image Editing
von: Huang, Mingzhen, et al.
Veröffentlicht: (2024)
von: Huang, Mingzhen, et al.
Veröffentlicht: (2024)
Edit3K: Universal Representation Learning for Video Editing Components
von: Gu, Xin, et al.
Veröffentlicht: (2024)
von: Gu, Xin, et al.
Veröffentlicht: (2024)
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
Video4Edit: Viewing Image Editing as a Degenerate Temporal Process
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
OphEdit: Training-Free Text-Guided Editing of Ophthalmic Surgical Videos
von: Jangir, Ritul, et al.
Veröffentlicht: (2026)
von: Jangir, Ritul, et al.
Veröffentlicht: (2026)
ExpertEdit: Learning Skill-Aware Motion Editing from Expert Videos
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2026)
von: Somayazulu, Arjun, et al.
Veröffentlicht: (2026)
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
von: Couairon, Paul, et al.
Veröffentlicht: (2023)
von: Couairon, Paul, et al.
Veröffentlicht: (2023)
MLV-Edit: Towards Consistent and Highly Efficient Editing for Minute-Level Videos
von: Cao, Yangyi, et al.
Veröffentlicht: (2026)
von: Cao, Yangyi, et al.
Veröffentlicht: (2026)
6D Object Pose Tracking in Internet Videos for Robotic Manipulation
von: Ponimatkin, Georgy, et al.
Veröffentlicht: (2025)
von: Ponimatkin, Georgy, et al.
Veröffentlicht: (2025)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
AlbedoEdit: Unified Instance-Level Video Editing with Albedo Guidance
von: Zhou, Xilong, et al.
Veröffentlicht: (2026)
von: Zhou, Xilong, et al.
Veröffentlicht: (2026)
FlexiEdit: Frequency-Aware Latent Refinement for Enhanced Non-Rigid Editing
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
MoEdit: On Learning Quantity Perception for Multi-object Image Editing
von: Li, Yanfeng, et al.
Veröffentlicht: (2025)
von: Li, Yanfeng, et al.
Veröffentlicht: (2025)
Edit-Your-Interest: Efficient Video Editing via Feature Most-Similar Propagation
von: Zuo, Yi, et al.
Veröffentlicht: (2025)
von: Zuo, Yi, et al.
Veröffentlicht: (2025)
UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
von: Bai, Jianhong, et al.
Veröffentlicht: (2024)
von: Bai, Jianhong, et al.
Veröffentlicht: (2024)
Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing
von: Zuo, Yi, et al.
Veröffentlicht: (2024)
von: Zuo, Yi, et al.
Veröffentlicht: (2024)
AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner
von: Zheng, Haojie, et al.
Veröffentlicht: (2025)
von: Zheng, Haojie, et al.
Veröffentlicht: (2025)
EditCtrl: Disentangled Local and Global Control for Real-Time Generative Video Editing
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ResidualViT for Efficient Temporally Dense Video Encoding
von: Soldan, Mattia, et al.
Veröffentlicht: (2025) -
Discovering Divergent Representations between Text-to-Image Models
von: Dunlap, Lisa, et al.
Veröffentlicht: (2025) -
Improving Personalized Search with Regularized Low-Rank Parameter Updates
von: Ryan, Fiona, et al.
Veröffentlicht: (2025) -
Generative Timelines for Instructed Visual Assembly
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024) -
Adapting Dual-encoder Vision-language Models for Paraphrased Retrieval
von: Cheng, Jiacheng, et al.
Veröffentlicht: (2024)