O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuqing, Wang, Junjie, Liu, Lin, Chu, Ruihang, Zhang, Xiaopeng, Tian, Qi, Yang, Yujiu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023)
von: Wang, Tan, et al.
Veröffentlicht: (2023)
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
DisCo3D: Distilling Multi-View Consistency for 3D Scene Editing
von: Chi, Yufeng, et al.
Veröffentlicht: (2025)
von: Chi, Yufeng, et al.
Veröffentlicht: (2025)
GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver
von: Chen, Yuqing, et al.
Veröffentlicht: (2026)
von: Chen, Yuqing, et al.
Veröffentlicht: (2026)
Velocity-Space 3D Asset Editing
von: Liu, Hao, et al.
Veröffentlicht: (2026)
von: Liu, Hao, et al.
Veröffentlicht: (2026)
DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents
von: Xu, Yilun, et al.
Veröffentlicht: (2024)
von: Xu, Yilun, et al.
Veröffentlicht: (2024)
PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video Reasoning
von: Ding, Yang, et al.
Veröffentlicht: (2025)
von: Ding, Yang, et al.
Veröffentlicht: (2025)
AnyCap Project: A Unified Framework, Dataset, and Benchmark for Controllable Omni-modal Captioning
von: Ren, Yiming, et al.
Veröffentlicht: (2025)
von: Ren, Yiming, et al.
Veröffentlicht: (2025)
Video-Zero: Self-Evolution Video Understanding
von: Zhang, Ruixu, et al.
Veröffentlicht: (2026)
von: Zhang, Ruixu, et al.
Veröffentlicht: (2026)
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
CogniEdit: Dense Gradient Flow Optimization for Fine-Grained Image Editing
von: Li, Yan, et al.
Veröffentlicht: (2025)
von: Li, Yan, et al.
Veröffentlicht: (2025)
DreamVE: Unified Instruction-based Image and Video Editing
von: Xia, Bin, et al.
Veröffentlicht: (2025)
von: Xia, Bin, et al.
Veröffentlicht: (2025)
Realistic and Controllable 3D Gaussian-Guided Object Editing for Driving Video Generation
von: Li, Jiusi, et al.
Veröffentlicht: (2025)
von: Li, Jiusi, et al.
Veröffentlicht: (2025)
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
DisCo-FLoc: Semantic-Free Floorplan Localization via $SE(2)$-Aware Contrastive Disambiguation
von: Zhong, Ping, et al.
Veröffentlicht: (2026)
von: Zhong, Ping, et al.
Veröffentlicht: (2026)
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance
von: Xu, Haohang, et al.
Veröffentlicht: (2026)
von: Xu, Haohang, et al.
Veröffentlicht: (2026)
DisCo-Layout: Disentangling and Coordinating Semantic and Physical Refinement in a Multi-Agent Framework for 3D Indoor Layout Synthesis
von: Gao, Jialin, et al.
Veröffentlicht: (2025)
von: Gao, Jialin, et al.
Veröffentlicht: (2025)
ImgEdit: A Unified Image Editing Dataset and Benchmark
von: Ye, Yang, et al.
Veröffentlicht: (2025)
von: Ye, Yang, et al.
Veröffentlicht: (2025)
VideoCoF: Unified Video Editing with Temporal Reasoner
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
Nav-$R^2$ Dual-Relation Reasoning for Generalizable Open-Vocabulary Object-Goal Navigation
von: Xiang, Wentao, et al.
Veröffentlicht: (2025)
von: Xiang, Wentao, et al.
Veröffentlicht: (2025)
AlbedoEdit: Unified Instance-Level Video Editing with Albedo Guidance
von: Zhou, Xilong, et al.
Veröffentlicht: (2026)
von: Zhou, Xilong, et al.
Veröffentlicht: (2026)
Omni-Video 2: Scaling MLLM-Conditioned Diffusion for Unified Video Generation and Editing
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
GaussianEditor: Editing 3D Gaussians Delicately with Text Instructions
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing
von: Liu, Lin, et al.
Veröffentlicht: (2026)
von: Liu, Lin, et al.
Veröffentlicht: (2026)
UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
von: Bai, Jianhong, et al.
Veröffentlicht: (2024)
von: Bai, Jianhong, et al.
Veröffentlicht: (2024)
AdaViewPlanner: Adapting Video Diffusion Models for Viewpoint Planning in 4D Scenes
von: Li, Yu, et al.
Veröffentlicht: (2025)
von: Li, Yu, et al.
Veröffentlicht: (2025)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
EditCtrl: Disentangled Local and Global Control for Real-Time Generative Video Editing
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
Rolling Shutter Correction with Intermediate Distortion Flow Estimation
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
von: Zhu, Hongyang, et al.
Veröffentlicht: (2025)
MLV-Edit: Towards Consistent and Highly Efficient Editing for Minute-Level Videos
von: Cao, Yangyi, et al.
Veröffentlicht: (2026)
von: Cao, Yangyi, et al.
Veröffentlicht: (2026)
UniEdit-I: Training-free Image Editing for Unified VLM via Iterative Understanding, Editing and Verifying
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
TextMaster: A Unified Framework for Realistic Text Editing via Glyph-Style Dual-Control
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
CrimEdit: Controllable Editing for Counterfactual Object Removal, Insertion, and Movement
von: Jeon, Boseong, et al.
Veröffentlicht: (2025)
von: Jeon, Boseong, et al.
Veröffentlicht: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
von: Lin, Yuxuan, et al.
Veröffentlicht: (2025)
von: Lin, Yuxuan, et al.
Veröffentlicht: (2025)
LoVoRA: Text-guided and Mask-free Video Object Removal and Addition with Learnable Object-aware Localization
von: Xiao, Zhihan, et al.
Veröffentlicht: (2025)
von: Xiao, Zhihan, et al.
Veröffentlicht: (2025)
Edit As You Wish: Video Caption Editing with Multi-grained User Control
von: Yao, Linli, et al.
Veröffentlicht: (2023)
von: Yao, Linli, et al.
Veröffentlicht: (2023)
AdaEdit: Adaptive Temporal and Channel Modulation for Flow-Based Image Editing
von: Li, Guandong, et al.
Veröffentlicht: (2026)
von: Li, Guandong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023) -
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025) -
DisCo3D: Distilling Multi-View Consistency for 3D Scene Editing
von: Chi, Yufeng, et al.
Veröffentlicht: (2025) -
GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver
von: Chen, Yuqing, et al.
Veröffentlicht: (2026) -
Velocity-Space 3D Asset Editing
von: Liu, Hao, et al.
Veröffentlicht: (2026)