UNIC: Unified In-Context Video Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Zixuan, He, Xuanhua, Liu, Quande, Wang, Qiulin, Wang, Xintao, Wan, Pengfei, Zhang, Di, Gai, Kun, Chen, Qifeng, Luo, Wenhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers
von: He, Xuanhua, et al.
Veröffentlicht: (2025)
von: He, Xuanhua, et al.
Veröffentlicht: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025)
von: Wei, Cong, et al.
Veröffentlicht: (2025)
VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning
von: Cai, Minghong, et al.
Veröffentlicht: (2025)
von: Cai, Minghong, et al.
Veröffentlicht: (2025)
Visual-Aware CoT: Achieving High-Fidelity Visual Consistency in Unified Models
von: Ye, Zixuan, et al.
Veröffentlicht: (2025)
von: Ye, Zixuan, et al.
Veröffentlicht: (2025)
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
StyleMaster: Stylize Your Video with Artistic Generation and Translation
von: Ye, Zixuan, et al.
Veröffentlicht: (2024)
von: Ye, Zixuan, et al.
Veröffentlicht: (2024)
UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation
von: Xu, Yiyan, et al.
Veröffentlicht: (2026)
von: Xu, Yiyan, et al.
Veröffentlicht: (2026)
A Survey of Interactive Generative Video
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
A Reason-then-Describe Instruction Interpreter for Controllable Video Generation
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025)
von: Du, Shian, et al.
Veröffentlicht: (2025)
VINO: A Unified Visual Generator with Interleaved OmniModal Context
von: Chen, Junyi, et al.
Veröffentlicht: (2026)
von: Chen, Junyi, et al.
Veröffentlicht: (2026)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
In-Context Audio Control of Video Diffusion Transformers
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
SketchVideo: Sketch-based Video Generation and Editing
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
von: Liu, Feng-Lin, et al.
Veröffentlicht: (2025)
ContextFlow: Training-Free Video Object Editing via Adaptive Context Enrichment
von: Chen, Yiyang, et al.
Veröffentlicht: (2025)
von: Chen, Yiyang, et al.
Veröffentlicht: (2025)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
von: Luo, Yawen, et al.
Veröffentlicht: (2026)
von: Luo, Yawen, et al.
Veröffentlicht: (2026)
Position: Interactive Generative Video as Next-Generation Game Engine
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
von: He, Haoran, et al.
Veröffentlicht: (2025)
von: He, Haoran, et al.
Veröffentlicht: (2025)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
von: Luo, Yawen, et al.
Veröffentlicht: (2025)
von: Luo, Yawen, et al.
Veröffentlicht: (2025)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
GameGen-X: Interactive Open-world Game Video Generation
von: Che, Haoxuan, et al.
Veröffentlicht: (2024)
von: Che, Haoxuan, et al.
Veröffentlicht: (2024)
RelightMaster: Precise Video Relighting with Multi-plane Light Images
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
von: Bian, Weikang, et al.
Veröffentlicht: (2025)
InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation
von: Rao, Zhefan, et al.
Veröffentlicht: (2026)
von: Rao, Zhefan, et al.
Veröffentlicht: (2026)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
UNIC: Learning Unified Multimodal Extrinsic Contact Estimation
von: Xu, Zhengtong, et al.
Veröffentlicht: (2026)
von: Xu, Zhengtong, et al.
Veröffentlicht: (2026)
GameFactory: Creating New Games with Generative Interactive Videos
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
Geometry-Aware Implicit Memory for Video World Models
von: Wei, Zhengxuan, et al.
Veröffentlicht: (2026)
von: Wei, Zhengxuan, et al.
Veröffentlicht: (2026)
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation
von: Duan, Lunhao, et al.
Veröffentlicht: (2024)
von: Duan, Lunhao, et al.
Veröffentlicht: (2024)
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
SemanticGen: Video Generation in Semantic Space
von: Bai, Jianhong, et al.
Veröffentlicht: (2025)
von: Bai, Jianhong, et al.
Veröffentlicht: (2025)
PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025)
von: Du, Shian, et al.
Veröffentlicht: (2025)
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
von: Ma, Yue, et al.
Veröffentlicht: (2023)
von: Ma, Yue, et al.
Veröffentlicht: (2023)
FilmWeaver: Weaving Consistent Multi-Shot Videos with Cache-Guided Autoregressive Diffusion
von: Luo, Xiangyang, et al.
Veröffentlicht: (2025)
von: Luo, Xiangyang, et al.
Veröffentlicht: (2025)
DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory
von: Yang, Zhenhao, et al.
Veröffentlicht: (2026)
von: Yang, Zhenhao, et al.
Veröffentlicht: (2026)
Improving Video Generation with Human Feedback
von: Liu, Jie, et al.
Veröffentlicht: (2025)
von: Liu, Jie, et al.
Veröffentlicht: (2025)
VideoTetris: Towards Compositional Text-to-Video Generation
von: Tian, Ye, et al.
Veröffentlicht: (2024)
von: Tian, Ye, et al.
Veröffentlicht: (2024)
Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control
von: Fu, Xiao, et al.
Veröffentlicht: (2025)
von: Fu, Xiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers
von: He, Xuanhua, et al.
Veröffentlicht: (2025) -
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025) -
VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning
von: Cai, Minghong, et al.
Veröffentlicht: (2025) -
Visual-Aware CoT: Achieving High-Fidelity Visual Consistency in Unified Models
von: Ye, Zixuan, et al.
Veröffentlicht: (2025) -
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
von: Ju, Xuan, et al.
Veröffentlicht: (2025)