MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zichen, Yu, Yue, Ouyang, Hao, Wang, Qiuyu, Ma, Shuailei, Cheng, Ka Leong, Wang, Wen, Bai, Qingyan, Zhang, Yuxuan, Zeng, Yanhong, Li, Yixuan, Zhu, Xing, Shen, Yujun, Chen, Qifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MagicQuill: An Intelligent Interactive Image Editing System
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
von: Bai, Qingyan, et al.
Veröffentlicht: (2025)
von: Bai, Qingyan, et al.
Veröffentlicht: (2025)
Edicho: Consistent Image Editing in the Wild
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
Calligrapher: Freestyle Text Image Customization
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
Learning Naturally Aggregated Appearance for Efficient 3D Editing
von: Cheng, Ka Leong, et al.
Veröffentlicht: (2023)
von: Cheng, Ka Leong, et al.
Veröffentlicht: (2023)
Real-time 3D-aware Portrait Editing from a Single Image
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
von: Bai, Qingyan, et al.
Veröffentlicht: (2024)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
von: Ouyang, Hao, et al.
Veröffentlicht: (2023)
von: Ouyang, Hao, et al.
Veröffentlicht: (2023)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
DepthLab: From Partial to Complete
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
Advancing Open-source World Models
von: Robbyant Team, et al.
Veröffentlicht: (2026)
von: Robbyant Team, et al.
Veröffentlicht: (2026)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
von: Lu, Yunhong, et al.
Veröffentlicht: (2025)
von: Lu, Yunhong, et al.
Veröffentlicht: (2025)
MangaNinja: Line Art Colorization with Precise Reference Following
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
AniDoc: Animation Creation Made Easier
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
MagicStick: Controllable Video Editing via Control Handle Transformations
von: Ma, Yue, et al.
Veröffentlicht: (2023)
von: Ma, Yue, et al.
Veröffentlicht: (2023)
InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2024)
Towards Degradation-Robust Reconstruction in Generalizable NeRF
von: Park, Chan Ho, et al.
Veröffentlicht: (2024)
von: Park, Chan Ho, et al.
Veröffentlicht: (2024)
NextQuill: Causal Preference Modeling for Enhancing LLM Personalization
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
Framer: Interactive Frame Interpolation
von: Wang, Wen, et al.
Veröffentlicht: (2024)
von: Wang, Wen, et al.
Veröffentlicht: (2024)
SPICE: A Synergistic, Precise, Iterative, and Customizable Image Editing Workflow
von: Tang, Kenan, et al.
Veröffentlicht: (2025)
von: Tang, Kenan, et al.
Veröffentlicht: (2025)
Probing Visual Planning in Image Editing Models
von: Zhou, Zhimu, et al.
Veröffentlicht: (2026)
von: Zhou, Zhimu, et al.
Veröffentlicht: (2026)
CueNet: Robust Audio-Visual Speaker Extraction through Cross-Modal Cue Mining and Interaction
von: Wang, Jiadong, et al.
Veröffentlicht: (2026)
von: Wang, Jiadong, et al.
Veröffentlicht: (2026)
StructXLIP: Enhancing Vision-language Models with Multimodal Structural Cues
von: Ruan, Zanxi, et al.
Veröffentlicht: (2026)
von: Ruan, Zanxi, et al.
Veröffentlicht: (2026)
Molecular Engineering of the Nano‐Bio Interface for Programmable and Precision Protein Delivery
von: Tianyu Ma, et al.
Veröffentlicht: (2026)
von: Tianyu Ma, et al.
Veröffentlicht: (2026)
SpA2V: Harnessing Spatial Auditory Cues for Audio-driven Spatially-aware Video Generation
von: Pham, Kien T., et al.
Veröffentlicht: (2025)
von: Pham, Kien T., et al.
Veröffentlicht: (2025)
Distribution of the Quill Mite Bubophilus asiobius Parasitizing Western Palaearctic Owls of the Genus Asio
von: Zbigniew Kwieciński, et al.
Veröffentlicht: (2025)
von: Zbigniew Kwieciński, et al.
Veröffentlicht: (2025)
LORE: Latent Optimization for Precise Semantic Control in Rectified Flow-based Image Editing
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2025)
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2025)
High-Fidelity GAN Inversion for Image Attribute Editing
von: Wang, Tengfei, et al.
Veröffentlicht: (2021)
von: Wang, Tengfei, et al.
Veröffentlicht: (2021)
DiffMagicFace: Identity Consistent Facial Editing of Real Videos
von: Yin, Huanghao, et al.
Veröffentlicht: (2026)
von: Yin, Huanghao, et al.
Veröffentlicht: (2026)
The 21-cm signature of X-ray heated halos around galaxies during cosmic dawn
von: Leong, Ka-Hou, et al.
Veröffentlicht: (2025)
von: Leong, Ka-Hou, et al.
Veröffentlicht: (2025)
Magic of quantum hypergraph states
von: Chen, Junjie, et al.
Veröffentlicht: (2023)
von: Chen, Junjie, et al.
Veröffentlicht: (2023)
Seeing and Hearing: Open-domain Visual-Audio Generation with Diffusion Latent Aligners
von: Xing, Yazhou, et al.
Veröffentlicht: (2024)
von: Xing, Yazhou, et al.
Veröffentlicht: (2024)
Diffusion-Based Visual Art Creation: A Survey and New Perspectives
von: Wang, Bingyuan, et al.
Veröffentlicht: (2024)
von: Wang, Bingyuan, et al.
Veröffentlicht: (2024)
PaintBench: Deterministic Evaluation of Precise Visual Editing
von: Xu, Kai, et al.
Veröffentlicht: (2026)
von: Xu, Kai, et al.
Veröffentlicht: (2026)
MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
DragDiffusion: Harnessing Diffusion Models for Interactive Point-based Image Editing
von: Shi, Yujun, et al.
Veröffentlicht: (2023)
von: Shi, Yujun, et al.
Veröffentlicht: (2023)
Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization
von: Zhang, Zhiwang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiwang, et al.
Veröffentlicht: (2025)
Learning Visual Generative Priors without Text
von: Ma, Shuailei, et al.
Veröffentlicht: (2024)
von: Ma, Shuailei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MagicQuill: An Intelligent Interactive Image Editing System
von: Liu, Zichen, et al.
Veröffentlicht: (2024) -
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
von: Bai, Qingyan, et al.
Veröffentlicht: (2025) -
Edicho: Consistent Image Editing in the Wild
von: Bai, Qingyan, et al.
Veröffentlicht: (2024) -
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
von: Wang, Hanlin, et al.
Veröffentlicht: (2025) -
Calligrapher: Freestyle Text Image Customization
von: Ma, Yue, et al.
Veröffentlicht: (2025)