GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zewei, Liu, Huan, Chen, Jun, Xu, Xiangyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
von: Xu, Xiangyu, et al.
Veröffentlicht: (2024)
von: Xu, Xiangyu, et al.
Veröffentlicht: (2024)
Instant3D: Instant Text-to-3D Generation
von: Li, Ming, et al.
Veröffentlicht: (2023)
von: Li, Ming, et al.
Veröffentlicht: (2023)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
von: Chiu, Pin-Yen, et al.
Veröffentlicht: (2025)
von: Chiu, Pin-Yen, et al.
Veröffentlicht: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
von: Liu, Ziyuan, et al.
Veröffentlicht: (2026)
von: Liu, Ziyuan, et al.
Veröffentlicht: (2026)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
Cross-Scenario Deraining Adaptation with Unpaired Data: Superpixel Structural Priors and Multi-Stage Pseudo-Rain Synthesis
von: Zhao, Kangbo, et al.
Veröffentlicht: (2026)
von: Zhao, Kangbo, et al.
Veröffentlicht: (2026)
ToonAging: Face Re-Aging upon Artistic Portrait Style Transfer
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
von: Singer, Assaf, et al.
Veröffentlicht: (2025)
von: Singer, Assaf, et al.
Veröffentlicht: (2025)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
von: Girdhar, Rohit, et al.
Veröffentlicht: (2023)
von: Girdhar, Rohit, et al.
Veröffentlicht: (2023)
Instruction-Driven 3D Facial Expression Generation and Transition
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
von: Bilecen, Bahri Batuhan, et al.
Veröffentlicht: (2024)
von: Bilecen, Bahri Batuhan, et al.
Veröffentlicht: (2024)
Magic Insert: Style-Aware Drag-and-Drop
von: Ruiz, Nataniel, et al.
Veröffentlicht: (2024)
von: Ruiz, Nataniel, et al.
Veröffentlicht: (2024)
DragVideo: Interactive Drag-style Video Editing
von: Deng, Yufan, et al.
Veröffentlicht: (2023)
von: Deng, Yufan, et al.
Veröffentlicht: (2023)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic Propagation
von: Liu, Haofeng, et al.
Veröffentlicht: (2024)
von: Liu, Haofeng, et al.
Veröffentlicht: (2024)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
von: Flynn, John, et al.
Veröffentlicht: (2026)
von: Flynn, John, et al.
Veröffentlicht: (2026)
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
A Survey on 3D Gaussian Splatting
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
Cert-LAS: Toward Certified Model Ownership Verification for Text-to-Image Diffusion Models via Layer-Adaptive Smoothing
von: Qi, Leyi, et al.
Veröffentlicht: (2026)
von: Qi, Leyi, et al.
Veröffentlicht: (2026)
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
von: Dai, Yudi, et al.
Veröffentlicht: (2024)
von: Dai, Yudi, et al.
Veröffentlicht: (2024)
Seeing World Dynamics in a Nutshell
von: Shen, Qiuhong, et al.
Veröffentlicht: (2025)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2025)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
von: Guo, Xin, et al.
Veröffentlicht: (2025)
von: Guo, Xin, et al.
Veröffentlicht: (2025)
Diffusion Model-Based Video Editing: A Survey
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
DragFlow: Unleashing DiT Priors with Region Based Supervision for Drag Editing
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
von: Zakharov, Sergey, et al.
Veröffentlicht: (2024)
von: Zakharov, Sergey, et al.
Veröffentlicht: (2024)
Reproducing DragDiffusion: Interactive Point-Based Editing with Diffusion Models
von: Subhan, Ali, et al.
Veröffentlicht: (2026)
von: Subhan, Ali, et al.
Veröffentlicht: (2026)
Lester: rotoscope animation through video object segmentation and tracking
von: Tous, Ruben
Veröffentlicht: (2024)
von: Tous, Ruben
Veröffentlicht: (2024)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
von: Sar, Ayan, et al.
Veröffentlicht: (2025)
von: Sar, Ayan, et al.
Veröffentlicht: (2025)
Extreme Compression of Adaptive Neural Images
von: Hoshikawa, Leo, et al.
Veröffentlicht: (2024)
von: Hoshikawa, Leo, et al.
Veröffentlicht: (2024)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
Freehand Sketch Generation from Mechanical Components
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
von: Liao, Zhichao, et al.
Veröffentlicht: (2024)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
von: Lin, Jiantao, et al.
Veröffentlicht: (2025)
von: Lin, Jiantao, et al.
Veröffentlicht: (2025)
Drag Your Gaussian: Effective Drag-Based Editing with Score Distillation for 3D Gaussian Splatting
von: Qu, Yansong, et al.
Veröffentlicht: (2025)
von: Qu, Yansong, et al.
Veröffentlicht: (2025)
Improving Generative Adversarial Network Generalization for Facial Expression Synthesis
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
von: Akram, Arbish, et al.
Veröffentlicht: (2026)
ParamsDrag: Interactive Parameter Space Exploration via Image-Space Dragging
von: Li, Guan, et al.
Veröffentlicht: (2024)
von: Li, Guan, et al.
Veröffentlicht: (2024)
DiffUHaul: A Training-Free Method for Object Dragging in Images
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
von: Avrahami, Omri, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
von: Xu, Xiangyu, et al.
Veröffentlicht: (2024) -
Instant3D: Instant Text-to-3D Generation
von: Li, Ming, et al.
Veröffentlicht: (2023) -
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
von: Sun, Zeyi, et al.
Veröffentlicht: (2024) -
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
von: Kim, Bumsoo, et al.
Veröffentlicht: (2024) -
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
von: Chiu, Pin-Yen, et al.
Veröffentlicht: (2025)