LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Siyi, Chen, Weiming, Tang, Yushun, He, Zhihai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent Bias Alignment for High-Fidelity Diffusion Inversion in Real-World Image Reconstruction and Manipulation
by: Chen, Weiming, et al.
Published: (2026)
by: Chen, Weiming, et al.
Published: (2026)
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024)
by: Jiang, Liyao, et al.
Published: (2024)
LEAD: Latent Realignment for Human Motion Diffusion
by: Andreou, Nefeli, et al.
Published: (2024)
by: Andreou, Nefeli, et al.
Published: (2024)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
by: Hong, Seokhyeon, et al.
Published: (2025)
by: Hong, Seokhyeon, et al.
Published: (2025)
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
by: Xu, Yu, et al.
Published: (2025)
by: Xu, Yu, et al.
Published: (2025)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
by: Wu, Guanjun, et al.
Published: (2025)
by: Wu, Guanjun, et al.
Published: (2025)
EditRoom: LLM-parameterized Graph Diffusion for Composable 3D Room Layout Editing
by: Zheng, Kaizhi, et al.
Published: (2024)
by: Zheng, Kaizhi, et al.
Published: (2024)
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
by: Wu, Zhennan, et al.
Published: (2024)
by: Wu, Zhennan, et al.
Published: (2024)
Learning Latent Representations for Image Translation using Frequency Distributed CycleGAN
by: Nigam, Shivangi, et al.
Published: (2025)
by: Nigam, Shivangi, et al.
Published: (2025)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
DragPoser: Motion Reconstruction from Variable Sparse Tracking Signals via Latent Space Optimization
by: Ponton, Jose Luis, et al.
Published: (2024)
by: Ponton, Jose Luis, et al.
Published: (2024)
NeRF-Insert: 3D Local Editing with Multimodal Control Signals
by: Sabat, Benet Oriol, et al.
Published: (2024)
by: Sabat, Benet Oriol, et al.
Published: (2024)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
by: Chen, Weiming, et al.
Published: (2025)
by: Chen, Weiming, et al.
Published: (2025)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
by: Shah, Ansh, et al.
Published: (2024)
by: Shah, Ansh, et al.
Published: (2024)
End-to-End Training for Unified Tokenization and Latent Denoising
by: Duggal, Shivam, et al.
Published: (2026)
by: Duggal, Shivam, et al.
Published: (2026)
Consistency^2: Consistent and Fast 3D Painting with Latent Consistency Models
by: Wang, Tianfu, et al.
Published: (2024)
by: Wang, Tianfu, et al.
Published: (2024)
GraphicsDreamer: Image to 3D Generation with Physical Consistency
by: Chen, Pei, et al.
Published: (2024)
by: Chen, Pei, et al.
Published: (2024)
MV-S2V: Multi-View Subject-Consistent Video Generation
by: Song, Ziyang, et al.
Published: (2026)
by: Song, Ziyang, et al.
Published: (2026)
Lazy Diffusion Transformer for Interactive Image Editing
by: Nitzan, Yotam, et al.
Published: (2024)
by: Nitzan, Yotam, et al.
Published: (2024)
FaceParts: Segmentation and Editing of Gaussian Splatting
by: Zapała, Tymoteusz, et al.
Published: (2026)
by: Zapała, Tymoteusz, et al.
Published: (2026)
Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer
by: Yin, Zixin, et al.
Published: (2025)
by: Yin, Zixin, et al.
Published: (2025)
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
Key-Locked Rank One Editing for Text-to-Image Personalization
by: Tewel, Yoad, et al.
Published: (2023)
by: Tewel, Yoad, et al.
Published: (2023)
Frankenstein: Generating Semantic-Compositional 3D Scenes in One Tri-Plane
by: Yan, Han, et al.
Published: (2024)
by: Yan, Han, et al.
Published: (2024)
Materialist: Physically Based Editing Using Single-Image Inverse Rendering
by: Wang, Lezhong, et al.
Published: (2025)
by: Wang, Lezhong, et al.
Published: (2025)
Exploration and Improvement of Nerf-based 3D Scene Editing Techniques
by: Fang, Shun, et al.
Published: (2024)
by: Fang, Shun, et al.
Published: (2024)
Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors
by: Tong, Mutian, et al.
Published: (2025)
by: Tong, Mutian, et al.
Published: (2025)
3DiFACE: Synthesizing and Editing Holistic 3D Facial Animation
by: Thambiraja, Balamurugan, et al.
Published: (2025)
by: Thambiraja, Balamurugan, et al.
Published: (2025)
VectorGym: A Multitask Benchmark for SVG Code Generation, Sketching, and Editing
by: Rodriguez, Juan, et al.
Published: (2026)
by: Rodriguez, Juan, et al.
Published: (2026)
CASIM: Composite Aware Semantic Injection for Text to Motion Generation
by: Chang, Che-Jui, et al.
Published: (2025)
by: Chang, Che-Jui, et al.
Published: (2025)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)
by: Daiya, Divyanshu, et al.
Published: (2024)
Learning to Edit Visual Programs with Self-Supervision
by: Jones, R. Kenny, et al.
Published: (2024)
by: Jones, R. Kenny, et al.
Published: (2024)
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
by: Xu, Tian-Xing, et al.
Published: (2025)
by: Xu, Tian-Xing, et al.
Published: (2025)
V-LASIK: Consistent Glasses-Removal from Videos Using Synthetic Data
by: Shalev-Arkushin, Rotem, et al.
Published: (2024)
by: Shalev-Arkushin, Rotem, et al.
Published: (2024)
DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos
by: Hu, Wenbo, et al.
Published: (2024)
by: Hu, Wenbo, et al.
Published: (2024)
GarmentCrafter: Progressive Novel View Synthesis for Single-View 3D Garment Reconstruction and Editing
by: Wang, Yuanhao, et al.
Published: (2025)
by: Wang, Yuanhao, et al.
Published: (2025)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
by: Ye, Weicai, et al.
Published: (2024)
by: Ye, Weicai, et al.
Published: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
ReplaceAnything3D:Text-Guided 3D Scene Editing with Compositional Neural Radiance Fields
by: Bartrum, Edward, et al.
Published: (2024)
by: Bartrum, Edward, et al.
Published: (2024)
Taming Diffusion Probabilistic Models for Character Control
by: Chen, Rui, et al.
Published: (2024)
by: Chen, Rui, et al.
Published: (2024)
Similar Items
-
Latent Bias Alignment for High-Fidelity Diffusion Inversion in Real-World Image Reconstruction and Manipulation
by: Chen, Weiming, et al.
Published: (2026) -
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024) -
LEAD: Latent Realignment for Human Motion Diffusion
by: Andreou, Nefeli, et al.
Published: (2024) -
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
by: Hong, Seokhyeon, et al.
Published: (2025) -
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
by: Xu, Yu, et al.
Published: (2025)