Unified Editing of Panorama, 3D Scenes, and Videos Through Disentangled Self-Attention Injection
Fuente:
arXiv
Saved in:
| Main Authors: | Kwon, Gihyun, Park, Jangho, Ye, Jong Chul |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ED-NeRF: Efficient Text-Guided Editing of 3D Scene with Latent Space NeRF
by: Park, Jangho, et al.
Published: (2023)
by: Park, Jangho, et al.
Published: (2023)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023)
by: Nam, Hyelin, et al.
Published: (2023)
FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching
by: Park, Jangho, et al.
Published: (2026)
by: Park, Jangho, et al.
Published: (2026)
Zero4D: Training-Free 4D Video Generation From Single Video Using Off-the-Shelf Video Diffusion
by: Park, Jangho, et al.
Published: (2025)
by: Park, Jangho, et al.
Published: (2025)
TweedieMix: Improving Multi-Concept Fusion for Diffusion-based Image/Video Generation
by: Kwon, Gihyun, et al.
Published: (2024)
by: Kwon, Gihyun, et al.
Published: (2024)
Unpaired Image-to-Image Translation via Neural Schrödinger Bridge
by: Kim, Beomsu, et al.
Published: (2023)
by: Kim, Beomsu, et al.
Published: (2023)
DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
Geometric 4D Stitching for Grounded 4D Generation
by: Park, Sunwoo, et al.
Published: (2026)
by: Park, Sunwoo, et al.
Published: (2026)
Solving Video Inverse Problems Using Image Diffusion Models
by: Kwon, Taesung, et al.
Published: (2024)
by: Kwon, Taesung, et al.
Published: (2024)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
by: Jeong, Hyeonho, et al.
Published: (2023)
by: Jeong, Hyeonho, et al.
Published: (2023)
VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion Models
by: Kwon, Taesung, et al.
Published: (2024)
by: Kwon, Taesung, et al.
Published: (2024)
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation
by: Jeong, Hyeonho, et al.
Published: (2025)
by: Jeong, Hyeonho, et al.
Published: (2025)
Patch-wise Graph Contrastive Learning for Image Translation
by: Jung, Chanyong, et al.
Published: (2023)
by: Jung, Chanyong, et al.
Published: (2023)
Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models
by: Kwon, Taesung, et al.
Published: (2026)
by: Kwon, Taesung, et al.
Published: (2026)
CRePE: Curved Ray Expectation Positional Encoding for Unified-Camera-Controlled Video Generation
by: Jin, Seonghyun, et al.
Published: (2026)
by: Jin, Seonghyun, et al.
Published: (2026)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
by: Lee, Dohun, et al.
Published: (2024)
by: Lee, Dohun, et al.
Published: (2024)
Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
by: Kwon, Gihyun, et al.
Published: (2024)
by: Kwon, Gihyun, et al.
Published: (2024)
ViBiDSampler: Enhancing Video Interpolation Using Bidirectional Diffusion Sampler
by: Yang, Serin, et al.
Published: (2024)
by: Yang, Serin, et al.
Published: (2024)
Training-Free Reward-Guided Image Editing via Trajectory Optimal Control
by: Chang, Jinho, et al.
Published: (2025)
by: Chang, Jinho, et al.
Published: (2025)
EditSplat: Multi-View Fusion and Attention-Guided Optimization for View-Consistent 3D Scene Editing with 3D Gaussian Splatting
by: Lee, Dong In, et al.
Published: (2024)
by: Lee, Dong In, et al.
Published: (2024)
FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing
by: Kim, Jeongsol, et al.
Published: (2025)
by: Kim, Jeongsol, et al.
Published: (2025)
DreamSampler: Unifying Diffusion Sampling and Score Distillation for Image Manipulation
by: Kim, Jeongsol, et al.
Published: (2024)
by: Kim, Jeongsol, et al.
Published: (2024)
FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction
by: Jin, Seonghyun, et al.
Published: (2026)
by: Jin, Seonghyun, et al.
Published: (2026)
Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution
by: Kim, Bryan Sangwoo, et al.
Published: (2026)
by: Kim, Bryan Sangwoo, et al.
Published: (2026)
FlowLPS: Langevin-Proximal Sampling for Flow-based Inverse Problem Solvers
by: Park, Jonghyun, et al.
Published: (2025)
by: Park, Jonghyun, et al.
Published: (2025)
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
by: Lee, Dohun, et al.
Published: (2025)
by: Lee, Dohun, et al.
Published: (2025)
Self-Guided Generation of Minority Samples Using Diffusion Models
by: Um, Soobin, et al.
Published: (2024)
by: Um, Soobin, et al.
Published: (2024)
TripleFDS: Triple Feature Disentanglement and Synthesis for Scene Text Editing
by: Bao, Yuchen, et al.
Published: (2025)
by: Bao, Yuchen, et al.
Published: (2025)
CellPainTR: Generalizable Representation Learning for Cross-Dataset Cell Painting Analysis
by: Caruzzo, Cedric, et al.
Published: (2025)
by: Caruzzo, Cedric, et al.
Published: (2025)
S3Editor: A Sparse Semantic-Disentangled Self-Training Framework for Face Video Editing
by: Wang, Guangzhi, et al.
Published: (2024)
by: Wang, Guangzhi, et al.
Published: (2024)
Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
by: Lee, Dohun, et al.
Published: (2026)
by: Lee, Dohun, et al.
Published: (2026)
OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation
by: Kim, Kwanyoung, et al.
Published: (2024)
by: Kim, Kwanyoung, et al.
Published: (2024)
Inference-Time Diffusion Model Distillation
by: Park, Geon Yeong, et al.
Published: (2024)
by: Park, Geon Yeong, et al.
Published: (2024)
Edit-As-Act: Goal-Regressive Planning for Open-Vocabulary 3D Indoor Scene Editing
by: Noh, Seongrae, et al.
Published: (2026)
by: Noh, Seongrae, et al.
Published: (2026)
Spectral Motion Alignment for Video Motion Transfer using Diffusion Models
by: Park, Geon Yeong, et al.
Published: (2024)
by: Park, Geon Yeong, et al.
Published: (2024)
Re-Attentional Controllable Video Diffusion Editing
by: Wang, Yuanzhi, et al.
Published: (2024)
by: Wang, Yuanzhi, et al.
Published: (2024)
Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion
by: Kang, Xueyang, et al.
Published: (2025)
by: Kang, Xueyang, et al.
Published: (2025)
LatentEditor: Text Driven Local Editing of 3D Scenes
by: Khalid, Umar, et al.
Published: (2023)
by: Khalid, Umar, et al.
Published: (2023)
Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models
by: Lee, Jeongjae, et al.
Published: (2026)
by: Lee, Jeongjae, et al.
Published: (2026)
Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM
by: Kim, Jaemin, et al.
Published: (2024)
by: Kim, Jaemin, et al.
Published: (2024)
Similar Items
-
ED-NeRF: Efficient Text-Guided Editing of 3D Scene with Latent Space NeRF
by: Park, Jangho, et al.
Published: (2023) -
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023) -
FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching
by: Park, Jangho, et al.
Published: (2026) -
Zero4D: Training-Free 4D Video Generation From Single Video Using Off-the-Shelf Video Diffusion
by: Park, Jangho, et al.
Published: (2025) -
TweedieMix: Improving Multi-Concept Fusion for Diffusion-based Image/Video Generation
by: Kwon, Gihyun, et al.
Published: (2024)