Taming Rectified Flow for Inversion and Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jiangshan, Pu, Junfu, Qi, Zhongang, Guo, Jiayi, Ma, Yue, Huang, Nisha, Chen, Yuxin, Li, Xiu, Shan, Ying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
Adams Bashforth Moulton Solver for Inversion and Editing in Rectified Flow
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?
von: Chen, Yuxin, et al.
Veröffentlicht: (2024)
von: Chen, Yuxin, et al.
Veröffentlicht: (2024)
OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video
von: Pu, Junfu, et al.
Veröffentlicht: (2026)
von: Pu, Junfu, et al.
Veröffentlicht: (2026)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)
von: Liu, Ye, et al.
Veröffentlicht: (2025)
SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
FireFlow: Fast Inversion of Rectified Flow for Image Semantic Editing
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing
von: Chen, Ze, et al.
Veröffentlicht: (2026)
von: Chen, Ze, et al.
Veröffentlicht: (2026)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
GRA: Detecting Oriented Objects through Group-wise Rotating and Attention
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
Optimal Transport for Rectified Flow Image Editing: Unifying Inversion-Based and Direct Methods
von: Lupascu, Marian, et al.
Veröffentlicht: (2025)
von: Lupascu, Marian, et al.
Veröffentlicht: (2025)
ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing
von: Huang, Nisha, et al.
Veröffentlicht: (2025)
von: Huang, Nisha, et al.
Veröffentlicht: (2025)
Elastic Diffusion Transformer
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
TReFT: Taming Rectified Flow Models For One-Step Image Translation
von: Li, Shengqian, et al.
Veröffentlicht: (2025)
von: Li, Shengqian, et al.
Veröffentlicht: (2025)
ARC-Chapter: Structuring Hour-Long Videos into Navigable Chapters and Hierarchical Summaries
von: Pu, Junfu, et al.
Veröffentlicht: (2025)
von: Pu, Junfu, et al.
Veröffentlicht: (2025)
Inversion-Free Style Transfer with Dual Rectified Flows
von: Deng, Yingying, et al.
Veröffentlicht: (2025)
von: Deng, Yingying, et al.
Veröffentlicht: (2025)
Semantic Image Inversion and Editing using Rectified Stochastic Differential Equations
von: Rout, Litu, et al.
Veröffentlicht: (2024)
von: Rout, Litu, et al.
Veröffentlicht: (2024)
Free Lunch for Stabilizing Rectified Flow Inversion
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
Efficient Rectified Flow for Image Fusion
von: Wang, Zirui, et al.
Veröffentlicht: (2025)
von: Wang, Zirui, et al.
Veröffentlicht: (2025)
PreciseCache: Precise Feature Caching for Efficient and High-fidelity Video Generation
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
SynopGround: A Large-Scale Dataset for Multi-Paragraph Video Grounding from TV Dramas and Synopses
von: Tan, Chaolei, et al.
Veröffentlicht: (2024)
von: Tan, Chaolei, et al.
Veröffentlicht: (2024)
DNAEdit: Direct Noise Alignment for Text-Guided Rectified Flow Editing
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding
von: Liu, Ye, et al.
Veröffentlicht: (2024)
von: Liu, Ye, et al.
Veröffentlicht: (2024)
Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
Mono2Stereo: A Benchmark and Empirical Study for Stereo Conversion
von: Yu, Songsong, et al.
Veröffentlicht: (2025)
von: Yu, Songsong, et al.
Veröffentlicht: (2025)
Delta Rectified Flow Sampling for Text-to-Image Editing
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
UniEdit-Flow: Unleashing Inversion and Editing in the Era of Flow Models
von: Jiao, Guanlong, et al.
Veröffentlicht: (2025)
von: Jiao, Guanlong, et al.
Veröffentlicht: (2025)
FluxSpace: Disentangled Semantic Editing in Rectified Flow Transformers
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
EA-VTR: Event-Aware Video-Text Retrieval
von: Ma, Zongyang, et al.
Veröffentlicht: (2024)
von: Ma, Zongyang, et al.
Veröffentlicht: (2024)
FlowBypass: Rectified Flow Trajectory Bypass for Training-Free Image Editing
von: Han, Menglin, et al.
Veröffentlicht: (2026)
von: Han, Menglin, et al.
Veröffentlicht: (2026)
Taming Flow-based I2V Models for Creative Video Editing
von: Kong, Xianghao, et al.
Veröffentlicht: (2025)
von: Kong, Xianghao, et al.
Veröffentlicht: (2025)
DOGR: Towards Versatile Visual Document Grounding and Referring
von: Zhou, Yinan, et al.
Veröffentlicht: (2024)
von: Zhou, Yinan, et al.
Veröffentlicht: (2024)
Taming Preference Mode Collapse via Directional Decoupling Alignment in Diffusion Reinforcement Learning
von: Chen, Chubin, et al.
Veröffentlicht: (2025)
von: Chen, Chubin, et al.
Veröffentlicht: (2025)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
FREE-Edit: Using Editing-aware Injection in Rectified Flow Models for Zero-shot Image-Driven Video Editing
von: Li, Maomao, et al.
Veröffentlicht: (2026)
von: Li, Maomao, et al.
Veröffentlicht: (2026)
SplitFlow: Flow Decomposition for Inversion-Free Text-to-Image Editing
von: Yoon, Sung-Hoon, et al.
Veröffentlicht: (2025)
von: Yoon, Sung-Hoon, et al.
Veröffentlicht: (2025)
Easy3E: Feed-Forward 3D Asset Editing via Rectified Voxel Flow
von: Hu, Shimin, et al.
Veröffentlicht: (2026)
von: Hu, Shimin, et al.
Veröffentlicht: (2026)
PosterLLaVa: Constructing a Unified Multi-modal Layout Generator with LLM
von: Yang, Tao, et al.
Veröffentlicht: (2024)
von: Yang, Tao, et al.
Veröffentlicht: (2024)
BlobCtrl: Taming Controllable Blob for Element-level Image Editing
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024) -
Adams Bashforth Moulton Solver for Inversion and Editing in Rectified Flow
von: Ma, Yongjia, et al.
Veröffentlicht: (2025) -
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?
von: Chen, Yuxin, et al.
Veröffentlicht: (2024) -
OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video
von: Pu, Junfu, et al.
Veröffentlicht: (2026) -
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)