ReFlex: Text-Guided Editing of Real Images in Rectified Flow via Mid-Step Feature Extraction and Attention Adaptation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jimyeong, Park, Jungwon, Song, Yeji, Kwak, Nojun, Rhee, Wonjong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
Selectively Informative Description can Reduce Undesired Embedding Entanglements in Text-to-Image Personalization
von: Kim, Jimyeong, et al.
Veröffentlicht: (2024)
von: Kim, Jimyeong, et al.
Veröffentlicht: (2024)
When Confidence Misleads: Suffix Anchoring and Anchor-Proximity Confidence Modulation for Diffusion Language Models
von: Park, Jungwon, et al.
Veröffentlicht: (2026)
von: Park, Jungwon, et al.
Veröffentlicht: (2026)
Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation
von: Ko, Jungmin, et al.
Veröffentlicht: (2026)
von: Ko, Jungmin, et al.
Veröffentlicht: (2026)
Soft Head Selection for Injecting ICL-Derived Task Embeddings
von: Park, Jungwon, et al.
Veröffentlicht: (2025)
von: Park, Jungwon, et al.
Veröffentlicht: (2025)
Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models
von: Park, Jungwon, et al.
Veröffentlicht: (2024)
von: Park, Jungwon, et al.
Veröffentlicht: (2024)
Evaluating Feature Attribution Methods for Electrocardiogram
von: Suh, Jangwon, et al.
Veröffentlicht: (2022)
von: Suh, Jangwon, et al.
Veröffentlicht: (2022)
Selective Aggregation of Attention Maps Improves Diffusion-Based Visual Interpretation
von: Park, Jungwon, et al.
Veröffentlicht: (2026)
von: Park, Jungwon, et al.
Veröffentlicht: (2026)
DOS: Directional Object Separation in Text Embeddings for Multi-Object Image Generation
von: Byun, Dongnam, et al.
Veröffentlicht: (2025)
von: Byun, Dongnam, et al.
Veröffentlicht: (2025)
Task-Specific Preconditioner for Cross-Domain Few-Shot Learning
von: Kang, Suhyun, et al.
Veröffentlicht: (2024)
von: Kang, Suhyun, et al.
Veröffentlicht: (2024)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
von: Song, Yeji, et al.
Veröffentlicht: (2025)
von: Song, Yeji, et al.
Veröffentlicht: (2025)
InstantEdit: Text-Guided Few-Step Image Editing with Piecewise Rectified Flow
von: Gong, Yiming, et al.
Veröffentlicht: (2025)
von: Gong, Yiming, et al.
Veröffentlicht: (2025)
Towards a Better Evaluation of Out-of-Domain Generalization
von: Hwang, Duhun, et al.
Veröffentlicht: (2024)
von: Hwang, Duhun, et al.
Veröffentlicht: (2024)
Targeted Data Protection for Diffusion Model by Matching Training Trajectory
von: Lee, Hojun, et al.
Veröffentlicht: (2025)
von: Lee, Hojun, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Generation Based Nurse Observation Extraction
von: Hwang, Kyomin, et al.
Veröffentlicht: (2026)
von: Hwang, Kyomin, et al.
Veröffentlicht: (2026)
Enhancing Contrastive Learning with Efficient Combinatorial Positive Pairing
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
Delta Rectified Flow Sampling for Text-to-Image Editing
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models
von: Lee, Junhoo, et al.
Veröffentlicht: (2026)
von: Lee, Junhoo, et al.
Veröffentlicht: (2026)
DNAEdit: Direct Noise Alignment for Text-Guided Rectified Flow Editing
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
von: Song, Jookyung, et al.
Veröffentlicht: (2025)
von: Song, Jookyung, et al.
Veröffentlicht: (2025)
A Benchmark Suite for Evaluating Neural Mutual Information Estimators on Unstructured Datasets
von: Lee, Kyungeun, et al.
Veröffentlicht: (2024)
von: Lee, Kyungeun, et al.
Veröffentlicht: (2024)
Conservative Generator, Progressive Discriminator: Coordination of Adversaries in Few-shot Incremental Image Synthesis
von: Kong, Chaerin, et al.
Veröffentlicht: (2022)
von: Kong, Chaerin, et al.
Veröffentlicht: (2022)
Understanding Differential Transformer Unchains Pretrained Self-Attentions
von: Kong, Chaerin, et al.
Veröffentlicht: (2025)
von: Kong, Chaerin, et al.
Veröffentlicht: (2025)
Causal Interpretation of Sparse Autoencoder Features in Vision
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
RFM-Editing: Rectified Flow Matching for Text-guided Audio Editing
von: Gao, Liting, et al.
Veröffentlicht: (2025)
von: Gao, Liting, et al.
Veröffentlicht: (2025)
Improving Forward Compatibility in Class Incremental Learning by Increasing Representation Rank and Feature Richness
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
SketcherX: AI-Driven Interactive Robotic drawing with Diffusion model and Vectorization Techniques
von: Song, Jookyung, et al.
Veröffentlicht: (2024)
von: Song, Jookyung, et al.
Veröffentlicht: (2024)
Toward Structural Multimodal Representations: Specialization, Selection, and Sparsification via Mixture-of-Experts
von: Choi, Hahyeon, et al.
Veröffentlicht: (2026)
von: Choi, Hahyeon, et al.
Veröffentlicht: (2026)
Enhancing Retrieval-Augmented Audio Captioning with Generation-Assisted Multimodal Querying and Progressive Learning
von: Changin, Choi, et al.
Veröffentlicht: (2024)
von: Changin, Choi, et al.
Veröffentlicht: (2024)
An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
von: Kim, Wonkyun, et al.
Veröffentlicht: (2024)
von: Kim, Wonkyun, et al.
Veröffentlicht: (2024)
MERLIN: Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank Pipeline
von: Han, Donghoon, et al.
Veröffentlicht: (2024)
von: Han, Donghoon, et al.
Veröffentlicht: (2024)
FireFlow: Fast Inversion of Rectified Flow for Image Semantic Editing
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
Taming Rectified Flow for Inversion and Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
FlowBypass: Rectified Flow Trajectory Bypass for Training-Free Image Editing
von: Han, Menglin, et al.
Veröffentlicht: (2026)
von: Han, Menglin, et al.
Veröffentlicht: (2026)
ReFlow-TTS: A Rectified Flow Model for High-fidelity Text-to-Speech
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
The Role of Teacher Calibration in Knowledge Distillation
von: Kim, Suyoung, et al.
Veröffentlicht: (2025)
von: Kim, Suyoung, et al.
Veröffentlicht: (2025)
Unveiling Key Aspects of Fine-Tuning in Sentence Embeddings: A Representation Rank Analysis
von: Jung, Euna, et al.
Veröffentlicht: (2024)
von: Jung, Euna, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024) -
Selectively Informative Description can Reduce Undesired Embedding Entanglements in Text-to-Image Personalization
von: Kim, Jimyeong, et al.
Veröffentlicht: (2024) -
When Confidence Misleads: Suffix Anchoring and Anchor-Proximity Confidence Modulation for Diffusion Language Models
von: Park, Jungwon, et al.
Veröffentlicht: (2026) -
Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation
von: Ko, Jungmin, et al.
Veröffentlicht: (2026) -
Soft Head Selection for Injecting ICL-Derived Task Embeddings
von: Park, Jungwon, et al.
Veröffentlicht: (2025)