InstructRL4Pix: Training Diffusion for Image Editing by Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Tiancheng, Liu, Jinxiu, Chen, Huajun, Liu, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InstructPix2NeRF: Instructed 3D Portrait Editing from a Single Image
von: Li, Jianhui, et al.
Veröffentlicht: (2023)
von: Li, Jianhui, et al.
Veröffentlicht: (2023)
InstructAny2Pix: Flexible Visual Editing via Multimodal Instruction Following
von: Li, Shufan, et al.
Veröffentlicht: (2023)
von: Li, Shufan, et al.
Veröffentlicht: (2023)
ReasonPix2Pix: Instruction Reasoning Dataset for Advanced Image Editing
von: Jin, Ying, et al.
Veröffentlicht: (2024)
von: Jin, Ying, et al.
Veröffentlicht: (2024)
InstructGIE: Towards Generalizable Image Editing
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
ThinkRL-Edit: Thinking in Reinforcement Learning for Reasoning-Centric Image Editing
von: Li, Hengjia, et al.
Veröffentlicht: (2026)
von: Li, Hengjia, et al.
Veröffentlicht: (2026)
PixArt-Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation
von: Chen, Junsong, et al.
Veröffentlicht: (2024)
von: Chen, Junsong, et al.
Veröffentlicht: (2024)
PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
von: Chen, Junsong, et al.
Veröffentlicht: (2023)
von: Chen, Junsong, et al.
Veröffentlicht: (2023)
Training-free Geometric Image Editing on Diffusion Models
von: Zhu, Hanshen, et al.
Veröffentlicht: (2025)
von: Zhu, Hanshen, et al.
Veröffentlicht: (2025)
PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
von: Zheng, Haitian, et al.
Veröffentlicht: (2025)
von: Zheng, Haitian, et al.
Veröffentlicht: (2025)
DiT4Edit: Diffusion Transformer for Image Editing
von: Feng, Kunyu, et al.
Veröffentlicht: (2024)
von: Feng, Kunyu, et al.
Veröffentlicht: (2024)
InstructBrush: Learning Attention-based Instruction Optimization for Image Editing
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2024)
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2024)
Instruct-CLIP: Improving Instruction-Guided Image Editing with Automated Data Refinement Using Contrastive Learning
von: Chen, Sherry X., et al.
Veröffentlicht: (2025)
von: Chen, Sherry X., et al.
Veröffentlicht: (2025)
Learning Feature-Preserving Portrait Editing from Generated Pairs
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
SRU-Pix2Pix: A Fusion-Driven Generator Network for Medical Image Translation with Few-Shot Learning
von: Qiu, Xihe, et al.
Veröffentlicht: (2026)
von: Qiu, Xihe, et al.
Veröffentlicht: (2026)
Light Future: Multimodal Action Frame Prediction via InstructPix2Pix
von: Zhong, Zesen, et al.
Veröffentlicht: (2025)
von: Zhong, Zesen, et al.
Veröffentlicht: (2025)
PixLens: A Novel Framework for Disentangled Evaluation in Diffusion-Based Image Editing with Object Detection + SAM
von: Stefanache, Stefan, et al.
Veröffentlicht: (2024)
von: Stefanache, Stefan, et al.
Veröffentlicht: (2024)
InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing
von: Yu, Haoran, et al.
Veröffentlicht: (2025)
von: Yu, Haoran, et al.
Veröffentlicht: (2025)
Leveraging Verifier-Based Reinforcement Learning in Image Editing
von: Guo, Hanzhong, et al.
Veröffentlicht: (2026)
von: Guo, Hanzhong, et al.
Veröffentlicht: (2026)
PixNerd: Pixel Neural Field Diffusion
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
TIGER: Text-Instructed 3D Gaussian Retrieval and Coherent Editing
von: Xu, Teng, et al.
Veröffentlicht: (2024)
von: Xu, Teng, et al.
Veröffentlicht: (2024)
Empowering Visual Creativity: A Vision-Language Assistant to Image Editing Recommendations
von: Shen, Tiancheng, et al.
Veröffentlicht: (2024)
von: Shen, Tiancheng, et al.
Veröffentlicht: (2024)
Image Editing As Programs with Diffusion Models
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
MIRG-RL: Multi-Image Reasoning and Grounding with Reinforcement Learning
von: Zheng, Lihao, et al.
Veröffentlicht: (2025)
von: Zheng, Lihao, et al.
Veröffentlicht: (2025)
RL-ScanIQA: Reinforcement-Learned Scanpaths for Blind 360°Image Quality Assessment
von: Wang, Yujia, et al.
Veröffentlicht: (2026)
von: Wang, Yujia, et al.
Veröffentlicht: (2026)
InstructX: Towards Unified Visual Editing with MLLM Guidance
von: Mou, Chong, et al.
Veröffentlicht: (2025)
von: Mou, Chong, et al.
Veröffentlicht: (2025)
Disentangling Instruction Influence in Diffusion Transformers for Parallel Multi-Instruction-Guided Image Editing
von: Liu, Hui, et al.
Veröffentlicht: (2025)
von: Liu, Hui, et al.
Veröffentlicht: (2025)
Diffusion Model-Based Image Editing: A Survey
von: Huang, Yi, et al.
Veröffentlicht: (2024)
von: Huang, Yi, et al.
Veröffentlicht: (2024)
Consistent Image Layout Editing with Diffusion Models
von: Xia, Tao, et al.
Veröffentlicht: (2025)
von: Xia, Tao, et al.
Veröffentlicht: (2025)
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
InstructEdit: Instruction-based Knowledge Editing for Large Language Models
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
Dual-Schedule Inversion: Training- and Tuning-Free Inversion for Real Image Editing
von: Huang, Jiancheng, et al.
Veröffentlicht: (2024)
von: Huang, Jiancheng, et al.
Veröffentlicht: (2024)
GeoPix: Multi-Modal Large Language Model for Pixel-level Image Understanding in Remote Sensing
von: Ou, Ruizhe, et al.
Veröffentlicht: (2025)
von: Ou, Ruizhe, et al.
Veröffentlicht: (2025)
Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback
von: Li, Zongjian, et al.
Veröffentlicht: (2025)
von: Li, Zongjian, et al.
Veröffentlicht: (2025)
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality
von: Ren, Baochang, et al.
Veröffentlicht: (2025)
von: Ren, Baochang, et al.
Veröffentlicht: (2025)
UniEdit-I: Training-free Image Editing for Unified VLM via Iterative Understanding, Editing and Verifying
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
Hiding Images in Diffusion Models by Editing Learned Score Functions
von: Chen, Haoyu, et al.
Veröffentlicht: (2025)
von: Chen, Haoyu, et al.
Veröffentlicht: (2025)
InstructVEdit: A Holistic Approach for Instructional Video Editing
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
GIDE: Unlocking Diffusion LLMs for Precise Training-Free Image Editing
von: Zhu, Zifeng, et al.
Veröffentlicht: (2026)
von: Zhu, Zifeng, et al.
Veröffentlicht: (2026)
Mapping New Realities: Ground Truth Image Creation with Pix2Pix Image-to-Image Translation
von: Li, Zhenglin, et al.
Veröffentlicht: (2024)
von: Li, Zhenglin, et al.
Veröffentlicht: (2024)
High-Fidelity Diffusion-based Image Editing
von: Hou, Chen, et al.
Veröffentlicht: (2023)
von: Hou, Chen, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
InstructPix2NeRF: Instructed 3D Portrait Editing from a Single Image
von: Li, Jianhui, et al.
Veröffentlicht: (2023) -
InstructAny2Pix: Flexible Visual Editing via Multimodal Instruction Following
von: Li, Shufan, et al.
Veröffentlicht: (2023) -
ReasonPix2Pix: Instruction Reasoning Dataset for Advanced Image Editing
von: Jin, Ying, et al.
Veröffentlicht: (2024) -
InstructGIE: Towards Generalizable Image Editing
von: Meng, Zichong, et al.
Veröffentlicht: (2024) -
ThinkRL-Edit: Thinking in Reinforcement Learning for Reasoning-Centric Image Editing
von: Li, Hengjia, et al.
Veröffentlicht: (2026)