PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Zewei, Duan, Zheng-Peng, Zhang, Jianxing, Guo, Chun-Le, Liu, Siyu, Chun, Hyungju, Park, Hyunhee, Liu, Zikun, Li, Chongyi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FaceMe: Robust Blind Face Restoration with Personal Identification
by: Liu, Siyu, et al.
Published: (2025)
by: Liu, Siyu, et al.
Published: (2025)
Iterative Predictor-Critic Code Decoding for Real-World Image Dehazing
by: Fu, Jiayi, et al.
Published: (2025)
by: Fu, Jiayi, et al.
Published: (2025)
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
by: Duan, Zheng-Peng, et al.
Published: (2024)
by: Duan, Zheng-Peng, et al.
Published: (2024)
Restore Anything with Masks: Leveraging Mask Image Modeling for Blind All-in-One Image Restoration
by: Qin, Chu-Jie, et al.
Published: (2024)
by: Qin, Chu-Jie, et al.
Published: (2024)
RetouchIQ: MLLM Agents for Instruction-Based Image Retouching with Generalist Reward
by: Wu, Qiucheng, et al.
Published: (2026)
by: Wu, Qiucheng, et al.
Published: (2026)
A Diffusion-Based Framework for Occluded Object Movement
by: Duan, Zheng-Peng, et al.
Published: (2025)
by: Duan, Zheng-Peng, et al.
Published: (2025)
EnsIR: An Ensemble Algorithm for Image Restoration via Gaussian Mixture Models
by: Sun, Shangquan, et al.
Published: (2024)
by: Sun, Shangquan, et al.
Published: (2024)
InstantRetouch: Personalized Image Retouching without Test-time Fine-tuning Using an Asymmetric Auto-Encoder
by: Weldengus, Temesgen Muruts, et al.
Published: (2025)
by: Weldengus, Temesgen Muruts, et al.
Published: (2025)
PerPilot: Personalizing VLM-based Mobile Agents via Memory and Exploration
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Improving Reconstruction of Representation Autoencoder
by: Liu, Siyu, et al.
Published: (2026)
by: Liu, Siyu, et al.
Published: (2026)
Agentic Retoucher for Text-To-Image Generation
by: Shen, Shaocheng, et al.
Published: (2026)
by: Shen, Shaocheng, et al.
Published: (2026)
Time-Aware One Step Diffusion Network for Real-World Image Super-Resolution
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
DetectAnyLLM: Towards Generalizable and Robust Detection of Machine-Generated Text Across Domains and Models
by: Fu, Jiachen, et al.
Published: (2025)
by: Fu, Jiachen, et al.
Published: (2025)
DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution
by: Duan, Zheng-Peng, et al.
Published: (2025)
by: Duan, Zheng-Peng, et al.
Published: (2025)
Lighting Every Darkness with 3DGS: Fast Training and Real-Time Rendering for HDR View Synthesis
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
Synergistic Multiscale Detail Refinement via Intrinsic Supervision for Underwater Image Enhancement
by: Zhang, Dehuan, et al.
Published: (2023)
by: Zhang, Dehuan, et al.
Published: (2023)
RetouchLLM: Training-free Code-based Image Retouching with Vision Language Models
by: Ye-Bin, Moon, et al.
Published: (2025)
by: Ye-Bin, Moon, et al.
Published: (2025)
CameraMaster: Unified Camera Semantic-Parameter Control for Photography Retouching
by: Yang, Qirui, et al.
Published: (2025)
by: Yang, Qirui, et al.
Published: (2025)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
by: Kim, Kangyeol, et al.
Published: (2024)
by: Kim, Kangyeol, et al.
Published: (2024)
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
by: Xue, Bowen, et al.
Published: (2025)
by: Xue, Bowen, et al.
Published: (2025)
Taming Lookup Tables for Efficient Image Retouching
by: Yang, Sidi, et al.
Published: (2024)
by: Yang, Sidi, et al.
Published: (2024)
Draw Like an Artist: Complex Scene Generation with Diffusion Model via Composition, Painting, and Retouching
by: Liu, Minghao, et al.
Published: (2024)
by: Liu, Minghao, et al.
Published: (2024)
Content-Adaptive Image Retouching Guided by Attribute-Based Text Representation
by: Zhu, Hancheng, et al.
Published: (2025)
by: Zhu, Hancheng, et al.
Published: (2025)
Multi-modal Agent Tuning: Building a VLM-Driven Agent for Efficient Tool Usage
by: Gao, Zhi, et al.
Published: (2024)
by: Gao, Zhi, et al.
Published: (2024)
Quantifying Signal-to-Noise Ratio in Neural Latent Trajectories via Fisher Information
by: Jeon, Hyungju, et al.
Published: (2024)
by: Jeon, Hyungju, et al.
Published: (2024)
VTinker: Guided Flow Upsampling and Texture Mapping for High-Resolution Video Frame Interpolation
by: Wu, Chenyang, et al.
Published: (2025)
by: Wu, Chenyang, et al.
Published: (2025)
UltraLED: Learning to See Everything in Ultra-High Dynamic Range Scenes
by: Meng, Yuang, et al.
Published: (2025)
by: Meng, Yuang, et al.
Published: (2025)
MoFRR: Mixture of Diffusion Models for Face Retouching Restoration
by: Liu, Jiaxin, et al.
Published: (2025)
by: Liu, Jiaxin, et al.
Published: (2025)
PUGAN: Physical Model-Guided Underwater Image Enhancement Using GAN with Dual-Discriminators
by: Cong, Runmin, et al.
Published: (2023)
by: Cong, Runmin, et al.
Published: (2023)
P‐55: An Optimizing Finger Separation Method with Machine Learning Algorithm used In‐Cell Capacitive Touch Panel
by: Ching-Yao Chao, et al.
Published: (2024)
by: Ching-Yao Chao, et al.
Published: (2024)
Type-R: Automatically Retouching Typos for Text-to-Image Generation
by: Shimoda, Wataru, et al.
Published: (2024)
by: Shimoda, Wataru, et al.
Published: (2024)
VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching
by: Guo, Yihong, et al.
Published: (2026)
by: Guo, Yihong, et al.
Published: (2026)
MUSE-VL: Modeling Unified VLM through Semantic Discrete Encoding
by: Xie, Rongchang, et al.
Published: (2024)
by: Xie, Rongchang, et al.
Published: (2024)
PhotoArtAgent: Intelligent Photo Retouching with Language Model-Based Artist Agents
by: Chen, Haoyu, et al.
Published: (2025)
by: Chen, Haoyu, et al.
Published: (2025)
Touch-R1: Reinforcing Touch Reasoning in MLLMs
by: Lai, Yingxin, et al.
Published: (2026)
by: Lai, Yingxin, et al.
Published: (2026)
My Words Imply Your Opinion: Reader Agent-based Propagation Enhancement for Personalized Implicit Emotion Analysis
by: Liao, Jian, et al.
Published: (2024)
by: Liao, Jian, et al.
Published: (2024)
An augmented reality‐facilitated question‐prompt‐interaction‐evaluation approach to fostering students' case‐handling competence in technical and vocational education
by: Chun‐Chun Chang
Published: (2024)
by: Chun‐Chun Chang
Published: (2024)
Enhancing Zero-shot Personalized Image Aesthetics Assessment with Profile-aware Multimodal LLM
by: Wang, Chun, et al.
Published: (2026)
by: Wang, Chun, et al.
Published: (2026)
SiCL: Silhouette-Driven Contrastive Learning for Unsupervised Person Re-Identification with Clothes Change
by: Li, Mingkun, et al.
Published: (2023)
by: Li, Mingkun, et al.
Published: (2023)
MonetGPT: Solving Puzzles Enhances MLLMs' Image Retouching Skills
by: Dutt, Niladri Shekhar, et al.
Published: (2025)
by: Dutt, Niladri Shekhar, et al.
Published: (2025)
Similar Items
-
FaceMe: Robust Blind Face Restoration with Personal Identification
by: Liu, Siyu, et al.
Published: (2025) -
Iterative Predictor-Critic Code Decoding for Real-World Image Dehazing
by: Fu, Jiayi, et al.
Published: (2025) -
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
by: Duan, Zheng-Peng, et al.
Published: (2024) -
Restore Anything with Masks: Leveraging Mask Image Modeling for Blind All-in-One Image Restoration
by: Qin, Chu-Jie, et al.
Published: (2024) -
RetouchIQ: MLLM Agents for Instruction-Based Image Retouching with Generalist Reward
by: Wu, Qiucheng, et al.
Published: (2026)