PhotoArtAgent: Intelligent Photo Retouching with Language Model-Based Artist Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Haoyu, Tao, Keda, Wang, Yizao, Wang, Xinlei, Zhu, Lei, Gu, Jinjin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent
by: Lin, Yunlong, et al.
Published: (2025)
by: Lin, Yunlong, et al.
Published: (2025)
VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching
by: Guo, Yihong, et al.
Published: (2026)
by: Guo, Yihong, et al.
Published: (2026)
RetouchIQ: MLLM Agents for Instruction-Based Image Retouching with Generalist Reward
by: Wu, Qiucheng, et al.
Published: (2026)
by: Wu, Qiucheng, et al.
Published: (2026)
PhotoAgent: Agentic Photo Editing with Exploratory Visual Aesthetic Planning
by: Yao, Mingde, et al.
Published: (2026)
by: Yao, Mingde, et al.
Published: (2026)
Overcoming False Illusions in Real-World Face Restoration with Multi-Modal Guided Diffusion Model
by: Tao, Keda, et al.
Published: (2024)
by: Tao, Keda, et al.
Published: (2024)
PhotoFramer: Multi-modal Image Composition Instruction
by: You, Zhiyuan, et al.
Published: (2025)
by: You, Zhiyuan, et al.
Published: (2025)
RestoreAgent: Autonomous Image Restoration Agent via Multimodal Large Language Models
by: Chen, Haoyu, et al.
Published: (2024)
by: Chen, Haoyu, et al.
Published: (2024)
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
by: Yu, Fanghua, et al.
Published: (2024)
by: Yu, Fanghua, et al.
Published: (2024)
Active Perception Agent for Omnimodal Audio-Video Understanding
by: Tao, Keda, et al.
Published: (2025)
by: Tao, Keda, et al.
Published: (2025)
PhotoAgent: A Robotic Photographer with Spatial and Aesthetic Understanding
by: Che, Lirong, et al.
Published: (2026)
by: Che, Lirong, et al.
Published: (2026)
RetouchLLM: Training-free Code-based Image Retouching with Vision Language Models
by: Ye-Bin, Moon, et al.
Published: (2025)
by: Ye-Bin, Moon, et al.
Published: (2025)
AnimeAgent: Is the Multi-Agent via Image-to-Video models a Good Disney Storytelling Artist?
by: Yan, Hailong, et al.
Published: (2026)
by: Yan, Hailong, et al.
Published: (2026)
Draw Like an Artist: Complex Scene Generation with Diffusion Model via Composition, Painting, and Retouching
by: Liu, Minghao, et al.
Published: (2024)
by: Liu, Minghao, et al.
Published: (2024)
PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
by: Chang, Zewei, et al.
Published: (2025)
by: Chang, Zewei, et al.
Published: (2025)
Unpaired Photo-realistic Image Deraining with Energy-informed Diffusion Model
by: Wen, Yuanbo, et al.
Published: (2024)
by: Wen, Yuanbo, et al.
Published: (2024)
Position: Agentic Systems Constitute a Key Component of Next-Generation Intelligent Image Processing
by: Gu, Jinjin
Published: (2025)
by: Gu, Jinjin
Published: (2025)
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
by: Duan, Zheng-Peng, et al.
Published: (2024)
by: Duan, Zheng-Peng, et al.
Published: (2024)
Label-guided Facial Retouching Reversion
by: Zhao, Guanhua, et al.
Published: (2024)
by: Zhao, Guanhua, et al.
Published: (2024)
JarvisEvo: Towards a Self-Evolving Photo Editing Agent with Synergistic Editor-Evaluator Optimization
by: Lin, Yunlong, et al.
Published: (2025)
by: Lin, Yunlong, et al.
Published: (2025)
FRRffusion: Unveiling Authenticity with Diffusion-Based Face Retouching Reversal
by: Xing, Fengchuang, et al.
Published: (2024)
by: Xing, Fengchuang, et al.
Published: (2024)
GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing
by: Wang, Zhenyu, et al.
Published: (2024)
by: Wang, Zhenyu, et al.
Published: (2024)
LucidFlux: Caption-Free Photo-Realistic Image Restoration via a Large-Scale Diffusion Transformer
by: Fei, Song, et al.
Published: (2025)
by: Fei, Song, et al.
Published: (2025)
DP^2-VL: Private Photo Dataset Protection by Data Poisoning for Vision-Language Models
by: Miao, Hongyi, et al.
Published: (2026)
by: Miao, Hongyi, et al.
Published: (2026)
ArtUV: Artist-style UV Unwrapping
by: Chen, Yuguang, et al.
Published: (2025)
by: Chen, Yuguang, et al.
Published: (2025)
Plug-and-Play 1.x-Bit KV Cache Quantization for Video Large Language Models
by: Tao, Keda, et al.
Published: (2025)
by: Tao, Keda, et al.
Published: (2025)
Text-to-Image Diffusion Models are Great Sketch-Photo Matchmakers
by: Koley, Subhadeep, et al.
Published: (2024)
by: Koley, Subhadeep, et al.
Published: (2024)
LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion
by: Xue, Clara, et al.
Published: (2026)
by: Xue, Clara, et al.
Published: (2026)
OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models
by: Tao, Keda, et al.
Published: (2025)
by: Tao, Keda, et al.
Published: (2025)
Photo-Realistic Image Restoration in the Wild with Controlled Vision-Language Models
by: Luo, Ziwei, et al.
Published: (2024)
by: Luo, Ziwei, et al.
Published: (2024)
Content-Adaptive Image Retouching Guided by Attribute-Based Text Representation
by: Zhu, Hancheng, et al.
Published: (2025)
by: Zhu, Hancheng, et al.
Published: (2025)
Controllable and Gradual Facial Blemishes Retouching via Physics-Based Modelling
by: Shuai, Chenhao, et al.
Published: (2024)
by: Shuai, Chenhao, et al.
Published: (2024)
HoliTom: Holistic Token Merging for Fast Video Large Language Models
by: Shao, Kele, et al.
Published: (2025)
by: Shao, Kele, et al.
Published: (2025)
Self-Supervised Selective-Guided Diffusion Model for Old-Photo Face Restoration
by: Li, Wenjie, et al.
Published: (2025)
by: Li, Wenjie, et al.
Published: (2025)
PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval
by: Xu, Tianyi, et al.
Published: (2026)
by: Xu, Tianyi, et al.
Published: (2026)
TalkPhoto: A Versatile Training-Free Conversational Assistant for Intelligent Image Editing
by: Hu, Yujie, et al.
Published: (2026)
by: Hu, Yujie, et al.
Published: (2026)
DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models
by: Tao, Keda, et al.
Published: (2024)
by: Tao, Keda, et al.
Published: (2024)
Photo Dating by Facial Age Aggregation
by: Paplham, Jakub, et al.
Published: (2025)
by: Paplham, Jakub, et al.
Published: (2025)
Removing Reflections from RAW Photos
by: Kee, Eric, et al.
Published: (2024)
by: Kee, Eric, et al.
Published: (2024)
LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules
by: Yang, Cheng, et al.
Published: (2025)
by: Yang, Cheng, et al.
Published: (2025)
Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs
by: Zhang, Kejia, et al.
Published: (2025)
by: Zhang, Kejia, et al.
Published: (2025)
Similar Items
-
JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent
by: Lin, Yunlong, et al.
Published: (2025) -
VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching
by: Guo, Yihong, et al.
Published: (2026) -
RetouchIQ: MLLM Agents for Instruction-Based Image Retouching with Generalist Reward
by: Wu, Qiucheng, et al.
Published: (2026) -
PhotoAgent: Agentic Photo Editing with Exploratory Visual Aesthetic Planning
by: Yao, Mingde, et al.
Published: (2026) -
Overcoming False Illusions in Real-World Face Restoration with Multi-Modal Guided Diffusion Model
by: Tao, Keda, et al.
Published: (2024)