Agentic Retoucher for Text-To-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Shen, Shaocheng, Liang, Jianfeng, Cai, Chunlei, Geng, Cong, Duan, Huiyu, Zhang, Xiaoyun, Hu, Qiang, Zhai, Guangtao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks
di: An, Yanru, et al.
Pubblicazione: (2025)
di: An, Yanru, et al.
Pubblicazione: (2025)
MoA-VR: A Mixture-of-Agents System Towards All-in-One Video Restoration
di: Liu, Lu, et al.
Pubblicazione: (2025)
di: Liu, Lu, et al.
Pubblicazione: (2025)
Life-IQA: Boosting Blind Image Quality Assessment through GCN-enhanced Layer Interaction and MoE-based Feature Decoupling
di: Tang, Long, et al.
Pubblicazione: (2025)
di: Tang, Long, et al.
Pubblicazione: (2025)
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
di: Fu, Kang, et al.
Pubblicazione: (2026)
di: Fu, Kang, et al.
Pubblicazione: (2026)
F-Bench: Rethinking Human Preference Evaluation Metrics for Benchmarking Face Generation, Customization, and Restoration
di: Liu, Lu, et al.
Pubblicazione: (2024)
di: Liu, Lu, et al.
Pubblicazione: (2024)
D$^2$-VR: Degradation-Robust and Distilled Video Restoration with Synergistic Optimization Strategy
di: Liang, Jianfeng, et al.
Pubblicazione: (2026)
di: Liang, Jianfeng, et al.
Pubblicazione: (2026)
UniProcessor: A Text-induced Unified Low-level Image Processor
di: Duan, Huiyu, et al.
Pubblicazione: (2024)
di: Duan, Huiyu, et al.
Pubblicazione: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
Quality Assessment for AI Generated Images with Instruction Tuning
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
GeoX-Bench: Benchmarking Cross-View Geo-Localization and Pose Estimation Capabilities of Large Multimodal Models
di: Zheng, Yushuo, et al.
Pubblicazione: (2025)
di: Zheng, Yushuo, et al.
Pubblicazione: (2025)
TIT-Score: Evaluating Long-Prompt Based Text-to-Image Alignment via Text-to-Image-to-Text Consistency
di: Wang, Juntong, et al.
Pubblicazione: (2025)
di: Wang, Juntong, et al.
Pubblicazione: (2025)
Taming Lookup Tables for Efficient Image Retouching
di: Yang, Sidi, et al.
Pubblicazione: (2024)
di: Yang, Sidi, et al.
Pubblicazione: (2024)
How is Visual Attention Influenced by Text Guidance? Database and Model
di: Sun, Yinan, et al.
Pubblicazione: (2024)
di: Sun, Yinan, et al.
Pubblicazione: (2024)
Vision-Language Consistency Guided Multi-modal Prompt Learning for Blind AI Generated Image Quality Assessment
di: Fu, Jun, et al.
Pubblicazione: (2024)
di: Fu, Jun, et al.
Pubblicazione: (2024)
AIGV-Assessor: Benchmarking and Evaluating the Perceptual Quality of Text-to-Video Generation with LMM
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
di: Wang, Jiarui, et al.
Pubblicazione: (2024)
MegaFusion: Extend Diffusion Models towards Higher-resolution Image Generation without Further Tuning
di: Wu, Haoning, et al.
Pubblicazione: (2024)
di: Wu, Haoning, et al.
Pubblicazione: (2024)
ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images
di: Li, Xinyue, et al.
Pubblicazione: (2026)
di: Li, Xinyue, et al.
Pubblicazione: (2026)
LL-Bench: Rethinking Low-Level Vision Evaluation in the Era of Large-Scale Generative Models
di: Liu, Lu, et al.
Pubblicazione: (2026)
di: Liu, Lu, et al.
Pubblicazione: (2026)
Medical Manifestation-Aware De-Identification
di: Tian, Yuan, et al.
Pubblicazione: (2024)
di: Tian, Yuan, et al.
Pubblicazione: (2024)
Generative Medical Image Anonymization Based on Latent Code Projection and Optimization
di: Li, Huiyu, et al.
Pubblicazione: (2025)
di: Li, Huiyu, et al.
Pubblicazione: (2025)
DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models
di: Wang, Juntong, et al.
Pubblicazione: (2026)
di: Wang, Juntong, et al.
Pubblicazione: (2026)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
di: Sampaio, Georgia Gabriela, et al.
Pubblicazione: (2024)
di: Sampaio, Georgia Gabriela, et al.
Pubblicazione: (2024)
IA-T2I: Internet-Augmented Text-to-Image Generation
di: Li, Chuanhao, et al.
Pubblicazione: (2025)
di: Li, Chuanhao, et al.
Pubblicazione: (2025)
TDVE-Assessor: Benchmarking and Evaluating the Quality of Text-Driven Video Editing with LMMs
di: Wang, Juntong, et al.
Pubblicazione: (2025)
di: Wang, Juntong, et al.
Pubblicazione: (2025)
FineVQ: Fine-Grained User Generated Content Video Quality Assessment
di: Duan, Huiyu, et al.
Pubblicazione: (2024)
di: Duan, Huiyu, et al.
Pubblicazione: (2024)
Multi-scale HSV Color Feature Embedding for High-fidelity NIR-to-RGB Spectrum Translation
di: Zhai, Huiyu, et al.
Pubblicazione: (2024)
di: Zhai, Huiyu, et al.
Pubblicazione: (2024)
NTIRE 2025 challenge on Text to Image Generation Model Quality Assessment
di: Han, Shuhao, et al.
Pubblicazione: (2025)
di: Han, Shuhao, et al.
Pubblicazione: (2025)
OBI-Bench: Can LMMs Aid in Study of Ancient Script on Oracle Bones?
di: Chen, Zijian, et al.
Pubblicazione: (2024)
di: Chen, Zijian, et al.
Pubblicazione: (2024)
Interactive Visual Assessment for Text-to-Image Generation Models
di: Mi, Xiaoyue, et al.
Pubblicazione: (2024)
di: Mi, Xiaoyue, et al.
Pubblicazione: (2024)
Decoupling Perception and Calibration: Label-Efficient Image Quality Assessment Framework
di: Li, Xinyue, et al.
Pubblicazione: (2026)
di: Li, Xinyue, et al.
Pubblicazione: (2026)
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
di: Hua, Shengyi, et al.
Pubblicazione: (2025)
di: Hua, Shengyi, et al.
Pubblicazione: (2025)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
di: Chen, Muxi, et al.
Pubblicazione: (2024)
di: Chen, Muxi, et al.
Pubblicazione: (2024)
CIDER: A Causal Cure for Brand-Obsessed Text-to-Image Models
di: Shen, Fangjian, et al.
Pubblicazione: (2025)
di: Shen, Fangjian, et al.
Pubblicazione: (2025)
LMM4LMM: Benchmarking and Evaluating Large-multimodal Image Generation with LMMs
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
Steering and Rectifying Latent Representation Manifolds in Frozen Multi-modal LLMs for Video Anomaly Detection
di: Cai, Zhaolin, et al.
Pubblicazione: (2026)
di: Cai, Zhaolin, et al.
Pubblicazione: (2026)
ManipShield: A Unified Framework for Image Manipulation Detection, Localization and Explanation
di: Xu, Zitong, et al.
Pubblicazione: (2025)
di: Xu, Zitong, et al.
Pubblicazione: (2025)
Exploration of Reproducible Generated Image Detection
di: Duan, Yihang
Pubblicazione: (2025)
di: Duan, Yihang
Pubblicazione: (2025)
Type-R: Automatically Retouching Typos for Text-to-Image Generation
di: Shimoda, Wataru, et al.
Pubblicazione: (2024)
di: Shimoda, Wataru, et al.
Pubblicazione: (2024)
HARIVO: Harnessing Text-to-Image Models for Video Generation
di: Kwon, Mingi, et al.
Pubblicazione: (2024)
di: Kwon, Mingi, et al.
Pubblicazione: (2024)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
di: Duan, Zhongjie, et al.
Pubblicazione: (2024)
di: Duan, Zhongjie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks
di: An, Yanru, et al.
Pubblicazione: (2025) -
MoA-VR: A Mixture-of-Agents System Towards All-in-One Video Restoration
di: Liu, Lu, et al.
Pubblicazione: (2025) -
Life-IQA: Boosting Blind Image Quality Assessment through GCN-enhanced Layer Interaction and MoE-based Feature Decoupling
di: Tang, Long, et al.
Pubblicazione: (2025) -
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
di: Fu, Kang, et al.
Pubblicazione: (2026) -
F-Bench: Rethinking Human Preference Evaluation Metrics for Benchmarking Face Generation, Customization, and Restoration
di: Liu, Lu, et al.
Pubblicazione: (2024)