Depicting Beyond Scores: Advancing Image Quality Assessment through Multi-modal Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | You, Zhiyuan, Li, Zheyuan, Gu, Jinjin, Yin, Zhenfei, Xue, Tianfan, Dong, Chao |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Descriptive Image Quality Assessment with A Large-scale Multi-modal Dataset
by: You, Zhiyuan, et al.
Published: (2024)
by: You, Zhiyuan, et al.
Published: (2024)
Teaching Large Language Models to Regress Accurate Image Quality Scores using Score Distribution
by: You, Zhiyuan, et al.
Published: (2025)
by: You, Zhiyuan, et al.
Published: (2025)
PhotoFramer: Multi-modal Image Composition Instruction
by: You, Zhiyuan, et al.
Published: (2025)
by: You, Zhiyuan, et al.
Published: (2025)
Revisiting the Generalization Problem of Low-level Vision Models Through the Lens of Image Deraining
by: Hu, Jinfan, et al.
Published: (2025)
by: Hu, Jinfan, et al.
Published: (2025)
Harnessing Diffusion-Yielded Score Priors for Image Restoration
by: Lin, Xinqi, et al.
Published: (2025)
by: Lin, Xinqi, et al.
Published: (2025)
An Intelligent Agentic System for Complex Image Restoration Problems
by: Zhu, Kaiwen, et al.
Published: (2024)
by: Zhu, Kaiwen, et al.
Published: (2024)
Interpreting Low-level Vision Models with Causal Effect Maps
by: Hu, Jinfan, et al.
Published: (2024)
by: Hu, Jinfan, et al.
Published: (2024)
UniCon: Unidirectional Information Flow for Effective Control of Large-Scale Diffusion Models
by: Yu, Fanghua, et al.
Published: (2025)
by: Yu, Fanghua, et al.
Published: (2025)
How far have we gone in Generative Image Restoration? A study on its capability, limitations and evaluation practices
by: Yin, Xiang, et al.
Published: (2026)
by: Yin, Xiang, et al.
Published: (2026)
DA-VAE: Plug-in Latent Compression for Diffusion via Detail Alignment
by: Cai, Xin, et al.
Published: (2026)
by: Cai, Xin, et al.
Published: (2026)
DeQA-Doc: Adapting DeQA-Score to Document Image Quality Assessment
by: Gao, Junjie, et al.
Published: (2025)
by: Gao, Junjie, et al.
Published: (2025)
Position: Evaluation of Visual Processing Should Be Human-Centered, Not Metric-Centered
by: Hu, Jinfan, et al.
Published: (2026)
by: Hu, Jinfan, et al.
Published: (2026)
PhoCoLens: Photorealistic and Consistent Reconstruction in Lensless Imaging
by: Cai, Xin, et al.
Published: (2024)
by: Cai, Xin, et al.
Published: (2024)
LM4LV: A Frozen Large Language Model for Low-level Vision Tasks
by: Zheng, Boyang, et al.
Published: (2024)
by: Zheng, Boyang, et al.
Published: (2024)
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
by: Yu, Fanghua, et al.
Published: (2024)
by: Yu, Fanghua, et al.
Published: (2024)
Reasoning as Representation: Rethinking Visual Reinforcement Learning in Image Quality Assessment
by: Zhao, Shijie, et al.
Published: (2025)
by: Zhao, Shijie, et al.
Published: (2025)
Q-Doc: Benchmarking Document Image Quality Assessment Capabilities in Multi-modal Large Language Models
by: Huang, Jiaxi, et al.
Published: (2025)
by: Huang, Jiaxi, et al.
Published: (2025)
Large Multi-modality Model Assisted AI-Generated Image Quality Assessment
by: Wang, Puyi, et al.
Published: (2024)
by: Wang, Puyi, et al.
Published: (2024)
Position: Agentic Systems Constitute a Key Component of Next-Generation Intelligent Image Processing
by: Gu, Jinjin
Published: (2025)
by: Gu, Jinjin
Published: (2025)
AutoDIR: Automatic All-in-One Image Restoration with Latent Diffusion
by: Jiang, Yitong, et al.
Published: (2023)
by: Jiang, Yitong, et al.
Published: (2023)
PhotoAgent: Agentic Photo Editing with Exploratory Visual Aesthetic Planning
by: Yao, Mingde, et al.
Published: (2026)
by: Yao, Mingde, et al.
Published: (2026)
Vision Language Modeling of Content, Distortion and Appearance for Image Quality Assessment
by: Zhou, Fei, et al.
Published: (2024)
by: Zhou, Fei, et al.
Published: (2024)
UltraFusion: Ultra High Dynamic Imaging using Exposure Fusion
by: Chen, Zixuan, et al.
Published: (2025)
by: Chen, Zixuan, et al.
Published: (2025)
A Physics-Informed Blur Learning Framework for Imaging Systems
by: Chen, Liqun, et al.
Published: (2025)
by: Chen, Liqun, et al.
Published: (2025)
AdaptiveISP: Learning an Adaptive Image Signal Processor for Object Detection
by: Wang, Yujin, et al.
Published: (2024)
by: Wang, Yujin, et al.
Published: (2024)
Incorporating Clinical Guidelines through Adapting Multi-modal Large Language Model for Prostate Cancer PI-RADS Scoring
by: Zhang, Tiantian, et al.
Published: (2024)
by: Zhang, Tiantian, et al.
Published: (2024)
Redefining Quality Criteria and Distance-Aware Score Modeling for Image Editing Assessment
by: Zhang, Xinjie, et al.
Published: (2026)
by: Zhang, Xinjie, et al.
Published: (2026)
Q-Ground: Image Quality Grounding with Large Multi-modality Models
by: Chen, Chaofeng, et al.
Published: (2024)
by: Chen, Chaofeng, et al.
Published: (2024)
Vision-Language Consistency Guided Multi-modal Prompt Learning for Blind AI Generated Image Quality Assessment
by: Fu, Jun, et al.
Published: (2024)
by: Fu, Jun, et al.
Published: (2024)
Segmentation Quality and Volumetric Accuracy in Medical Imaging
by: Zhang, Zheyuan, et al.
Published: (2024)
by: Zhang, Zheyuan, et al.
Published: (2024)
Candidate Set Re-ranking for Composed Image Retrieval with Dual Multi-modal Encoder
by: Liu, Zheyuan, et al.
Published: (2023)
by: Liu, Zheyuan, et al.
Published: (2023)
Analytic Score Optimization for Multi Dimension Video Quality Assessment
by: Lin, Boda, et al.
Published: (2026)
by: Lin, Boda, et al.
Published: (2026)
Beyond Score Changes: Adversarial Attack on No-Reference Image Quality Assessment from Two Perspectives
by: Yang, Chenxi, et al.
Published: (2024)
by: Yang, Chenxi, et al.
Published: (2024)
Multi-modal Learnable Queries for Image Aesthetics Assessment
by: Xiong, Zhiwei, et al.
Published: (2024)
by: Xiong, Zhiwei, et al.
Published: (2024)
MP5: A Multi-modal Open-ended Embodied System in Minecraft via Active Perception
by: Qin, Yiran, et al.
Published: (2023)
by: Qin, Yiran, et al.
Published: (2023)
AcT2I: Evaluating and Improving Action Depiction in Text-to-Image Models
by: Malaviya, Vatsal, et al.
Published: (2025)
by: Malaviya, Vatsal, et al.
Published: (2025)
Assessment of Multimodal Large Language Models in Alignment with Human Values
by: Shi, Zhelun, et al.
Published: (2024)
by: Shi, Zhelun, et al.
Published: (2024)
DiffIER: Optimizing Diffusion Models with Iterative Error Reduction
by: Chen, Ao, et al.
Published: (2025)
by: Chen, Ao, et al.
Published: (2025)
A Preliminary Exploration Towards General Image Restoration
by: Kong, Xiangtao, et al.
Published: (2024)
by: Kong, Xiangtao, et al.
Published: (2024)
PolarFree: Polarization-based Reflection-free Imaging
by: Yao, Mingde, et al.
Published: (2025)
by: Yao, Mingde, et al.
Published: (2025)
Similar Items
-
Enhancing Descriptive Image Quality Assessment with A Large-scale Multi-modal Dataset
by: You, Zhiyuan, et al.
Published: (2024) -
Teaching Large Language Models to Regress Accurate Image Quality Scores using Score Distribution
by: You, Zhiyuan, et al.
Published: (2025) -
PhotoFramer: Multi-modal Image Composition Instruction
by: You, Zhiyuan, et al.
Published: (2025) -
Revisiting the Generalization Problem of Low-level Vision Models Through the Lens of Image Deraining
by: Hu, Jinfan, et al.
Published: (2025) -
Harnessing Diffusion-Yielded Score Priors for Image Restoration
by: Lin, Xinqi, et al.
Published: (2025)