Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Zhuoxuan, Zhang, Jian, Yuan, Xinbin, Jiang, Peng-Tao, Chen, Wenxiang, Tang, Bowen, Yao, Lujian, Wang, Qiyuan, Chen, Jinwen, Li, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding
by: Fang, Pengcheng, et al.
Published: (2026)
by: Fang, Pengcheng, et al.
Published: (2026)
Efficient and Accurate Image Provenance Analysis: A Scalable Pipeline for Large-scale Images
by: Lai, Jiewei, et al.
Published: (2025)
by: Lai, Jiewei, et al.
Published: (2025)
fMRI Exploration of Visual Quality Assessment
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
Video Quality Assessment for Resolution Cross-Over in Live Sports
by: Zhu, Jingwen, et al.
Published: (2025)
by: Zhu, Jingwen, et al.
Published: (2025)
Scaling Audio-Visual Quality Assessment Dataset via Crowdsourcing
by: Yang, Renyu, et al.
Published: (2026)
by: Yang, Renyu, et al.
Published: (2026)
DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
Bringing Textual Prompt to AI-Generated Image Quality Assessment
by: Qu, Bowen, et al.
Published: (2024)
by: Qu, Bowen, et al.
Published: (2024)
MSPT: A Lightweight Face Image Quality Assessment Method with Multi-stage Progressive Training
by: Xiao, Xiongwei, et al.
Published: (2025)
by: Xiao, Xiongwei, et al.
Published: (2025)
SarcasmMiner: A Dual-Track Post-Training Framework for Robust Audio-Visual Sarcasm Reasoning
by: Li, Zhu, et al.
Published: (2026)
by: Li, Zhu, et al.
Published: (2026)
HistLLM: A Unified Framework for LLM-Based Multimodal Recommendation with User History Encoding and Compression
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
UNQA: Unified No-Reference Quality Assessment for Audio, Image, Video, and Audio-Visual Content
by: Cao, Yuqin, et al.
Published: (2024)
by: Cao, Yuqin, et al.
Published: (2024)
Robust Symbolic Reasoning for Visual Narratives via Hierarchical and Semantically Normalized Knowledge Graphs
by: Chen, Yi-Chun
Published: (2025)
by: Chen, Yi-Chun
Published: (2025)
Video Quality Assessment with Texture Information Fusion for Streaming Applications
by: Menon, Vignesh V, et al.
Published: (2023)
by: Menon, Vignesh V, et al.
Published: (2023)
Predictive Sampling for Efficient Pairwise Subjective Image Quality Assessment
by: Mohammadi, Shima, et al.
Published: (2023)
by: Mohammadi, Shima, et al.
Published: (2023)
Learning Quality from Complexity and Structure: A Feature-Fused XGBoost Model for Video Quality Assessment
by: Premkumar, Amritha, et al.
Published: (2025)
by: Premkumar, Amritha, et al.
Published: (2025)
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning
by: Bai, Hayes, et al.
Published: (2026)
by: Bai, Hayes, et al.
Published: (2026)
AcoustEmo: Open-Vocabulary Emotion Reasoning via Utterance-Aware Acoustic Q-Former
by: Zhang, Liyun, et al.
Published: (2026)
by: Zhang, Liyun, et al.
Published: (2026)
In-place Double Stimulus Methodology for Subjective Assessment of High Quality Images
by: Mohammadi, Shima, et al.
Published: (2025)
by: Mohammadi, Shima, et al.
Published: (2025)
NeRF-QA: Neural Radiance Fields Quality Assessment Database
by: Martin, Pedro, et al.
Published: (2023)
by: Martin, Pedro, et al.
Published: (2023)
Blind Image Quality Assessment Using Visual Neuron Model and Visual Attention Mechanism
by: Hua-Wen Chang, et al.
Published: (2026)
by: Hua-Wen Chang, et al.
Published: (2026)
Photography Perspective Composition: Towards Aesthetic Perspective Recommendation
by: Yao, Lujian, et al.
Published: (2025)
by: Yao, Lujian, et al.
Published: (2025)
ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
Study of Subjective and Objective Quality Assessment of Mobile Cloud Gaming Videos
by: Saha, Avinab, et al.
Published: (2023)
by: Saha, Avinab, et al.
Published: (2023)
SFQA: A Comprehensive Perceptual Quality Assessment Dataset for Singing Face Generation
by: Gao, Zhilin, et al.
Published: (2026)
by: Gao, Zhilin, et al.
Published: (2026)
NeRF View Synthesis: Subjective Quality Assessment and Objective Metrics Evaluation
by: Martin, Pedro, et al.
Published: (2024)
by: Martin, Pedro, et al.
Published: (2024)
Deep Bi-directional Attention Network for Image Super-Resolution Quality Assessment
by: Li, Yixiao, et al.
Published: (2024)
by: Li, Yixiao, et al.
Published: (2024)
Subjective Quality Assessment of Dynamic 3D Meshes in Virtual Reality Environment
by: Nguyen, Duc V., et al.
Published: (2026)
by: Nguyen, Duc V., et al.
Published: (2026)
Perceptual Quality Assessment of Octree-RAHT Encoded 3D Point Clouds
by: Duan, Dongshuai, et al.
Published: (2024)
by: Duan, Dongshuai, et al.
Published: (2024)
SVLA: A Unified Speech-Vision-Language Assistant with Multimodal Reasoning and Speech Generation
by: Huynh, Ngoc Dung, et al.
Published: (2025)
by: Huynh, Ngoc Dung, et al.
Published: (2025)
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
by: Wan, Siqi, et al.
Published: (2025)
by: Wan, Siqi, et al.
Published: (2025)
PointPCA: Point Cloud Objective Quality Assessment Using PCA-Based Descriptors
by: Alexiou, Evangelos, et al.
Published: (2021)
by: Alexiou, Evangelos, et al.
Published: (2021)
Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning
by: Yuan, Xinbin, et al.
Published: (2025)
by: Yuan, Xinbin, et al.
Published: (2025)
Exploring Rich Subjective Quality Information for Image Quality Assessment in the Wild
by: Min, Xiongkuo, et al.
Published: (2024)
by: Min, Xiongkuo, et al.
Published: (2024)
Perceptual Visual Quality Assessment: Principles, Methods, and Future Directions
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
Ges-QA: A Multidimensional Quality Assessment Dataset for Audio-to-3D Gesture Generation
by: Gao, Zhilin, et al.
Published: (2025)
by: Gao, Zhilin, et al.
Published: (2025)
Comparative Study of Subjective Video Quality Assessment Test Methods in Crowdsourcing for Varied Use Cases
by: Naderi, Babak, et al.
Published: (2025)
by: Naderi, Babak, et al.
Published: (2025)
Content-Adaptive Rate-Quality Curve Prediction Model in Media Processing System
by: Yin, Shibo, et al.
Published: (2024)
by: Yin, Shibo, et al.
Published: (2024)
The Interspeech 2026 Audio Reasoning Challenge: Evaluating Reasoning Process Quality for Audio Reasoning Models and Agents
by: Ma, Ziyang, et al.
Published: (2026)
by: Ma, Ziyang, et al.
Published: (2026)
OpenAVS: Training-Free Open-Vocabulary Audio Visual Segmentation with Foundational Models
by: Chen, Shengkai, et al.
Published: (2025)
by: Chen, Shengkai, et al.
Published: (2025)
Similar Items
-
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
by: Lu, Yiting, et al.
Published: (2025) -
MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding
by: Fang, Pengcheng, et al.
Published: (2026) -
Efficient and Accurate Image Provenance Analysis: A Scalable Pipeline for Large-scale Images
by: Lai, Jiewei, et al.
Published: (2025) -
fMRI Exploration of Visual Quality Assessment
by: Zhang, Yiming, et al.
Published: (2024) -
Video Quality Assessment for Resolution Cross-Over in Live Sports
by: Zhu, Jingwen, et al.
Published: (2025)