Few-Shot Image Quality Assessment via Adaptation of Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Xudong, Huang, Zihao, Zhang, Yan, Shen, Yunhang, Li, Ke, Zheng, Xiawu, Cao, Liujuan, Ji, Rongrong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Feature Denoising Diffusion Model for Blind Image Quality Assessment
di: Li, Xudong, et al.
Pubblicazione: (2024)
di: Li, Xudong, et al.
Pubblicazione: (2024)
Multi-Modal Prompt Learning on Blind Image Quality Assessment
di: Pan, Wensheng, et al.
Pubblicazione: (2024)
di: Pan, Wensheng, et al.
Pubblicazione: (2024)
Q-DeepSight: Incentivizing Thinking with Images for Image Quality Assessment and Refinement
di: Li, Xudong, et al.
Pubblicazione: (2026)
di: Li, Xudong, et al.
Pubblicazione: (2026)
Adaptive Feature Selection for No-Reference Image Quality Assessment by Mitigating Semantic Noise Sensitivity
di: Li, Xudong, et al.
Pubblicazione: (2023)
di: Li, Xudong, et al.
Pubblicazione: (2023)
Contrastive Local Manifold Learning for No-Reference Image Quality Assessment
di: Huang, Zihao, et al.
Pubblicazione: (2024)
di: Huang, Zihao, et al.
Pubblicazione: (2024)
Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs
di: Li, Xudong, et al.
Pubblicazione: (2025)
di: Li, Xudong, et al.
Pubblicazione: (2025)
VEGA: Learning Interleaved Image-Text Comprehension in Vision-Language Large Models
di: Zhou, Chenyu, et al.
Pubblicazione: (2024)
di: Zhou, Chenyu, et al.
Pubblicazione: (2024)
Pseudo-Label Quality Decoupling and Correction for Semi-Supervised Instance Segmentation
di: Lin, Jianghang, et al.
Pubblicazione: (2025)
di: Lin, Jianghang, et al.
Pubblicazione: (2025)
What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation
di: Lin, Jianghang, et al.
Pubblicazione: (2025)
di: Lin, Jianghang, et al.
Pubblicazione: (2025)
Depth-Guided Semi-Supervised Instance Segmentation
di: Chen, Xin, et al.
Pubblicazione: (2024)
di: Chen, Xin, et al.
Pubblicazione: (2024)
Cantor: Inspiring Multimodal Chain-of-Thought of MLLM
di: Gao, Timin, et al.
Pubblicazione: (2024)
di: Gao, Timin, et al.
Pubblicazione: (2024)
AnomalyPainter: Vision-Language-Diffusion Synergy for Zero-Shot Realistic and Diverse Industrial Anomaly Synthesis
di: Lai, Zhangyu, et al.
Pubblicazione: (2025)
di: Lai, Zhangyu, et al.
Pubblicazione: (2025)
HUWSOD: Holistic Self-training for Unified Weakly Supervised Object Detection
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial Inspection
di: Lu, Yilin, et al.
Pubblicazione: (2025)
di: Lu, Yilin, et al.
Pubblicazione: (2025)
Purifying, Labeling, and Utilizing: A High-Quality Pipeline for Small Object Detection
di: Wang, Siwei, et al.
Pubblicazione: (2025)
di: Wang, Siwei, et al.
Pubblicazione: (2025)
Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
di: Luo, Gen, et al.
Pubblicazione: (2024)
di: Luo, Gen, et al.
Pubblicazione: (2024)
Cluster-Aware Prompt Ensemble Learning for Few-Shot Vision-Language Model Adaptation
di: Chen, Zhi, et al.
Pubblicazione: (2025)
di: Chen, Zhi, et al.
Pubblicazione: (2025)
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
di: Sun, Zhen, et al.
Pubblicazione: (2025)
di: Sun, Zhen, et al.
Pubblicazione: (2025)
From Objects to Events: Unlocking Complex Visual Understanding in Object Detectors via LLM-guided Symbolic Reasoning
di: Zeng, Yuhui, et al.
Pubblicazione: (2025)
di: Zeng, Yuhui, et al.
Pubblicazione: (2025)
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation
di: Ke, Shuyan, et al.
Pubblicazione: (2026)
di: Ke, Shuyan, et al.
Pubblicazione: (2026)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
DeOcc-1-to-3: 3D De-Occlusion from a Single Image via Self-Supervised Multi-View Diffusion
di: Qu, Yansong, et al.
Pubblicazione: (2025)
di: Qu, Yansong, et al.
Pubblicazione: (2025)
Scale Contrastive Learning with Selective Attentions for Blind Image Quality Assessment
di: Hu, Runze, et al.
Pubblicazione: (2024)
di: Hu, Runze, et al.
Pubblicazione: (2024)
Solving the Catastrophic Forgetting Problem in Generalized Category Discovery
di: Cao, Xinzi, et al.
Pubblicazione: (2025)
di: Cao, Xinzi, et al.
Pubblicazione: (2025)
FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression
di: Tong, Bo, et al.
Pubblicazione: (2024)
di: Tong, Bo, et al.
Pubblicazione: (2024)
FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
di: Shen, You, et al.
Pubblicazione: (2025)
di: Shen, You, et al.
Pubblicazione: (2025)
MIHBench: Benchmarking and Mitigating Multi-Image Hallucinations in Multimodal Large Language Models
di: Li, Jiale, et al.
Pubblicazione: (2025)
di: Li, Jiale, et al.
Pubblicazione: (2025)
HRSAM: Efficient Interactive Segmentation in High-Resolution Images
di: Huang, You, et al.
Pubblicazione: (2024)
di: Huang, You, et al.
Pubblicazione: (2024)
Unleashing MLLMs on the Edge: A Unified Framework for Cross-Modal ReID via Adaptive SVD Distillation
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
Can Unified Generation and Understanding Models Maintain Semantic Equivalence Across Different Output Modalities?
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
di: Fu, Chaoyou, et al.
Pubblicazione: (2023)
di: Fu, Chaoyou, et al.
Pubblicazione: (2023)
Complementary Subspace Low-Rank Adaptation of Vision-Language Models for Few-Shot Classification
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
PartFormer: Awakening Latent Diverse Representation from Vision Transformer for Object Re-Identification
di: Tan, Lei, et al.
Pubblicazione: (2024)
di: Tan, Lei, et al.
Pubblicazione: (2024)
Semi-Supervised Few-Shot Adaptation of Vision-Language Models
di: Silva-Rodríguez, Julio, et al.
Pubblicazione: (2026)
di: Silva-Rodríguez, Julio, et al.
Pubblicazione: (2026)
Low-Rank Few-Shot Adaptation of Vision-Language Models
di: Zanella, Maxime, et al.
Pubblicazione: (2024)
di: Zanella, Maxime, et al.
Pubblicazione: (2024)
SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning
di: Huang, Haoyu, et al.
Pubblicazione: (2026)
di: Huang, Haoyu, et al.
Pubblicazione: (2026)
An Efficient and Mixed Heterogeneous Model for Image Restoration
di: Gu, Yubin, et al.
Pubblicazione: (2025)
di: Gu, Yubin, et al.
Pubblicazione: (2025)
GS-Bias: Global-Spatial Bias Learner for Single-Image Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
Prototype-Based Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
di: Huang, You, et al.
Pubblicazione: (2025)
di: Huang, You, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Feature Denoising Diffusion Model for Blind Image Quality Assessment
di: Li, Xudong, et al.
Pubblicazione: (2024) -
Multi-Modal Prompt Learning on Blind Image Quality Assessment
di: Pan, Wensheng, et al.
Pubblicazione: (2024) -
Q-DeepSight: Incentivizing Thinking with Images for Image Quality Assessment and Refinement
di: Li, Xudong, et al.
Pubblicazione: (2026) -
Adaptive Feature Selection for No-Reference Image Quality Assessment by Mitigating Semantic Noise Sensitivity
di: Li, Xudong, et al.
Pubblicazione: (2023) -
Contrastive Local Manifold Learning for No-Reference Image Quality Assessment
di: Huang, Zihao, et al.
Pubblicazione: (2024)