Multi-modal Learnable Queries for Image Aesthetics Assessment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiong, Zhiwei, Zhang, Yunfan, Shen, Zhiqi, Ren, Peiran, Yu, Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Image Aesthetics Assessment via Learnable Queries
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
Generalizable Non-Line-of-Sight Imaging with Learnable Physical Priors
von: Sun, Shida, et al.
Veröffentlicht: (2024)
von: Sun, Shida, et al.
Veröffentlicht: (2024)
Style-Consistent 3D Indoor Scene Synthesis with Decoupled Objects
von: Zhang, Yunfan, et al.
Veröffentlicht: (2024)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2024)
UNIAA: A Unified Multi-modal Image Aesthetic Assessment Baseline and Benchmark
von: Zhou, Zhaokun, et al.
Veröffentlicht: (2024)
von: Zhou, Zhaokun, et al.
Veröffentlicht: (2024)
InsTex: Indoor Scenes Stylized Texture Synthesis
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
Advancing Aesthetic Image Generation via Composition Transfer
von: Zou, Kai, et al.
Veröffentlicht: (2026)
von: Zou, Kai, et al.
Veröffentlicht: (2026)
AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics Perception
von: Huang, Yipo, et al.
Veröffentlicht: (2024)
von: Huang, Yipo, et al.
Veröffentlicht: (2024)
Image Aesthetics Assessment using Multi Channel Convolutional Neural Networks
von: Doshi, Nishi, et al.
Veröffentlicht: (2019)
von: Doshi, Nishi, et al.
Veröffentlicht: (2019)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
von: Wang, Hu, et al.
Veröffentlicht: (2023)
von: Wang, Hu, et al.
Veröffentlicht: (2023)
UNK-VQA: A Dataset and a Probe into the Abstention Ability of Multi-modal Large Models
von: Guo, Yangyang, et al.
Veröffentlicht: (2023)
von: Guo, Yangyang, et al.
Veröffentlicht: (2023)
Learning to Look before Learning to Like: Incorporating Human Visual Cognition into Aesthetic Quality Assessment
von: Yu, Liwen, et al.
Veröffentlicht: (2026)
von: Yu, Liwen, et al.
Veröffentlicht: (2026)
Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
von: Ye, Hualin, et al.
Veröffentlicht: (2025)
von: Ye, Hualin, et al.
Veröffentlicht: (2025)
UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries
von: Zhu, Yijie, et al.
Veröffentlicht: (2025)
von: Zhu, Yijie, et al.
Veröffentlicht: (2025)
From Concepts to Judgments: Interpretable Image Aesthetic Assessment
von: Liu, Xiao-Chang, et al.
Veröffentlicht: (2026)
von: Liu, Xiao-Chang, et al.
Veröffentlicht: (2026)
Large Multi-modality Model Assisted AI-Generated Image Quality Assessment
von: Wang, Puyi, et al.
Veröffentlicht: (2024)
von: Wang, Puyi, et al.
Veröffentlicht: (2024)
QTrack: Query-Driven Reasoning for Multi-modal MOT
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
MultiPull: Detailing Signed Distance Functions by Pulling Multi-Level Queries at Multi-Step
von: Noda, Takeshi, et al.
Veröffentlicht: (2024)
von: Noda, Takeshi, et al.
Veröffentlicht: (2024)
LQ-Adapter: ViT-Adapter with Learnable Queries for Gallbladder Cancer Detection from Ultrasound Image
von: Madan, Chetan, et al.
Veröffentlicht: (2024)
von: Madan, Chetan, et al.
Veröffentlicht: (2024)
MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence
von: Liu, Chonghan, et al.
Veröffentlicht: (2025)
von: Liu, Chonghan, et al.
Veröffentlicht: (2025)
RBench-V: A Primary Assessment for Visual Reasoning Models with Multi-modal Outputs
von: Guo, Meng-Hao, et al.
Veröffentlicht: (2025)
von: Guo, Meng-Hao, et al.
Veröffentlicht: (2025)
An Order-Complexity Aesthetic Assessment Model for Aesthetic-aware Music Recommendation
von: Jin, Xin, et al.
Veröffentlicht: (2024)
von: Jin, Xin, et al.
Veröffentlicht: (2024)
Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
von: Behrad, Fatemeh, et al.
Veröffentlicht: (2025)
von: Behrad, Fatemeh, et al.
Veröffentlicht: (2025)
A$^3$: Towards Advertising Aesthetic Assessment
von: Ji, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaiyuan, et al.
Veröffentlicht: (2026)
Fine-grained Image Aesthetic Assessment: Learning Discriminative Scores from Relative Ranks
von: Yang, Zhichao, et al.
Veröffentlicht: (2026)
von: Yang, Zhichao, et al.
Veröffentlicht: (2026)
Micro-gesture Online Recognition using Learnable Query Points
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding
von: Cao, Shuo, et al.
Veröffentlicht: (2025)
von: Cao, Shuo, et al.
Veröffentlicht: (2025)
UniQA: Unified Vision-Language Pre-training for Image Quality and Aesthetic Assessment
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
Beyond Editing Pairs: Fine-Grained Instructional Image Editing via Multi-Scale Learnable Regions
von: Ma, Chenrui, et al.
Veröffentlicht: (2025)
von: Ma, Chenrui, et al.
Veröffentlicht: (2025)
Transformer-empowered Multi-modal Item Embedding for Enhanced Image Search in E-Commerce
von: Liu, Chang, et al.
Veröffentlicht: (2023)
von: Liu, Chang, et al.
Veröffentlicht: (2023)
ForeSea: AI Forensic Search with Multi-modal Queries for Video Surveillance
von: Park, Hyojin, et al.
Veröffentlicht: (2026)
von: Park, Hyojin, et al.
Veröffentlicht: (2026)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
Pixel-Aware Stable Diffusion for Realistic Image Super-resolution and Personalized Stylization
von: Yang, Tao, et al.
Veröffentlicht: (2023)
von: Yang, Tao, et al.
Veröffentlicht: (2023)
Bridging Cognitive Gap: Hierarchical Description Learning for Artistic Image Aesthetics Assessment
von: Liu, Henglin, et al.
Veröffentlicht: (2025)
von: Liu, Henglin, et al.
Veröffentlicht: (2025)
BoQ: A Place is Worth a Bag of Learnable Queries
von: Ali-Bey, Amar, et al.
Veröffentlicht: (2024)
von: Ali-Bey, Amar, et al.
Veröffentlicht: (2024)
SA-IQA: Redefining Image Quality Assessment for Spatial Aesthetics with Multi-Dimensional Rewards
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
Enhancing Descriptive Image Quality Assessment with A Large-scale Multi-modal Dataset
von: You, Zhiyuan, et al.
Veröffentlicht: (2024)
von: You, Zhiyuan, et al.
Veröffentlicht: (2024)
EarthBridge: A Solution for 4th Multi-modal Aerial View Image Challenge Translation Track
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2026)
BackFlip: The Impact of Local and Global Data Augmentations on Artistic Image Aesthetic Assessment
von: Strafforello, Ombretta, et al.
Veröffentlicht: (2024)
von: Strafforello, Ombretta, et al.
Veröffentlicht: (2024)
Large Multi-modal Models Can Interpret Features in Large Multi-modal Models
von: Zhang, Kaichen, et al.
Veröffentlicht: (2024)
von: Zhang, Kaichen, et al.
Veröffentlicht: (2024)
HumanAesExpert: Advancing a Multi-Modality Foundation Model for Human Image Aesthetic Assessment
von: Liao, Zhichao, et al.
Veröffentlicht: (2025)
von: Liao, Zhichao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Image Aesthetics Assessment via Learnable Queries
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023) -
Generalizable Non-Line-of-Sight Imaging with Learnable Physical Priors
von: Sun, Shida, et al.
Veröffentlicht: (2024) -
Style-Consistent 3D Indoor Scene Synthesis with Decoupled Objects
von: Zhang, Yunfan, et al.
Veröffentlicht: (2024) -
UNIAA: A Unified Multi-modal Image Aesthetic Assessment Baseline and Benchmark
von: Zhou, Zhaokun, et al.
Veröffentlicht: (2024) -
InsTex: Indoor Scenes Stylized Texture Synthesis
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)