Modeling Image-Caption Rating from Comparative Judgments
Fuente:
arXiv
Salvato in:
| Autori principali: | Minni, Kezia, Zhang, Qiang, Khan, Monoshiz Mahbub, Yu, Zhe |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Modeling Art Evaluations from Comparative Judgments: A Deep Learning Approach to Predicting Aesthetic Preferences
di: Bethi, Manoj Reddy, et al.
Pubblicazione: (2026)
di: Bethi, Manoj Reddy, et al.
Pubblicazione: (2026)
Pretrained Image-Text Models are Secretly Video Captioners
di: Zhang, Chunhui, et al.
Pubblicazione: (2025)
di: Zhang, Chunhui, et al.
Pubblicazione: (2025)
Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments
di: Depaolini, Icaro Re, et al.
Pubblicazione: (2026)
di: Depaolini, Icaro Re, et al.
Pubblicazione: (2026)
Differentially Private Representation Learning via Image Captioning
di: Sander, Tom, et al.
Pubblicazione: (2024)
di: Sander, Tom, et al.
Pubblicazione: (2024)
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023)
di: Lei, Shiye, et al.
Pubblicazione: (2023)
Pixels to Prose: Understanding the art of Image Captioning
di: Singh, Hrishikesh, et al.
Pubblicazione: (2024)
di: Singh, Hrishikesh, et al.
Pubblicazione: (2024)
Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models
di: Lai, Zhengfeng, et al.
Pubblicazione: (2024)
di: Lai, Zhengfeng, et al.
Pubblicazione: (2024)
Linear Alignment of Vision-language Models for Image Captioning
di: Paischer, Fabian, et al.
Pubblicazione: (2023)
di: Paischer, Fabian, et al.
Pubblicazione: (2023)
Image-Caption Encoding for Improving Zero-Shot Generalization
di: Yu, Eric Yang, et al.
Pubblicazione: (2024)
di: Yu, Eric Yang, et al.
Pubblicazione: (2024)
ICC: Quantifying Image Caption Concreteness for Multimodal Dataset Curation
di: Yanuka, Moran, et al.
Pubblicazione: (2024)
di: Yanuka, Moran, et al.
Pubblicazione: (2024)
STAND: Semantic Anchoring Constraint with Dual-Granularity Disambiguation for Remote Sensing Image Change Captioning
di: Gong, Yanpei, et al.
Pubblicazione: (2026)
di: Gong, Yanpei, et al.
Pubblicazione: (2026)
Generalizable Geometric Image Caption Synthesis
di: Xin, Yue, et al.
Pubblicazione: (2025)
di: Xin, Yue, et al.
Pubblicazione: (2025)
Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge
di: Dognin, Pierre, et al.
Pubblicazione: (2020)
di: Dognin, Pierre, et al.
Pubblicazione: (2020)
Leveraging Pre-trained CNNs for Efficient Feature Extraction in Rice Leaf Disease Classification
di: Sobuj, Md. Shohanur Islam, et al.
Pubblicazione: (2024)
di: Sobuj, Md. Shohanur Islam, et al.
Pubblicazione: (2024)
Large VLM-based Stylized Sports Captioning
di: Dhar, Sauptik, et al.
Pubblicazione: (2025)
di: Dhar, Sauptik, et al.
Pubblicazione: (2025)
PICS: Pipeline for Image Captioning and Search
di: Rosario, Grant, et al.
Pubblicazione: (2024)
di: Rosario, Grant, et al.
Pubblicazione: (2024)
Hyperdimensional Cross-Modal Alignment of Frozen Language and Image Models for Efficient Image Captioning
di: Dalvi, Abhishek, et al.
Pubblicazione: (2026)
di: Dalvi, Abhishek, et al.
Pubblicazione: (2026)
Emergent Natural Language with Communication Games for Improving Image Captioning Capabilities without Additional Data
di: Dutta, Parag, et al.
Pubblicazione: (2025)
di: Dutta, Parag, et al.
Pubblicazione: (2025)
Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
BioCAP: Exploiting Synthetic Captions Beyond Labels in Biological Foundation Models
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
Enhancing Diffusion Model Stability for Image Restoration via Gradient Management
di: Wu, Hongjie, et al.
Pubblicazione: (2025)
di: Wu, Hongjie, et al.
Pubblicazione: (2025)
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers
di: Yang, Chenglin, et al.
Pubblicazione: (2023)
di: Yang, Chenglin, et al.
Pubblicazione: (2023)
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
di: Deria, Ankan, et al.
Pubblicazione: (2025)
di: Deria, Ankan, et al.
Pubblicazione: (2025)
Image Clustering via the Principle of Rate Reduction in the Age of Pretrained Models
di: Chu, Tianzhe, et al.
Pubblicazione: (2023)
di: Chu, Tianzhe, et al.
Pubblicazione: (2023)
From Pixels to Prose: A Large Dataset of Dense Image Captions
di: Singla, Vasu, et al.
Pubblicazione: (2024)
di: Singla, Vasu, et al.
Pubblicazione: (2024)
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
Comparative Analysis of Deep Learning Models for Perception in Autonomous Vehicles
di: Khan, Jalal
Pubblicazione: (2025)
di: Khan, Jalal
Pubblicazione: (2025)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
di: Luo, Jianjie, et al.
Pubblicazione: (2024)
di: Luo, Jianjie, et al.
Pubblicazione: (2024)
Good Representation, Better Explanation: Role of Convolutional Neural Networks in Transformer-Based Remote Sensing Image Captioning
di: Das, Swadhin, et al.
Pubblicazione: (2025)
di: Das, Swadhin, et al.
Pubblicazione: (2025)
Diff-3DCap: Shape Captioning with Diffusion Models
di: Shu, Zhenyu, et al.
Pubblicazione: (2025)
di: Shu, Zhenyu, et al.
Pubblicazione: (2025)
SynC: Synthetic Image Caption Dataset Refinement with One-to-many Mapping for Zero-shot Image Captioning
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
PM2: A New Prompting Multi-modal Model Paradigm for Few-shot Medical Image Classification
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
GNN-ViTCap: GNN-Enhanced Multiple Instance Learning with Vision Transformers for Whole Slide Image Classification and Captioning
di: Raju, S M Taslim Uddin, et al.
Pubblicazione: (2025)
di: Raju, S M Taslim Uddin, et al.
Pubblicazione: (2025)
Learning to Rank Caption Chains for Video-Text Alignment
di: Blume, Ansel, et al.
Pubblicazione: (2026)
di: Blume, Ansel, et al.
Pubblicazione: (2026)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
di: Lai, Zhengfeng, et al.
Pubblicazione: (2023)
di: Lai, Zhengfeng, et al.
Pubblicazione: (2023)
Brazilian Portuguese Image Captioning with Transformers: A Study on Cross-Native-Translated Dataset
di: Bromonschenkel, Gabriel, et al.
Pubblicazione: (2026)
di: Bromonschenkel, Gabriel, et al.
Pubblicazione: (2026)
Temporal Object Captioning for Street Scene Videos from LiDAR Tracks
di: Gopinathan, Vignesh, et al.
Pubblicazione: (2025)
di: Gopinathan, Vignesh, et al.
Pubblicazione: (2025)
Semi-Supervised Image Captioning Considering Wasserstein Graph Matching
di: Yang, Yang
Pubblicazione: (2024)
di: Yang, Yang
Pubblicazione: (2024)
AICAttack: Adversarial Image Captioning Attack with Attention-Based Optimization
di: Li, Jiyao, et al.
Pubblicazione: (2024)
di: Li, Jiyao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Modeling Art Evaluations from Comparative Judgments: A Deep Learning Approach to Predicting Aesthetic Preferences
di: Bethi, Manoj Reddy, et al.
Pubblicazione: (2026) -
Pretrained Image-Text Models are Secretly Video Captioners
di: Zhang, Chunhui, et al.
Pubblicazione: (2025) -
Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments
di: Depaolini, Icaro Re, et al.
Pubblicazione: (2026) -
Differentially Private Representation Learning via Image Captioning
di: Sander, Tom, et al.
Pubblicazione: (2024) -
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023)