Redemption Score: A Multi-Modal Evaluation Framework for Image Captioning via Distributional, Perceptual, and Linguistic Signal Triangulation
Fuente:
arXiv
Saved in:
| Main Authors: | Dahal, Ashim, Ghimire, Ankit, Murad, Saydul Akbar, Rahimi, Nick |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Lingual Cyber Threat Detection in Tweets/X Using ML, DL, and LLM: A Comparative Analysis
by: Murad, Saydul Akbar, et al.
Published: (2025)
by: Murad, Saydul Akbar, et al.
Published: (2025)
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity
by: Murad, Saydul Akbar, et al.
Published: (2025)
by: Murad, Saydul Akbar, et al.
Published: (2025)
POVQA: Preference-Optimized Video Question Answering with Rationales for Data Efficiency
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
Efficiency Bottlenecks of Convolutional Kolmogorov-Arnold Networks: A Comprehensive Scrutiny with ImageNet, AlexNet, LeNet and Tabular Classification
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
Embedding Shift Dissection on CLIP: Effects of Augmentations on VLM's Representation Learning
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
Heuristical Comparison of Vision Transformers Against Convolutional Neural Networks for Semantic Segmentation on Remote Sensing Imagery
by: Dahal, Ashim, et al.
Published: (2024)
by: Dahal, Ashim, et al.
Published: (2024)
Unveiling Thoughts: A Review of Advancements in EEG Brain Signal Decoding into Text
by: Murad, Saydul Akbar, et al.
Published: (2024)
by: Murad, Saydul Akbar, et al.
Published: (2024)
A Depth-Aware Comparative Study of Euclidean and Hyperbolic Graph Neural Networks on Bitcoin Transaction Systems
by: Ghimire, Ankit, et al.
Published: (2026)
by: Ghimire, Ankit, et al.
Published: (2026)
A Fusion of context-aware based BanglaBERT and Two-Layer Stacked LSTM Framework for Multi-Label Cyberbullying Detection
by: Raquib, Mirza, et al.
Published: (2026)
by: Raquib, Mirza, et al.
Published: (2026)
A Unified BERT-CNN-BiLSTM Framework for Simultaneous Headline Classification and Sentiment Analysis of Bangla News
by: Raquib, Mirza, et al.
Published: (2025)
by: Raquib, Mirza, et al.
Published: (2025)
Adaptive Anchor Policies for Efficient 4D Gaussian Streaming
by: Dahal, Ashim, et al.
Published: (2026)
by: Dahal, Ashim, et al.
Published: (2026)
Gamma2Patterns: Deep Cognitive Attention Region Identification and Gamma-Alpha Pattern Analysis
by: Jahan, Sobhana, et al.
Published: (2026)
by: Jahan, Sobhana, et al.
Published: (2026)
MpoxSLDNet: A Novel CNN Model for Detecting Monkeypox Lesions and Performance Comparison with Pre-trained Models
by: Dihan, Fatema Jannat, et al.
Published: (2024)
by: Dihan, Fatema Jannat, et al.
Published: (2024)
Multi-Head Attention based interaction-aware architecture for Bangla Handwritten Character Recognition: Introducing a Primary Dataset
by: Raquib, Mirza, et al.
Published: (2026)
by: Raquib, Mirza, et al.
Published: (2026)
SPECS: Specificity-Enhanced CLIP-Score for Long Image Caption Evaluation
by: Chen, Xiaofu, et al.
Published: (2025)
by: Chen, Xiaofu, et al.
Published: (2025)
Linguistically Informed Multimodal Fusion for Vietnamese Scene-Text Image Captioning: Dataset, Graph Framework, and Phonological Attention
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
Robust and Real-Time Bangladeshi Currency Recognition: A Dual-Stream MobileNet and EfficientNet Approach
by: Subreena, et al.
Published: (2026)
by: Subreena, et al.
Published: (2026)
ScaleCap: Inference-Time Scalable Image Captioning via Dual-Modality Debiasing
by: Xing, Long, et al.
Published: (2025)
by: Xing, Long, et al.
Published: (2025)
Beyond Perplexity: Multi-dimensional Safety Evaluation of LLM Compression
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
GONE: Structural Knowledge Unlearning via Neighborhood-Expanded Distribution Shaping
by: Dahal, Chahana, et al.
Published: (2026)
by: Dahal, Chahana, et al.
Published: (2026)
Text-Only Training for Image Captioning with Retrieval Augmentation and Modality Gap Correction
by: Fonseca, Rui, et al.
Published: (2025)
by: Fonseca, Rui, et al.
Published: (2025)
Analysis of Zero Day Attack Detection Using MLP and XAI
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
Image Captioning via Compact Bidirectional Architecture
by: Song, Zijie, et al.
Published: (2022)
by: Song, Zijie, et al.
Published: (2022)
MSD-Score: Multi-Scale Distributional Scoring for Reference-Free Image Caption Evaluation
by: Kan, Shichao, et al.
Published: (2026)
by: Kan, Shichao, et al.
Published: (2026)
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
by: Gupta, Ashim, et al.
Published: (2025)
by: Gupta, Ashim, et al.
Published: (2025)
Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
by: Gupta, Ashim, et al.
Published: (2025)
by: Gupta, Ashim, et al.
Published: (2025)
CAF-Score: Calibrating CLAP with LALMs for Reference-free Audio Captioning Evaluation
by: Lee, Insung, et al.
Published: (2026)
by: Lee, Insung, et al.
Published: (2026)
GEMA-Score: Granular Explainable Multi-Agent Scoring Framework for Radiology Report Evaluation
by: Zhang, Zhenxuan, et al.
Published: (2025)
by: Zhang, Zhenxuan, et al.
Published: (2025)
CIC: A Framework for Culturally-Aware Image Captioning
by: Yun, Youngsik, et al.
Published: (2024)
by: Yun, Youngsik, et al.
Published: (2024)
Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts
by: Nooralahzadeh, Farhad, et al.
Published: (2026)
by: Nooralahzadeh, Farhad, et al.
Published: (2026)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
by: Cheng, Myra, et al.
Published: (2024)
by: Cheng, Myra, et al.
Published: (2024)
Altogether: Image Captioning via Re-aligning Alt-text
by: Xu, Hu, et al.
Published: (2024)
by: Xu, Hu, et al.
Published: (2024)
Towards a Linguistic Evaluation of Narratives: A Quantitative Stylistic Framework
by: Maisto, Alessandro
Published: (2026)
by: Maisto, Alessandro
Published: (2026)
Can LLM Agents Identify Spoken Dialects like a Linguist?
by: Bystrich, Tobias, et al.
Published: (2026)
by: Bystrich, Tobias, et al.
Published: (2026)
LinguaSynth: Heterogeneous Linguistic Signals for News Classification
by: Zhang, Duo, et al.
Published: (2025)
by: Zhang, Duo, et al.
Published: (2025)
On the Wings of Imagination: Conflicting Script-based Multi-role Framework for Humor Caption Generation
by: Shang, Wenbo, et al.
Published: (2026)
by: Shang, Wenbo, et al.
Published: (2026)
EXPERT: An Explainable Image Captioning Evaluation Metric with Structured Explanations
by: Kim, Hyunjong, et al.
Published: (2025)
by: Kim, Hyunjong, et al.
Published: (2025)
Multi-Modal Language Models as Text-to-Image Model Evaluators
by: Chen, Jiahui, et al.
Published: (2025)
by: Chen, Jiahui, et al.
Published: (2025)
MMMModal -- Multi-Images Multi-Audio Multi-turn Multi-Modal
by: Zolkepli, Husein, et al.
Published: (2024)
by: Zolkepli, Husein, et al.
Published: (2024)
Agent-Driven Corpus Linguistics: A Framework for Autonomous Linguistic Discovery
by: Yu, Jia, et al.
Published: (2026)
by: Yu, Jia, et al.
Published: (2026)
Similar Items
-
Multi-Lingual Cyber Threat Detection in Tweets/X Using ML, DL, and LLM: A Comparative Analysis
by: Murad, Saydul Akbar, et al.
Published: (2025) -
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity
by: Murad, Saydul Akbar, et al.
Published: (2025) -
POVQA: Preference-Optimized Video Question Answering with Rationales for Data Efficiency
by: Dahal, Ashim, et al.
Published: (2025) -
Efficiency Bottlenecks of Convolutional Kolmogorov-Arnold Networks: A Comprehensive Scrutiny with ImageNet, AlexNet, LeNet and Tabular Classification
by: Dahal, Ashim, et al.
Published: (2025) -
Embedding Shift Dissection on CLIP: Effects of Augmentations on VLM's Representation Learning
by: Dahal, Ashim, et al.
Published: (2025)