See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Azeez, Mohammad Anas, Deria, Ankan, Siddiqui, Zohaib Hasan, Dukre, Adinath Madhavrao, Ali, Rafiq, Atito, Sara, Xie, Yutong, Razzak, Imran |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Atypical Mitosis Classification with DenseNet121: Stain-Aware Augmentation and Hybrid Loss for Domain Generalization
by: Dukre, Adinath, et al.
Published: (2025)
by: Dukre, Adinath, et al.
Published: (2025)
MedMO: Grounding and Understanding Multimodal Large Language Model for Medical Images
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
by: Kolli, Govinda, et al.
Published: (2026)
by: Kolli, Govinda, et al.
Published: (2026)
TuLaBM: Tumor-Biased Latent Bridge Matching for Contrast-Enhanced MRI Synthesis
by: Rege, Atharva, et al.
Published: (2026)
by: Rege, Atharva, et al.
Published: (2026)
Truth, Trust, and Trouble: Medical AI on the Edge
by: Azeez, Mohammad Anas, et al.
Published: (2025)
by: Azeez, Mohammad Anas, et al.
Published: (2025)
ChildGuard: A Specialized Dataset for Combatting Child-Targeted Hate Speech
by: Kashyap, Gautam Siddharth, et al.
Published: (2025)
by: Kashyap, Gautam Siddharth, et al.
Published: (2025)
LLMs on a Budget? Say HOLA
by: Siddiqui, Zohaib Hasan, et al.
Published: (2025)
by: Siddiqui, Zohaib Hasan, et al.
Published: (2025)
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
by: Tang, Feilong, et al.
Published: (2025)
by: Tang, Feilong, et al.
Published: (2025)
Can We Predict Your Next Move Without Breaking Your Privacy?
by: Soni, Arpita, et al.
Published: (2025)
by: Soni, Arpita, et al.
Published: (2025)
Seeing to Ground: Visual Attention for Hallucination-Resilient MDLLMs
by: Narnaware, Vishal, et al.
Published: (2026)
by: Narnaware, Vishal, et al.
Published: (2026)
Attention-Driven Framework for Non-Rigid Medical Image Registration
by: Iqbal, Muhammad Zafar, et al.
Published: (2026)
by: Iqbal, Muhammad Zafar, et al.
Published: (2026)
Do Large Language Models Speak All Languages Equally? A Comparative Study in Low-Resource Settings
by: Hasan, Md. Arid, et al.
Published: (2024)
by: Hasan, Md. Arid, et al.
Published: (2024)
A Machine Learning Approach to Predict Biological Age and its Longitudinal Drivers
by: Dunbayeva, Nazira, et al.
Published: (2025)
by: Dunbayeva, Nazira, et al.
Published: (2025)
MuGa-VTON: Multi-Garment Virtual Try-On via Diffusion Transformers with Prompt Customization
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training
by: Kumar, Komal, et al.
Published: (2026)
by: Kumar, Komal, et al.
Published: (2026)
PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
by: Alex, Tony, et al.
Published: (2025)
by: Alex, Tony, et al.
Published: (2025)
Fairness Evaluation for Uplift Modeling in the Absence of Ground Truth
by: Kadioglu, Serdar, et al.
Published: (2024)
by: Kadioglu, Serdar, et al.
Published: (2024)
DeepChest: Dynamic Gradient-Free Task Weighting for Effective Multi-Task Learning in Chest X-ray Classification
by: Mohamed, Youssef, et al.
Published: (2025)
by: Mohamed, Youssef, et al.
Published: (2025)
Fair Orientations: Proportionality and Equitability
by: Sun, Ankang, et al.
Published: (2026)
by: Sun, Ankang, et al.
Published: (2026)
Adenocarcinoma Segmentation Using Pre-trained Swin-UNet with Parallel Cross-Attention for Multi-Domain Imaging
by: Qayyum, Abdul, et al.
Published: (2024)
by: Qayyum, Abdul, et al.
Published: (2024)
Retinal Lipidomics Associations as Candidate Biomarkers for Cardiovascular Health
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
The Eye as a Window to Systemic Health: A Survey of Retinal Imaging from Classical Techniques to Oculomics
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
Reducing Tool Hallucination via Reliability Alignment
by: Xu, Hongshen, et al.
Published: (2024)
by: Xu, Hongshen, et al.
Published: (2024)
Region Guided Attention Network for Retinal Vessel Segmentation
by: Javed, Syed, et al.
Published: (2024)
by: Javed, Syed, et al.
Published: (2024)
Listen Then See: Video Alignment with Speaker Attention
by: Agrawal, Aviral, et al.
Published: (2024)
by: Agrawal, Aviral, et al.
Published: (2024)
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
by: Matys, Piotr, et al.
Published: (2025)
by: Matys, Piotr, et al.
Published: (2025)
CMSA-Net: Causal Multi-scale Aggregation with Adaptive Multi-source Reference for Video Polyp Segmentation
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
FairGRPO: Fair Reinforcement Learning for Equitable Clinical Reasoning
by: Dai, Shiqi, et al.
Published: (2025)
by: Dai, Shiqi, et al.
Published: (2025)
Better to Ask in English: Evaluation of Large Language Models on English, Low-resource and Cross-Lingual Settings
by: Dey, Krishno, et al.
Published: (2024)
by: Dey, Krishno, et al.
Published: (2024)
Ground Truths
Published: (2024)
Published: (2024)
LMBF-Net: A Lightweight Multipath Bidirectional Focal Attention Network for Multifeatures Segmentation
by: Khan, Tariq M, et al.
Published: (2024)
by: Khan, Tariq M, et al.
Published: (2024)
You Only Speak Once to See
by: Yang, Wenhao, et al.
Published: (2024)
by: Yang, Wenhao, et al.
Published: (2024)
Calm-Whisper: Reduce Whisper Hallucination On Non-Speech By Calming Crazy Heads Down
by: Wang, Yingzhi, et al.
Published: (2025)
by: Wang, Yingzhi, et al.
Published: (2025)
A Robust Algorithm for Contactless Fingerprint Enhancement and Matching
by: Siddiqui, Mahrukh, et al.
Published: (2024)
by: Siddiqui, Mahrukh, et al.
Published: (2024)
Fair and Equitable Benefit-sharing in International Law
by: Morgera, Elisa
Published: (2024)
by: Morgera, Elisa
Published: (2024)
Reallocating Attention Across Layers to Reduce Multimodal Hallucination
by: Lu, Haolang, et al.
Published: (2025)
by: Lu, Haolang, et al.
Published: (2025)
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs
by: Zhang, Xiaofeng, et al.
Published: (2024)
by: Zhang, Xiaofeng, et al.
Published: (2024)
Amino acid profile of eighteen isolates of different edible macrofungal species
by: S. Azeez
Published: (2020)
by: S. Azeez
Published: (2020)
Similar Items
-
Robust Atypical Mitosis Classification with DenseNet121: Stain-Aware Augmentation and Hybrid Loss for Domain Generalization
by: Dukre, Adinath, et al.
Published: (2025) -
MedMO: Grounding and Understanding Multimodal Large Language Model for Medical Images
by: Deria, Ankan, et al.
Published: (2026) -
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
by: Deria, Ankan, et al.
Published: (2025) -
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
by: Kolli, Govinda, et al.
Published: (2026) -
TuLaBM: Tumor-Biased Latent Bridge Matching for Contrast-Enhanced MRI Synthesis
by: Rege, Atharva, et al.
Published: (2026)