Logit-Attention Divergence: Mitigating Position Bias in Multi-Image Retrieval via Attention-Guided Calibration
Fuente:
arXiv
Saved in:
| Main Authors: | Xian, Mingtao, Yang, Yifeng, Gu, Qinying, Wang, Xinbing, Ye, Nanyang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$Δ\mathrm{Energy}$: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
TINS: Test-time ID-prototype-separated Negative Semantics Learning for OOD Detection
by: Yang, Yifeng, et al.
Published: (2026)
by: Yang, Yifeng, et al.
Published: (2026)
Less is More: Masking Elements in Image Condition Features Avoids Content Leakages in Style Transfer Diffusion Models
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
OODD: Test-time Out-of-Distribution Detection with Dynamic Dictionary
by: Yang, Yifeng, et al.
Published: (2025)
by: Yang, Yifeng, et al.
Published: (2025)
Bayesian Cross-Modal Alignment Learning for Few-Shot Out-of-Distribution Generalization
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
LLaVA-MLB: Mitigating and Leveraging Attention Bias for Training-Free Video LLMs
by: Shen, Leqi, et al.
Published: (2025)
by: Shen, Leqi, et al.
Published: (2025)
Attention Calibration for Disentangled Text-to-Image Personalization
by: Zhang, Yanbing, et al.
Published: (2024)
by: Zhang, Yanbing, et al.
Published: (2024)
G-NAS: Generalizable Neural Architecture Search for Single Domain Generalization Object Detection
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Looking Back and Forth: Cross-Image Attention Calibration and Attentive Preference Learning for Multi-Image Hallucination Mitigation
by: Yang, Xiaochen, et al.
Published: (2026)
by: Yang, Xiaochen, et al.
Published: (2026)
PNAS-MOT: Multi-Modal Object Tracking with Pareto Neural Architecture Search
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
MIST: Mitigating Intersectional Bias with Disentangled Cross-Attention Editing in Text-to-Image Diffusion Models
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
AttentionGS: Towards Initialization-Free 3D Gaussian Splatting via Structural Attention
by: Liu, Ziao, et al.
Published: (2025)
by: Liu, Ziao, et al.
Published: (2025)
Cross-Modal Attention Calibration for LVLM Hallucination Mitigation
by: Li, Jiaming, et al.
Published: (2025)
by: Li, Jiaming, et al.
Published: (2025)
Identifying and Mitigating Position Bias of Multi-image Vision-Language Models
by: Tian, Xinyu, et al.
Published: (2025)
by: Tian, Xinyu, et al.
Published: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
by: Zhu, Younan, et al.
Published: (2025)
by: Zhu, Younan, et al.
Published: (2025)
Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens
by: Zheng, Haohan, et al.
Published: (2025)
by: Zheng, Haohan, et al.
Published: (2025)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
by: Li, Qiming, et al.
Published: (2026)
by: Li, Qiming, et al.
Published: (2026)
Visual Position Prompt for MLLM based Visual Grounding
by: Tang, Wei, et al.
Published: (2025)
by: Tang, Wei, et al.
Published: (2025)
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
by: Green, Michael, et al.
Published: (2025)
by: Green, Michael, et al.
Published: (2025)
Generalized Logit Adjustment: Calibrating Fine-tuned Models by Removing Label Bias in Foundation Models
by: Zhu, Beier, et al.
Published: (2023)
by: Zhu, Beier, et al.
Published: (2023)
Efficient Masked Image Compression with Position-Indexed Self-Attention
by: Dai, Chengjie, et al.
Published: (2025)
by: Dai, Chengjie, et al.
Published: (2025)
Multi-party Collaborative Attention Control for Image Customization
by: Yang, Han, et al.
Published: (2025)
by: Yang, Han, et al.
Published: (2025)
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
by: Nozawa, Yuji, et al.
Published: (2025)
by: Nozawa, Yuji, et al.
Published: (2025)
Tracing and Mitigating Hallucinations in Multimodal LLMs via Dynamic Attention Localization
by: Yang, Tiancheng, et al.
Published: (2025)
by: Yang, Tiancheng, et al.
Published: (2025)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
by: Fazli, Mehrdad, et al.
Published: (2025)
by: Fazli, Mehrdad, et al.
Published: (2025)
LIME: Localized Image Editing via Attention Regularization in Diffusion Models
by: Simsar, Enis, et al.
Published: (2023)
by: Simsar, Enis, et al.
Published: (2023)
Mitigating Low-Frequency Bias: Feature Recalibration and Frequency Attention Regularization for Adversarial Robustness
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
by: Wang, Qing, et al.
Published: (2025)
by: Wang, Qing, et al.
Published: (2025)
Multi-View Large Reconstruction Model via Geometry-Aware Positional Encoding and Attention
by: Li, Mengfei, et al.
Published: (2024)
by: Li, Mengfei, et al.
Published: (2024)
BiasMap: Leveraging Cross-Attentions to Discover and Mitigate Hidden Social Biases in Text-to-Image Generation
by: Chakraborty, Rajatsubhra, et al.
Published: (2025)
by: Chakraborty, Rajatsubhra, et al.
Published: (2025)
Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation
by: Yang, Chen
Published: (2026)
by: Yang, Chen
Published: (2026)
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
by: Zhao, Jianfei, et al.
Published: (2025)
by: Zhao, Jianfei, et al.
Published: (2025)
Resolving Ambiguity in Composed Image Retrieval via Calibrated Interaction
by: Tran, Amsisan, et al.
Published: (2026)
by: Tran, Amsisan, et al.
Published: (2026)
Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation
by: Chen, Beitao, et al.
Published: (2025)
by: Chen, Beitao, et al.
Published: (2025)
Robust Message Embedding via Attention Flow-Based Steganography
by: Ye, Huayuan, et al.
Published: (2024)
by: Ye, Huayuan, et al.
Published: (2024)
Towards Robust Federated Learning via Logits Calibration on Non-IID Data
by: Qiao, Yu, et al.
Published: (2024)
by: Qiao, Yu, et al.
Published: (2024)
LumiNet: Perception-Driven Knowledge Distillation via Statistical Logit Calibration
by: Hossain, Md. Ismail, et al.
Published: (2023)
by: Hossain, Md. Ismail, et al.
Published: (2023)
FracDetNet: Advanced Fracture Detection via Dual-Focus Attention and Multi-scale Calibration in Medical X-ray Imaging
by: Sun, Yuyang, et al.
Published: (2025)
by: Sun, Yuyang, et al.
Published: (2025)
Gradient-Attention Guided Dual-Masking Synergetic Framework for Robust Text-based Person Retrieval
by: Zheng, Tianlu, et al.
Published: (2025)
by: Zheng, Tianlu, et al.
Published: (2025)
Similar Items
-
$Δ\mathrm{Energy}$: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization
by: Zhu, Lin, et al.
Published: (2025) -
TINS: Test-time ID-prototype-separated Negative Semantics Learning for OOD Detection
by: Yang, Yifeng, et al.
Published: (2026) -
Less is More: Masking Elements in Image Condition Features Avoids Content Leakages in Style Transfer Diffusion Models
by: Zhu, Lin, et al.
Published: (2025) -
OODD: Test-time Out-of-Distribution Detection with Dynamic Dictionary
by: Yang, Yifeng, et al.
Published: (2025) -
Bayesian Cross-Modal Alignment Learning for Few-Shot Out-of-Distribution Generalization
by: Zhu, Lin, et al.
Published: (2025)