Treble Counterfactual VLMs: A Causal Approach to Hallucination
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Shawn, Qu, Jiashu, Zhou, Yuxiao, Qin, Yuehan, Yang, Tiankai, Zhao, Yue |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DPU: Dynamic Prototype Updating for Multimodal Out-of-Distribution Detection
por: Li, Shawn, et al.
Publicado: (2024)
por: Li, Shawn, et al.
Publicado: (2024)
CheXPO: Preference Optimization for Chest X-ray VLMs with Counterfactual Rationale
por: Liang, Xiao, et al.
Publicado: (2025)
por: Liang, Xiao, et al.
Publicado: (2025)
Tone Matters: The Impact of Linguistic Tone on Hallucination in VLMs
por: Hong, Weihao, et al.
Publicado: (2026)
por: Hong, Weihao, et al.
Publicado: (2026)
DASH: Detection and Assessment of Systematic Hallucinations of VLMs
por: Augustin, Maximilian, et al.
Publicado: (2025)
por: Augustin, Maximilian, et al.
Publicado: (2025)
Latent Causal Modeling for 3D Brain MRI Counterfactuals
por: Peng, Wei, et al.
Publicado: (2024)
por: Peng, Wei, et al.
Publicado: (2024)
Cross-modal Causal Intervention for Alzheimer's Disease Prediction
por: Jin, Yutao, et al.
Publicado: (2025)
por: Jin, Yutao, et al.
Publicado: (2025)
Taming Hallucinations: Boosting MLLMs' Video Understanding via Counterfactual Video Generation
por: Huang, Zhe, et al.
Publicado: (2025)
por: Huang, Zhe, et al.
Publicado: (2025)
Causally Steered Diffusion for Automated Video Counterfactual Generation
por: Spyrou, Nikos, et al.
Publicado: (2025)
por: Spyrou, Nikos, et al.
Publicado: (2025)
HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning
por: Lymperaiou, Maria, et al.
Publicado: (2025)
por: Lymperaiou, Maria, et al.
Publicado: (2025)
SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters
por: Shah, Arya, et al.
Publicado: (2026)
por: Shah, Arya, et al.
Publicado: (2026)
Causally-Grounded Dual-Path Attention Intervention for Object Hallucination Mitigation in LVLMs
por: Yu, Liu, et al.
Publicado: (2025)
por: Yu, Liu, et al.
Publicado: (2025)
UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation
por: Zhao, Xiangyu, et al.
Publicado: (2024)
por: Zhao, Xiangyu, et al.
Publicado: (2024)
A Cognitive Paradigm Approach to Probe the Perception-Reasoning Interface in VLMs
por: Vaishnav, Mohit, et al.
Publicado: (2025)
por: Vaishnav, Mohit, et al.
Publicado: (2025)
Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs
por: Lin, Chenchen, et al.
Publicado: (2026)
por: Lin, Chenchen, et al.
Publicado: (2026)
CMOOD: Concept-based Multi-label OOD Detection
por: Liu, Zhendong, et al.
Publicado: (2024)
por: Liu, Zhendong, et al.
Publicado: (2024)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
por: Tong, Lei, et al.
Publicado: (2025)
por: Tong, Lei, et al.
Publicado: (2025)
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
por: Clark, Christopher, et al.
Publicado: (2026)
por: Clark, Christopher, et al.
Publicado: (2026)
Causal Decoding for Hallucination-Resistant Multimodal Large Language Models
por: Tan, Shiwei, et al.
Publicado: (2026)
por: Tan, Shiwei, et al.
Publicado: (2026)
LogicGaze: Benchmarking Causal Consistency in Visual Narratives via Counterfactual Verification
por: Driscoll, Rory, et al.
Publicado: (2026)
por: Driscoll, Rory, et al.
Publicado: (2026)
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
por: Li, Shuo, et al.
Publicado: (2024)
por: Li, Shuo, et al.
Publicado: (2024)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
por: Sun, Han, et al.
Publicado: (2026)
por: Sun, Han, et al.
Publicado: (2026)
What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-modal Models
por: Kim, Junho, et al.
Publicado: (2024)
por: Kim, Junho, et al.
Publicado: (2024)
VLMs have Tunnel Vision: Evaluating Nonlocal Visual Reasoning in Leading VLMs
por: Berman, Shmuel, et al.
Publicado: (2025)
por: Berman, Shmuel, et al.
Publicado: (2025)
Vision-Language Introspection: Mitigating Overconfident Hallucinations in MLLMs via Interpretable Bi-Causal Steering
por: Liu, Shuliang, et al.
Publicado: (2026)
por: Liu, Shuliang, et al.
Publicado: (2026)
Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination
por: Li, Xinzhuo, et al.
Publicado: (2025)
por: Li, Xinzhuo, et al.
Publicado: (2025)
Improved Noise Schedule for Diffusion Training
por: Hang, Tiankai, et al.
Publicado: (2024)
por: Hang, Tiankai, et al.
Publicado: (2024)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
por: Luo, Kun, et al.
Publicado: (2026)
por: Luo, Kun, et al.
Publicado: (2026)
Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving
por: Lian, Weitong, et al.
Publicado: (2026)
por: Lian, Weitong, et al.
Publicado: (2026)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
por: Li, Yiwei, et al.
Publicado: (2026)
por: Li, Yiwei, et al.
Publicado: (2026)
Two Causally Related Needles in a Video Haystack
por: Li, Miaoyu, et al.
Publicado: (2025)
por: Li, Miaoyu, et al.
Publicado: (2025)
VACoT: Rethinking Visual Data Augmentation with VLMs
por: Xu, Zhengzhuo, et al.
Publicado: (2025)
por: Xu, Zhengzhuo, et al.
Publicado: (2025)
Decoding the Pulse of Reasoning VLMs in Multi-Image Understanding Tasks
por: Li, Chenjun
Publicado: (2026)
por: Li, Chenjun
Publicado: (2026)
Focusing by Contrastive Attention: Enhancing VLMs' Visual Reasoning
por: Ge, Yuyao, et al.
Publicado: (2025)
por: Ge, Yuyao, et al.
Publicado: (2025)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
por: Yin, Jianghao, et al.
Publicado: (2026)
por: Yin, Jianghao, et al.
Publicado: (2026)
Listener-Rewarded Thinking in VLMs for Image Preferences
por: Gambashidze, Alexander, et al.
Publicado: (2025)
por: Gambashidze, Alexander, et al.
Publicado: (2025)
CounterVid: Counterfactual Video Generation for Mitigating Action and Temporal Hallucinations in Video-Language Models
por: Poppi, Tobia, et al.
Publicado: (2026)
por: Poppi, Tobia, et al.
Publicado: (2026)
SpinBench: Perspective and Rotation as a Lens on Spatial Reasoning in VLMs
por: Zhang, Yuyou, et al.
Publicado: (2025)
por: Zhang, Yuyou, et al.
Publicado: (2025)
Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge
por: Lu, Shuai, et al.
Publicado: (2026)
por: Lu, Shuai, et al.
Publicado: (2026)
Replace in Translation: Boost Concept Alignment in Counterfactual Text-to-Image
por: Li, Sifan, et al.
Publicado: (2025)
por: Li, Sifan, et al.
Publicado: (2025)
TPC: Cross-Temporal Prediction Connection for Vision-Language Model Hallucination Reduction
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Ejemplares similares
-
DPU: Dynamic Prototype Updating for Multimodal Out-of-Distribution Detection
por: Li, Shawn, et al.
Publicado: (2024) -
CheXPO: Preference Optimization for Chest X-ray VLMs with Counterfactual Rationale
por: Liang, Xiao, et al.
Publicado: (2025) -
Tone Matters: The Impact of Linguistic Tone on Hallucination in VLMs
por: Hong, Weihao, et al.
Publicado: (2026) -
DASH: Detection and Assessment of Systematic Hallucinations of VLMs
por: Augustin, Maximilian, et al.
Publicado: (2025) -
Latent Causal Modeling for 3D Brain MRI Counterfactuals
por: Peng, Wei, et al.
Publicado: (2024)