SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters
Fuente:
arXiv
Guardado en:
| Autores principales: | Shah, Arya, Mishra, Deepali, Silpasuwanchai, Chaklam |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models
por: Shah, Arya, et al.
Publicado: (2026)
por: Shah, Arya, et al.
Publicado: (2026)
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation
por: Shah, Arya, et al.
Publicado: (2026)
por: Shah, Arya, et al.
Publicado: (2026)
Barriers in Integrating Medical Visual Question Answering into Radiology Workflows: A Scoping Review and Clinicians' Insights
por: Mishra, Deepali, et al.
Publicado: (2025)
por: Mishra, Deepali, et al.
Publicado: (2025)
To See or To Please: Uncovering Visual Sycophancy and Split Beliefs in VLMs
por: Hong, Rui, et al.
Publicado: (2026)
por: Hong, Rui, et al.
Publicado: (2026)
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
por: Li, Shuo, et al.
Publicado: (2024)
por: Li, Shuo, et al.
Publicado: (2024)
Benchmarking and Mitigating Sycophancy in Medical Vision Language Models
por: Xu, Juangui, et al.
Publicado: (2025)
por: Xu, Juangui, et al.
Publicado: (2025)
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
por: Yuan, Botai, et al.
Publicado: (2025)
por: Yuan, Botai, et al.
Publicado: (2025)
To Agree or To Be Right? The Grounding-Sycophancy Tradeoff in Medical Vision-Language Models
por: Aranya, OFM Riaz Rahman, et al.
Publicado: (2026)
por: Aranya, OFM Riaz Rahman, et al.
Publicado: (2026)
Edge Reliability Gap in Vision-Language Models: Quantifying Failure Modes of Compressed VLMs Under Visual Corruption
por: Erol, Mehmet Kaan
Publicado: (2026)
por: Erol, Mehmet Kaan
Publicado: (2026)
From Overload to Convergence: Supporting Multi-Issue Human-AI Negotiation with Bayesian Visualization
por: Parmar, Mehul, et al.
Publicado: (2026)
por: Parmar, Mehul, et al.
Publicado: (2026)
PAS : Prelim Attention Score for Detecting Object Hallucinations in Large Vision--Language Models
por: Hoang-Xuan, Nhat, et al.
Publicado: (2025)
por: Hoang-Xuan, Nhat, et al.
Publicado: (2025)
Treble Counterfactual VLMs: A Causal Approach to Hallucination
por: Li, Shawn, et al.
Publicado: (2025)
por: Li, Shawn, et al.
Publicado: (2025)
DASH: Detection and Assessment of Systematic Hallucinations of VLMs
por: Augustin, Maximilian, et al.
Publicado: (2025)
por: Augustin, Maximilian, et al.
Publicado: (2025)
Evaluating Vision Language Models (VLMs) for Radiology: A Comprehensive Analysis
por: Li, Frank, et al.
Publicado: (2025)
por: Li, Frank, et al.
Publicado: (2025)
VLMs have Tunnel Vision: Evaluating Nonlocal Visual Reasoning in Leading VLMs
por: Berman, Shmuel, et al.
Publicado: (2025)
por: Berman, Shmuel, et al.
Publicado: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
por: Shah, Arya, et al.
Publicado: (2025)
por: Shah, Arya, et al.
Publicado: (2025)
Tone Matters: The Impact of Linguistic Tone on Hallucination in VLMs
por: Hong, Weihao, et al.
Publicado: (2026)
por: Hong, Weihao, et al.
Publicado: (2026)
Mixed Signals: Decoding VLMs' Reasoning and Underlying Bias in Vision-Language Conflict
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
por: Pezeshkpour, Pouya, et al.
Publicado: (2025)
DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs
por: Kukreja, Dikshant, et al.
Publicado: (2026)
por: Kukreja, Dikshant, et al.
Publicado: (2026)
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding
por: Clark, Christopher, et al.
Publicado: (2026)
por: Clark, Christopher, et al.
Publicado: (2026)
Review of Hallucination Understanding in Large Language and Vision Models
por: Ho, Zhengyi, et al.
Publicado: (2025)
por: Ho, Zhengyi, et al.
Publicado: (2025)
When Big Models Train Small Ones: Label-Free Model Parity Alignment for Efficient Visual Question Answering using Small VLMs
por: Penamakuri, Abhirama Subramanyam, et al.
Publicado: (2025)
por: Penamakuri, Abhirama Subramanyam, et al.
Publicado: (2025)
Towards Lossless Ultimate Vision Token Compression for VLMs
por: Zheng, Dehua, et al.
Publicado: (2025)
por: Zheng, Dehua, et al.
Publicado: (2025)
Revealing Multi-View Hallucination in Large Vision-Language Models
por: Park, Wooje, et al.
Publicado: (2026)
por: Park, Wooje, et al.
Publicado: (2026)
Can Vision-Language Models be a Good Guesser? Exploring VLMs for Times and Location Reasoning
por: Zhang, Gengyuan, et al.
Publicado: (2023)
por: Zhang, Gengyuan, et al.
Publicado: (2023)
Can Generalist Vision Language Models (VLMs) Rival Specialist Medical VLMs? Benchmarking and Strategic Insights
por: Zhong, Yuan, et al.
Publicado: (2025)
por: Zhong, Yuan, et al.
Publicado: (2025)
Enhancing Subsequent Video Retrieval via Vision-Language Models (VLMs)
por: Duan, Yicheng, et al.
Publicado: (2025)
por: Duan, Yicheng, et al.
Publicado: (2025)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
por: Luo, Kun, et al.
Publicado: (2026)
por: Luo, Kun, et al.
Publicado: (2026)
Your Vision-Language Model Can't Even Count to 20: Exposing the Failures of VLMs in Compositional Counting
por: Guo, Xuyang, et al.
Publicado: (2025)
por: Guo, Xuyang, et al.
Publicado: (2025)
NanoVLMs: How small can we go and still make coherent Vision Language Models?
por: Agarwalla, Mukund, et al.
Publicado: (2025)
por: Agarwalla, Mukund, et al.
Publicado: (2025)
Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models
por: Zhang, Chengsheng, et al.
Publicado: (2026)
por: Zhang, Chengsheng, et al.
Publicado: (2026)
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models
por: Liu, Yufang, et al.
Publicado: (2024)
por: Liu, Yufang, et al.
Publicado: (2024)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
por: Zhang, Yuanhong, et al.
Publicado: (2026)
por: Zhang, Yuanhong, et al.
Publicado: (2026)
Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models
por: Seth, Ashish, et al.
Publicado: (2024)
por: Seth, Ashish, et al.
Publicado: (2024)
Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
por: Huo, Fushuo, et al.
Publicado: (2024)
por: Huo, Fushuo, et al.
Publicado: (2024)
Unified Spatio-Temporal Token Scoring for Efficient Video VLMs
por: Zhang, Jianrui, et al.
Publicado: (2026)
por: Zhang, Jianrui, et al.
Publicado: (2026)
Mitigating Open-Vocabulary Caption Hallucinations
por: Ben-Kish, Assaf, et al.
Publicado: (2023)
por: Ben-Kish, Assaf, et al.
Publicado: (2023)
Multi-Object Hallucination in Vision-Language Models
por: Chen, Xuweiyi, et al.
Publicado: (2024)
por: Chen, Xuweiyi, et al.
Publicado: (2024)
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models
por: Yakun, Cui, et al.
Publicado: (2026)
por: Yakun, Cui, et al.
Publicado: (2026)
Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Models
por: Yan, Bei, et al.
Publicado: (2024)
por: Yan, Bei, et al.
Publicado: (2024)
Ejemplares similares
-
Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models
por: Shah, Arya, et al.
Publicado: (2026) -
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation
por: Shah, Arya, et al.
Publicado: (2026) -
Barriers in Integrating Medical Visual Question Answering into Radiology Workflows: A Scoping Review and Clinicians' Insights
por: Mishra, Deepali, et al.
Publicado: (2025) -
To See or To Please: Uncovering Visual Sycophancy and Split Beliefs in VLMs
por: Hong, Rui, et al.
Publicado: (2026) -
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
por: Li, Shuo, et al.
Publicado: (2024)