Understanding and Rectifying Safety Perception Distortion in VLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zou, Xiaohan, Kang, Jian, Kesidis, George, Lin, Lu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TechING: Towards Real World Technical Image Understanding via VLMs
von: Nadeem, Tafazzul, et al.
Veröffentlicht: (2026)
von: Nadeem, Tafazzul, et al.
Veröffentlicht: (2026)
VLMGuard-R1: Proactive Safety Alignment for VLMs via Reasoning-Driven Prompt Optimization
von: Chen, Menglan, et al.
Veröffentlicht: (2025)
von: Chen, Menglan, et al.
Veröffentlicht: (2025)
ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness
von: Liang, Yijun, et al.
Veröffentlicht: (2025)
von: Liang, Yijun, et al.
Veröffentlicht: (2025)
On the Perception Bottleneck of VLMs for Chart Understanding
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
von: Li, Shuo, et al.
Veröffentlicht: (2024)
von: Li, Shuo, et al.
Veröffentlicht: (2024)
Leveraging NTPs for Efficient Hallucination Detection in VLMs
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
Sensitivity of Generative VLMs to Semantically and Lexically Altered Prompts
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
Manager: Aggregating Insights from Unimodal Experts in Two-Tower VLMs and MLLMs
von: Xu, Xiao, et al.
Veröffentlicht: (2025)
von: Xu, Xiao, et al.
Veröffentlicht: (2025)
Bidirectional Long-Range Parser for Sequential Data Understanding
von: Leotescu, George, et al.
Veröffentlicht: (2024)
von: Leotescu, George, et al.
Veröffentlicht: (2024)
Robustness of Structured Data Extraction from Perspectively Distorted Documents
von: Nakada, Hyakka, et al.
Veröffentlicht: (2025)
von: Nakada, Hyakka, et al.
Veröffentlicht: (2025)
Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?
von: Park, Simon, et al.
Veröffentlicht: (2025)
von: Park, Simon, et al.
Veröffentlicht: (2025)
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
PerLA: Perceptive 3D Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
Can World Models Benefit VLMs for World Dynamics?
von: Zhang, Kevin, et al.
Veröffentlicht: (2025)
von: Zhang, Kevin, et al.
Veröffentlicht: (2025)
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
Toward Inherently Robust VLMs Against Visual Perception Attacks
von: MohajerAnsari, Pedram, et al.
Veröffentlicht: (2025)
von: MohajerAnsari, Pedram, et al.
Veröffentlicht: (2025)
Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs
von: Pan, Zhiyu, et al.
Veröffentlicht: (2026)
von: Pan, Zhiyu, et al.
Veröffentlicht: (2026)
VLSU: Mapping the Limits of Joint Multimodal Understanding for AI Safety
von: Palaskar, Shruti, et al.
Veröffentlicht: (2025)
von: Palaskar, Shruti, et al.
Veröffentlicht: (2025)
Hierarchical Safety Realignment: Lightweight Restoration of Safety in Pruned Large Vision-Language Models
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs
von: Daxberger, Erik, et al.
Veröffentlicht: (2025)
von: Daxberger, Erik, et al.
Veröffentlicht: (2025)
Fine-tuning MLLMs Without Forgetting Is Easier Than You Think
von: Li, He, et al.
Veröffentlicht: (2026)
von: Li, He, et al.
Veröffentlicht: (2026)
Unraveling the Truth: Do VLMs really Understand Charts? A Deep Dive into Consistency and Robustness
von: Mukhopadhyay, Srija, et al.
Veröffentlicht: (2024)
von: Mukhopadhyay, Srija, et al.
Veröffentlicht: (2024)
Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models
von: Karamcheti, Siddharth, et al.
Veröffentlicht: (2024)
von: Karamcheti, Siddharth, et al.
Veröffentlicht: (2024)
CIVET: Systematic Evaluation of Understanding in VLMs
von: Rizzoli, Massimo, et al.
Veröffentlicht: (2025)
von: Rizzoli, Massimo, et al.
Veröffentlicht: (2025)
Temporal Preference Optimization for Long-Form Video Understanding
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Towards Efficient Vision-Language Tuning: More Information Density, More Generalizability
von: Hao, Tianxiang, et al.
Veröffentlicht: (2023)
von: Hao, Tianxiang, et al.
Veröffentlicht: (2023)
Evaluating and Advancing Multimodal Large Language Models in Perception Ability Lens
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
Improving Language Understanding from Screenshots
von: Gao, Tianyu, et al.
Veröffentlicht: (2024)
von: Gao, Tianyu, et al.
Veröffentlicht: (2024)
ETA: Evaluating Then Aligning Safety of Vision Language Models at Inference Time
von: Ding, Yi, et al.
Veröffentlicht: (2024)
von: Ding, Yi, et al.
Veröffentlicht: (2024)
LongPerceptualThoughts: Distilling System-2 Reasoning for System-1 Perception
von: Liao, Yuan-Hong, et al.
Veröffentlicht: (2025)
von: Liao, Yuan-Hong, et al.
Veröffentlicht: (2025)
VisMin: Visual Minimal-Change Understanding
von: Awal, Rabiul, et al.
Veröffentlicht: (2024)
von: Awal, Rabiul, et al.
Veröffentlicht: (2024)
Quantize-then-Rectify: Efficient VQ-VAE Training
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry
von: Wu, Changti, et al.
Veröffentlicht: (2025)
von: Wu, Changti, et al.
Veröffentlicht: (2025)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
DocAtlas: Multilingual Document Understanding Across 80+ Languages
von: Heakl, Ahmed, et al.
Veröffentlicht: (2026)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2026)
Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding
von: Liu, Xiangyue, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyue, et al.
Veröffentlicht: (2026)
Do VLMs Have a Moral Backbone? A Study on the Fragile Morality of Vision-Language Models
von: Liu, Zhining, et al.
Veröffentlicht: (2026)
von: Liu, Zhining, et al.
Veröffentlicht: (2026)
Video DataFlywheel: Resolving the Impossible Data Trinity in Video-Language Understanding
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Optimizing GPT for Video Understanding: Zero-Shot Performance and Prompt Engineering
von: Beliaev, Mark, et al.
Veröffentlicht: (2025)
von: Beliaev, Mark, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TechING: Towards Real World Technical Image Understanding via VLMs
von: Nadeem, Tafazzul, et al.
Veröffentlicht: (2026) -
VLMGuard-R1: Proactive Safety Alignment for VLMs via Reasoning-Driven Prompt Optimization
von: Chen, Menglan, et al.
Veröffentlicht: (2025) -
ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness
von: Liang, Yijun, et al.
Veröffentlicht: (2025) -
On the Perception Bottleneck of VLMs for Chart Understanding
von: Liu, Junteng, et al.
Veröffentlicht: (2025) -
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
von: Li, Shuo, et al.
Veröffentlicht: (2024)