Visible Yet Unreadable: A Systematic Blind Spot of Vision Language Models Across Writing Systems
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Jie, Xu, Ting, Deng, Gelei, Hu, Runyi, Qiu, Han, Zhang, Tianwei, Guo, Qing, Tsang, Ivor |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Unifying Watermarking via Dimension-Aware Mapping
par: Meng, Jiale, et autres
Publié: (2026)
par: Meng, Jiale, et autres
Publié: (2026)
SceneTAP: Scene-Coherent Typographic Adversarial Planner against Vision-Language Models in Real-World Environments
par: Cao, Yue, et autres
Publié: (2024)
par: Cao, Yue, et autres
Publié: (2024)
SuperMark: Robust and Training-free Image Watermarking via Diffusion-based Super-Resolution
par: Hu, Runyi, et autres
Publié: (2024)
par: Hu, Runyi, et autres
Publié: (2024)
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
par: Hu, Runyi, et autres
Publié: (2025)
par: Hu, Runyi, et autres
Publié: (2025)
When Audio and Text Disagree: Revealing Text Bias in Large Audio-Language Models
par: Wang, Cheng, et autres
Publié: (2025)
par: Wang, Cheng, et autres
Publié: (2025)
When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models
par: Ou, Haoran, et autres
Publié: (2025)
par: Ou, Haoran, et autres
Publié: (2025)
Beyond Retrieval: Improving Evidence Quality for LLM-based Multimodal Fact-Checking
par: Ou, Haoran, et autres
Publié: (2025)
par: Ou, Haoran, et autres
Publié: (2025)
IRCopilot: Automated Incident Response with Large Language Models
par: Lin, Xihuan, et autres
Publié: (2025)
par: Lin, Xihuan, et autres
Publié: (2025)
Robust-Wide: Robust Watermarking against Instruction-driven Image Editing
par: Hu, Runyi, et autres
Publié: (2024)
par: Hu, Runyi, et autres
Publié: (2024)
Mask Image Watermarking
par: Hu, Runyi, et autres
Publié: (2025)
par: Hu, Runyi, et autres
Publié: (2025)
Boosting Transferability in Vision-Language Attacks via Diversification along the Intersection Region of Adversarial Trajectory
par: Gao, Sensen, et autres
Publié: (2024)
par: Gao, Sensen, et autres
Publié: (2024)
What Makes a Good LLM Agent for Real-world Penetration Testing?
par: Deng, Gelei, et autres
Publié: (2026)
par: Deng, Gelei, et autres
Publié: (2026)
MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots
par: Deng, Gelei, et autres
Publié: (2023)
par: Deng, Gelei, et autres
Publié: (2023)
Readable Twins of Unreadable Models
par: Pancerz, Krzysztof, et autres
Publié: (2025)
par: Pancerz, Krzysztof, et autres
Publié: (2025)
Oedipus: LLM-enchanced Reasoning CAPTCHA Solver
par: Deng, Gelei, et autres
Publié: (2024)
par: Deng, Gelei, et autres
Publié: (2024)
Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models
par: Cui, Chenhang, et autres
Publié: (2024)
par: Cui, Chenhang, et autres
Publié: (2024)
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
par: Ou, Haoran, et autres
Publié: (2026)
par: Ou, Haoran, et autres
Publié: (2026)
Advancing Analytic Class-Incremental Learning through Vision-Language Calibration
par: Zhao, Binyu, et autres
Publié: (2026)
par: Zhao, Binyu, et autres
Publié: (2026)
Image-Based Geolocation Using Large Vision-Language Models
par: Liu, Yi, et autres
Publié: (2024)
par: Liu, Yi, et autres
Publié: (2024)
TRACE: Structure-Aware Character Encoding for Robust and Generalizable Document Watermarking
par: Meng, Jiale, et autres
Publié: (2026)
par: Meng, Jiale, et autres
Publié: (2026)
Collaborative Group-Aware Hashing for Fast Recommender Systems
par: Zhang, Yan, et autres
Publié: (2025)
par: Zhang, Yan, et autres
Publié: (2025)
Enhancing Model Defense Against Jailbreaks with Proactive Safety Reasoning
par: Yang, Xianglin, et autres
Publié: (2025)
par: Yang, Xianglin, et autres
Publié: (2025)
Turning Bias into Bugs: Bandit-Guided Style Manipulation Attacks on LLM Judges
par: Yang, Xianglin, et autres
Publié: (2026)
par: Yang, Xianglin, et autres
Publié: (2026)
HC$^2$L: Hybrid and Cooperative Contrastive Learning for Cross-lingual Spoken Language Understanding
par: Xing, Bowen, et autres
Publié: (2024)
par: Xing, Bowen, et autres
Publié: (2024)
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
par: Zhang, Yechao, et autres
Publié: (2026)
par: Zhang, Yechao, et autres
Publié: (2026)
SPHERE: Unveiling Spatial Blind Spots in Vision-Language Models Through Hierarchical Evaluation
par: Zhang, Wenyu, et autres
Publié: (2024)
par: Zhang, Wenyu, et autres
Publié: (2024)
Vision Language Models Are Not (Yet) Spelling Correctors
par: Liang, Junhong, et autres
Publié: (2025)
par: Liang, Junhong, et autres
Publié: (2025)
GenderCARE: A Comprehensive Framework for Assessing and Reducing Gender Bias in Large Language Models
par: Tang, Kunsheng, et autres
Publié: (2024)
par: Tang, Kunsheng, et autres
Publié: (2024)
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
par: Yang, Ruozhao, et autres
Publié: (2025)
par: Yang, Ruozhao, et autres
Publié: (2025)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
par: Deng, Gelei, et autres
Publié: (2024)
par: Deng, Gelei, et autres
Publié: (2024)
AutoEG: Exploiting Known Third-Party Vulnerabilities in Black-Box Web Applications
par: Yang, Ruozhao, et autres
Publié: (2026)
par: Yang, Ruozhao, et autres
Publié: (2026)
Making Interview Multilingualism Visible: Transnational Chinese Language Teacher Identity Construction
par: Chengwen Yuan, et autres
Publié: (2026)
par: Chengwen Yuan, et autres
Publié: (2026)
Unveiling AI's Blind Spots: An Oracle for In-Domain, Out-of-Domain, and Adversarial Errors
par: Han, Shuangpeng, et autres
Publié: (2024)
par: Han, Shuangpeng, et autres
Publié: (2024)
Semantic-Aligned Adversarial Evolution Triangle for High-Transferability Vision-Language Attack
par: Jia, Xiaojun, et autres
Publié: (2024)
par: Jia, Xiaojun, et autres
Publié: (2024)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
par: Khan, Mohammed Safi Ur Rahman, et autres
Publié: (2026)
par: Khan, Mohammed Safi Ur Rahman, et autres
Publié: (2026)
Blind Spots
par: Drucker, Johanna
Publié: (2009)
par: Drucker, Johanna
Publié: (2009)
LAFR: Efficient Diffusion-based Blind Face Restoration via Latent Codebook Alignment Adapter
par: Li, Runyi, et autres
Publié: (2025)
par: Li, Runyi, et autres
Publié: (2025)
Illuminating Blind Spots of Language Models with Targeted Agent-in-the-Loop Synthetic Data
par: Lippmann, Philip, et autres
Publié: (2024)
par: Lippmann, Philip, et autres
Publié: (2024)
ColorBlindnessEval: Can Vision-Language Models Pass Color Blindness Tests?
par: Ling, Zijian, et autres
Publié: (2025)
par: Ling, Zijian, et autres
Publié: (2025)
Transductive Reward Inference on Graph
par: Qu, Bohao, et autres
Publié: (2024)
par: Qu, Bohao, et autres
Publié: (2024)
Documents similaires
-
Unifying Watermarking via Dimension-Aware Mapping
par: Meng, Jiale, et autres
Publié: (2026) -
SceneTAP: Scene-Coherent Typographic Adversarial Planner against Vision-Language Models in Real-World Environments
par: Cao, Yue, et autres
Publié: (2024) -
SuperMark: Robust and Training-free Image Watermarking via Diffusion-based Super-Resolution
par: Hu, Runyi, et autres
Publié: (2024) -
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
par: Hu, Runyi, et autres
Publié: (2025) -
When Audio and Text Disagree: Revealing Text Bias in Large Audio-Language Models
par: Wang, Cheng, et autres
Publié: (2025)