High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | He, Mengqi, Tian, Xinyu, Shen, Xin, Ni, Jinhong, Zou, Shu, Yang, Zhaoyuan, Zhang, Jing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Making High-Level AI Design Decisions Explicit Using a Binary Stream System-Designation Approach
di: Mossbridge, Julia
Pubblicazione: (2024)
di: Mossbridge, Julia
Pubblicazione: (2024)
A Review of Pseudo-Labeling for Computer Vision
di: Kage, Patrick, et al.
Pubblicazione: (2024)
di: Kage, Patrick, et al.
Pubblicazione: (2024)
Synthetic Photography Detection: A Visual Guidance for Identifying Synthetic Images Created by AI
di: Mathys, Melanie, et al.
Pubblicazione: (2024)
di: Mathys, Melanie, et al.
Pubblicazione: (2024)
CerberusDet: Unified Multi-Dataset Object Detection
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
Threats and Opportunities in AI-generated Images for Armed Forces
di: Meier, Raphael
Pubblicazione: (2025)
di: Meier, Raphael
Pubblicazione: (2025)
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
di: Ren, Yumeng, et al.
Pubblicazione: (2025)
di: Ren, Yumeng, et al.
Pubblicazione: (2025)
Super-additive Cooperation in Language Model Agents
di: Tonini, Filippo, et al.
Pubblicazione: (2025)
di: Tonini, Filippo, et al.
Pubblicazione: (2025)
Eye-gaze Guided Multi-modal Alignment for Medical Representation Learning
di: Ma, Chong, et al.
Pubblicazione: (2024)
di: Ma, Chong, et al.
Pubblicazione: (2024)
Synthetic Image Generation in Cyber Influence Operations: An Emergent Threat?
di: Mathys, Melanie, et al.
Pubblicazione: (2024)
di: Mathys, Melanie, et al.
Pubblicazione: (2024)
The Continuity Layer: Why Intelligence Needs an Architecture for What It Carries Forward
di: Tanguturi, Samuel Sameer
Pubblicazione: (2026)
di: Tanguturi, Samuel Sameer
Pubblicazione: (2026)
Explaining Any ML Model? -- On Goals and Capabilities of XAI
di: Renftle, Moritz, et al.
Pubblicazione: (2022)
di: Renftle, Moritz, et al.
Pubblicazione: (2022)
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images
di: Slika, Bouthaina, et al.
Pubblicazione: (2023)
di: Slika, Bouthaina, et al.
Pubblicazione: (2023)
Context-Aware Full Body Anonymization using Text-to-Image Diffusion Models
di: Zwick, Pascal, et al.
Pubblicazione: (2024)
di: Zwick, Pascal, et al.
Pubblicazione: (2024)
Sparse vs Contiguous Adversarial Pixel Perturbations in Multimodal Models: An Empirical Analysis
di: Botocan, Cristian-Alexandru, et al.
Pubblicazione: (2024)
di: Botocan, Cristian-Alexandru, et al.
Pubblicazione: (2024)
Measuring proximity to standard planes during fetal brain ultrasound scanning
di: Di Vece, Chiara, et al.
Pubblicazione: (2024)
di: Di Vece, Chiara, et al.
Pubblicazione: (2024)
Quaternion Convolutional Neural Networks: Current Advances and Future Directions
di: Altamirano-Gomez, Gerardo, et al.
Pubblicazione: (2023)
di: Altamirano-Gomez, Gerardo, et al.
Pubblicazione: (2023)
Beyond Specialization: Assessing the Capabilities of MLLMs in Age and Gender Estimation
di: Kuprashevich, Maksim, et al.
Pubblicazione: (2024)
di: Kuprashevich, Maksim, et al.
Pubblicazione: (2024)
Position: Tensor Networks are a Valuable Asset for Green AI
di: Memmel, Eva, et al.
Pubblicazione: (2022)
di: Memmel, Eva, et al.
Pubblicazione: (2022)
What Does 'Human-Centred AI' Mean?
di: Guest, Olivia
Pubblicazione: (2025)
di: Guest, Olivia
Pubblicazione: (2025)
What's my role? Modelling responsibility for AI-based safety-critical systems
di: Ryan, Philippa, et al.
Pubblicazione: (2023)
di: Ryan, Philippa, et al.
Pubblicazione: (2023)
Towards Friendly AI: A Comprehensive Review and New Perspectives on Human-AI Alignment
di: Sun, Qiyang, et al.
Pubblicazione: (2024)
di: Sun, Qiyang, et al.
Pubblicazione: (2024)
Provocations from the Humanities for Generative AI Research
di: Klein, Lauren, et al.
Pubblicazione: (2025)
di: Klein, Lauren, et al.
Pubblicazione: (2025)
Data Feminism for AI
di: Klein, Lauren, et al.
Pubblicazione: (2024)
di: Klein, Lauren, et al.
Pubblicazione: (2024)
Visual Orientalism in the AI Era: From West-East Binaries to English-Language Centrism
di: Zhao, Zhilong, et al.
Pubblicazione: (2025)
di: Zhao, Zhilong, et al.
Pubblicazione: (2025)
FUTURE-AI: International consensus guideline for trustworthy and deployable artificial intelligence in healthcare
di: Lekadir, Karim, et al.
Pubblicazione: (2023)
di: Lekadir, Karim, et al.
Pubblicazione: (2023)
Pictures Of MIDI: Controlled Music Generation via Graphical Prompts for Image-Based Diffusion Inpainting
di: Hawley, Scott H.
Pubblicazione: (2024)
di: Hawley, Scott H.
Pubblicazione: (2024)
The Landscape of Generative AI in Information Systems: A Synthesis of Secondary Reviews and Research Agendas
di: Jarzębowicz, Aleksander, et al.
Pubblicazione: (2026)
di: Jarzębowicz, Aleksander, et al.
Pubblicazione: (2026)
SIDEs: Separating Idealization from Deceptive Explanations in xAI
di: Sullivan, Emily
Pubblicazione: (2024)
di: Sullivan, Emily
Pubblicazione: (2024)
Event-based Solutions for Human-centered Applications: A Comprehensive Review
di: Adra, Mira, et al.
Pubblicazione: (2025)
di: Adra, Mira, et al.
Pubblicazione: (2025)
GACL: Graph Attention Collaborative Learning for Temporal QoS Prediction
di: Hu, Shengxiang, et al.
Pubblicazione: (2024)
di: Hu, Shengxiang, et al.
Pubblicazione: (2024)
GuardSec: A Multi-Modal Web Platform for Real-Time Digital Fraud Detection, Entity Verification, and Connection Security Analysis in the African Context
di: Bansimba, Gilda Rech, et al.
Pubblicazione: (2026)
di: Bansimba, Gilda Rech, et al.
Pubblicazione: (2026)
The Power of Absence: Thinking with Archival Theory in Algorithmic Design
di: Sherman, Jihan, et al.
Pubblicazione: (2024)
di: Sherman, Jihan, et al.
Pubblicazione: (2024)
Cyber Humanism in Education: Reclaiming Agency through AI and Learning Sciences
di: Adorni, Giovanni
Pubblicazione: (2025)
di: Adorni, Giovanni
Pubblicazione: (2025)
StereoCrafter: Diffusion-based Generation of Long and High-fidelity Stereoscopic 3D from Monocular Videos
di: Zhao, Sijie, et al.
Pubblicazione: (2024)
di: Zhao, Sijie, et al.
Pubblicazione: (2024)
COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coherence Tomography Image Classification
di: Rivera, Mariano, et al.
Pubblicazione: (2025)
di: Rivera, Mariano, et al.
Pubblicazione: (2025)
Antagonistic AI
di: Cai, Alice, et al.
Pubblicazione: (2024)
di: Cai, Alice, et al.
Pubblicazione: (2024)
Proto-FG3D: Prototype-based Interpretable Fine-Grained 3D Shape Classification
di: Ma, Shuxian, et al.
Pubblicazione: (2025)
di: Ma, Shuxian, et al.
Pubblicazione: (2025)
Lookism: The overlooked bias in computer vision
di: Gulati, Aditya, et al.
Pubblicazione: (2024)
di: Gulati, Aditya, et al.
Pubblicazione: (2024)
AzSLD: Azerbaijani Sign Language Dataset for Fingerspelling, Word, and Sentence Translation with Baseline Software
di: Alishzade, Nigar, et al.
Pubblicazione: (2024)
di: Alishzade, Nigar, et al.
Pubblicazione: (2024)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
di: Nguyen, Tan, et al.
Pubblicazione: (2024)
di: Nguyen, Tan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Making High-Level AI Design Decisions Explicit Using a Binary Stream System-Designation Approach
di: Mossbridge, Julia
Pubblicazione: (2024) -
A Review of Pseudo-Labeling for Computer Vision
di: Kage, Patrick, et al.
Pubblicazione: (2024) -
Synthetic Photography Detection: A Visual Guidance for Identifying Synthetic Images Created by AI
di: Mathys, Melanie, et al.
Pubblicazione: (2024) -
CerberusDet: Unified Multi-Dataset Object Detection
di: Tolstykh, Irina, et al.
Pubblicazione: (2024) -
Threats and Opportunities in AI-generated Images for Armed Forces
di: Meier, Raphael
Pubblicazione: (2025)