SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Peri, Raghuveer, Jayanthi, Sai Muralidhar, Ronanki, Srikanth, Bhatia, Anshu, Mundnich, Karel, Dingliwal, Saket, Das, Nilaksh, Hou, Zejiang, Huybrechts, Goeric, Vishnubhotla, Srikanth, Garcia-Romero, Daniel, Srinivasan, Sundararajan, Han, Kyu J, Kirchhoff, Katrin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark
by: Huybrechts, Goeric, et al.
Published: (2025)
by: Huybrechts, Goeric, et al.
Published: (2025)
Zero-resource Speech Translation and Recognition with LLMs
by: Mundnich, Karel, et al.
Published: (2024)
by: Mundnich, Karel, et al.
Published: (2024)
SpeechVerse: A Large-scale Generalizable Audio Language Model
by: Das, Nilaksh, et al.
Published: (2024)
by: Das, Nilaksh, et al.
Published: (2024)
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR
by: Huybrechts, Goeric, et al.
Published: (2023)
by: Huybrechts, Goeric, et al.
Published: (2023)
Sequential Editing for Lifelong Training of Speech Recognition Models
by: Kulshreshtha, Devang, et al.
Published: (2024)
by: Kulshreshtha, Devang, et al.
Published: (2024)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
by: Min, Do June, et al.
Published: (2024)
by: Min, Do June, et al.
Published: (2024)
Robust Multimodal Safety via Conditional Decoding
by: Kumar, Anurag, et al.
Published: (2026)
by: Kumar, Anurag, et al.
Published: (2026)
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents
by: Choi, Daewon, et al.
Published: (2026)
by: Choi, Daewon, et al.
Published: (2026)
ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling
by: Song, Woomin, et al.
Published: (2026)
by: Song, Woomin, et al.
Published: (2026)
Accelerated Test-Time Scaling with Model-Free Speculative Sampling
by: Song, Woomin, et al.
Published: (2025)
by: Song, Woomin, et al.
Published: (2025)
Compress, Gather, and Recompute: REFORMing Long-Context Processing in Transformers
by: Song, Woomin, et al.
Published: (2025)
by: Song, Woomin, et al.
Published: (2025)
Think Clearly: Improving Reasoning via Redundant Token Pruning
by: Choi, Daewon, et al.
Published: (2025)
by: Choi, Daewon, et al.
Published: (2025)
CriSPO: Multi-Aspect Critique-Suggestion-guided Automatic Prompt Optimization for Text Generation
by: He, Han, et al.
Published: (2024)
by: He, Han, et al.
Published: (2024)
Low-Degree Testing Over Grids
by: Amireddy, Prashanth, et al.
Published: (2023)
by: Amireddy, Prashanth, et al.
Published: (2023)
Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
by: Jeoung, Sullam, et al.
Published: (2024)
by: Jeoung, Sullam, et al.
Published: (2024)
Reinforcement Learning as a Parsimonious Alternative to Prediction Cascades: A Case Study on Image Segmentation
by: Srikishan, Bharat, et al.
Published: (2024)
by: Srikishan, Bharat, et al.
Published: (2024)
Algorithmic Guarantees for Distilling Supervised and Offline RL Datasets
by: Gupta, Aaryan, et al.
Published: (2025)
by: Gupta, Aaryan, et al.
Published: (2025)
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
by: Singh, Sagalpreet, et al.
Published: (2025)
by: Singh, Sagalpreet, et al.
Published: (2025)
Salient Information Prompting to Steer Content in Prompt-based Abstractive Summarization
by: Xu, Lei, et al.
Published: (2024)
by: Xu, Lei, et al.
Published: (2024)
Psychological Centrality
by: Srikanth, Janani
Published: (2025)
by: Srikanth, Janani
Published: (2025)
AI Knowledge Recommender as a Digital Learning Companion
by: Srikanth, Madabhushi
Published: (2026)
by: Srikanth, Madabhushi
Published: (2026)
An Equivalent form of Twin Prime Conjecture connected with a sequence of arithmetic progressions
by: Cherukupally, Srikanth
Published: (2026)
by: Cherukupally, Srikanth
Published: (2026)
Reinforcement Learning Based Escape Route Generation in Low Visibility Environments
by: Srikanth, Hari
Published: (2024)
by: Srikanth, Hari
Published: (2024)
ClarifAI: Enhancing AI Interpretability and Transparency through Case-Based Reasoning and Ontology-Driven Approach for Improved Decision-Making
by: Vemula, Srikanth
Published: (2025)
by: Vemula, Srikanth
Published: (2025)
Linear Function Approximation as a Computationally Efficient Method to Solve Classical Reinforcement Learning Challenges
by: Srikanth, Hari
Published: (2024)
by: Srikanth, Hari
Published: (2024)
On the size of $\{a: 1\leq a<n, n|a^2-1, a|n^2-1\}$ for number $n$
by: Cherukupally, Srikanth
Published: (2026)
by: Cherukupally, Srikanth
Published: (2026)
Systematic Risk and Accounting Determinants: An Empirical Assessment in the Indian Stock Market
by: Srikanth Parthasarathy
Published: (2019)
by: Srikanth Parthasarathy
Published: (2019)
The Algebraic Cost of a Boolean Sum
by: Orzel, Ian, et al.
Published: (2025)
by: Orzel, Ian, et al.
Published: (2025)
One‐Sided Schmitt Trigger‐Based 11T Carbon Nanotube Field Effect Transistor—Based Static Random‐Access Memory Cell for Modern IoT Embedded Devices at 32 nm Technology
by: Srinivasan Jayanthi, et al.
Published: (2025)
by: Srinivasan Jayanthi, et al.
Published: (2025)
Consumer-Driven Contract Testing: A Foundation for Reliable, High-Velocity Microservices Delivery
by: Srikanth Chakravarthy Vankayala
Published: (2022)
by: Srikanth Chakravarthy Vankayala
Published: (2022)
On a cyclic structure of generators modulo primes
by: Ch, Srikanth, et al.
Published: (2026)
by: Ch, Srikanth, et al.
Published: (2026)
The Taguchi method for optimizing nonlinear pulse propagation in optical fibers
by: Adity, et al.
Published: (2026)
by: Adity, et al.
Published: (2026)
Lepton flavor violation in the Majorana and Dirac scotogenic models
by: Hundi, Raghavendra Srikanth
Published: (2025)
by: Hundi, Raghavendra Srikanth
Published: (2025)
Commutative algebra inspired by modularity lifting
by: Iyengar, Srikanth B.
Published: (2025)
by: Iyengar, Srikanth B.
Published: (2025)
FRACTAL: Fine-Grained Scoring from Aggregate Text Labels
by: Makhija, Yukti, et al.
Published: (2024)
by: Makhija, Yukti, et al.
Published: (2024)
LLP-Bench: A Large Scale Tabular Benchmark for Learning from Label Proportions
by: Brahmbhatt, Anand, et al.
Published: (2023)
by: Brahmbhatt, Anand, et al.
Published: (2023)
Aggregating Data for Optimal and Private Learning
by: Agarwal, Sushant, et al.
Published: (2024)
by: Agarwal, Sushant, et al.
Published: (2024)
Improving Lip-synchrony in Direct Audio-Visual Speech-to-Speech Translation
by: Goncalves, Lucas, et al.
Published: (2024)
by: Goncalves, Lucas, et al.
Published: (2024)
A Near-Optimal Polynomial Distance Lemma Over Boolean Slices
by: Amireddy, Prashanth, et al.
Published: (2025)
by: Amireddy, Prashanth, et al.
Published: (2025)
New Bounds for the Ideal Proof System in Positive Characteristic
by: Behera, Amik Raj, et al.
Published: (2025)
by: Behera, Amik Raj, et al.
Published: (2025)
Similar Items
-
Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark
by: Huybrechts, Goeric, et al.
Published: (2025) -
Zero-resource Speech Translation and Recognition with LLMs
by: Mundnich, Karel, et al.
Published: (2024) -
SpeechVerse: A Large-scale Generalizable Audio Language Model
by: Das, Nilaksh, et al.
Published: (2024) -
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR
by: Huybrechts, Goeric, et al.
Published: (2023) -
Sequential Editing for Lifelong Training of Speech Recognition Models
by: Kulshreshtha, Devang, et al.
Published: (2024)