Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhagwatkar, Rishika, Nayak, Shravan, Bayat, Reza, Roger, Alexis, Kaplan, Daniel Z, Bashivan, Pouya, Rish, Irina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Adversarial Robustness of Discrete Image Tokenizers
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2026)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2026)
CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
Towards ethical multimodal systems
von: Roger, Alexis, et al.
Veröffentlicht: (2023)
von: Roger, Alexis, et al.
Veröffentlicht: (2023)
A Guide to Robust Generalization: The Impact of Architecture, Pre-training, and Optimization Strategy
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
Geometry of naturalistic object representations in recurrent neural network models of working memory
von: Lei, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Lei, Xiaoxuan, et al.
Veröffentlicht: (2024)
Building spatial world models from sparse transitional episodic memories
von: He, Zizhan, et al.
Veröffentlicht: (2025)
von: He, Zizhan, et al.
Veröffentlicht: (2025)
Image Tiling for High-Resolution Reasoning: Balancing Local Detail with Global Context
von: de Margerie, Anatole Jacquin, et al.
Veröffentlicht: (2025)
von: de Margerie, Anatole Jacquin, et al.
Veröffentlicht: (2025)
CHIRP: A Fine-Grained Benchmark for Open-Ended Response Evaluation in Vision-Language Models
von: Roger, Alexis, et al.
Veröffentlicht: (2025)
von: Roger, Alexis, et al.
Veröffentlicht: (2025)
Small Vocabularies, Big Gains: Pretraining and Tokenization in Time Series Models
von: Roger, Alexis, et al.
Veröffentlicht: (2025)
von: Roger, Alexis, et al.
Veröffentlicht: (2025)
LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series
von: Roger, Alexis, et al.
Veröffentlicht: (2026)
von: Roger, Alexis, et al.
Veröffentlicht: (2026)
Caption This, Reason That: VLMs Caught in the Middle
von: Weng, Zihan, et al.
Veröffentlicht: (2025)
von: Weng, Zihan, et al.
Veröffentlicht: (2025)
IWISDM: Assessing instruction following in multimodal models at scale
von: Lei, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Lei, Xiaoxuan, et al.
Veröffentlicht: (2024)
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
Influence Functions for Efficient Data Selection in Reasoning
von: Humane, Prateek, et al.
Veröffentlicht: (2025)
von: Humane, Prateek, et al.
Veröffentlicht: (2025)
From Proof to Program: Characterizing Tool-Induced Reasoning Hallucinations in Large Language Models
von: Bayat, Farima Fatahi, et al.
Veröffentlicht: (2025)
von: Bayat, Farima Fatahi, et al.
Veröffentlicht: (2025)
Enhancing Context Through Contrast
von: Ambilduke, Kshitij, et al.
Veröffentlicht: (2024)
von: Ambilduke, Kshitij, et al.
Veröffentlicht: (2024)
Akkumulierte Prekarität
von: Bayat, Reza
Veröffentlicht: (2025)
von: Bayat, Reza
Veröffentlicht: (2025)
Adversarial Robustness of Partitioned Quantum Classifiers
von: Kananian, Pouya, et al.
Veröffentlicht: (2025)
von: Kananian, Pouya, et al.
Veröffentlicht: (2025)
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
von: Rasul, Kashif, et al.
Veröffentlicht: (2023)
von: Rasul, Kashif, et al.
Veröffentlicht: (2023)
Closed-Loop Bidirectional Prompting for Adversarial Robustness of Vision Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
Adversarial Robustness in Distributed Quantum Machine Learning
von: Kananian, Pouya, et al.
Veröffentlicht: (2025)
von: Kananian, Pouya, et al.
Veröffentlicht: (2025)
Discovering Failure Modes in Vision-Language Models using RL
von: Jain, Kanishk, et al.
Veröffentlicht: (2026)
von: Jain, Kanishk, et al.
Veröffentlicht: (2026)
Random Initialization Can't Catch Up: The Advantage of Language Model Transfer for Time Series Forecasting
von: Riachi, Roland, et al.
Veröffentlicht: (2025)
von: Riachi, Roland, et al.
Veröffentlicht: (2025)
Enhancing Efficiency in Vision Transformer Networks: Design Techniques and Insights
von: Heidari, Moein, et al.
Veröffentlicht: (2024)
von: Heidari, Moein, et al.
Veröffentlicht: (2024)
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
von: Pasand, Ali Saheb, et al.
Veröffentlicht: (2026)
von: Pasand, Ali Saheb, et al.
Veröffentlicht: (2026)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
von: Legate, Gwen, et al.
Veröffentlicht: (2025)
von: Legate, Gwen, et al.
Veröffentlicht: (2025)
VFA: Vision Frequency Analysis of Foundation Models and Human
von: Darvishi-Bayazi, Mohammad-Javad, et al.
Veröffentlicht: (2024)
von: Darvishi-Bayazi, Mohammad-Javad, et al.
Veröffentlicht: (2024)
Adversarial Prompt Tuning for Vision-Language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
Adversarial Prompt Distillation for Vision-Language Models
von: Luo, Lin, et al.
Veröffentlicht: (2024)
von: Luo, Lin, et al.
Veröffentlicht: (2024)
NAP-Tuning: Neural Augmented Prompt Tuning for Adversarially Robust Vision-Language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
Evolution-based Region Adversarial Prompt Learning for Robustness Enhancement in Vision-Language Models
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
Dropout Prompt Learning: Towards Robust and Adaptive Vision-Language Models
von: Chen, Biao, et al.
Veröffentlicht: (2025)
von: Chen, Biao, et al.
Veröffentlicht: (2025)
Towards LLMs Robustness to Changes in Prompt Format Styles
von: Ngweta, Lilian, et al.
Veröffentlicht: (2025)
von: Ngweta, Lilian, et al.
Veröffentlicht: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
von: Li, Lin, et al.
Veröffentlicht: (2024)
von: Li, Lin, et al.
Veröffentlicht: (2024)
Benchmarking Vision Language Models for Cultural Understanding
von: Nayak, Shravan, et al.
Veröffentlicht: (2024)
von: Nayak, Shravan, et al.
Veröffentlicht: (2024)
R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning
von: Sheng, Lijun, et al.
Veröffentlicht: (2025)
von: Sheng, Lijun, et al.
Veröffentlicht: (2025)
SC-FDMA as a Delay-Doppler Domain Modulation Technique
von: Farhang, Arman, et al.
Veröffentlicht: (2024)
von: Farhang, Arman, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
On the Adversarial Robustness of Discrete Image Tokenizers
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2026) -
CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025) -
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025) -
Towards ethical multimodal systems
von: Roger, Alexis, et al.
Veröffentlicht: (2023) -
A Guide to Robust Generalization: The Impact of Architecture, Pre-training, and Optimization Strategy
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)