Universal Adversarial Perturbations for Vision-Language Pre-trained Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Peng-Fei, Huang, Zi, Bai, Guangdong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Surrealistic-like Image Generation with Vision-Language Models
by: Ayten, Elif, et al.
Published: (2024)
by: Ayten, Elif, et al.
Published: (2024)
VACoDe: Visual Augmented Contrastive Decoding
by: Kim, Sihyeon, et al.
Published: (2024)
by: Kim, Sihyeon, et al.
Published: (2024)
Effective and Robust Adversarial Training against Data and Label Corruptions
by: Zhang, Peng-Fei, et al.
Published: (2024)
by: Zhang, Peng-Fei, et al.
Published: (2024)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
by: Zhang, Junbin, et al.
Published: (2022)
by: Zhang, Junbin, et al.
Published: (2022)
From Language Models to Practical Self-Improving Computer Agents
by: Sheng, Alex
Published: (2024)
by: Sheng, Alex
Published: (2024)
MultiHateClip: A Multilingual Benchmark Dataset for Hateful Video Detection on YouTube and Bilibili
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
ECG-FM: An Open Electrocardiogram Foundation Model
by: McKeen, Kaden, et al.
Published: (2024)
by: McKeen, Kaden, et al.
Published: (2024)
Understanding Input Selectivity in Mamba: Impact on Approximation Power, Memorization, and Associative Recall Capacity
by: Huang, Ningyuan, et al.
Published: (2025)
by: Huang, Ningyuan, et al.
Published: (2025)
Enhancing PyKEEN with Multiple Negative Sampling Solutions for Knowledge Graph Embedding Models
by: d'Amato, Claudia, et al.
Published: (2025)
by: d'Amato, Claudia, et al.
Published: (2025)
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
by: Zhao, Zelin, et al.
Published: (2023)
by: Zhao, Zelin, et al.
Published: (2023)
Hilbert-Geo: Solving Solid Geometric Problems by Neural-Symbolic Reasoning
by: Xu, Ruoran, et al.
Published: (2026)
by: Xu, Ruoran, et al.
Published: (2026)
Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors
by: Zhang, Junbin, et al.
Published: (2026)
by: Zhang, Junbin, et al.
Published: (2026)
Feature Relevancy, Necessity and Usefulness: Complexity and Algorithms
by: Capdevielle, Tomás, et al.
Published: (2025)
by: Capdevielle, Tomás, et al.
Published: (2025)
A Theoretical Framework for Adaptive Utility-Weighted Benchmarking
by: Waggoner, Philip
Published: (2026)
by: Waggoner, Philip
Published: (2026)
A Taxonomy of Omnicidal Futures Involving Artificial Intelligence
by: Critch, Andrew, et al.
Published: (2025)
by: Critch, Andrew, et al.
Published: (2025)
Driving down Poisson error can offset classification error in clinical tasks
by: Delahunt, Charles B., et al.
Published: (2024)
by: Delahunt, Charles B., et al.
Published: (2024)
Machine Learning for Physical Simulation Challenge Results and Retrospective Analysis: Power Grid Use Case
by: Leyli-Abadi, Milad, et al.
Published: (2025)
by: Leyli-Abadi, Milad, et al.
Published: (2025)
Generative inpainting of incomplete Euclidean distance matrices of trajectories generated by a fractional Brownian motion
by: Lobashev, Alexander, et al.
Published: (2024)
by: Lobashev, Alexander, et al.
Published: (2024)
Deploying Large Language Models With Retrieval Augmented Generation
by: Prabhune, Sonal, et al.
Published: (2024)
by: Prabhune, Sonal, et al.
Published: (2024)
Complex Facial Expression Recognition Using Deep Knowledge Distillation of Basic Features
by: Maiden, Angus, et al.
Published: (2023)
by: Maiden, Angus, et al.
Published: (2023)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
by: Tu, Songjun, et al.
Published: (2025)
by: Tu, Songjun, et al.
Published: (2025)
Hierarchical Refinement of Universal Multimodal Attacks on Vision-Language Models
by: Zhang, Peng-Fei, et al.
Published: (2026)
by: Zhang, Peng-Fei, et al.
Published: (2026)
Compressible Softmax-Attended Language under Incompressible Attention
by: Lee, Wonsuk
Published: (2026)
by: Lee, Wonsuk
Published: (2026)
Intelligence as Computation
by: Brock, Oliver
Published: (2024)
by: Brock, Oliver
Published: (2024)
EncQA: Benchmarking Vision-Language Models on Visual Encodings for Charts
by: Mukherjee, Kushin, et al.
Published: (2025)
by: Mukherjee, Kushin, et al.
Published: (2025)
MAA: Meticulous Adversarial Attack against Vision-Language Pre-trained Models
by: Zhang, Peng-Fei, et al.
Published: (2025)
by: Zhang, Peng-Fei, et al.
Published: (2025)
Towards Interpretable Visual Decoding with Attention to Brain Representations
by: Feng, Pinyuan, et al.
Published: (2025)
by: Feng, Pinyuan, et al.
Published: (2025)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
A ZeNN architecture to avoid the Gaussian trap
by: Carvalho, Luís, et al.
Published: (2025)
by: Carvalho, Luís, et al.
Published: (2025)
The Many Challenges of Human-Like Agents in Virtual Game Environments
by: Swiechowski, Maciej, et al.
Published: (2025)
by: Swiechowski, Maciej, et al.
Published: (2025)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
by: Bian, Yijun, et al.
Published: (2024)
by: Bian, Yijun, et al.
Published: (2024)
Quantifying Behavioral Dissimilarity Between Mathematical Expressions
by: Mežnar, Sebastian, et al.
Published: (2024)
by: Mežnar, Sebastian, et al.
Published: (2024)
Approximating Discrimination Within Models When Faced With Several Non-Binary Sensitive Attributes
by: Bian, Yijun, et al.
Published: (2024)
by: Bian, Yijun, et al.
Published: (2024)
Sparse vs Contiguous Adversarial Pixel Perturbations in Multimodal Models: An Empirical Analysis
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
STAC: Leveraging Spatio-Temporal Data Associations For Efficient Cross-Camera Streaming and Analytics
by: Gupta, Ragini, et al.
Published: (2024)
by: Gupta, Ragini, et al.
Published: (2024)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
by: He, Mengqi, et al.
Published: (2025)
by: He, Mengqi, et al.
Published: (2025)
ATEX-CF: Attack-Informed Counterfactual Explanations for Graph Neural Networks
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
by: Filus, Katarzyna, et al.
Published: (2025)
by: Filus, Katarzyna, et al.
Published: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
by: Abhishek, Alok, et al.
Published: (2026)
by: Abhishek, Alok, et al.
Published: (2026)
Similar Items
-
Surrealistic-like Image Generation with Vision-Language Models
by: Ayten, Elif, et al.
Published: (2024) -
VACoDe: Visual Augmented Contrastive Decoding
by: Kim, Sihyeon, et al.
Published: (2024) -
Effective and Robust Adversarial Training against Data and Label Corruptions
by: Zhang, Peng-Fei, et al.
Published: (2024) -
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
by: Zhang, Junbin, et al.
Published: (2022) -
From Language Models to Practical Self-Improving Computer Agents
by: Sheng, Alex
Published: (2024)