On the Reliability of Vision-Language Models Under Adversarial Frequency-Domain Perturbations
Fuente:
arXiv
Saved in:
| Main Authors: | Vice, Jordan, Akhtar, Naveed, Gao, Yansong, Hartley, Richard, Mian, Ajmal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Manipulating and Mitigating Generative Model Biases without Retraining
by: Vice, Jordan, et al.
Published: (2024)
by: Vice, Jordan, et al.
Published: (2024)
Exploring Bias in over 100 Text-to-Image Generative Models
by: Vice, Jordan, et al.
Published: (2025)
by: Vice, Jordan, et al.
Published: (2025)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
by: Vice, Jordan, et al.
Published: (2024)
by: Vice, Jordan, et al.
Published: (2024)
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
by: Vice, Jordan, et al.
Published: (2024)
by: Vice, Jordan, et al.
Published: (2024)
CymbaDiff: Structured Spatial Diffusion for Sketch-based 3D Semantic Urban Scene Generation
by: Liang, Li, et al.
Published: (2025)
by: Liang, Li, et al.
Published: (2025)
Mitigating Memorization in Text-to-Image Diffusion via Region-Aware Prompt Augmentation and Multimodal Copy Detection
by: Chen, Yunzhuo, et al.
Published: (2026)
by: Chen, Yunzhuo, et al.
Published: (2026)
Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
by: Liang, Li, et al.
Published: (2025)
by: Liang, Li, et al.
Published: (2025)
Deeper Diffusion Models Amplify Bias
by: Hakemi, Shahin, et al.
Published: (2025)
by: Hakemi, Shahin, et al.
Published: (2025)
Single-weight Model Editing for Post-hoc Spurious Correlation Neutralization
by: Hakemi, Shahin, et al.
Published: (2025)
by: Hakemi, Shahin, et al.
Published: (2025)
DRBD-Mamba for Robust and Efficient Brain Tumor Segmentation with Analytical Insights
by: Ali, Danish, et al.
Published: (2025)
by: Ali, Danish, et al.
Published: (2025)
D3Seg: Dependency-Aware Diffusion for Brain Tumor Segmentation with Missing Modalities
by: Ali, Danish, et al.
Published: (2026)
by: Ali, Danish, et al.
Published: (2026)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
by: Li, Congliang, et al.
Published: (2022)
by: Li, Congliang, et al.
Published: (2022)
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
by: Yang, Peiyu, et al.
Published: (2026)
by: Yang, Peiyu, et al.
Published: (2026)
Dynamic watermarks in images generated by diffusion models
by: Chen, Yunzhuo, et al.
Published: (2025)
by: Chen, Yunzhuo, et al.
Published: (2025)
Deepfake Detection with Spatio-Temporal Consistency and Attention
by: Chen, Yunzhuo, et al.
Published: (2025)
by: Chen, Yunzhuo, et al.
Published: (2025)
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
by: Ibrahim, Muhammad, et al.
Published: (2025)
by: Ibrahim, Muhammad, et al.
Published: (2025)
Efficient Diffusion Models for Vision: A Survey
by: Ulhaq, Anwaar, et al.
Published: (2022)
by: Ulhaq, Anwaar, et al.
Published: (2022)
SCTransNet: Spatial-channel Cross Transformer Network for Infrared Small Target Detection
by: Yuan, Shuai, et al.
Published: (2024)
by: Yuan, Shuai, et al.
Published: (2024)
Suitability of KANs for Computer Vision: A preliminary investigation
by: Azam, Basim, et al.
Published: (2024)
by: Azam, Basim, et al.
Published: (2024)
Artificial intelligence techniques in inherited retinal diseases: A review
by: Trinh, Han, et al.
Published: (2024)
by: Trinh, Han, et al.
Published: (2024)
Context-guided Responsible Data Augmentation with Diffusion Models
by: Islam, Khawar, et al.
Published: (2025)
by: Islam, Khawar, et al.
Published: (2025)
Domain-invariant Prototypes for Semantic Segmentation
by: Yang, Zhengeng, et al.
Published: (2022)
by: Yang, Zhengeng, et al.
Published: (2022)
Modeling Human Skeleton Joint Dynamics for Fall Detection
by: Zahan, Sania, et al.
Published: (2025)
by: Zahan, Sania, et al.
Published: (2025)
Sample-agnostic Adversarial Perturbation for Vision-Language Pre-training Models
by: Zheng, Haonan, et al.
Published: (2024)
by: Zheng, Haonan, et al.
Published: (2024)
Rethinking the Need for Source Models: Source-Free Domain Adaptation from Scratch Guided by a Vision-Language Model
by: Bingtao, Zhou, et al.
Published: (2026)
by: Bingtao, Zhou, et al.
Published: (2026)
Implicit Neural Representation-Based Continuous Single Image Super-Resolution: An Empirical Benchmark
by: Nasir, Tayyab, et al.
Published: (2026)
by: Nasir, Tayyab, et al.
Published: (2026)
Spatial-Frequency Discriminability for Revealing Adversarial Perturbations
by: Wang, Chao, et al.
Published: (2023)
by: Wang, Chao, et al.
Published: (2023)
Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control
by: Le, Nhat, et al.
Published: (2026)
by: Le, Nhat, et al.
Published: (2026)
SDFA: Structure Aware Discriminative Feature Aggregation for Efficient Human Fall Detection in Video
by: Zahan, Sania, et al.
Published: (2025)
by: Zahan, Sania, et al.
Published: (2025)
Occlusion-aware Text-Image-Point Cloud Pretraining for Open-World 3D Object Recognition
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
Class-Partitioned VQ-VAE and Latent Flow Matching for Point Cloud Scene Generation
by: Edirimuni, Dasith de Silva, et al.
Published: (2026)
by: Edirimuni, Dasith de Silva, et al.
Published: (2026)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
by: Guo, Zihao, et al.
Published: (2026)
by: Guo, Zihao, et al.
Published: (2026)
Temporally Consistent Referring Video Object Segmentation with Hybrid Memory
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
by: Jiang, Jiantong, et al.
Published: (2024)
by: Jiang, Jiantong, et al.
Published: (2024)
FLaTEC: Frequency-Disentangled Latent Triplanes for Efficient Compression of LiDAR Point Clouds
by: Zhang, Xiaoge, et al.
Published: (2025)
by: Zhang, Xiaoge, et al.
Published: (2025)
Plug-and-Play Interpretable Responsible Text-to-Image Generation via Dual-Space Multi-facet Concept Control
by: Azam, Basim, et al.
Published: (2025)
by: Azam, Basim, et al.
Published: (2025)
Improving Model Generalization by On-manifold Adversarial Augmentation in the Frequency Domain
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
Integrating Frequency-Domain Representations with Low-Rank Adaptation in Vision-Language Models
by: Khan, Md Azim, et al.
Published: (2025)
by: Khan, Md Azim, et al.
Published: (2025)
One Perturbation is Enough: On Generating Universal Adversarial Perturbations against Vision-Language Pre-training Models
by: Fang, Hao, et al.
Published: (2024)
by: Fang, Hao, et al.
Published: (2024)
Evaluating Adversarial Robustness in the Spatial Frequency Domain
by: Liao, Keng-Hsin, et al.
Published: (2024)
by: Liao, Keng-Hsin, et al.
Published: (2024)
Similar Items
-
Manipulating and Mitigating Generative Model Biases without Retraining
by: Vice, Jordan, et al.
Published: (2024) -
Exploring Bias in over 100 Text-to-Image Generative Models
by: Vice, Jordan, et al.
Published: (2025) -
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
by: Vice, Jordan, et al.
Published: (2024) -
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
by: Vice, Jordan, et al.
Published: (2024) -
CymbaDiff: Structured Spatial Diffusion for Sketch-based 3D Semantic Urban Scene Generation
by: Liang, Li, et al.
Published: (2025)