PG-Attack: A Precision-Guided Adversarial Attack Framework Against Vision Foundation Models for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Jiyuan, Chen, Zhaoyu, Jiang, Kaixun, Guo, Haijing, Gao, Shuyong, Zhang, Wenqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024)
Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking
von: Fu, Jiyuan, et al.
Veröffentlicht: (2026)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2026)
Boosting the Transferability of Adversarial Attacks with Global Momentum Initialization
von: Wang, Jiafeng, et al.
Veröffentlicht: (2022)
von: Wang, Jiafeng, et al.
Veröffentlicht: (2022)
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
A Multi-task Adversarial Attack Against Face Authentication
von: Wang, Hanrui, et al.
Veröffentlicht: (2024)
von: Wang, Hanrui, et al.
Veröffentlicht: (2024)
Boosting Adversarial Transferability with Spatial Adversarial Alignment
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2025)
Hierarchical Refinement of Universal Multimodal Attacks on Vision-Language Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
Clean Image May be Dangerous: Data Poisoning Attacks Against Deep Hashing
von: Li, Shuai, et al.
Veröffentlicht: (2025)
von: Li, Shuai, et al.
Veröffentlicht: (2025)
VideoPure: Diffusion-based Adversarial Purification for Video Recognition
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025)
Exploring the Adversarial Robustness of Face Forgery Detection with Decision-based Black-box Attacks
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2023)
von: Chen, Zhaoyu, et al.
Veröffentlicht: (2023)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
von: Huang, Zhijian, et al.
Veröffentlicht: (2024)
von: Huang, Zhijian, et al.
Veröffentlicht: (2024)
Improving Adversarial Transferability with Neighbourhood Gradient Information
von: Guo, Haijing, et al.
Veröffentlicht: (2024)
von: Guo, Haijing, et al.
Veröffentlicht: (2024)
ROI-Guided Point Cloud Geometry Compression Towards Human and Machine Vision
von: Liang, Xie, et al.
Veröffentlicht: (2025)
von: Liang, Xie, et al.
Veröffentlicht: (2025)
BadCM: Invisible Backdoor Attack Against Cross-Modal Learning
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
From Attack to Protection: Leveraging Watermarking Attack Network for Advanced Add-on Watermarking
von: Nam, Seung-Hun, et al.
Veröffentlicht: (2020)
von: Nam, Seung-Hun, et al.
Veröffentlicht: (2020)
A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
A Preprocessing Framework for Video Machine Vision under Compression
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Waymo-3DSkelMo: A Multi-Agent 3D Skeletal Motion Dataset for Pedestrian Interaction Modeling in Autonomous Driving
von: Zhu, Guangxun, et al.
Veröffentlicht: (2025)
von: Zhu, Guangxun, et al.
Veröffentlicht: (2025)
Visual Adversarial Attack on Vision-Language Models for Autonomous Driving
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
RSAgent: Learning to Reason and Act for Text-Guided Segmentation via Multi-Turn Tool Invocations
von: He, Xingqi, et al.
Veröffentlicht: (2025)
von: He, Xingqi, et al.
Veröffentlicht: (2025)
Attack on Scene Flow using Point Clouds
von: Oskouie, Haniyeh Ehsani, et al.
Veröffentlicht: (2024)
von: Oskouie, Haniyeh Ehsani, et al.
Veröffentlicht: (2024)
Prompt-Aware Adaptive Elastic Weight Consolidation for Continual Learning in Medical Vision-Language Models
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
Comparing the Robustness of Modern No-Reference Image- and Video-Quality Metrics to Adversarial Attacks
von: Antsiferova, Anastasia, et al.
Veröffentlicht: (2023)
von: Antsiferova, Anastasia, et al.
Veröffentlicht: (2023)
Dynamic Semantic-Aware Correlation Modeling for UAV Tracking
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models
von: Huang, Yuhang, et al.
Veröffentlicht: (2024)
von: Huang, Yuhang, et al.
Veröffentlicht: (2024)
DPC: Dual-Prompt Collaboration for Tuning Vision-Language Models
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
Semi-supervised Chinese Poem-to-Painting Generation via Cycle-consistent Adversarial Networks
von: Lu, Zhengyang, et al.
Veröffentlicht: (2024)
von: Lu, Zhengyang, et al.
Veröffentlicht: (2024)
Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing
von: Chen, Yaru, et al.
Veröffentlicht: (2025)
von: Chen, Yaru, et al.
Veröffentlicht: (2025)
Towards Transferable Attacks Against Vision-LLMs in Autonomous Driving with Typography
von: Chung, Nhat, et al.
Veröffentlicht: (2024)
von: Chung, Nhat, et al.
Veröffentlicht: (2024)
TimeLoc: A Unified End-to-End Framework for Precise Timestamp Localization in Long Videos
von: Zhang, Chen-Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Chen-Lin, et al.
Veröffentlicht: (2025)
RoWSFormer: A Robust Watermarking Framework with Swin Transformer for Enhanced Geometric Attack Resilience
von: Chen, Weitong, et al.
Veröffentlicht: (2024)
von: Chen, Weitong, et al.
Veröffentlicht: (2024)
Towards Real-World Adverse Weather Image Restoration: Enhancing Clearness and Semantics with Vision-Language Models
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2026)
von: Cai, Qi, et al.
Veröffentlicht: (2026)
HeGraphAdapter: Tuning Multi-Modal Vision-Language Models with Heterogeneous Graph Adapter
von: Zhao, Yumiao, et al.
Veröffentlicht: (2024)
von: Zhao, Yumiao, et al.
Veröffentlicht: (2024)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
von: Mandelli, Sara, et al.
Veröffentlicht: (2024)
von: Mandelli, Sara, et al.
Veröffentlicht: (2024)
POINTS1.5: Building a Vision-Language Model towards Real World Applications
von: Liu, Yuan, et al.
Veröffentlicht: (2024)
von: Liu, Yuan, et al.
Veröffentlicht: (2024)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2023)
von: Min, Chen, et al.
Veröffentlicht: (2023)
Q-Bench: A Benchmark for General-Purpose Foundation Models on Low-level Vision
von: Wu, Haoning, et al.
Veröffentlicht: (2023)
von: Wu, Haoning, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction
von: Fu, Jiyuan, et al.
Veröffentlicht: (2024) -
Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment
von: Jiang, Kaixun, et al.
Veröffentlicht: (2025) -
VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking
von: Fu, Jiyuan, et al.
Veröffentlicht: (2026) -
Boosting the Transferability of Adversarial Attacks with Global Momentum Initialization
von: Wang, Jiafeng, et al.
Veröffentlicht: (2022) -
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
von: Xu, Jingning, et al.
Veröffentlicht: (2026)