Breaking the Illusion of Security via Interpretation: Interpretable Vision Transformer Systems under Attack
Fuente:
arXiv
Salvato in:
| Autori principali: | Abdukhamidov, Eldor, Abuhamad, Mohammed, Woo, Simon S., Kim, Hyoungshick, Abuhmed, Tamer |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Attacking interpretable NLP systems
di: Abdukhamidov, Eldor, et al.
Pubblicazione: (2025)
di: Abdukhamidov, Eldor, et al.
Pubblicazione: (2025)
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
di: Lapin, Mykyta, et al.
Pubblicazione: (2025)
di: Lapin, Mykyta, et al.
Pubblicazione: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks
di: Feng, Zhou, et al.
Pubblicazione: (2025)
di: Feng, Zhou, et al.
Pubblicazione: (2025)
Structured Contrastive Learning for Interpretable Latent Representations
di: Shen, Zhengyang, et al.
Pubblicazione: (2025)
di: Shen, Zhengyang, et al.
Pubblicazione: (2025)
U-SEG: Uncertainty in SEGmentation -- A systematic multi-variable exploration
di: Smith, Michael, et al.
Pubblicazione: (2026)
di: Smith, Michael, et al.
Pubblicazione: (2026)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
di: Wu, Jason, et al.
Pubblicazione: (2026)
di: Wu, Jason, et al.
Pubblicazione: (2026)
Canonical Space Representation for 4D Panoptic Segmentation of Articulated Objects
di: Gomes, Manuel, et al.
Pubblicazione: (2025)
di: Gomes, Manuel, et al.
Pubblicazione: (2025)
CASE: Contrastive Activation for Saliency Estimation
di: Williamson, Dane, et al.
Pubblicazione: (2025)
di: Williamson, Dane, et al.
Pubblicazione: (2025)
BenthiCat: An opti-acoustic dataset for advancing benthic classification and habitat mapping
di: Rajani, Hayat, et al.
Pubblicazione: (2025)
di: Rajani, Hayat, et al.
Pubblicazione: (2025)
From Articles to Canopies: Knowledge-Driven Pseudo-Labelling for Tree Species Classification using LLM Experts
di: Romaszewski, Michał, et al.
Pubblicazione: (2026)
di: Romaszewski, Michał, et al.
Pubblicazione: (2026)
Leveraging Causal Reasoning Method for Explaining Medical Image Segmentation Models
di: Jiang, Limai, et al.
Pubblicazione: (2026)
di: Jiang, Limai, et al.
Pubblicazione: (2026)
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
di: Li, Tianle, et al.
Pubblicazione: (2025)
di: Li, Tianle, et al.
Pubblicazione: (2025)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
di: Du, Guanchen, et al.
Pubblicazione: (2025)
di: Du, Guanchen, et al.
Pubblicazione: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
di: Fang, Biyi, et al.
Pubblicazione: (2025)
di: Fang, Biyi, et al.
Pubblicazione: (2025)
PhysVid: Physics Aware Local Conditioning for Generative Video Models
di: Pathak, Saurabh, et al.
Pubblicazione: (2026)
di: Pathak, Saurabh, et al.
Pubblicazione: (2026)
Combining Euclidean Alignment and Data Augmentation for BCI decoding
di: Rodrigues, Gustavo H., et al.
Pubblicazione: (2024)
di: Rodrigues, Gustavo H., et al.
Pubblicazione: (2024)
How Can One Choose the Best CAM-Based Explainability Method for a CNN Model?
di: Costa, Daniel da Silva, et al.
Pubblicazione: (2026)
di: Costa, Daniel da Silva, et al.
Pubblicazione: (2026)
SpectralCA: Bi-Directional Cross-Attention for Next-Generation UAV Hyperspectral Vision
di: Brovko, D. V.
Pubblicazione: (2025)
di: Brovko, D. V.
Pubblicazione: (2025)
Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
di: Perera, Amal S., et al.
Pubblicazione: (2025)
di: Perera, Amal S., et al.
Pubblicazione: (2025)
Risk-Calibrated Bayesian Streaming Intrusion Detection with SRE-Aligned Decisions
di: Youssef, Michel
Pubblicazione: (2025)
di: Youssef, Michel
Pubblicazione: (2025)
Implementing Adaptations for Vision AutoRegressive Model
di: Shaikh, Kaif, et al.
Pubblicazione: (2025)
di: Shaikh, Kaif, et al.
Pubblicazione: (2025)
Localizing Adversarial Attacks To Produces More Imperceptible Noise
di: Reddy, Pavan, et al.
Pubblicazione: (2025)
di: Reddy, Pavan, et al.
Pubblicazione: (2025)
Nonparametric Partial Disentanglement via Mechanism Sparsity: Sparse Actions, Interventions and Sparse Temporal Dependencies
di: Lachapelle, Sébastien, et al.
Pubblicazione: (2024)
di: Lachapelle, Sébastien, et al.
Pubblicazione: (2024)
Adapting SAM with Dynamic Similarity Graphs for Few-Shot Parameter-Efficient Small Dense Object Detection: A Case Study of Chickpea Pods in Field Conditions
di: Jiang, Xintong, et al.
Pubblicazione: (2025)
di: Jiang, Xintong, et al.
Pubblicazione: (2025)
TUMLS: Trustful Fully Unsupervised Multi-Level Segmentation for Whole Slide Images of Histology
di: Rehamnia, Walid, et al.
Pubblicazione: (2025)
di: Rehamnia, Walid, et al.
Pubblicazione: (2025)
Neural Reasoning Networks: Efficient Interpretable Neural Networks With Automatic Textual Explanations
di: Carrow, Stephen, et al.
Pubblicazione: (2024)
di: Carrow, Stephen, et al.
Pubblicazione: (2024)
Butter: Frequency Consistency and Hierarchical Fusion for Autonomous Driving Object Detection
di: Lin, Xiaojian, et al.
Pubblicazione: (2025)
di: Lin, Xiaojian, et al.
Pubblicazione: (2025)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
Interpreting Structured Perturbations in Image Protection Methods for Diffusion Models
di: Martin, Michael R., et al.
Pubblicazione: (2025)
di: Martin, Michael R., et al.
Pubblicazione: (2025)
The Spotlight Resonance Method: Resolving the Alignment of Embedded Activations
di: Bird, George
Pubblicazione: (2025)
di: Bird, George
Pubblicazione: (2025)
AGOP-IxG: A Gradient Covariance Filter for Local Feature Attribution on Tabular Data, with a Controlled Benchmark
di: Katakam, Raj Kiran Gupta
Pubblicazione: (2026)
di: Katakam, Raj Kiran Gupta
Pubblicazione: (2026)
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
di: Dutta, Utsav, et al.
Pubblicazione: (2026)
di: Dutta, Utsav, et al.
Pubblicazione: (2026)
Mamba base PKD for efficient knowledge compression
di: Medina, José, et al.
Pubblicazione: (2025)
di: Medina, José, et al.
Pubblicazione: (2025)
Subspace Geometry Governs Catastrophic Forgetting in Low-Rank Adaptation
di: Steele, Brady
Pubblicazione: (2026)
di: Steele, Brady
Pubblicazione: (2026)
ZClassifier: Temperature Tuning and Manifold Approximation via KL Divergence on Logit Space
di: Yong, Shim Soon
Pubblicazione: (2025)
di: Yong, Shim Soon
Pubblicazione: (2025)
Feature emergence via margin maximization: case studies in algebraic tasks
di: Morwani, Depen, et al.
Pubblicazione: (2023)
di: Morwani, Depen, et al.
Pubblicazione: (2023)
Remaining Useful Life Estimation for Turbofan Engines: A Comparative Study of Classical, CNN, and LSTM Approaches
di: Goel, Astitva, et al.
Pubblicazione: (2026)
di: Goel, Astitva, et al.
Pubblicazione: (2026)
Learning to Land Anywhere: Transferable Generative Models for Aircraft Trajectories
di: Larsen, Olav Finne Praesteng, et al.
Pubblicazione: (2025)
di: Larsen, Olav Finne Praesteng, et al.
Pubblicazione: (2025)
SATORIS-N: Spectral Analysis based Traffic Observation Recovery via Informed Subspaces and Nuclear-norm minimization
di: Mohanty, Sampad, et al.
Pubblicazione: (2026)
di: Mohanty, Sampad, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Attacking interpretable NLP systems
di: Abdukhamidov, Eldor, et al.
Pubblicazione: (2025) -
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
di: Lapin, Mykyta, et al.
Pubblicazione: (2025) -
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024) -
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks
di: Feng, Zhou, et al.
Pubblicazione: (2025) -
Structured Contrastive Learning for Interpretable Latent Representations
di: Shen, Zhengyang, et al.
Pubblicazione: (2025)