Integrating uncertainty quantification into randomized smoothing based robustness guarantees
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Däubener, Sina, Maag, Kira, Krueger, David, Fischer, Asja |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Detecting Adversarial Attacks in Semantic Segmentation via Uncertainty Estimation: A Deep Analysis
von: Maag, Kira, et al.
Veröffentlicht: (2024)
von: Maag, Kira, et al.
Veröffentlicht: (2024)
Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks
von: Pizarro, Matías, et al.
Veröffentlicht: (2026)
von: Pizarro, Matías, et al.
Veröffentlicht: (2026)
DistriBlock: Identifying adversarial audio samples by leveraging characteristics of the output distribution
von: Pizarro, Matías, et al.
Veröffentlicht: (2023)
von: Pizarro, Matías, et al.
Veröffentlicht: (2023)
Lightweight Model Attribution and Detection of Synthetic Speech via Audio Residual Fingerprints
von: Pizarro, Matías, et al.
Veröffentlicht: (2024)
von: Pizarro, Matías, et al.
Veröffentlicht: (2024)
An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees
von: Tran, Hoang, et al.
Veröffentlicht: (2026)
von: Tran, Hoang, et al.
Veröffentlicht: (2026)
Near-Optimal differentially private low-rank trace regression with guaranteed private initialization
von: Zha, Mengyue
Veröffentlicht: (2024)
von: Zha, Mengyue
Veröffentlicht: (2024)
Understanding In-Context Learning of Linear Models in Transformers Through an Adversarial Lens
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
Prompt Obfuscation for Large Language Models
von: Pape, David, et al.
Veröffentlicht: (2024)
von: Pape, David, et al.
Veröffentlicht: (2024)
Nemesis: Noise-randomized Encryption with Modular Efficiency and Secure Integration in Machine Learning Systems
von: Zhao, Dongfang
Veröffentlicht: (2024)
von: Zhao, Dongfang
Veröffentlicht: (2024)
AlertBERT: A noise-robust alert grouping framework for simultaneous cyber attacks
von: Karner, Lukas, et al.
Veröffentlicht: (2026)
von: Karner, Lukas, et al.
Veröffentlicht: (2026)
Learning to Forget using Hypernetworks
von: Rangel, Jose Miguel Lara, et al.
Veröffentlicht: (2024)
von: Rangel, Jose Miguel Lara, et al.
Veröffentlicht: (2024)
AI-Generated Faces in the Real World: A Large-Scale Case Study of Twitter Profile Images
von: Ricker, Jonas, et al.
Veröffentlicht: (2024)
von: Ricker, Jonas, et al.
Veröffentlicht: (2024)
Maximize margins for robust splicing detection
von: de Kergunic, Julien Simon, et al.
Veröffentlicht: (2025)
von: de Kergunic, Julien Simon, et al.
Veröffentlicht: (2025)
A PUF-Based Approach for Copy Protection of Intellectual Property in Neural Network Models
von: Dorfmeister, Daniel, et al.
Veröffentlicht: (2026)
von: Dorfmeister, Daniel, et al.
Veröffentlicht: (2026)
Learning diverse attacks on large language models for robust red-teaming and safety tuning
von: Lee, Seanie, et al.
Veröffentlicht: (2024)
von: Lee, Seanie, et al.
Veröffentlicht: (2024)
ACE: A Security Architecture for LLM-Integrated App Systems
von: Li, Evan, et al.
Veröffentlicht: (2025)
von: Li, Evan, et al.
Veröffentlicht: (2025)
Securing Federated Learning against Backdoor Threats with Foundation Model Integration
von: Bi, Xiaohuan, et al.
Veröffentlicht: (2024)
von: Bi, Xiaohuan, et al.
Veröffentlicht: (2024)
Explainable Malware Detection through Integrated Graph Reduction and Learning Techniques
von: Mohammadian, Hesamodin, et al.
Veröffentlicht: (2024)
von: Mohammadian, Hesamodin, et al.
Veröffentlicht: (2024)
Integrating Explainable AI for Effective Malware Detection in Encrypted Network Traffic
von: Zeleke, Sileshi Nibret, et al.
Veröffentlicht: (2025)
von: Zeleke, Sileshi Nibret, et al.
Veröffentlicht: (2025)
Integrating Feature Attention and Temporal Modeling for Collaborative Financial Risk Assessment
von: Yao, Yue, et al.
Veröffentlicht: (2025)
von: Yao, Yue, et al.
Veröffentlicht: (2025)
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2024)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2024)
Uncertainty quantification by block bootstrap for differentially private stochastic gradient descent
von: Dette, Holger, et al.
Veröffentlicht: (2024)
von: Dette, Holger, et al.
Veröffentlicht: (2024)
Differentially Private Publication of Electricity Time Series Data in Smart Grids
von: Shaham, Sina, et al.
Veröffentlicht: (2024)
von: Shaham, Sina, et al.
Veröffentlicht: (2024)
PickleBall: Secure Deserialization of Pickle-based Machine Learning Models (Extended Report)
von: Kellas, Andreas D., et al.
Veröffentlicht: (2025)
von: Kellas, Andreas D., et al.
Veröffentlicht: (2025)
WaterMax: breaking the LLM watermark detectability-robustness-quality trade-off
von: Giboulot, Eva, et al.
Veröffentlicht: (2024)
von: Giboulot, Eva, et al.
Veröffentlicht: (2024)
Nonideality-aware training makes memristive networks more robust to adversarial attacks
von: Joksas, Dovydas, et al.
Veröffentlicht: (2024)
von: Joksas, Dovydas, et al.
Veröffentlicht: (2024)
F-RBA: A Federated Learning-based Framework for Risk-based Authentication
von: Fereidouni, Hamidreza, et al.
Veröffentlicht: (2024)
von: Fereidouni, Hamidreza, et al.
Veröffentlicht: (2024)
Design Patterns for Securing LLM Agents against Prompt Injections
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2025)
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2025)
BadSampler: Harnessing the Power of Catastrophic Forgetting to Poison Byzantine-robust Federated Learning
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
Malware Detection based on API calls
von: Fellicious, Christofer, et al.
Veröffentlicht: (2025)
von: Fellicious, Christofer, et al.
Veröffentlicht: (2025)
Privacy Vulnerabilities in Marginals-based Synthetic Data
von: Golob, Steven, et al.
Veröffentlicht: (2024)
von: Golob, Steven, et al.
Veröffentlicht: (2024)
Tiny, Hardware-Independent, Compression-based Classification
von: Meyers, Charles, et al.
Veröffentlicht: (2026)
von: Meyers, Charles, et al.
Veröffentlicht: (2026)
No More, No Less: Task Alignment in Terminal Agents
von: Mavali, Sina, et al.
Veröffentlicht: (2026)
von: Mavali, Sina, et al.
Veröffentlicht: (2026)
Weights Shuffling for Improving DPSGD in Transformer-based Models
von: Yang, Jungang, et al.
Veröffentlicht: (2024)
von: Yang, Jungang, et al.
Veröffentlicht: (2024)
Entity-based Reinforcement Learning for Autonomous Cyber Defence
von: Thompson, Isaac Symes, et al.
Veröffentlicht: (2024)
von: Thompson, Isaac Symes, et al.
Veröffentlicht: (2024)
Adaptive Probe-based Steering for Robust LLM Jailbreaking
von: Chen, Junxi, et al.
Veröffentlicht: (2026)
von: Chen, Junxi, et al.
Veröffentlicht: (2026)
Immersion and Invariance-based Coding for Privacy-Preserving Federated Learning
von: Hayati, Haleh, et al.
Veröffentlicht: (2024)
von: Hayati, Haleh, et al.
Veröffentlicht: (2024)
User authentication system based on human exhaled breath physics
von: Karunanethy, Mukesh, et al.
Veröffentlicht: (2024)
von: Karunanethy, Mukesh, et al.
Veröffentlicht: (2024)
On the Generalizability of Machine Learning-based Ransomware Detection in Block Storage
von: Reategui, Nicolas, et al.
Veröffentlicht: (2024)
von: Reategui, Nicolas, et al.
Veröffentlicht: (2024)
Embedding-based classifiers can detect prompt injection attacks
von: Ayub, Md. Ahsan, et al.
Veröffentlicht: (2024)
von: Ayub, Md. Ahsan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Detecting Adversarial Attacks in Semantic Segmentation via Uncertainty Estimation: A Deep Analysis
von: Maag, Kira, et al.
Veröffentlicht: (2024) -
Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks
von: Pizarro, Matías, et al.
Veröffentlicht: (2026) -
DistriBlock: Identifying adversarial audio samples by leveraging characteristics of the output distribution
von: Pizarro, Matías, et al.
Veröffentlicht: (2023) -
Lightweight Model Attribution and Detection of Synthetic Speech via Audio Residual Fingerprints
von: Pizarro, Matías, et al.
Veröffentlicht: (2024) -
An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees
von: Tran, Hoang, et al.
Veröffentlicht: (2026)