Protecting against simultaneous data poisoning attacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alex, Neel, Siddiqui, Shoaib Ahmed, Sanyal, Amartya, Krueger, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Position: Capability Control Should be a Separate Goal From Alignment
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2026)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2026)
Blockwise Self-Supervised Learning at Scale
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
Keeping up with dynamic attackers: Certifying robustness to adaptive online data poisoning
von: Bose, Avinandan, et al.
Veröffentlicht: (2025)
von: Bose, Avinandan, et al.
Veröffentlicht: (2025)
A deeper look at depth pruning of LLMs
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
Exploring the design space of deep-learning-based weather forecasting systems
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
Provably effective detection of effective data poisoning attacks
von: Gallagher, Jonathan, et al.
Veröffentlicht: (2025)
von: Gallagher, Jonathan, et al.
Veröffentlicht: (2025)
Less Noise, Same Certificate: Retain Sensitivity for Unlearning
von: Heinzler, Carolin, et al.
Veröffentlicht: (2026)
von: Heinzler, Carolin, et al.
Veröffentlicht: (2026)
Online Learning and Unlearning
von: Hu, Yaxi, et al.
Veröffentlicht: (2025)
von: Hu, Yaxi, et al.
Veröffentlicht: (2025)
Language Generation with Replay: A Learning-Theoretic View of Model Collapse
von: Racca, Giorgio, et al.
Veröffentlicht: (2026)
von: Racca, Giorgio, et al.
Veröffentlicht: (2026)
An Iterative Algorithm for Differentially Private $k$-PCA with Adaptive Noise
von: Düngler, Johanna, et al.
Veröffentlicht: (2025)
von: Düngler, Johanna, et al.
Veröffentlicht: (2025)
On the Growth of Mistakes in Differentially Private Online Learning: A Lower Bound Perspective
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2024)
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2024)
From Dormant to Deleted: Tamper-Resistant Unlearning Through Weight-Space Regularization
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2025)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2025)
The Topological Trouble With Transformers
von: Mozer, Michael C., et al.
Veröffentlicht: (2026)
von: Mozer, Michael C., et al.
Veröffentlicht: (2026)
LoRA and Privacy: When Random Projections Help (and When They Don't)
von: Hu, Yaxi, et al.
Veröffentlicht: (2026)
von: Hu, Yaxi, et al.
Veröffentlicht: (2026)
Adaptive Sampling and Clipping for Private Worst-Case Group Optimization
von: Cairney-Leeming, Max, et al.
Veröffentlicht: (2026)
von: Cairney-Leeming, Max, et al.
Veröffentlicht: (2026)
Provable Privacy with Non-Private Pre-Processing
von: Hu, Yaxi, et al.
Veröffentlicht: (2024)
von: Hu, Yaxi, et al.
Veröffentlicht: (2024)
Learning in an Echo Chamber: Online Learning with Replay Adversary
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2025)
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2025)
Permissive Information-Flow Analysis for Large Language Models
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
Provable unlearning in topic modeling and downstream tasks
von: Wei, Stanley, et al.
Veröffentlicht: (2024)
von: Wei, Stanley, et al.
Veröffentlicht: (2024)
Differentially Private Steering for Large Language Model Alignment
von: Goel, Anmol, et al.
Veröffentlicht: (2025)
von: Goel, Anmol, et al.
Veröffentlicht: (2025)
Accuracy on the wrong line: On the pitfalls of noisy data for out-of-distribution generalisation
von: Sanyal, Amartya, et al.
Veröffentlicht: (2024)
von: Sanyal, Amartya, et al.
Veröffentlicht: (2024)
The Role of Learning Algorithms in Collective Action
von: Ben-Dov, Omri, et al.
Veröffentlicht: (2024)
von: Ben-Dov, Omri, et al.
Veröffentlicht: (2024)
Fairness for the People, by the People: Minority Collective Action
von: Ben-Dov, Omri, et al.
Veröffentlicht: (2025)
von: Ben-Dov, Omri, et al.
Veröffentlicht: (2025)
On Evaluating LLMs' Capabilities as Functional Approximators: A Bayesian Perspective
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs
von: Shysheya, Aliaksandra, et al.
Veröffentlicht: (2025)
von: Shysheya, Aliaksandra, et al.
Veröffentlicht: (2025)
AlertBERT: A noise-robust alert grouping framework for simultaneous cyber attacks
von: Karner, Lukas, et al.
Veröffentlicht: (2026)
von: Karner, Lukas, et al.
Veröffentlicht: (2026)
A GAN-based data poisoning framework against anomaly detection in vertical federated learning
von: Chen, Xiaolin, et al.
Veröffentlicht: (2024)
von: Chen, Xiaolin, et al.
Veröffentlicht: (2024)
Correlation inference attacks against machine learning models
von: Creţu, Ana-Maria, et al.
Veröffentlicht: (2021)
von: Creţu, Ana-Maria, et al.
Veröffentlicht: (2021)
The poison of dimensionality
von: Hoang, Lê-Nguyên
Veröffentlicht: (2024)
von: Hoang, Lê-Nguyên
Veröffentlicht: (2024)
Evaluating randomized smoothing as a defense against adversarial attacks in trajectory prediction
von: Schumann, Julian F., et al.
Veröffentlicht: (2026)
von: Schumann, Julian F., et al.
Veröffentlicht: (2026)
Noisy Neighbors: Efficient membership inference attacks against LLMs
von: Galli, Filippo, et al.
Veröffentlicht: (2024)
von: Galli, Filippo, et al.
Veröffentlicht: (2024)
Fixed-point graph convolutional networks against adversarial attacks
von: Khan, Shakib, et al.
Veröffentlicht: (2025)
von: Khan, Shakib, et al.
Veröffentlicht: (2025)
On the use of neurosymbolic AI for defending against cyber attacks
von: Grov, Gudmund, et al.
Veröffentlicht: (2024)
von: Grov, Gudmund, et al.
Veröffentlicht: (2024)
Corrective Machine Unlearning
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
von: Goel, Shashwat, et al.
Veröffentlicht: (2024)
Modular Jets for Supervised Pipelines: Diagnosing Mirage vs Identifiability
von: Sanyal, Suman
Veröffentlicht: (2025)
von: Sanyal, Suman
Veröffentlicht: (2025)
Perception Learning: A Formal Separation of Sensory Representation Learning from Decision Learning
von: Sanyal, Suman
Veröffentlicht: (2025)
von: Sanyal, Suman
Veröffentlicht: (2025)
Robust Mixture Learning when Outliers Overwhelm Small Groups
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2024)
von: Dmitriev, Daniil, et al.
Veröffentlicht: (2024)
Competition is the key: A Game Theoretic Causal Discovery Approach
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
Concept activation vectors: a unifying view and adversarial attacks
von: Schnoor, Ekkehard, et al.
Veröffentlicht: (2025)
von: Schnoor, Ekkehard, et al.
Veröffentlicht: (2025)
Comparing privacy notions for protection against reconstruction attacks in machine learning
von: Biswas, Sayan, et al.
Veröffentlicht: (2025)
von: Biswas, Sayan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Position: Capability Control Should be a Separate Goal From Alignment
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2026) -
Blockwise Self-Supervised Learning at Scale
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023) -
Keeping up with dynamic attackers: Certifying robustness to adaptive online data poisoning
von: Bose, Avinandan, et al.
Veröffentlicht: (2025) -
A deeper look at depth pruning of LLMs
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024) -
Exploring the design space of deep-learning-based weather forecasting systems
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)