Saved in:
| Main Author: | Santos-Grueiro, Igor |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.08449 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Privacy Leakage in Split Learning
by: Qiu, Xinchi, et al.
Published: (2023)
by: Qiu, Xinchi, et al.
Published: (2023)
Beyond TVLA: Anderson-Darling Leakage Assessment for Neural Network Side-Channel Leakage Detection
by: Mikulec, Ján, et al.
Published: (2026)
by: Mikulec, Ján, et al.
Published: (2026)
CTIGuardian: A Few-Shot Framework for Mitigating Privacy Leakage in Fine-Tuned LLMs
by: Arachchige, Shashie Dilhara Batan, et al.
Published: (2025)
by: Arachchige, Shashie Dilhara Batan, et al.
Published: (2025)
Training on Fake Labels: Mitigating Label Leakage in Split Learning via Secure Dimension Transformation
by: Jiang, Yukun, et al.
Published: (2024)
by: Jiang, Yukun, et al.
Published: (2024)
FLARE: A Wireless Side-Channel Fingerprinting Attack on Federated Learning
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
Improving IoT Intrusion Detection Through SMOTE-Based Oversampling and Extended Multi-Model Evaluation on Side-Channel Power Data
by: Shahzad, Muhammad Khuram, et al.
Published: (2026)
by: Shahzad, Muhammad Khuram, et al.
Published: (2026)
Shape and Substance: Dual-Layer Side-Channel Attacks on Local Vision-Language Models
by: Hadad, Eyal, et al.
Published: (2026)
by: Hadad, Eyal, et al.
Published: (2026)
Precise Extraction of Deep Learning Models via Side-Channel Attacks on Edge/Endpoint Devices
by: Lee, Younghan, et al.
Published: (2024)
by: Lee, Younghan, et al.
Published: (2024)
When Graph Structure Becomes a Liability: A Critical Re-Evaluation of Graph Neural Networks for Bitcoin Fraud Detection under Temporal Distribution Shift
by: Maganti, Saket
Published: (2026)
by: Maganti, Saket
Published: (2026)
Mitigating the Structural Bias in Graph Adversarial Defenses
by: Fang, Junyuan, et al.
Published: (2025)
by: Fang, Junyuan, et al.
Published: (2025)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
by: Nie, Yuzhou, et al.
Published: (2024)
by: Nie, Yuzhou, et al.
Published: (2024)
From Data Leak to Secret Misses: The Impact of Data Leakage on Secret Detection Models
by: Soltaniani, Farnaz, et al.
Published: (2026)
by: Soltaniani, Farnaz, et al.
Published: (2026)
LeakSealer: A Semisupervised Defense for LLMs Against Prompt Injection and Leakage Attacks
by: Panebianco, Francesco, et al.
Published: (2025)
by: Panebianco, Francesco, et al.
Published: (2025)
When Speculation Spills Secrets: Side Channels via Speculative Decoding In LLMs
by: Wei, Jiankun, et al.
Published: (2024)
by: Wei, Jiankun, et al.
Published: (2024)
Position: Retire the "Positive Backdoor" Label -- Secret Alignment Requires Strict and Systematic Evaluation
by: Li, Jianwei, et al.
Published: (2026)
by: Li, Jianwei, et al.
Published: (2026)
Client-Side Patching against Backdoor Attacks in Federated Learning
by: Molina-Coronado, Borja
Published: (2024)
by: Molina-Coronado, Borja
Published: (2024)
Theoretical Analysis of Privacy Leakage in Trustworthy Federated Learning: A Perspective from Linear Algebra and Optimization Theory
by: Zhang, Xiaojin, et al.
Published: (2024)
by: Zhang, Xiaojin, et al.
Published: (2024)
The Dark Side of Digital Twins: Adversarial Attacks on AI-Driven Water Forecasting
by: Homaei, Mohammadhossein, et al.
Published: (2025)
by: Homaei, Mohammadhossein, et al.
Published: (2025)
DESIGN: Encrypted GNN Inference via Server-Side Input Graph Pruning
by: Zhao, Kaixiang, et al.
Published: (2025)
by: Zhao, Kaixiang, et al.
Published: (2025)
Mitigating Many-Shot Jailbreaking
by: Ackerman, Christopher M., et al.
Published: (2025)
by: Ackerman, Christopher M., et al.
Published: (2025)
One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries
by: Zloczower, Itay, et al.
Published: (2026)
by: Zloczower, Itay, et al.
Published: (2026)
Prompt-Induced Over-Generation as Denial-of-Service: A Black-Box Attack-Side Benchmark
by: Manu, et al.
Published: (2025)
by: Manu, et al.
Published: (2025)
Stateless Yet Not Forgetful: Implicit Memory as a Hidden Channel in LLMs
by: Salem, Ahmed, et al.
Published: (2026)
by: Salem, Ahmed, et al.
Published: (2026)
Topological Signatures of Adversaries in Multimodal Alignments
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
Fair Finetuning Mitigates Distribution Inference Attacks
by: Naidu, Rakshit
Published: (2026)
by: Naidu, Rakshit
Published: (2026)
Rethinking Pruning for Backdoor Mitigation: An Optimization Perspective
by: Li, Nan, et al.
Published: (2024)
by: Li, Nan, et al.
Published: (2024)
Jailbreaking and Mitigation of Vulnerabilities in Large Language Models
by: Peng, Benji, et al.
Published: (2024)
by: Peng, Benji, et al.
Published: (2024)
Ghost in the Context: Measuring Policy-Carriage Failures in Decision-Time Assembly
by: Santos-Grueiro, Igor
Published: (2026)
by: Santos-Grueiro, Igor
Published: (2026)
No More, No Less: Task Alignment in Terminal Agents
by: Mavali, Sina, et al.
Published: (2026)
by: Mavali, Sina, et al.
Published: (2026)
Leveraging RAG for Training-Free Alignment of LLMs
by: Halloran, John T.
Published: (2026)
by: Halloran, John T.
Published: (2026)
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
by: Ren, Zhiyao, et al.
Published: (2025)
by: Ren, Zhiyao, et al.
Published: (2025)
Adaptive PII Mitigation Framework for Large Language Models
by: Asthana, Shubhi, et al.
Published: (2025)
by: Asthana, Shubhi, et al.
Published: (2025)
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
by: Min, Rui, et al.
Published: (2024)
by: Min, Rui, et al.
Published: (2024)
Mitigating Deep Reinforcement Learning Backdoors in the Neural Activation Space
by: Vyas, Sanyam, et al.
Published: (2024)
by: Vyas, Sanyam, et al.
Published: (2024)
Mutual Information Guided Backdoor Mitigation for Pre-trained Encoders
by: Han, Tingxu, et al.
Published: (2024)
by: Han, Tingxu, et al.
Published: (2024)
A Cryptographic Perspective on Mitigation vs. Detection in Machine Learning
by: Gluch, Greg, et al.
Published: (2025)
by: Gluch, Greg, et al.
Published: (2025)
SALT: Steering Activations towards Leakage-free Thinking in Chain of Thought
by: Batra, Shourya, et al.
Published: (2025)
by: Batra, Shourya, et al.
Published: (2025)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
by: Liao, Zeyi, et al.
Published: (2024)
by: Liao, Zeyi, et al.
Published: (2024)
Improved Algorithms for Differentially Private Language Model Alignment
by: Chen, Keyu, et al.
Published: (2025)
by: Chen, Keyu, et al.
Published: (2025)
Similar Items
-
Evaluating Privacy Leakage in Split Learning
by: Qiu, Xinchi, et al.
Published: (2023) -
Beyond TVLA: Anderson-Darling Leakage Assessment for Neural Network Side-Channel Leakage Detection
by: Mikulec, Ján, et al.
Published: (2026) -
CTIGuardian: A Few-Shot Framework for Mitigating Privacy Leakage in Fine-Tuned LLMs
by: Arachchige, Shashie Dilhara Batan, et al.
Published: (2025) -
Training on Fake Labels: Mitigating Label Leakage in Split Learning via Secure Dimension Transformation
by: Jiang, Yukun, et al.
Published: (2024) -
FLARE: A Wireless Side-Channel Fingerprinting Attack on Federated Learning
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)