Catastrophic Overfitting: A Potential Blessing in Disguise
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Mengnan, Zhang, Lihe, Kong, Yuqiu, Yin, Baocai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unveiling the Backdoor Mechanism Hidden Behind Catastrophic Overfitting in Fast Adversarial Training
by: Zhao, Mengnan, et al.
Published: (2026)
by: Zhao, Mengnan, et al.
Published: (2026)
CoreUnlearn: Rethinking Concept Unlearning through Disentangled Component-Level Erasure in Text-guided Diffusion Models
by: Zhao, Mengnan, et al.
Published: (2026)
by: Zhao, Mengnan, et al.
Published: (2026)
Mitigating Error Amplification in Fast Adversarial Training
by: Zhao, Mengnan, et al.
Published: (2026)
by: Zhao, Mengnan, et al.
Published: (2026)
Disguised Copyright Infringement of Latent Diffusion Models
by: Lu, Yiwei, et al.
Published: (2024)
by: Lu, Yiwei, et al.
Published: (2024)
The Surprising Harmfulness of Benign Overfitting for Adversarial Robustness
by: Hao, Yifan, et al.
Published: (2024)
by: Hao, Yifan, et al.
Published: (2024)
Separable Multi-Concept Erasure from Diffusion Models
by: Zhao, Mengnan, et al.
Published: (2024)
by: Zhao, Mengnan, et al.
Published: (2024)
Membership Inference Attacks Beyond Overfitting
by: Khalil, Mona, et al.
Published: (2025)
by: Khalil, Mona, et al.
Published: (2025)
COBRA: Catastrophic Bit-flip Reliability Analysis of State-Space Models
by: Das, Sanjay, et al.
Published: (2025)
by: Das, Sanjay, et al.
Published: (2025)
Hiding in Plain Sight: Disguising Data Stealing Attacks in Federated Learning
by: Garov, Kostadin, et al.
Published: (2023)
by: Garov, Kostadin, et al.
Published: (2023)
Jailbreaking the Non-Transferable Barrier via Test-Time Data Disguising
by: Xiang, Yongli, et al.
Published: (2025)
by: Xiang, Yongli, et al.
Published: (2025)
Channel-Level Semantic Perturbations: Unlearnable Examples for Diverse Training Paradigms
by: Wang, Bo, et al.
Published: (2026)
by: Wang, Bo, et al.
Published: (2026)
How Catastrophic is Your LLM? Certifying Risk in Conversation
by: Wang, Chengxiao, et al.
Published: (2025)
by: Wang, Chengxiao, et al.
Published: (2025)
A Novel GPT-Based Framework for Anomaly Detection in System Logs
by: Zhang, Zeng, et al.
Published: (2025)
by: Zhang, Zeng, et al.
Published: (2025)
BadSampler: Harnessing the Power of Catastrophic Forgetting to Poison Byzantine-robust Federated Learning
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
Private and Communication-Efficient Federated Learning based on Differentially Private Sketches
by: Zhang, Meifan, et al.
Published: (2024)
by: Zhang, Meifan, et al.
Published: (2024)
AttnDiff: Attention-based Differential Fingerprinting for Large Language Models
by: Zhang, Haobo, et al.
Published: (2026)
by: Zhang, Haobo, et al.
Published: (2026)
Differentially Private Optimization for Non-Decomposable Objective Functions
by: Kong, Weiwei, et al.
Published: (2023)
by: Kong, Weiwei, et al.
Published: (2023)
Metamorphic Malware Evolution: The Potential and Peril of Large Language Models
by: Madani, Pooria
Published: (2024)
by: Madani, Pooria
Published: (2024)
Data-centric NLP Backdoor Defense from the Lens of Memorization
by: Wang, Zhenting, et al.
Published: (2024)
by: Wang, Zhenting, et al.
Published: (2024)
Does Training with Synthetic Data Truly Protect Privacy?
by: Zhao, Yunpeng, et al.
Published: (2025)
by: Zhao, Yunpeng, et al.
Published: (2025)
MERLOT: A Distilled LLM-based Mixture-of-Experts Framework for Scalable Encrypted Traffic Classification
by: Chen, Yuxuan, et al.
Published: (2024)
by: Chen, Yuxuan, et al.
Published: (2024)
CLASP: Training-Free LLM-Assisted Source Code Watermarking via Semantic-Preserving Transformations
by: Xu, Rui, et al.
Published: (2025)
by: Xu, Rui, et al.
Published: (2025)
Asymmetric Bias in Text-to-Image Generation with Adversarial Attacks
by: Shahgir, Haz Sameen, et al.
Published: (2023)
by: Shahgir, Haz Sameen, et al.
Published: (2023)
PGD-Imp: Rethinking and Unleashing Potential of Classic PGD with Dual Strategies for Imperceptible Adversarial Attacks
by: Li, Jin, et al.
Published: (2024)
by: Li, Jin, et al.
Published: (2024)
Protecting Copyright of Medical Pre-trained Language Models: Training-Free Backdoor Model Watermarking
by: Kong, Cong, et al.
Published: (2024)
by: Kong, Cong, et al.
Published: (2024)
MambaITD: An Efficient Cross-Modal Mamba Network for Insider Threat Detection
by: Kong, Kaichuan, et al.
Published: (2025)
by: Kong, Kaichuan, et al.
Published: (2025)
Efficient and Verifiable Privacy-Preserving Convolutional Computation for CNN Inference with Untrusted Clouds
by: Lu, Jinyu, et al.
Published: (2025)
by: Lu, Jinyu, et al.
Published: (2025)
Enhancing Continual Learning for Software Vulnerability Prediction: Addressing Catastrophic Forgetting via Hybrid-Confidence-Aware Selective Replay for Temporal LLM Fine-Tuning
by: Dou, Xuhui, et al.
Published: (2026)
by: Dou, Xuhui, et al.
Published: (2026)
Catastrophic Cyber Capabilities Benchmark (3CB): Robustly Evaluating LLM Agent Cyber Offense Capabilities
by: Anurin, Andrey, et al.
Published: (2024)
by: Anurin, Andrey, et al.
Published: (2024)
Zero-Knowledge Proof-based Verifiable Decentralized Machine Learning in Communication Network: A Comprehensive Survey
by: Xing, Zhibo, et al.
Published: (2023)
by: Xing, Zhibo, et al.
Published: (2023)
DPIS: An Enhanced Mechanism for Differentially Private SGD with Importance Sampling
by: Wei, Jianxin, et al.
Published: (2022)
by: Wei, Jianxin, et al.
Published: (2022)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
by: Yin, Chenlong, et al.
Published: (2026)
by: Yin, Chenlong, et al.
Published: (2026)
A Channel-Triggered Backdoor Attack on Wireless Semantic Image Reconstruction
by: Wan, Jialin, et al.
Published: (2025)
by: Wan, Jialin, et al.
Published: (2025)
Revisiting Gradient Pruning: A Dual Realization for Defending against Gradient Attacks
by: Xue, Lulu, et al.
Published: (2024)
by: Xue, Lulu, et al.
Published: (2024)
Nemesis: Noise-randomized Encryption with Modular Efficiency and Secure Integration in Machine Learning Systems
by: Zhao, Dongfang
Published: (2024)
by: Zhao, Dongfang
Published: (2024)
Selective Pre-training for Private Fine-tuning
by: Yu, Da, et al.
Published: (2023)
by: Yu, Da, et al.
Published: (2023)
Towards Efficient Target-Level Machine Unlearning Based on Essential Graph
by: Xu, Heng, et al.
Published: (2024)
by: Xu, Heng, et al.
Published: (2024)
Spatiotemporal-Aware Bit-Flip Injection on DNN-based Advanced Driver Assistance Systems (extended version)
by: Zhao, Taibiao, et al.
Published: (2026)
by: Zhao, Taibiao, et al.
Published: (2026)
Insufficient Statistics Perturbation: Stable Estimators for Private Least Squares
by: Brown, Gavin, et al.
Published: (2024)
by: Brown, Gavin, et al.
Published: (2024)
Differentially Private Relational Learning with Entity-level Privacy Guarantees
by: Huang, Yinan, et al.
Published: (2025)
by: Huang, Yinan, et al.
Published: (2025)
Similar Items
-
Unveiling the Backdoor Mechanism Hidden Behind Catastrophic Overfitting in Fast Adversarial Training
by: Zhao, Mengnan, et al.
Published: (2026) -
CoreUnlearn: Rethinking Concept Unlearning through Disentangled Component-Level Erasure in Text-guided Diffusion Models
by: Zhao, Mengnan, et al.
Published: (2026) -
Mitigating Error Amplification in Fast Adversarial Training
by: Zhao, Mengnan, et al.
Published: (2026) -
Disguised Copyright Infringement of Latent Diffusion Models
by: Lu, Yiwei, et al.
Published: (2024) -
The Surprising Harmfulness of Benign Overfitting for Adversarial Robustness
by: Hao, Yifan, et al.
Published: (2024)