Trojan Cleansing with Neural Collapse
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Xihe, Fields, Greg, Jandali, Yaman, Javidi, Tara, Koushanfar, Farinaz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MergeGuard: Efficient Thwarting of Trojan Attacks in Machine Learning Models
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2025)
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2025)
Optimizing Privacy-Preserving Primitives to Support LLM-Scale Applications
by: Jandali, Yaman, et al.
Published: (2025)
by: Jandali, Yaman, et al.
Published: (2025)
SWaRL: Safeguard Code Watermarking via Reinforcement Learning
by: Javidnia, Neusha, et al.
Published: (2026)
by: Javidnia, Neusha, et al.
Published: (2026)
Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign
by: Zhang, Ruisi, et al.
Published: (2025)
by: Zhang, Ruisi, et al.
Published: (2025)
Trojans in Artificial Intelligence (TrojAI) Final Report
by: Reese, Kristopher W., et al.
Published: (2026)
by: Reese, Kristopher W., et al.
Published: (2026)
SSL-Cleanse: Trojan Detection and Mitigation in Self-Supervised Learning
by: Zheng, Mengxin, et al.
Published: (2023)
by: Zheng, Mengxin, et al.
Published: (2023)
LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning
by: Spracklen, Joseph, et al.
Published: (2026)
by: Spracklen, Joseph, et al.
Published: (2026)
Props for Machine-Learning Security
by: Juels, Ari, et al.
Published: (2024)
by: Juels, Ari, et al.
Published: (2024)
EmMark: Robust Watermarks for IP Protection of Embedded Quantized Large Language Models
by: Zhang, Ruisi, et al.
Published: (2024)
by: Zhang, Ruisi, et al.
Published: (2024)
Watermarking Large Language Models and the Generated Content: Opportunities and Challenges
by: Zhang, Ruisi, et al.
Published: (2024)
by: Zhang, Ruisi, et al.
Published: (2024)
Token-Specific Watermarking with Enhanced Detectability and Semantic Coherence for Large Language Models
by: Huo, Mingjia, et al.
Published: (2024)
by: Huo, Mingjia, et al.
Published: (2024)
Diff-Cleanse: Identifying and Mitigating Backdoor Attacks in Diffusion Models
by: Hao, Jiang, et al.
Published: (2024)
by: Hao, Jiang, et al.
Published: (2024)
DROP: Poison Dilution via Knowledge Distillation for Federated Learning
by: Syros, Georgios, et al.
Published: (2025)
by: Syros, Georgios, et al.
Published: (2025)
ICtoken: An NFT for Hardware IP Protection
by: Balla, Shashank, et al.
Published: (2024)
by: Balla, Shashank, et al.
Published: (2024)
EPA: Neural Collapse Inspired Robust Out-of-Distribution Detector
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Game of Trojans: Adaptive Adversaries Against Output-based Trojaned-Model Detectors
by: Sahabandu, Dinuka, et al.
Published: (2024)
by: Sahabandu, Dinuka, et al.
Published: (2024)
Are You Using Reliable Graph Prompts? Trojan Prompt Attacks on Graph Neural Networks
by: Lin, Minhua, et al.
Published: (2024)
by: Lin, Minhua, et al.
Published: (2024)
Zero-Knowledge Proof Frameworks: A Systematic Survey
by: Sheybani, Nojan, et al.
Published: (2025)
by: Sheybani, Nojan, et al.
Published: (2025)
TrojanPuzzle: Covertly Poisoning Code-Suggestion Models
by: Aghakhani, Hojjat, et al.
Published: (2023)
by: Aghakhani, Hojjat, et al.
Published: (2023)
TrojanPraise: Jailbreak LLMs via Benign Fine-Tuning
by: Xie, Zhixin, et al.
Published: (2026)
by: Xie, Zhixin, et al.
Published: (2026)
Trojan Horse Hunt in Time Series Forecasting for Space Operations
by: Kotowski, Krzysztof, et al.
Published: (2025)
by: Kotowski, Krzysztof, et al.
Published: (2025)
Evil from Within: Machine Learning Backdoors through Hardware Trojans
by: Warnecke, Alexander, et al.
Published: (2023)
by: Warnecke, Alexander, et al.
Published: (2023)
TROJAN-GUARD: Hardware Trojans Detection Using GNN in RTL Designs
by: Thorat, Kiran, et al.
Published: (2025)
by: Thorat, Kiran, et al.
Published: (2025)
PAC to the Future: Zero-Knowledge Proofs of PAC Private Systems
by: Repetto, Guilhem, et al.
Published: (2026)
by: Repetto, Guilhem, et al.
Published: (2026)
Automated Physical Design Watermarking Leveraging Graph Neural Networks
by: Zhang, Ruisi, et al.
Published: (2024)
by: Zhang, Ruisi, et al.
Published: (2024)
Trojan horse hunt in deep forecasting models: Insights from the European Space Agency competition
by: Kotowski, Krzysztof, et al.
Published: (2026)
by: Kotowski, Krzysztof, et al.
Published: (2026)
CryptoGen: Secure Transformer Generation with Encrypted KV-Cache Reuse
by: Zhang, Hedong, et al.
Published: (2026)
by: Zhang, Hedong, et al.
Published: (2026)
On Trojan Signatures in Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
TrojanDam: Detection-Free Backdoor Defense in Federated Learning through Proactive Model Robustification utilizing OOD Data
by: Dai, Yanbo, et al.
Published: (2025)
by: Dai, Yanbo, et al.
Published: (2025)
Evasive Hardware Trojan through Adversarial Power Trace
by: Omidi, Behnam, et al.
Published: (2024)
by: Omidi, Behnam, et al.
Published: (2024)
An AI Architecture with the Capability to Classify and Explain Hardware Trojans
by: Whitten, Paul, et al.
Published: (2024)
by: Whitten, Paul, et al.
Published: (2024)
REMARK-LLM: A Robust and Efficient Watermarking Framework for Generative Large Language Models
by: Zhang, Ruisi, et al.
Published: (2023)
by: Zhang, Ruisi, et al.
Published: (2023)
Uncertainty-Aware Hardware Trojan Detection Using Multimodal Deep Learning
by: Vishwakarma, Rahul, et al.
Published: (2024)
by: Vishwakarma, Rahul, et al.
Published: (2024)
If You Don't Understand It, Don't Use It: Eliminating Trojans with Filters Between Layers
by: Hernandez, Adriano
Published: (2024)
by: Hernandez, Adriano
Published: (2024)
SpectralGuard: Detecting Memory Collapse Attacks in State Space Models
by: Bonetto, Davi
Published: (2026)
by: Bonetto, Davi
Published: (2026)
Neural Trojans
by: Liu, Yuntao, et al.
Published: (2017)
by: Liu, Yuntao, et al.
Published: (2017)
ICMarks: A Robust Watermarking Framework for Integrated Circuit Physical Design IP Protection
by: Zhang, Ruisi, et al.
Published: (2024)
by: Zhang, Ruisi, et al.
Published: (2024)
AttestLLM: Efficient Attestation Framework for Billion-scale On-device LLMs
by: Zhang, Ruisi, et al.
Published: (2025)
by: Zhang, Ruisi, et al.
Published: (2025)
TrojanForge: Generating Adversarial Hardware Trojan Examples Using Reinforcement Learning
by: Sarihi, Amin, et al.
Published: (2024)
by: Sarihi, Amin, et al.
Published: (2024)
TroLLoc: Logic Locking and Layout Hardening for IC Security Closure against Hardware Trojans
by: Wang, Fangzhou, et al.
Published: (2024)
by: Wang, Fangzhou, et al.
Published: (2024)
Similar Items
-
MergeGuard: Efficient Thwarting of Trojan Attacks in Machine Learning Models
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2025) -
Optimizing Privacy-Preserving Primitives to Support LLM-Scale Applications
by: Jandali, Yaman, et al.
Published: (2025) -
SWaRL: Safeguard Code Watermarking via Reinforcement Learning
by: Javidnia, Neusha, et al.
Published: (2026) -
Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign
by: Zhang, Ruisi, et al.
Published: (2025) -
Trojans in Artificial Intelligence (TrojAI) Final Report
by: Reese, Kristopher W., et al.
Published: (2026)