Guardado en:
| Autores principales: | Fogel, Ariel, Hofman, Omer, Cohen, Eilon, Vainshtein, Roman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.04653 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
When Scanners Lie: Evaluator Instability in LLM Red-Teaming
por: Erez, Lidor, et al.
Publicado: (2026)
por: Erez, Lidor, et al.
Publicado: (2026)
Insights and Current Gaps in Open-Source LLM Vulnerability Scanners: A Comparative Analysis
por: Brokman, Jonathan, et al.
Publicado: (2024)
por: Brokman, Jonathan, et al.
Publicado: (2024)
Compromising Embodied Agents with Contextual Backdoor Attacks
por: Liu, Aishan, et al.
Publicado: (2024)
por: Liu, Aishan, et al.
Publicado: (2024)
Erased but Not Forgotten: How Backdoors Compromise Concept Erasure
por: Braun, Tobias, et al.
Publicado: (2025)
por: Braun, Tobias, et al.
Publicado: (2025)
DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense
por: You, Ziyang, et al.
Publicado: (2026)
por: You, Ziyang, et al.
Publicado: (2026)
Model Supply Chain Poisoning: Backdooring Pre-trained Models via Embedding Indistinguishability
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Adversarial Intent is a Latent Variable: Stateful Trust Inference for Securing Multimodal Agentic RAG
por: Singh, Inderjeet, et al.
Publicado: (2026)
por: Singh, Inderjeet, et al.
Publicado: (2026)
Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain
por: Boisvert, Léo, et al.
Publicado: (2025)
por: Boisvert, Léo, et al.
Publicado: (2025)
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
por: Xiang, Zhen, et al.
Publicado: (2024)
por: Xiang, Zhen, et al.
Publicado: (2024)
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
por: Zhang, Boyang, et al.
Publicado: (2024)
por: Zhang, Boyang, et al.
Publicado: (2024)
ChatBug: A Common Vulnerability of Aligned LLMs Induced by Chat Templates
por: Jiang, Fengqing, et al.
Publicado: (2024)
por: Jiang, Fengqing, et al.
Publicado: (2024)
STAR: Detecting Inference-time Backdoors in LLM Reasoning via State-Transition Amplification Ratio
por: Park, Seong-Gyu, et al.
Publicado: (2026)
por: Park, Seong-Gyu, et al.
Publicado: (2026)
MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience Retrieval
por: Srivastava, Saksham Sahai, et al.
Publicado: (2025)
por: Srivastava, Saksham Sahai, et al.
Publicado: (2025)
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs
por: Guo, Zhen, et al.
Publicado: (2025)
por: Guo, Zhen, et al.
Publicado: (2025)
Cross-LLM Generalization of Behavioral Backdoor Detection in AI Agent Supply Chains
por: Sanna, Arun Chowdary
Publicado: (2025)
por: Sanna, Arun Chowdary
Publicado: (2025)
Detecting Backdoor Attacks via Similarity in Semantic Communication Systems
por: Wei, Ziyang, et al.
Publicado: (2025)
por: Wei, Ziyang, et al.
Publicado: (2025)
Machine Learning Models Have a Supply Chain Problem
por: Meiklejohn, Sarah, et al.
Publicado: (2025)
por: Meiklejohn, Sarah, et al.
Publicado: (2025)
MRMMIA: Membership Inference Attacks on Memory in Chat Agents
por: Chen, Kai, et al.
Publicado: (2026)
por: Chen, Kai, et al.
Publicado: (2026)
Semi-Supervised Supply Chain Fraud Detection with Unsupervised Pre-Filtering
por: Moradi, Fatemeh, et al.
Publicado: (2025)
por: Moradi, Fatemeh, et al.
Publicado: (2025)
Detection of Compromised Functions in a Serverless Cloud Environment
por: Lavi, Danielle, et al.
Publicado: (2024)
por: Lavi, Danielle, et al.
Publicado: (2024)
Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
por: Wen, Yuxin, et al.
Publicado: (2024)
por: Wen, Yuxin, et al.
Publicado: (2024)
Fake or Compromised? Making Sense of Malicious Clients in Federated Learning
por: Mozaffari, Hamid, et al.
Publicado: (2024)
por: Mozaffari, Hamid, et al.
Publicado: (2024)
BoBa: Boosting Backdoor Detection through Data Distribution Inference in Federated Learning
por: Jiang, Zhengyuan, et al.
Publicado: (2024)
por: Jiang, Zhengyuan, et al.
Publicado: (2024)
FHE-Agent: Automating CKKS Configuration for Practical Encrypted Inference via an LLM-Guided Agentic Framework
por: Xu, Nuo, et al.
Publicado: (2025)
por: Xu, Nuo, et al.
Publicado: (2025)
End-to-End Anti-Backdoor Learning on Images and Time Series
por: Jiang, Yujing, et al.
Publicado: (2024)
por: Jiang, Yujing, et al.
Publicado: (2024)
SoK: Reducing the Vulnerability of Fine-tuned Language Models to Membership Inference Attacks
por: Amit, Guy, et al.
Publicado: (2024)
por: Amit, Guy, et al.
Publicado: (2024)
Detecting Compromised IoT Devices Using Autoencoders with Sequential Hypothesis Testing
por: Mainuddin, Md, et al.
Publicado: (2024)
por: Mainuddin, Md, et al.
Publicado: (2024)
Authority Backdoor: A Certifiable Backdoor Mechanism for Authoring DNNs
por: Yang, Han, et al.
Publicado: (2025)
por: Yang, Han, et al.
Publicado: (2025)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
por: Wu, Baoyuan, et al.
Publicado: (2024)
por: Wu, Baoyuan, et al.
Publicado: (2024)
BadTemplate: A Training-Free Backdoor Attack via Chat Template Against Large Language Models
por: Wang, Zihan, et al.
Publicado: (2026)
por: Wang, Zihan, et al.
Publicado: (2026)
DP-Dueling: Learning from Preference Feedback without Compromising User Privacy
por: Saha, Aadirupa, et al.
Publicado: (2024)
por: Saha, Aadirupa, et al.
Publicado: (2024)
Publishing Neural Networks in Drug Discovery Might Compromise Training Data Privacy
por: Krüger, Fabian P., et al.
Publicado: (2024)
por: Krüger, Fabian P., et al.
Publicado: (2024)
Backdoor Learning Curves: Explaining Backdoor Poisoning Beyond Influence Functions
por: Cinà, Antonio Emanuele, et al.
Publicado: (2021)
por: Cinà, Antonio Emanuele, et al.
Publicado: (2021)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
por: Draguns, Andis, et al.
Publicado: (2024)
por: Draguns, Andis, et al.
Publicado: (2024)
Hardware-Triggered Backdoors
por: Möller, Jonas, et al.
Publicado: (2026)
por: Möller, Jonas, et al.
Publicado: (2026)
Whispers in the Machine: Confidentiality in Agentic Systems
por: Evertz, Jonathan, et al.
Publicado: (2024)
por: Evertz, Jonathan, et al.
Publicado: (2024)
Hiding Backdoors within Event Sequence Data via Poisoning Attacks
por: Ermilova, Alina, et al.
Publicado: (2023)
por: Ermilova, Alina, et al.
Publicado: (2023)
DeDe: Detecting Backdoor Samples for SSL Encoders via Decoders
por: Hou, Sizai, et al.
Publicado: (2024)
por: Hou, Sizai, et al.
Publicado: (2024)
A Survey of Learning-Based Intrusion Detection Systems for In-Vehicle Network
por: Althunayyan, Muzun, et al.
Publicado: (2025)
por: Althunayyan, Muzun, et al.
Publicado: (2025)
Fusing Pruned and Backdoored Models: Optimal Transport-based Data-free Backdoor Mitigation
por: Lin, Weilin, et al.
Publicado: (2024)
por: Lin, Weilin, et al.
Publicado: (2024)
Ejemplares similares
-
When Scanners Lie: Evaluator Instability in LLM Red-Teaming
por: Erez, Lidor, et al.
Publicado: (2026) -
Insights and Current Gaps in Open-Source LLM Vulnerability Scanners: A Comparative Analysis
por: Brokman, Jonathan, et al.
Publicado: (2024) -
Compromising Embodied Agents with Contextual Backdoor Attacks
por: Liu, Aishan, et al.
Publicado: (2024) -
Erased but Not Forgotten: How Backdoors Compromise Concept Erasure
por: Braun, Tobias, et al.
Publicado: (2025) -
DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense
por: You, Ziyang, et al.
Publicado: (2026)