Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain
Fuente:
arXiv
Saved in:
| Main Authors: | Boisvert, Léo, Puri, Abhay, Evuru, Chandra Kiran Reddy, Sepahvand, Nazanin, Chapados, Nicolas, Cappart, Quentin, Lacoste, Alexandre, Dvijotham, Krishnamurthy Dj, Drouin, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DoomArena: A framework for Testing AI Agents Against Evolving Security Threats
by: Boisvert, Leo, et al.
Published: (2025)
by: Boisvert, Leo, et al.
Published: (2025)
WorkArena++: Towards Compositional Planning and Reasoning-based Common Knowledge Work Tasks
by: Boisvert, Léo, et al.
Published: (2024)
by: Boisvert, Léo, et al.
Published: (2024)
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
by: Bhagwatkar, Rishika, et al.
Published: (2025)
by: Bhagwatkar, Rishika, et al.
Published: (2025)
Towards a Generic Representation of Combinatorial Problems for Learning-Based Approaches
by: Boisvert, Léo, et al.
Published: (2024)
by: Boisvert, Léo, et al.
Published: (2024)
WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
by: Drouin, Alexandre, et al.
Published: (2024)
by: Drouin, Alexandre, et al.
Published: (2024)
Keeping up with dynamic attackers: Certifying robustness to adaptive online data poisoning
by: Bose, Avinandan, et al.
Published: (2025)
by: Bose, Avinandan, et al.
Published: (2025)
The BrowserGym Ecosystem for Web Agent Research
by: De Chezelles, Thibault Le Sellier, et al.
Published: (2024)
by: De Chezelles, Thibault Le Sellier, et al.
Published: (2024)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
InsightBench: Evaluating Business Analytics Agents Through Multi-Step Insight Generation
by: Sahu, Gaurav, et al.
Published: (2024)
by: Sahu, Gaurav, et al.
Published: (2024)
TACTiS-2: Better, Faster, Simpler Attentional Copulas for Multivariate Time Series
by: Ashok, Arjun, et al.
Published: (2023)
by: Ashok, Arjun, et al.
Published: (2023)
Adaptive Diffusion Denoised Smoothing : Certified Robustness via Randomized Smoothing with Differentially Private Guided Denoising Diffusion
by: Shpilevskiy, Frederick, et al.
Published: (2025)
by: Shpilevskiy, Frederick, et al.
Published: (2025)
RECAP: Retrieval-Augmented Audio Captioning
by: Ghosh, Sreyan, et al.
Published: (2023)
by: Ghosh, Sreyan, et al.
Published: (2023)
Efficient Error Certification for Physics-Informed Neural Networks
by: Eiras, Francisco, et al.
Published: (2023)
by: Eiras, Francisco, et al.
Published: (2023)
No, of Course I Can! Deeper Fine-Tuning Attacks That Bypass Token-Level Safety Mechanisms
by: Kazdan, Joshua, et al.
Published: (2025)
by: Kazdan, Joshua, et al.
Published: (2025)
Context is Key: A Benchmark for Forecasting with Essential Textual Information
by: Williams, Andrew Robert, et al.
Published: (2024)
by: Williams, Andrew Robert, et al.
Published: (2024)
CoDa: Constrained Generation based Data Augmentation for Low-Resource NLP
by: Evuru, Chandra Kiran Reddy, et al.
Published: (2024)
by: Evuru, Chandra Kiran Reddy, et al.
Published: (2024)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
by: Miculicich, Lesly, et al.
Published: (2025)
by: Miculicich, Lesly, et al.
Published: (2025)
LitLLM: A Toolkit for Scientific Literature Review
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
ReCLAP: Improving Zero Shot Audio Classification by Describing Sounds
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
CausalArmor: Efficient Indirect Prompt Injection Guardrails via Causal Attribution
by: Kim, Minbeom, et al.
Published: (2026)
by: Kim, Minbeom, et al.
Published: (2026)
FocusAgent: Simple Yet Effective Ways of Trimming the Large Context of Web Agents
by: Kerboua, Imene, et al.
Published: (2025)
by: Kerboua, Imene, et al.
Published: (2025)
Achieving the Tightest Relaxation of Sigmoids for Formal Verification
by: Chevalier, Samuel, et al.
Published: (2024)
by: Chevalier, Samuel, et al.
Published: (2024)
A Taxonomy for Design and Evaluation of Prompt-Based Natural Language Explanations
by: Nejadgholi, Isar, et al.
Published: (2025)
by: Nejadgholi, Isar, et al.
Published: (2025)
Sleeper Cell: Injecting Latent Malice Temporal Backdoors into Tool-Using LLMs
by: Pallakonda, Bhanu, et al.
Published: (2026)
by: Pallakonda, Bhanu, et al.
Published: (2026)
ASPIRE: Language-Guided Data Augmentation for Improving Robustness Against Spurious Correlations
by: Ghosh, Sreyan, et al.
Published: (2023)
by: Ghosh, Sreyan, et al.
Published: (2023)
Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
A Closer Look at the Limitations of Instruction Tuning
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
ProSE: Diffusion Priors for Speech Enhancement
by: Kumar, Sonal, et al.
Published: (2025)
by: Kumar, Sonal, et al.
Published: (2025)
Norm-Bounded Low-Rank Adaptation
by: Wang, Ruigang, et al.
Published: (2025)
by: Wang, Ruigang, et al.
Published: (2025)
Monotone, Bi-Lipschitz, and Polyak-Lojasiewicz Networks
by: Wang, Ruigang, et al.
Published: (2024)
by: Wang, Ruigang, et al.
Published: (2024)
Global Rewards in Multi-Agent Deep Reinforcement Learning for Autonomous Mobility on Demand Systems
by: Hoppe, Heiko, et al.
Published: (2023)
by: Hoppe, Heiko, et al.
Published: (2023)
Deep Learning for Data-Driven Districting-and-Routing
by: Ferraz, Arthur, et al.
Published: (2024)
by: Ferraz, Arthur, et al.
Published: (2024)
Influence of Zn 2+ and Oxygen Supply on Malic Acid Production and Growth of Aspergillus oryzae
by: Lukas Hartmann, et al.
Published: (2026)
by: Lukas Hartmann, et al.
Published: (2026)
Beyond Naïve Prompting: Strategies for Improved Context-aided Forecasting with LLMs
by: Ashok, Arjun, et al.
Published: (2025)
by: Ashok, Arjun, et al.
Published: (2025)
Data Selection for Transfer Unlearning
by: Sepahvand, Nazanin Mohammadi, et al.
Published: (2024)
by: Sepahvand, Nazanin Mohammadi, et al.
Published: (2024)
GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
MARCO: A Memory-Augmented Reinforcement Framework for Combinatorial Optimization
by: Garmendia, Andoni I., et al.
Published: (2024)
by: Garmendia, Andoni I., et al.
Published: (2024)
An Evaluation of "Choice" as a Selection Tool in the Field of Western History.
by: Boisvert, Marianne
Published: (1977)
by: Boisvert, Marianne
Published: (1977)
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
by: Madhusudhan, Nishanth, et al.
Published: (2026)
by: Madhusudhan, Nishanth, et al.
Published: (2026)
Similar Items
-
DoomArena: A framework for Testing AI Agents Against Evolving Security Threats
by: Boisvert, Leo, et al.
Published: (2025) -
WorkArena++: Towards Compositional Planning and Reasoning-based Common Knowledge Work Tasks
by: Boisvert, Léo, et al.
Published: (2024) -
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
by: Bhagwatkar, Rishika, et al.
Published: (2025) -
Towards a Generic Representation of Combinatorial Problems for Learning-Based Approaches
by: Boisvert, Léo, et al.
Published: (2024) -
WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
by: Drouin, Alexandre, et al.
Published: (2024)