A Rusty Link in the AI Supply Chain: Detecting Evil Configurations in Model Repositories
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Ziqi, Fu, Qian, Ding, Junchen, Deng, Gelei, Liu, Yi, Li, Yuekang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TombRaider: Entering the Vault of History to Jailbreak Large Language Models
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
IllusionCAPTCHA: A CAPTCHA based on Visual Illusion
von: Ding, Ziqi, et al.
Veröffentlicht: (2025)
von: Ding, Ziqi, et al.
Veröffentlicht: (2025)
Membership Inference Attacks Against Video Large Language Models
von: Song, Wei, et al.
Veröffentlicht: (2026)
von: Song, Wei, et al.
Veröffentlicht: (2026)
Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation
von: Jin, David, et al.
Veröffentlicht: (2025)
von: Jin, David, et al.
Veröffentlicht: (2025)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
SNARE: Adaptive Scenario Synthesis for Eliciting Overeager Behavior in Coding Agents
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
Credential Leakage in LLM Agent Skills: A Large-Scale Empirical Study
von: Chen, Zhihao, et al.
Veröffentlicht: (2026)
von: Chen, Zhihao, et al.
Veröffentlicht: (2026)
Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
von: Qu, Yubin, et al.
Veröffentlicht: (2026)
Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content
von: Guo, Ruoqi, et al.
Veröffentlicht: (2026)
von: Guo, Ruoqi, et al.
Veröffentlicht: (2026)
Enhancing Model Defense Against Jailbreaks with Proactive Safety Reasoning
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
ai.txt: A Domain-Specific Language for Guiding AI Interactions with the Internet
von: Li, Yuekang, et al.
Veröffentlicht: (2025)
von: Li, Yuekang, et al.
Veröffentlicht: (2025)
Efficient Detection of Toxic Prompts in Large Language Models
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
Agentic AI for Autonomous Defense in Software Supply Chain Security: Beyond Provenance to Vulnerability Mitigation
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
Surveying the Operational Cybersecurity and Supply Chain Threat Landscape when Developing and Deploying AI Systems
von: Smith, Michael R, et al.
Veröffentlicht: (2025)
von: Smith, Michael R, et al.
Veröffentlicht: (2025)
Investigating Security Implications of Automatically Generated Code on the Software Supply Chain
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
Secret Stealing Attacks on Local LLM Fine-Tuning through Supply-Chain Model Code Backdoors
von: Li, Zi, et al.
Veröffentlicht: (2026)
von: Li, Zi, et al.
Veröffentlicht: (2026)
SecRepoBench: Benchmarking Code Agents for Secure Code Completion in Real-World Repositories
von: Shen, Chihao, et al.
Veröffentlicht: (2025)
von: Shen, Chihao, et al.
Veröffentlicht: (2025)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models
von: Ou, Haoran, et al.
Veröffentlicht: (2025)
von: Ou, Haoran, et al.
Veröffentlicht: (2025)
Image-Based Geolocation Using Large Vision-Language Models
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
A Framework for Agricultural Food Supply Chain using Blockchain
von: N, Sudarssan
Veröffentlicht: (2024)
von: N, Sudarssan
Veröffentlicht: (2024)
SOK: A Taxonomy of Attack Vectors and Defense Strategies for Agentic Supply Chain Runtime
von: Jiang, Xiaochong, et al.
Veröffentlicht: (2026)
von: Jiang, Xiaochong, et al.
Veröffentlicht: (2026)
ShellForge: Adversarial Co-Evolution of Webshell Generation and Multi-View Detection for Robust Webshell Defense
von: Ding, Yizhong
Veröffentlicht: (2026)
von: Ding, Yizhong
Veröffentlicht: (2026)
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
Identifying the Supply Chain of AI for Trustworthiness and Risk Management in Critical Applications
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense
von: Liu, Cheng, et al.
Veröffentlicht: (2026)
von: Liu, Cheng, et al.
Veröffentlicht: (2026)
Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
A Survey on Blockchain-based Supply Chain Finance with Progress and Future directions
von: Luo, Zhengdong
Veröffentlicht: (2024)
von: Luo, Zhengdong
Veröffentlicht: (2024)
Prompt Injection attack against LLM-integrated Applications
von: Liu, Yi, et al.
Veröffentlicht: (2023)
von: Liu, Yi, et al.
Veröffentlicht: (2023)
Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models
von: Li, Xi, et al.
Veröffentlicht: (2024)
von: Li, Xi, et al.
Veröffentlicht: (2024)
CoTSRF: Utilize Chain of Thought as Stealthy and Robust Fingerprint of Large Language Models
von: Ren, Zhenzhen, et al.
Veröffentlicht: (2025)
von: Ren, Zhenzhen, et al.
Veröffentlicht: (2025)
Who's the Evil Twin? Differential Auditing for Undesired Behavior
von: Balappanawar, Ishwar, et al.
Veröffentlicht: (2025)
von: Balappanawar, Ishwar, et al.
Veröffentlicht: (2025)
TEESlice: Protecting Sensitive Neural Network Models in Trusted Execution Environments When Attackers have Pre-Trained Models
von: Li, Ding, et al.
Veröffentlicht: (2024)
von: Li, Ding, et al.
Veröffentlicht: (2024)
AIAuditTrack: A Framework for AI Security system
von: Luo, Zixun, et al.
Veröffentlicht: (2025)
von: Luo, Zixun, et al.
Veröffentlicht: (2025)
Large Language Model Supply Chain: Open Problems From the Security Perspective
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TombRaider: Entering the Vault of History to Jailbreak Large Language Models
von: Ding, Junchen, et al.
Veröffentlicht: (2025) -
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024) -
Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems
von: Qu, Yubin, et al.
Veröffentlicht: (2026) -
IllusionCAPTCHA: A CAPTCHA based on Visual Illusion
von: Ding, Ziqi, et al.
Veröffentlicht: (2025) -
Membership Inference Attacks Against Video Large Language Models
von: Song, Wei, et al.
Veröffentlicht: (2026)