A Biosecurity Agent for Lifecycle LLM Biosecurity Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meng, Meiyin, Zhang, Zaixi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generative AI for Biosciences: Emerging Threats and Roadmap to Biosecurity
von: Zhang, Zaixi, et al.
Veröffentlicht: (2025)
von: Zhang, Zaixi, et al.
Veröffentlicht: (2025)
Know Your Scientist: KYC as Biosecurity Infrastructure
von: Feldman, Jonathan, et al.
Veröffentlicht: (2026)
von: Feldman, Jonathan, et al.
Veröffentlicht: (2026)
Biosecurity-Aware AI: Agentic Risk Auditing of Soft Prompt Attacks on ESM-Based Variant Predictors
von: Zhan, Huixin
Veröffentlicht: (2025)
von: Zhan, Huixin
Veröffentlicht: (2025)
SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment
von: Lin, Xixun, et al.
Veröffentlicht: (2026)
von: Lin, Xixun, et al.
Veröffentlicht: (2026)
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
von: Zhang, Yixiang, et al.
Veröffentlicht: (2026)
von: Zhang, Yixiang, et al.
Veröffentlicht: (2026)
Automating Function-Level TARA for Automotive Full-Lifecycle Security
von: Yang, Yuqiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuqiao, et al.
Veröffentlicht: (2025)
GeneBreaker: Jailbreak Attacks against DNA Language Models with Pathogenicity Guidance
von: Zhang, Zaixi, et al.
Veröffentlicht: (2025)
von: Zhang, Zaixi, et al.
Veröffentlicht: (2025)
BioRefusalAudit: Auditing Biosecurity Refusal Depth Using General and Domain-Fine-Tuned Sparse Autoencoders
von: DeLeeuw, Caleb
Veröffentlicht: (2026)
von: DeLeeuw, Caleb
Veröffentlicht: (2026)
When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
von: Luo, Mingyu, et al.
Veröffentlicht: (2026)
von: Luo, Mingyu, et al.
Veröffentlicht: (2026)
DomainDynamics: Lifecycle-Aware Risk Timeline Construction for Domain Names
von: Chiba, Daiki, et al.
Veröffentlicht: (2024)
von: Chiba, Daiki, et al.
Veröffentlicht: (2024)
Security and Privacy Measurement on Chinese Consumer IoT Traffic based on Device Lifecycle
von: Jin, Chenghua, et al.
Veröffentlicht: (2025)
von: Jin, Chenghua, et al.
Veröffentlicht: (2025)
GuardianPWA: Enhancing Security Throughout the Progressive Web App Installation Lifecycle
von: Wang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Wang, Mengxiao, et al.
Veröffentlicht: (2025)
On Securing the Software Development Lifecycle in IoT RISC-V Trusted Execution Environments
von: Wilde, Annika, et al.
Veröffentlicht: (2026)
von: Wilde, Annika, et al.
Veröffentlicht: (2026)
Observable Channels, Not Just Storage: Evaluating Privacy Leakage in LLM Agent Pipelines
von: Huang, Tao, et al.
Veröffentlicht: (2026)
von: Huang, Tao, et al.
Veröffentlicht: (2026)
Atlas: A Framework for ML Lifecycle Provenance & Transparency
von: Spoczynski, Marcin, et al.
Veröffentlicht: (2025)
von: Spoczynski, Marcin, et al.
Veröffentlicht: (2025)
Agent Safety Alignment via Reinforcement Learning
von: Sha, Zeyang, et al.
Veröffentlicht: (2025)
von: Sha, Zeyang, et al.
Veröffentlicht: (2025)
PriMod4AI: Lifecycle-Aware Privacy Threat Modeling for AI Systems using LLM
von: Savaliya, Gautam, et al.
Veröffentlicht: (2026)
von: Savaliya, Gautam, et al.
Veröffentlicht: (2026)
SecMLOps: A Comprehensive Framework for Integrating Security Throughout the MLOps Lifecycle
von: Zhang, Xinrui, et al.
Veröffentlicht: (2026)
von: Zhang, Xinrui, et al.
Veröffentlicht: (2026)
Alleviating the Fear of Losing Alignment in LLM Fine-tuning
von: Yang, Kang, et al.
Veröffentlicht: (2025)
von: Yang, Kang, et al.
Veröffentlicht: (2025)
HackerSignal: A Large-Scale Multi-Source Dataset Linking Hacker Community Discourse to the CVE Vulnerability Lifecycle
von: Ampel, Benjamin M., et al.
Veröffentlicht: (2026)
von: Ampel, Benjamin M., et al.
Veröffentlicht: (2026)
Watermarking LLM Agent Trajectories
von: Meng, Wenlong, et al.
Veröffentlicht: (2026)
von: Meng, Wenlong, et al.
Veröffentlicht: (2026)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
von: Kong, Dezhang, et al.
Veröffentlicht: (2025)
von: Kong, Dezhang, et al.
Veröffentlicht: (2025)
Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
HarmRLVR: Weaponizing Verifiable Rewards for Harmful LLM Alignment
von: Liu, Yuexiao, et al.
Veröffentlicht: (2025)
von: Liu, Yuexiao, et al.
Veröffentlicht: (2025)
"Hello, is this Anna?": Unpacking the Lifecycle of Pig-Butchering Scams
von: Oak, Rajvardhan, et al.
Veröffentlicht: (2025)
von: Oak, Rajvardhan, et al.
Veröffentlicht: (2025)
ParikkhaChain: Blockchain-Based Result Processing and Privacy-Preserving Academic Record Management for the Complete Examination Lifecycle
von: Momin, Rabib Jahin Ibn, et al.
Veröffentlicht: (2026)
von: Momin, Rabib Jahin Ibn, et al.
Veröffentlicht: (2026)
Securing LLM Agents Need Intent-to-Execution Integrity
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
The Biosecurity Individual
von: Offizier, Frederike
Veröffentlicht: (2024)
von: Offizier, Frederike
Veröffentlicht: (2024)
When Skills Lie: Hidden-Comment Injection in LLM Agents
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
MemoPhishAgent: Memory-Augmented Multi-Modal LLM Agent for Phishing URL Detection
von: Chen, Xuan, et al.
Veröffentlicht: (2026)
von: Chen, Xuan, et al.
Veröffentlicht: (2026)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
von: Wang, Peiran, et al.
Veröffentlicht: (2026)
von: Wang, Peiran, et al.
Veröffentlicht: (2026)
PentestAgent: Incorporating LLM Agents to Automated Penetration Testing
von: Shen, Xiangmin, et al.
Veröffentlicht: (2024)
von: Shen, Xiangmin, et al.
Veröffentlicht: (2024)
Ghost in the Agent: Redefining Information Flow Tracking for LLM Agents
von: Cai, Yuandao, et al.
Veröffentlicht: (2026)
von: Cai, Yuandao, et al.
Veröffentlicht: (2026)
Exposing LLM User Privacy via Traffic Fingerprint Analysis: A Study of Privacy Risks in LLM Agent Interactions
von: Zhang, Yixiang, et al.
Veröffentlicht: (2025)
von: Zhang, Yixiang, et al.
Veröffentlicht: (2025)
Cognitive Control Architecture (CCA): A Lifecycle Supervision Framework for Robustly Aligned AI Agents
von: Liang, Zhibo, et al.
Veröffentlicht: (2025)
von: Liang, Zhibo, et al.
Veröffentlicht: (2025)
AgentGuard: An Attribute-Based Access Control Framework for Tool-Use LLM-Based Agent
von: Luo, Jiaqi, et al.
Veröffentlicht: (2026)
von: Luo, Jiaqi, et al.
Veröffentlicht: (2026)
Mining Characteristics of Vulnerable Smart Contracts Across Lifecycle Stages
von: Peng, Hongli, et al.
Veröffentlicht: (2025)
von: Peng, Hongli, et al.
Veröffentlicht: (2025)
Out of Sight, Still at Risk: The Lifecycle of Transitive Vulnerabilities in Maven
von: Przymus, Piotr, et al.
Veröffentlicht: (2025)
von: Przymus, Piotr, et al.
Veröffentlicht: (2025)
An LLM Agent-based Framework for Whaling Countermeasures
von: Miyamoto, Daisuke, et al.
Veröffentlicht: (2026)
von: Miyamoto, Daisuke, et al.
Veröffentlicht: (2026)
Enhancing Software Supply Chain Resilience: Strategy For Mitigating Software Supply Chain Security Risks And Ensuring Security Continuity In Development Lifecycle
von: Akinsola, Ahmed, et al.
Veröffentlicht: (2024)
von: Akinsola, Ahmed, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Generative AI for Biosciences: Emerging Threats and Roadmap to Biosecurity
von: Zhang, Zaixi, et al.
Veröffentlicht: (2025) -
Know Your Scientist: KYC as Biosecurity Infrastructure
von: Feldman, Jonathan, et al.
Veröffentlicht: (2026) -
Biosecurity-Aware AI: Agentic Risk Auditing of Soft Prompt Attacks on ESM-Based Variant Predictors
von: Zhan, Huixin
Veröffentlicht: (2025) -
SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment
von: Lin, Xixun, et al.
Veröffentlicht: (2026) -
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
von: Zhang, Yixiang, et al.
Veröffentlicht: (2026)