Enforcing Cybersecurity Constraints for LLM-driven Robot Agents for Online Transactions
Fuente:
arXiv
Saved in:
| Main Authors: | Shah, Shraddha Pradipbhai, Deshpande, Aditya Vilas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
by: Lin, Justin W., et al.
Published: (2025)
by: Lin, Justin W., et al.
Published: (2025)
Frontier AI's Impact on the Cybersecurity Landscape
by: Potter, Yujin, et al.
Published: (2025)
by: Potter, Yujin, et al.
Published: (2025)
A Review of Cybersecurity Incidents in the Food and Agriculture Sector
by: Kulkarni, Ajay, et al.
Published: (2024)
by: Kulkarni, Ajay, et al.
Published: (2024)
Artificial Intelligence in Cybersecurity: Building Resilient Cyber Diplomacy Frameworks
by: Stoltz, Michael
Published: (2024)
by: Stoltz, Michael
Published: (2024)
BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Reviewers?
by: Jiang, Fengqing, et al.
Published: (2025)
by: Jiang, Fengqing, et al.
Published: (2025)
Robustness and Cybersecurity in the EU Artificial Intelligence Act
by: Nolte, Henrik, et al.
Published: (2025)
by: Nolte, Henrik, et al.
Published: (2025)
The New Frontier of Cybersecurity: Emerging Threats and Innovations
by: Dave, Daksh, et al.
Published: (2023)
by: Dave, Daksh, et al.
Published: (2023)
AI-Driven Cyber Threat Intelligence Automation
by: Shah, Shrit, et al.
Published: (2024)
by: Shah, Shrit, et al.
Published: (2024)
To Patch or Not to Patch: Motivations, Challenges, and Implications for Cybersecurity
by: Nurse, Jason R. C.
Published: (2025)
by: Nurse, Jason R. C.
Published: (2025)
Game Theory Meets LLM and Agentic AI: Reimagining Cybersecurity for the Age of Intelligent Threats
by: Zhu, Quanyan
Published: (2025)
by: Zhu, Quanyan
Published: (2025)
Nuclear Deployed: Analyzing Catastrophic Risks in Decision-making of Autonomous LLM Agents
by: Xu, Rongwu, et al.
Published: (2025)
by: Xu, Rongwu, et al.
Published: (2025)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
by: Kulothungan, Vikram
Published: (2025)
by: Kulothungan, Vikram
Published: (2025)
Detecting Verbatim LLM Copy-Paste in Homework
by: Aiersilan, Aizierjiang
Published: (2026)
by: Aiersilan, Aizierjiang
Published: (2026)
RedTeamLLM: an Agentic AI framework for offensive security
by: Challita, Brian, et al.
Published: (2025)
by: Challita, Brian, et al.
Published: (2025)
How Well Can LLM Agents Simulate End-User Security and Privacy Attitudes and Behaviors?
by: Li, Yuxuan, et al.
Published: (2026)
by: Li, Yuxuan, et al.
Published: (2026)
Guardrails for trust, safety, and ethical development and deployment of Large Language Models (LLM)
by: Biswas, Anjanava, et al.
Published: (2026)
by: Biswas, Anjanava, et al.
Published: (2026)
To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack
by: Zhuo, Terry Yue, et al.
Published: (2026)
by: Zhuo, Terry Yue, et al.
Published: (2026)
A Whole New World: Creating a Parallel-Poisoned Web Only AI-Agents Can See
by: Zychlinski, Shaked
Published: (2025)
by: Zychlinski, Shaked
Published: (2025)
LENS-XAI: Redefining Lightweight and Explainable Network Security through Knowledge Distillation and Variational Autoencoders for Scalable Intrusion Detection in Cybersecurity
by: Yagiz, Muhammet Anil, et al.
Published: (2025)
by: Yagiz, Muhammet Anil, et al.
Published: (2025)
Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risks of Language Models
by: Zhang, Andy K., et al.
Published: (2024)
by: Zhang, Andy K., et al.
Published: (2024)
ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents
by: Lee, Seunghyun, et al.
Published: (2026)
by: Lee, Seunghyun, et al.
Published: (2026)
RedSage: A Cybersecurity Generalist LLM
by: Suryanto, Naufal, et al.
Published: (2026)
by: Suryanto, Naufal, et al.
Published: (2026)
SentinelSphere: Integrating AI-Powered Real-Time Threat Detection with Cybersecurity Awareness Training
by: Tantaroudas, Nikolaos D., et al.
Published: (2026)
by: Tantaroudas, Nikolaos D., et al.
Published: (2026)
Detecting Unsuccessful Students in Cybersecurity Exercises in Two Different Learning Environments
by: Švábenský, Valdemar, et al.
Published: (2024)
by: Švábenský, Valdemar, et al.
Published: (2024)
A Public Theory of Distillation Resistance via Constraint-Coupled Reasoning Architectures
by: Wei, Peng, et al.
Published: (2026)
by: Wei, Peng, et al.
Published: (2026)
The Task Shield: Enforcing Task Alignment to Defend Against Indirect Prompt Injection in LLM Agents
by: Jia, Feiran, et al.
Published: (2024)
by: Jia, Feiran, et al.
Published: (2024)
Relevance as a Vulnerability: How Web Retrieval Degrades Safety Alignment in LLM Agents
by: Nawal, Aditya, et al.
Published: (2026)
by: Nawal, Aditya, et al.
Published: (2026)
On the Suitability of LLM-Driven Agents for Dark Pattern Audits
by: Sun, Chen, et al.
Published: (2026)
by: Sun, Chen, et al.
Published: (2026)
How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
by: Zhou, Zhenhong, et al.
Published: (2024)
by: Zhou, Zhenhong, et al.
Published: (2024)
Can LLMs Infer Conversational Agent Users' Personality Traits from Chat History?
by: Cögendez, Derya, et al.
Published: (2026)
by: Cögendez, Derya, et al.
Published: (2026)
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
by: Lee, Michael S., et al.
Published: (2026)
by: Lee, Michael S., et al.
Published: (2026)
AI Agents Under EU Law
by: Nannini, Luca, et al.
Published: (2026)
by: Nannini, Luca, et al.
Published: (2026)
Dynamic Risk Assessments for Offensive Cybersecurity Agents
by: Wei, Boyi, et al.
Published: (2025)
by: Wei, Boyi, et al.
Published: (2025)
Enforcing Benign Trajectories: A Behavioral Firewall for Structured-Workflow AI Agents
by: Dang, Hung
Published: (2026)
by: Dang, Hung
Published: (2026)
Ask What Your Country Can Do For You: Towards a Public Red Teaming Model
by: Kennedy, Wm. Matthew, et al.
Published: (2025)
by: Kennedy, Wm. Matthew, et al.
Published: (2025)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
by: Palakodety, Shriphani
Published: (2025)
by: Palakodety, Shriphani
Published: (2025)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
by: Lin, Zhiqiang, et al.
Published: (2025)
by: Lin, Zhiqiang, et al.
Published: (2025)
Synthetic Data and Health Privacy
by: Abgrall, Gwénolé, et al.
Published: (2025)
by: Abgrall, Gwénolé, et al.
Published: (2025)
Playing Devil's Advocate: Unmasking Toxicity and Vulnerabilities in Large Vision-Language Models
by: Erol, Abdulkadir, et al.
Published: (2025)
by: Erol, Abdulkadir, et al.
Published: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
by: Bucknall, Ben, et al.
Published: (2025)
by: Bucknall, Ben, et al.
Published: (2025)
Similar Items
-
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
by: Lin, Justin W., et al.
Published: (2025) -
Frontier AI's Impact on the Cybersecurity Landscape
by: Potter, Yujin, et al.
Published: (2025) -
A Review of Cybersecurity Incidents in the Food and Agriculture Sector
by: Kulkarni, Ajay, et al.
Published: (2024) -
Artificial Intelligence in Cybersecurity: Building Resilient Cyber Diplomacy Frameworks
by: Stoltz, Michael
Published: (2024) -
BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Reviewers?
by: Jiang, Fengqing, et al.
Published: (2025)