Manifold of Failure: Behavioral Attraction Basins in Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Munshi, Sarthak, Bhatt, Manish, Narajala, Vineeth Sai, Habler, Idan, Al-Kahfah, Ammar, Huang, Ken, Gatto, Blake |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
ETDI: Mitigating Tool Squatting and Rug Pull Attacks in Model Context Protocol (MCP) by using OAuth-Enhanced Tool Definitions and Policy-Based Access Control
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
Enterprise-Grade Security for the Model Context Protocol (MCP): Frameworks and Mitigation Strategies
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
Large Empirical Case Study: Go-Explore adapted for AI Red Team Testing
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
Building A Secure Agentic AI Application Leveraging A2A Protocol
von: Habler, Idan, et al.
Veröffentlicht: (2025)
von: Habler, Idan, et al.
Veröffentlicht: (2025)
Agent Capability Negotiation and Binding Protocol (ACNBP)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
COALESCE: Economic and Security Dynamics of Skill-Based Task Outsourcing Among Team of Autonomous LLM Agents
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)
MAIF: Enforcing AI Trust and Provenance with an Artifact-Centric Agentic Paradigm
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
Agent Name Service (ANS): A Universal Directory for Secure AI Agent Discovery and Interoperability
von: Huang, Ken, et al.
Veröffentlicht: (2025)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
Adversarial Hubness Detector: Detecting Hubness Poisoning in Retrieval-Augmented Generation Systems
von: Habler, Idan, et al.
Veröffentlicht: (2026)
von: Habler, Idan, et al.
Veröffentlicht: (2026)
Securing Agentic AI: A Comprehensive Threat Model and Mitigation Framework for Generative AI Agents
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025)
A Novel Zero-Trust Identity Framework for Agentic AI: Decentralized Authentication and Fine-Grained Access Control
von: Huang, Ken, et al.
Veröffentlicht: (2025)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
LAAF: Logic-layer Automated Attack Framework A Systematic Red-Teaming Methodology for LPCI Vulnerabilities in Agentic Large Language Model Systems
von: Atta, Hammad, et al.
Veröffentlicht: (2026)
von: Atta, Hammad, et al.
Veröffentlicht: (2026)
Security Steerability is All You Need
von: Hazan, Itay, et al.
Veröffentlicht: (2025)
von: Hazan, Itay, et al.
Veröffentlicht: (2025)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
von: Bhatt, Manish
Veröffentlicht: (2026)
von: Bhatt, Manish
Veröffentlicht: (2026)
A2AS: Agentic AI Runtime Security and Self-Defense
von: Neelou, Eugene, et al.
Veröffentlicht: (2025)
von: Neelou, Eugene, et al.
Veröffentlicht: (2025)
Bhatt Conjectures: On Necessary-But-Not-Sufficient Benchmark Tautology for Human Like Reasoning
von: Bhatt, Manish
Veröffentlicht: (2025)
von: Bhatt, Manish
Veröffentlicht: (2025)
From Tool Orchestration to Code Execution: A Study of MCP Design Choices
von: Felendler, Yuval, et al.
Veröffentlicht: (2026)
von: Felendler, Yuval, et al.
Veröffentlicht: (2026)
AAGATE: A NIST AI RMF-Aligned Governance Platform for Agentic AI
von: Huang, Ken, et al.
Veröffentlicht: (2025)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
Tag&Tab: Pretraining Data Detection in Large Language Models Using Keyword-Based Membership Inference Attack
von: Antebi, Sagiv, et al.
Veröffentlicht: (2025)
von: Antebi, Sagiv, et al.
Veröffentlicht: (2025)
ACSE-Eval: Can LLMs threat model real-world cloud infrastructure?
von: Munshi, Sarthak, et al.
Veröffentlicht: (2025)
von: Munshi, Sarthak, et al.
Veröffentlicht: (2025)
Introduction to IoT
von: Ananna, Tajkia Nuri, et al.
Veröffentlicht: (2023)
von: Ananna, Tajkia Nuri, et al.
Veröffentlicht: (2023)
Logic layer Prompt Control Injection (LPCI): A Novel Security Vulnerability Class in Agentic Systems
von: Atta, Hammad, et al.
Veröffentlicht: (2025)
von: Atta, Hammad, et al.
Veröffentlicht: (2025)
Mind the Web: The Security of Web Use Agents
von: Shapira, Avishag, et al.
Veröffentlicht: (2025)
von: Shapira, Avishag, et al.
Veröffentlicht: (2025)
Towards Smart Healthcare: Challenges and Opportunities in IoT and ML
von: Saifuzzaman, Munshi, et al.
Veröffentlicht: (2023)
von: Saifuzzaman, Munshi, et al.
Veröffentlicht: (2023)
LexiMark: Robust Watermarking via Lexical Substitutions to Enhance Membership Verification of an LLM's Textual Training Data
von: German, Eyal, et al.
Veröffentlicht: (2025)
von: German, Eyal, et al.
Veröffentlicht: (2025)
Cyber security of OT networks: A tutorial and overview
von: Kapoor, Sarthak, et al.
Veröffentlicht: (2025)
von: Kapoor, Sarthak, et al.
Veröffentlicht: (2025)
Augmenting Parameter-Efficient Pre-trained Language Models with Large Language Models
von: Anand, Saurabh, et al.
Veröffentlicht: (2026)
von: Anand, Saurabh, et al.
Veröffentlicht: (2026)
Towards Practical Data-Dependent Memory-Hard Functions with Optimal Sustained Space Trade-offs in the Parallel Random Oracle Model
von: Blocki, Jeremiah, et al.
Veröffentlicht: (2025)
von: Blocki, Jeremiah, et al.
Veröffentlicht: (2025)
GPT in Sheep's Clothing: The Risk of Customized GPTs
von: Antebi, Sagiv, et al.
Veröffentlicht: (2024)
von: Antebi, Sagiv, et al.
Veröffentlicht: (2024)
Dependency-Aware Privacy for Multi-turn Agents
von: Anshumaan, Divyam, et al.
Veröffentlicht: (2026)
von: Anshumaan, Divyam, et al.
Veröffentlicht: (2026)
CYBERSECEVAL 3: Advancing the Evaluation of Cybersecurity Risks and Capabilities in Large Language Models
von: Wan, Shengye, et al.
Veröffentlicht: (2024)
von: Wan, Shengye, et al.
Veröffentlicht: (2024)
Generalized Quantum-assisted Digital Signature
von: Tarable, Alberto, et al.
Veröffentlicht: (2024)
von: Tarable, Alberto, et al.
Veröffentlicht: (2024)
Trustworthy Agentic AI Requires Deterministic Architectural Boundaries
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
Binary Diff Summarization using Large Language Models
von: Udeshi, Meet, et al.
Veröffentlicht: (2025)
von: Udeshi, Meet, et al.
Veröffentlicht: (2025)
Can LLMs be Scammed? A Baseline Measurement Study
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
Phishsense-1B: A Technical Perspective on an AI-Powered Phishing Detection Model
von: Blake, SE
Veröffentlicht: (2025)
von: Blake, SE
Veröffentlicht: (2025)
CyberSecEval 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language Models
von: Bhatt, Manish, et al.
Veröffentlicht: (2024)
von: Bhatt, Manish, et al.
Veröffentlicht: (2024)
Can Developers rely on LLMs for Secure IaC Development?
von: Firouzi, Ehsan, et al.
Veröffentlicht: (2026)
von: Firouzi, Ehsan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
von: Bhatt, Manish, et al.
Veröffentlicht: (2026) -
ETDI: Mitigating Tool Squatting and Rug Pull Attacks in Model Context Protocol (MCP) by using OAuth-Enhanced Tool Definitions and Policy-Based Access Control
von: Bhatt, Manish, et al.
Veröffentlicht: (2025) -
Enterprise-Grade Security for the Model Context Protocol (MCP): Frameworks and Mitigation Strategies
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025) -
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
von: Narajala, Vineeth Sai, et al.
Veröffentlicht: (2025) -
Large Empirical Case Study: Go-Explore adapted for AI Red Team Testing
von: Bhatt, Manish, et al.
Veröffentlicht: (2025)