Security-First AI: Foundations for Robust and Trustworthy Systems
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Tallam, Krti |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CyberSentinel: An Emergent Threat Detection System for AI Security
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Engineering Risk-Aware, Security-by-Design Frameworks for Assurance of Large-Scale Autonomous AI Models
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Embedding Trust at Scale: Physics-Aware Neural Watermarking for Secure and Verifiable Data Pipelines
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Operationalizing CaMeL: Strengthening LLM Defenses for Enterprise Deployment
von: Tallam, Krti, et al.
Veröffentlicht: (2025)
von: Tallam, Krti, et al.
Veröffentlicht: (2025)
The Cyber Immune System: Harnessing Adversarial Forces for Security Resilience
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Transforming Cyber Defense: Harnessing Agentic and Frontier AI for Proactive, Ethical Threat Intelligence
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Layered Mutability: Continuity and Governance in Persistent Self-Modifying Agents
von: Tallam, Krti
Veröffentlicht: (2026)
von: Tallam, Krti
Veröffentlicht: (2026)
On the Foundations of Trustworthy Artificial Intelligence
von: Dunham, TJ
Veröffentlicht: (2026)
von: Dunham, TJ
Veröffentlicht: (2026)
Removing Watermarks with Partial Regeneration using Semantic Information
von: Tallam, Krti, et al.
Veröffentlicht: (2025)
von: Tallam, Krti, et al.
Veröffentlicht: (2025)
Authorization Propagation in Multi-Agent AI Systems: Identity Governance as Infrastructure
von: Tallam, Krti
Veröffentlicht: (2026)
von: Tallam, Krti
Veröffentlicht: (2026)
Trustworthy AI-Generative Content for Intelligent Network Service: Robustness, Security, and Fairness
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
MCP Guardian: A Security-First Layer for Safeguarding MCP-Based AI System
von: Kumar, Sonu, et al.
Veröffentlicht: (2025)
von: Kumar, Sonu, et al.
Veröffentlicht: (2025)
Agentic AI for Cyber Resilience: A New Security Paradigm and Its System-Theoretic Foundations
von: Li, Tao, et al.
Veröffentlicht: (2025)
von: Li, Tao, et al.
Veröffentlicht: (2025)
Alignment, Agency and Autonomy in Frontier AI: A Systems Engineering Perspective
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
The Adaptive Arms Race: Redefining Robustness in AI Security
von: Tsingenopoulos, Ilias, et al.
Veröffentlicht: (2023)
von: Tsingenopoulos, Ilias, et al.
Veröffentlicht: (2023)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
Meta-Sealing: A Revolutionizing Integrity Assurance Protocol for Transparent, Tamper-Proof, and Trustworthy AI System
von: Krishnamoorthy, Mahesh Vaijainthymala
Veröffentlicht: (2024)
von: Krishnamoorthy, Mahesh Vaijainthymala
Veröffentlicht: (2024)
Llama-3.1-FoundationAI-SecurityLLM-8B-Instruct Technical Report
von: Weerawardhena, Sajana, et al.
Veröffentlicht: (2025)
von: Weerawardhena, Sajana, et al.
Veröffentlicht: (2025)
Llama-3.1-FoundationAI-SecurityLLM-Base-8B Technical Report
von: Kassianik, Paul, et al.
Veröffentlicht: (2025)
von: Kassianik, Paul, et al.
Veröffentlicht: (2025)
Towards Trustworthy AI: Secure Deepfake Detection using CNNs and Zero-Knowledge Proofs
von: Islam, H M Mohaimanul, et al.
Veröffentlicht: (2025)
von: Islam, H M Mohaimanul, et al.
Veröffentlicht: (2025)
Offensive Security for AI Systems: Concepts, Practices, and Applications
von: Harguess, Josh, et al.
Veröffentlicht: (2025)
von: Harguess, Josh, et al.
Veröffentlicht: (2025)
AI-Governed Agent Architecture for Web-Trustworthy Tokenization of Alternative Assets
von: Borjigin, Ailiya, et al.
Veröffentlicht: (2025)
von: Borjigin, Ailiya, et al.
Veröffentlicht: (2025)
Decoding the Black Box: Integrating Moral Imagination with Technical AI Governance
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
Securing AI Systems: A Guide to Known Attacks and Impacts
von: Kiribuchi, Naoto, et al.
Veröffentlicht: (2025)
von: Kiribuchi, Naoto, et al.
Veröffentlicht: (2025)
Fortify Your Foundations: Practical Privacy and Security for Foundation Model Deployments In The Cloud
von: Chrapek, Marcin, et al.
Veröffentlicht: (2024)
von: Chrapek, Marcin, et al.
Veröffentlicht: (2024)
MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models
von: Xu, Chejian, et al.
Veröffentlicht: (2025)
von: Xu, Chejian, et al.
Veröffentlicht: (2025)
Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
Security of AI Agents
von: He, Yifeng, et al.
Veröffentlicht: (2024)
von: He, Yifeng, et al.
Veröffentlicht: (2024)
Toward Trustworthy Agentic AI: A Multimodal Framework for Preventing Prompt Injection Attacks
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
AI Security Map: Holistic Organization of AI Security Technologies and Impacts on Stakeholders
von: Kato, Hiroya, et al.
Veröffentlicht: (2025)
von: Kato, Hiroya, et al.
Veröffentlicht: (2025)
From Autonomous Agents to Integrated Systems, A New Paradigm: Orchestrated Distributed Intelligence
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems
von: Liang, Eric
Veröffentlicht: (2026)
von: Liang, Eric
Veröffentlicht: (2026)
ATLAS: AI-Assisted Threat-to-Assertion Learning for System-on-Chip Security Verification
von: Tashdid, Ishraq, et al.
Veröffentlicht: (2026)
von: Tashdid, Ishraq, et al.
Veröffentlicht: (2026)
Trustworthy Distributed AI Systems: Robustness, Privacy, and Governance
von: Wei, Wenqi, et al.
Veröffentlicht: (2024)
von: Wei, Wenqi, et al.
Veröffentlicht: (2024)
Building Trustworthy Multimodal AI: A Review of Fairness, Transparency, and Ethics in Vision-Language Tasks
von: Saleh, Mohammad, et al.
Veröffentlicht: (2025)
von: Saleh, Mohammad, et al.
Veröffentlicht: (2025)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
von: Xing, Wenpeng, et al.
Veröffentlicht: (2025)
von: Xing, Wenpeng, et al.
Veröffentlicht: (2025)
Security of and by Generative AI platforms
von: Hayagreevan, Hari, et al.
Veröffentlicht: (2024)
von: Hayagreevan, Hari, et al.
Veröffentlicht: (2024)
The AI Security Pyramid of Pain
von: Ward, Chris M., et al.
Veröffentlicht: (2024)
von: Ward, Chris M., et al.
Veröffentlicht: (2024)
Secure Multiparty Generative AI
von: Shrestha, Manil, et al.
Veröffentlicht: (2024)
von: Shrestha, Manil, et al.
Veröffentlicht: (2024)
Bridging Privacy and Robustness for Trustworthy Machine Learning
von: Zhang, Xiaojin, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaojin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CyberSentinel: An Emergent Threat Detection System for AI Security
von: Tallam, Krti
Veröffentlicht: (2025) -
Engineering Risk-Aware, Security-by-Design Frameworks for Assurance of Large-Scale Autonomous AI Models
von: Tallam, Krti
Veröffentlicht: (2025) -
Embedding Trust at Scale: Physics-Aware Neural Watermarking for Secure and Verifiable Data Pipelines
von: Tallam, Krti
Veröffentlicht: (2025) -
Operationalizing CaMeL: Strengthening LLM Defenses for Enterprise Deployment
von: Tallam, Krti, et al.
Veröffentlicht: (2025) -
The Cyber Immune System: Harnessing Adversarial Forces for Security Resilience
von: Tallam, Krti
Veröffentlicht: (2025)