AVISE: Framework for Evaluating the Security of AI Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Lempinen, Mikko, Kemppainen, Joni, Raesalmi, Niklas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation
por: Roy, Joyjit, et al.
Publicado: (2026)
por: Roy, Joyjit, et al.
Publicado: (2026)
Towards Reliable and Practical LLM Security Evaluations via Bayesian Modelling
por: Llewellyn, Mary, et al.
Publicado: (2025)
por: Llewellyn, Mary, et al.
Publicado: (2025)
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
por: Heye, David, et al.
Publicado: (2026)
por: Heye, David, et al.
Publicado: (2026)
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
por: Uzor, GodsGift, et al.
Publicado: (2025)
por: Uzor, GodsGift, et al.
Publicado: (2025)
Swiss-Bench 003: Evaluating LLM Reliability and Adversarial Security for Swiss Regulatory Contexts
por: Uenal, Fatih
Publicado: (2026)
por: Uenal, Fatih
Publicado: (2026)
When RAG Chatbots Expose Their Backend: An Anonymized Case Study of Privacy and Security Risks in Patient-Facing Medical AI
por: Madrid-García, Alfredo, et al.
Publicado: (2026)
por: Madrid-García, Alfredo, et al.
Publicado: (2026)
In AI Sweet Harmony: Sociopragmatic Guardrail Bypasses and Evaluation-Awareness in OpenAI gpt-oss-20b
por: Durner, Nils
Publicado: (2025)
por: Durner, Nils
Publicado: (2025)
LLM Platform Security: Applying a Systematic Evaluation Framework to OpenAI's ChatGPT Plugins
por: Iqbal, Umar, et al.
Publicado: (2023)
por: Iqbal, Umar, et al.
Publicado: (2023)
BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards
por: Dorn, Diego, et al.
Publicado: (2024)
por: Dorn, Diego, et al.
Publicado: (2024)
Cognitive Control Architecture (CCA): A Lifecycle Supervision Framework for Robustly Aligned AI Agents
por: Liang, Zhibo, et al.
Publicado: (2025)
por: Liang, Zhibo, et al.
Publicado: (2025)
The Ethics of Interaction: Mitigating Security Threats in LLMs
por: Kumar, Ashutosh, et al.
Publicado: (2024)
por: Kumar, Ashutosh, et al.
Publicado: (2024)
RvB: Automating AI System Hardening via Iterative Red-Blue Games
por: Huang, Lige, et al.
Publicado: (2026)
por: Huang, Lige, et al.
Publicado: (2026)
LLM for SoC Security: A Paradigm Shift
por: Saha, Dipayan, et al.
Publicado: (2023)
por: Saha, Dipayan, et al.
Publicado: (2023)
A Survey on Agentic Security: Applications, Threats and Defenses
por: Shahriar, Asif, et al.
Publicado: (2025)
por: Shahriar, Asif, et al.
Publicado: (2025)
Text Embedding Inversion Security for Multilingual Language Models
por: Chen, Yiyi, et al.
Publicado: (2024)
por: Chen, Yiyi, et al.
Publicado: (2024)
From Threat Intelligence to Firewall Rules: Semantic Relations in Hybrid AI Agent and Expert System Architectures
por: Bonfanti, Chiara, et al.
Publicado: (2026)
por: Bonfanti, Chiara, et al.
Publicado: (2026)
Security and Privacy Challenges of Large Language Models: A Survey
por: Das, Badhan Chandra, et al.
Publicado: (2024)
por: Das, Badhan Chandra, et al.
Publicado: (2024)
Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments
por: Rigaki, Maria, et al.
Publicado: (2023)
por: Rigaki, Maria, et al.
Publicado: (2023)
Institutional Platform for Secure Self-Service Large Language Model Exploration
por: Bumgardner, V. K. Cody, et al.
Publicado: (2024)
por: Bumgardner, V. K. Cody, et al.
Publicado: (2024)
A High-Capacity and Secure Disambiguation Algorithm for Neural Linguistic Steganography
por: Feng, Yapei, et al.
Publicado: (2025)
por: Feng, Yapei, et al.
Publicado: (2025)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
por: Li, Xueyi, et al.
Publicado: (2026)
por: Li, Xueyi, et al.
Publicado: (2026)
Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model
por: Wu, Tianyi, et al.
Publicado: (2026)
por: Wu, Tianyi, et al.
Publicado: (2026)
From Vulnerabilities to Remediation: A Systematic Literature Review of LLMs in Code Security
por: Basic, Enna, et al.
Publicado: (2024)
por: Basic, Enna, et al.
Publicado: (2024)
Balancing Innovation and Privacy: Data Security Strategies in Natural Language Processing Applications
por: Liu, Shaobo, et al.
Publicado: (2024)
por: Liu, Shaobo, et al.
Publicado: (2024)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
por: Chen, Junkai, et al.
Publicado: (2025)
por: Chen, Junkai, et al.
Publicado: (2025)
EventHunter: Dynamic Clustering and Ranking of Security Events from Hacker Forum Discussions
por: Ech-Chammakhy, Yasir, et al.
Publicado: (2025)
por: Ech-Chammakhy, Yasir, et al.
Publicado: (2025)
Generative AI Security: Challenges and Countermeasures
por: Zhu, Banghua, et al.
Publicado: (2024)
por: Zhu, Banghua, et al.
Publicado: (2024)
Proof-of-Guardrail in AI Agents and What (Not) to Trust from It
por: Jin, Xisen, et al.
Publicado: (2026)
por: Jin, Xisen, et al.
Publicado: (2026)
RedacBench: Can AI Erase Your Secrets?
por: Jeon, Hyunjun, et al.
Publicado: (2026)
por: Jeon, Hyunjun, et al.
Publicado: (2026)
Analysis and prevention of AI-based phishing email attacks
por: Eze, Chibuike Samuel, et al.
Publicado: (2024)
por: Eze, Chibuike Samuel, et al.
Publicado: (2024)
Gandalf the Red: Adaptive Security for LLMs
por: Pfister, Niklas, et al.
Publicado: (2025)
por: Pfister, Niklas, et al.
Publicado: (2025)
VelLMes: A high-interaction AI-based deception framework
por: Sladić, Muris, et al.
Publicado: (2025)
por: Sladić, Muris, et al.
Publicado: (2025)
Modeling the Attack: Detecting AI-Generated Text by Quantifying Adversarial Perturbations
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
What Was Your Prompt? A Remote Keylogging Attack on AI Assistants
por: Weiss, Roy, et al.
Publicado: (2024)
por: Weiss, Roy, et al.
Publicado: (2024)
Code Vulnerability Detection Across Different Programming Languages with AI Models
por: Humran, Hael Abdulhakim Ali, et al.
Publicado: (2025)
por: Humran, Hael Abdulhakim Ali, et al.
Publicado: (2025)
PsySafe: A Comprehensive Framework for Psychological-based Attack, Defense, and Evaluation of Multi-agent System Safety
por: Zhang, Zaibin, et al.
Publicado: (2024)
por: Zhang, Zaibin, et al.
Publicado: (2024)
An Independent Safety Evaluation of Kimi K2.5
por: Yong, Zheng-Xin, et al.
Publicado: (2026)
por: Yong, Zheng-Xin, et al.
Publicado: (2026)
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
por: An, Hengyu, et al.
Publicado: (2026)
por: An, Hengyu, et al.
Publicado: (2026)
LATTICE: Evaluating Decision Support Utility of Crypto Agents
por: Chan, Aaron, et al.
Publicado: (2026)
por: Chan, Aaron, et al.
Publicado: (2026)
Ejemplares similares
-
PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024) -
AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation
por: Roy, Joyjit, et al.
Publicado: (2026) -
Towards Reliable and Practical LLM Security Evaluations via Bayesian Modelling
por: Llewellyn, Mary, et al.
Publicado: (2025) -
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
por: Heye, David, et al.
Publicado: (2026) -
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
por: Uzor, GodsGift, et al.
Publicado: (2025)