Demo: ViolentUTF as An Accessible Platform for Generative AI Red Teaming
Fuente:
arXiv
Guardado en:
| Autor principal: | Nguyen, Tam n. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ollabench: Evaluating LLMs' Reasoning for Human-centric Interdependent Cybersecurity
por: Nguyen, Tam n.
Publicado: (2024)
por: Nguyen, Tam n.
Publicado: (2024)
Red Teaming AI Red Teaming
por: Majumdar, Subhabrata, et al.
Publicado: (2025)
por: Majumdar, Subhabrata, et al.
Publicado: (2025)
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
por: Cai, Jiacheng, et al.
Publicado: (2024)
por: Cai, Jiacheng, et al.
Publicado: (2024)
Red Teaming Methodology for Design Obfuscation
por: Liu, Yuntao, et al.
Publicado: (2025)
por: Liu, Yuntao, et al.
Publicado: (2025)
Autonomous Adversary: Red-Teaming in the age of LLM
por: Mamun, Mohammad, et al.
Publicado: (2026)
por: Mamun, Mohammad, et al.
Publicado: (2026)
How Effective Are Publicly Accessible Deepfake Detection Tools? A Comparative Evaluation of Open-Source and Free-to-Use Platforms
por: Rettinger, Michael, et al.
Publicado: (2026)
por: Rettinger, Michael, et al.
Publicado: (2026)
A Red Teaming Framework for Evaluating Robustness of AI-enabled Security Orchestration, Automation, and Response Systems
por: Shaikh, Ayan Javeed, et al.
Publicado: (2026)
por: Shaikh, Ayan Javeed, et al.
Publicado: (2026)
Demo: SGCode: A Flexible Prompt-Optimizing System for Secure Generation of Code
por: Ton, Khiem, et al.
Publicado: (2024)
por: Ton, Khiem, et al.
Publicado: (2024)
Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours
por: Dheekonda, Raja Sekhar Rao, et al.
Publicado: (2026)
por: Dheekonda, Raja Sekhar Rao, et al.
Publicado: (2026)
From Firewalls to Frontiers: AI Red-Teaming is a Domain-Specific Evolution of Cyber Red-Teaming
por: Sinha, Anusha, et al.
Publicado: (2025)
por: Sinha, Anusha, et al.
Publicado: (2025)
Automated Progressive Red Teaming
por: Jiang, Bojian, et al.
Publicado: (2024)
por: Jiang, Bojian, et al.
Publicado: (2024)
Training a General Purpose Automated Red Teaming Model
por: Padmakumar, Aishwarya, et al.
Publicado: (2026)
por: Padmakumar, Aishwarya, et al.
Publicado: (2026)
BlackIce: A Containerized Red Teaming Toolkit for AI Security Testing
por: Kaplan, Caelin, et al.
Publicado: (2025)
por: Kaplan, Caelin, et al.
Publicado: (2025)
Red-Teaming LLM Multi-Agent Systems via Communication Attacks
por: He, Pengfei, et al.
Publicado: (2025)
por: He, Pengfei, et al.
Publicado: (2025)
Red Teaming Large Reasoning Models
por: Chen, Jiawei, et al.
Publicado: (2025)
por: Chen, Jiawei, et al.
Publicado: (2025)
From Promise to Peril: Rethinking Cybersecurity Red and Blue Teaming in the Age of LLMs
por: Abuadbba, Alsharif, et al.
Publicado: (2025)
por: Abuadbba, Alsharif, et al.
Publicado: (2025)
A Systematic Review of Algorithmic Red Teaming Methodologies for Assurance and Security of AI Applications
por: Srivastava, Shruti, et al.
Publicado: (2026)
por: Srivastava, Shruti, et al.
Publicado: (2026)
AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration
por: Zhou, Andy, et al.
Publicado: (2025)
por: Zhou, Andy, et al.
Publicado: (2025)
When Scanners Lie: Evaluator Instability in LLM Red-Teaming
por: Erez, Lidor, et al.
Publicado: (2026)
por: Erez, Lidor, et al.
Publicado: (2026)
SkillAttack: Automated Red Teaming of Agent Skills through Attack Path Refinement
por: Duan, Zenghao, et al.
Publicado: (2026)
por: Duan, Zenghao, et al.
Publicado: (2026)
Red Team Redemption: A Structured Comparison of Open-Source Tools for Adversary Emulation
por: Landauer, Max, et al.
Publicado: (2024)
por: Landauer, Max, et al.
Publicado: (2024)
RedTeamLLM: an Agentic AI framework for offensive security
por: Challita, Brian, et al.
Publicado: (2025)
por: Challita, Brian, et al.
Publicado: (2025)
PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
RedTWIZ: Diverse LLM Red Teaming via Adaptive Attack Planning
por: Horal, Artur, et al.
Publicado: (2025)
por: Horal, Artur, et al.
Publicado: (2025)
RedShell: A Generative AI-Based Approach to Ethical Hacking
por: Bessa, Ricardo, et al.
Publicado: (2026)
por: Bessa, Ricardo, et al.
Publicado: (2026)
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
por: Wang, Yanting, et al.
Publicado: (2026)
por: Wang, Yanting, et al.
Publicado: (2026)
GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models
por: Wang, Zilong, et al.
Publicado: (2025)
por: Wang, Zilong, et al.
Publicado: (2025)
Prompt Optimization and Evaluation for LLM Automated Red Teaming
por: Freenor, Michael, et al.
Publicado: (2025)
por: Freenor, Michael, et al.
Publicado: (2025)
Persona-Conditioned Adversarial Prompting (PCAP): Multi-Identity Red-Teaming for Enhanced Adversarial Prompt Discovery
por: Morasso, Cristian, et al.
Publicado: (2026)
por: Morasso, Cristian, et al.
Publicado: (2026)
IPI-proxy: An Intercepting Proxy for Red-Teaming Web-Browsing AI Agents Against Indirect Prompt Injection
por: Chia-Pei, et al.
Publicado: (2026)
por: Chia-Pei, et al.
Publicado: (2026)
Resource Consumption Red-Teaming for Large Vision-Language Models
por: Gao, Haoran, et al.
Publicado: (2025)
por: Gao, Haoran, et al.
Publicado: (2025)
A Red Teaming Roadmap Towards System-Level Safety
por: Wang, Zifan, et al.
Publicado: (2025)
por: Wang, Zifan, et al.
Publicado: (2025)
MonitoringBench: Semi-Automated Red-Teaming for Agent Monitoring
por: Jotautaitė, Monika, et al.
Publicado: (2026)
por: Jotautaitė, Monika, et al.
Publicado: (2026)
Demo: TOSense -- What Did You Just Agree to?
por: Chen, Xinzhang, et al.
Publicado: (2025)
por: Chen, Xinzhang, et al.
Publicado: (2025)
Red Teaming with Artificial Intelligence-Driven Cyberattacks: A Scoping Review
por: Al-Azzawi, Mays, et al.
Publicado: (2025)
por: Al-Azzawi, Mays, et al.
Publicado: (2025)
Proteus: A Self-Evolving Red Team for Agent Skill Ecosystems
por: Zhou, Zhaojiacheng
Publicado: (2026)
por: Zhou, Zhaojiacheng
Publicado: (2026)
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
por: He, Pengfei, et al.
Publicado: (2026)
por: He, Pengfei, et al.
Publicado: (2026)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
por: Yin, Chenlong, et al.
Publicado: (2026)
por: Yin, Chenlong, et al.
Publicado: (2026)
ContextualJailbreak: Evolutionary Red-Teaming via Simulated Conversational Priming
por: Béjar, Mario Rodríguez, et al.
Publicado: (2026)
por: Béjar, Mario Rodríguez, et al.
Publicado: (2026)
Large Empirical Case Study: Go-Explore adapted for AI Red Team Testing
por: Bhatt, Manish, et al.
Publicado: (2025)
por: Bhatt, Manish, et al.
Publicado: (2025)
Ejemplares similares
-
Ollabench: Evaluating LLMs' Reasoning for Human-centric Interdependent Cybersecurity
por: Nguyen, Tam n.
Publicado: (2024) -
Red Teaming AI Red Teaming
por: Majumdar, Subhabrata, et al.
Publicado: (2025) -
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
por: Cai, Jiacheng, et al.
Publicado: (2024) -
Red Teaming Methodology for Design Obfuscation
por: Liu, Yuntao, et al.
Publicado: (2025) -
Autonomous Adversary: Red-Teaming in the age of LLM
por: Mamun, Mohammad, et al.
Publicado: (2026)