EnIGMA: Interactive Tools Substantially Assist LM Agents in Finding Security Vulnerabilities
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Abramovich, Talor, Udeshi, Meet, Shao, Minghao, Lieret, Kilian, Xi, Haoran, Milner, Kimberly, Jancheska, Sofija, Yang, John, Jimenez, Carlos E., Khorrami, Farshad, Krishnamurthy, Prashanth, Dolan-Gavitt, Brendan, Shafique, Muhammad, Narasimhan, Karthik, Karri, Ramesh, Press, Ofir |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
NYU CTF Bench: A Scalable Open-Source Benchmark Dataset for Evaluating LLMs in Offensive Security
par: Shao, Minghao, et autres
Publié: (2024)
par: Shao, Minghao, et autres
Publié: (2024)
REMaQE: Reverse Engineering Math Equations from Executables
par: Udeshi, Meet, et autres
Publié: (2023)
par: Udeshi, Meet, et autres
Publié: (2023)
SaMOSA: Sandbox for Malware Orchestration and Side-Channel Analysis
par: Udeshi, Meet, et autres
Publié: (2025)
par: Udeshi, Meet, et autres
Publié: (2025)
An Empirical Evaluation of LLMs for Solving Offensive Security Challenges
par: Shao, Minghao, et autres
Publié: (2024)
par: Shao, Minghao, et autres
Publié: (2024)
D-CIPHER: Dynamic Collaborative Intelligent Multi-Agent System with Planner and Heterogeneous Executors for Offensive Security
par: Udeshi, Meet, et autres
Publié: (2025)
par: Udeshi, Meet, et autres
Publié: (2025)
CRAKEN: Cybersecurity LLM Agent with Knowledge-Based Execution
par: Shao, Minghao, et autres
Publié: (2025)
par: Shao, Minghao, et autres
Publié: (2025)
Ransomware 3.0: Self-Composing and LLM-Orchestrated
par: Raz, Md, et autres
Publié: (2025)
par: Raz, Md, et autres
Publié: (2025)
Safeguarding LLMs Against Misuse and AI-Driven Malware Using Steganographic Canaries
par: Raz, Md, et autres
Publié: (2026)
par: Raz, Md, et autres
Publié: (2026)
AI In Cybersecurity Education -- Scalable Agentic CTF Design Principles and Educational Outcomes
par: Xi, Haoran, et autres
Publié: (2026)
par: Xi, Haoran, et autres
Publié: (2026)
Enabling Deep Visibility into VxWorks-Based Embedded Controllers in Cyber-Physical Systems for Anomaly Detection
par: Krishnamurthy, Prashanth, et autres
Publié: (2025)
par: Krishnamurthy, Prashanth, et autres
Publié: (2025)
SCAMPER -- Synchrophasor Covert chAnnel for Malicious and Protective ERrands
par: Krishnamurthy, Prashanth, et autres
Publié: (2025)
par: Krishnamurthy, Prashanth, et autres
Publié: (2025)
Real-Time Multi-Modal Subcomponent-Level Measurements for Trustworthy System Monitoring and Malware Detection
par: Khorrami, Farshad, et autres
Publié: (2025)
par: Khorrami, Farshad, et autres
Publié: (2025)
Binary Diff Summarization using Large Language Models
par: Udeshi, Meet, et autres
Publié: (2025)
par: Udeshi, Meet, et autres
Publié: (2025)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
par: Rani, Nanda, et autres
Publié: (2026)
par: Rani, Nanda, et autres
Publié: (2026)
Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark
par: Shao, Minghao, et autres
Publié: (2025)
par: Shao, Minghao, et autres
Publié: (2025)
Tracking Real-time Anomalies in Cyber-Physical Systems Through Dynamic Behavioral Analysis
par: Krishnamurthy, Prashanth, et autres
Publié: (2024)
par: Krishnamurthy, Prashanth, et autres
Publié: (2024)
HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
par: Bhat, Vineet, et autres
Publié: (2024)
par: Bhat, Vineet, et autres
Publié: (2024)
From Trace to Line: LLM Agent for Real-World OSS Vulnerability Localization
par: Xi, Haoran, et autres
Publié: (2025)
par: Xi, Haoran, et autres
Publié: (2025)
MapleGrasp: Mask-guided Feature Pooling for Language-driven Efficient Robotic Grasping
par: Bhat, Vineet, et autres
Publié: (2025)
par: Bhat, Vineet, et autres
Publié: (2025)
SENTAUR: Security EnhaNced Trojan Assessment Using LLMs Against Undesirable Revisions
par: Bhandari, Jitendra, et autres
Publié: (2024)
par: Bhandari, Jitendra, et autres
Publié: (2024)
Data-Efficient System Identification via Lipschitz Neural Networks
par: Wei, Shiqing, et autres
Publié: (2024)
par: Wei, Shiqing, et autres
Publié: (2024)
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
par: Patel, Naman, et autres
Publié: (2025)
par: Patel, Naman, et autres
Publié: (2025)
Confidence-Aware Safe and Stable Control of Control-Affine Systems
par: Wei, Shiqing, et autres
Publié: (2024)
par: Wei, Shiqing, et autres
Publié: (2024)
Combining Switching Mechanism with Re-Initialization and Anomaly Detection for Resiliency of Cyber-Physical Systems
par: Fu, Hao, et autres
Publié: (2024)
par: Fu, Hao, et autres
Publié: (2024)
Prescribed-Time Stability Properties of Interconnected Systems
par: Krishnamurthy, Prashanth, et autres
Publié: (2024)
par: Krishnamurthy, Prashanth, et autres
Publié: (2024)
Learning a Better Control Barrier Function Under Uncertain Dynamics
par: Dai, Bolun, et autres
Publié: (2023)
par: Dai, Bolun, et autres
Publié: (2023)
Robust Neural Lyapunov Control for Nonlinear Systems With Quadratically Bounded Disturbances
par: Shiqing Wei, et autres
Publié: (2025)
par: Shiqing Wei, et autres
Publié: (2025)
Grounding LLMs For Robot Task Planning Using Closed-loop State Feedback
par: Bhat, Vineet, et autres
Publié: (2024)
par: Bhat, Vineet, et autres
Publié: (2024)
3D CAVLA: Leveraging Depth and 3D Context to Generalize Vision Language Action Models for Unseen Tasks
par: Bhat, Vineet, et autres
Publié: (2025)
par: Bhat, Vineet, et autres
Publié: (2025)
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
par: Vineet Bhat, et autres
Publié: (2025)
par: Vineet Bhat, et autres
Publié: (2025)
A Control Barrier Function-Constrained Model Predictive Control Framework for Safe Reinforcement Learning
par: Kaypak, Ali Umut, et autres
Publié: (2026)
par: Kaypak, Ali Umut, et autres
Publié: (2026)
SHIELD: A Host-Independent Framework for Ransomware Detection using Deep Filesystem Features
par: Raz, Md, et autres
Publié: (2025)
par: Raz, Md, et autres
Publié: (2025)
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
par: Yang, John, et autres
Publié: (2024)
par: Yang, John, et autres
Publié: (2024)
RESCORE: LLM-Driven Simulation Recovery in Control Systems Research Papers
par: Bhat, Vineet, et autres
Publié: (2026)
par: Bhat, Vineet, et autres
Publié: (2026)
Sailing Through Point Clouds: Safe Navigation Using Point Cloud Based Control Barrier Functions
par: Dai, Bolun, et autres
Publié: (2024)
par: Dai, Bolun, et autres
Publié: (2024)
Out-of-Distribution Detection with Overlap Index
par: Fu, Hao, et autres
Publié: (2024)
par: Fu, Hao, et autres
Publié: (2024)
An Upper Bound for the Distribution Overlap Index and Its Applications
par: Fu, Hao, et autres
Publié: (2022)
par: Fu, Hao, et autres
Publié: (2022)
Differentiable Optimization Based Time-Varying Control Barrier Functions for Dynamic Obstacle Avoidance
par: Dai, Bolun, et autres
Publié: (2023)
par: Dai, Bolun, et autres
Publié: (2023)
CLIPScope: Enhancing Zero-Shot OOD Detection with Bayesian Scoring
par: Fu, Hao, et autres
Publié: (2024)
par: Fu, Hao, et autres
Publié: (2024)
Proactive Hierarchical Control Barrier Function-Based Safety Prioritization in Close Human-Robot Interaction Scenarios
par: Maithani, Patanjali, et autres
Publié: (2025)
par: Maithani, Patanjali, et autres
Publié: (2025)
Documents similaires
-
NYU CTF Bench: A Scalable Open-Source Benchmark Dataset for Evaluating LLMs in Offensive Security
par: Shao, Minghao, et autres
Publié: (2024) -
REMaQE: Reverse Engineering Math Equations from Executables
par: Udeshi, Meet, et autres
Publié: (2023) -
SaMOSA: Sandbox for Malware Orchestration and Side-Channel Analysis
par: Udeshi, Meet, et autres
Publié: (2025) -
An Empirical Evaluation of LLMs for Solving Offensive Security Challenges
par: Shao, Minghao, et autres
Publié: (2024) -
D-CIPHER: Dynamic Collaborative Intelligent Multi-Agent System with Planner and Heterogeneous Executors for Offensive Security
par: Udeshi, Meet, et autres
Publié: (2025)