RedCodeAgent: Automatic Red-teaming Agent against Diverse Code Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Chengquan, Xie, Chulin, Yang, Yu, Chen, Zhaorun, Lin, Zinan, Davies, Xander, Gal, Yarin, Song, Dawn, Li, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
von: Guo, Chengquan, et al.
Veröffentlicht: (2025)
von: Guo, Chengquan, et al.
Veröffentlicht: (2025)
CodeAgent: Autonomous Communicative Agents for Code Review
von: Tang, Xunzhu, et al.
Veröffentlicht: (2024)
von: Tang, Xunzhu, et al.
Veröffentlicht: (2024)
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses
von: Lin, Jiahang, et al.
Veröffentlicht: (2026)
von: Lin, Jiahang, et al.
Veröffentlicht: (2026)
SolAgent: A Specialized Multi-Agent Framework for Solidity Code Generation
von: Chen, Wei, et al.
Veröffentlicht: (2026)
von: Chen, Wei, et al.
Veröffentlicht: (2026)
MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents
von: Steinberg, Jonathan, et al.
Veröffentlicht: (2026)
von: Steinberg, Jonathan, et al.
Veröffentlicht: (2026)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
ABTest: Behavior-Driven Testing for AI Coding Agents
von: Dai, Wuyang, et al.
Veröffentlicht: (2026)
von: Dai, Wuyang, et al.
Veröffentlicht: (2026)
Multi-Agent Code-Orchestrated Generation for Reliable Infrastructure-as-Code
von: Khan, Rana Nameer Hussain, et al.
Veröffentlicht: (2025)
von: Khan, Rana Nameer Hussain, et al.
Veröffentlicht: (2025)
GeoJSON Agents:A Multi-Agent LLM Architecture for Geospatial Analysis-Function Calling vs Code Generation
von: Luo, Qianqian, et al.
Veröffentlicht: (2025)
von: Luo, Qianqian, et al.
Veröffentlicht: (2025)
GraphCodeAgent: Dual Graph-Guided LLM Agent for Retrieval-Augmented Repo-Level Code Generation
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
DialogAgent: An Auto-engagement Agent for Code Question Answering Data Production
von: Liang, Xiaoyun, et al.
Veröffentlicht: (2024)
von: Liang, Xiaoyun, et al.
Veröffentlicht: (2024)
CodeCureAgent: Automatic Classification and Repair of Static Analysis Warnings
von: Joos, Pascal, et al.
Veröffentlicht: (2025)
von: Joos, Pascal, et al.
Veröffentlicht: (2025)
A Low-Code Approach for the Automatic Personalization of Conversational Agents
von: Conrardy, Aaron, et al.
Veröffentlicht: (2026)
von: Conrardy, Aaron, et al.
Veröffentlicht: (2026)
AI IDEs or Autonomous Agents? Measuring the Impact of Coding Agents on Software Development
von: Agarwal, Shyam, et al.
Veröffentlicht: (2026)
von: Agarwal, Shyam, et al.
Veröffentlicht: (2026)
SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion
von: Ma, George, et al.
Veröffentlicht: (2025)
von: Ma, George, et al.
Veröffentlicht: (2025)
Sema Code: Decoupling AI Coding Agents into Programmable, Embeddable Infrastructure
von: Wang, Huacan, et al.
Veröffentlicht: (2026)
von: Wang, Huacan, et al.
Veröffentlicht: (2026)
Fingerprinting AI Coding Agents on GitHub
von: Ghaleb, Taher A.
Veröffentlicht: (2026)
von: Ghaleb, Taher A.
Veröffentlicht: (2026)
Coding Agents Don't Know When to Act
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
von: He, Ping, et al.
Veröffentlicht: (2025)
von: He, Ping, et al.
Veröffentlicht: (2025)
RepoTransAgent: Multi-Agent LLM Framework for Repository-Aware Code Translation
von: Guan, Ziqi, et al.
Veröffentlicht: (2025)
von: Guan, Ziqi, et al.
Veröffentlicht: (2025)
Can Coding Agents Be General Agents?
von: Ivanov, Maksim, et al.
Veröffentlicht: (2026)
von: Ivanov, Maksim, et al.
Veröffentlicht: (2026)
Code Review Agent Benchmark
von: Zhang, Yuntong, et al.
Veröffentlicht: (2026)
von: Zhang, Yuntong, et al.
Veröffentlicht: (2026)
Decoding the Configuration of AI Coding Agents: Insights from Claude Code Projects
von: Santos, Helio Victor F., et al.
Veröffentlicht: (2025)
von: Santos, Helio Victor F., et al.
Veröffentlicht: (2025)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
Scaling Coding Agents via Atomic Skills
von: Ma, Yingwei, et al.
Veröffentlicht: (2026)
von: Ma, Yingwei, et al.
Veröffentlicht: (2026)
Knowledge-Guided Multi-Agent Framework for Application-Level Software Code Generation
von: Xiong, Qian, et al.
Veröffentlicht: (2025)
von: Xiong, Qian, et al.
Veröffentlicht: (2025)
SWE-Shepherd: Advancing PRMs for Reinforcing Code Agents
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
TypeScript Repository Indexing for Code Agent Retrieval
von: Pu, Junsong, et al.
Veröffentlicht: (2026)
von: Pu, Junsong, et al.
Veröffentlicht: (2026)
Agentic Much? Adoption of Coding Agents on GitHub
von: Robbes, Romain, et al.
Veröffentlicht: (2026)
von: Robbes, Romain, et al.
Veröffentlicht: (2026)
Do AI Agents Really Improve Code Readability?
von: Horikawa, Kyogo, et al.
Veröffentlicht: (2026)
von: Horikawa, Kyogo, et al.
Veröffentlicht: (2026)
LLM Assisted Coding with Metamorphic Specification Mutation Agent
von: Akhond, Mostafijur Rahman, et al.
Veröffentlicht: (2025)
von: Akhond, Mostafijur Rahman, et al.
Veröffentlicht: (2025)
SWE-Cycle: Benchmarking Code Agents across the Complete Issue Resolution Cycle
von: Guan, Hao, et al.
Veröffentlicht: (2026)
von: Guan, Hao, et al.
Veröffentlicht: (2026)
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
von: Mo, Wenjie Jacky, et al.
Veröffentlicht: (2025)
von: Mo, Wenjie Jacky, et al.
Veröffentlicht: (2025)
Workflows vs Agents for Code Translation
von: Gray, Henry, et al.
Veröffentlicht: (2025)
von: Gray, Henry, et al.
Veröffentlicht: (2025)
Bootstrapping Coding Agents: The Specification Is the Program
von: Monperrus, Martin
Veröffentlicht: (2026)
von: Monperrus, Martin
Veröffentlicht: (2026)
CodeCoR: An LLM-Based Self-Reflective Multi-Agent Framework for Code Generation
von: Pan, Ruwei, et al.
Veröffentlicht: (2025)
von: Pan, Ruwei, et al.
Veröffentlicht: (2025)
When is Generated Code Difficult to Comprehend? Assessing AI Agent Python Code Proficiency in the Wild
von: Temkulkiat, Nanthit, et al.
Veröffentlicht: (2026)
von: Temkulkiat, Nanthit, et al.
Veröffentlicht: (2026)
Red Teaming Program Repair Agents: When Correct Patches can Hide Vulnerabilities
von: Chen, Simin, et al.
Veröffentlicht: (2025)
von: Chen, Simin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024) -
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
von: Guo, Chengquan, et al.
Veröffentlicht: (2025) -
CodeAgent: Autonomous Communicative Agents for Code Review
von: Tang, Xunzhu, et al.
Veröffentlicht: (2024) -
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
von: Fu, Lingyue, et al.
Veröffentlicht: (2025) -
Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses
von: Lin, Jiahang, et al.
Veröffentlicht: (2026)