Saved in:
| Main Authors: | Mo, Wenjie Jacky, Liu, Qin, Wen, Xiaofei, Jung, Dongwon, Askari, Hadi, Zhou, Wenxuan, Zhao, Zhe, Chen, Muhao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.22063 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
by: Guo, Chengquan, et al.
Published: (2024)
by: Guo, Chengquan, et al.
Published: (2024)
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
by: Guo, Chengquan, et al.
Published: (2025)
by: Guo, Chengquan, et al.
Published: (2025)
RedCodeAgent: Automatic Red-teaming Agent against Diverse Code Agents
by: Guo, Chengquan, et al.
Published: (2025)
by: Guo, Chengquan, et al.
Published: (2025)
SynthCoder: A Synthetical Strategy to Tune LLMs for Code Completion
by: Yu, Dongjun, et al.
Published: (2025)
by: Yu, Dongjun, et al.
Published: (2025)
ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation
by: He, Minghua, et al.
Published: (2025)
by: He, Minghua, et al.
Published: (2025)
Automating API Documentation from Crowdsourced Knowledge
by: Kou, Bonan, et al.
Published: (2026)
by: Kou, Bonan, et al.
Published: (2026)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
by: Du, Yongkang, et al.
Published: (2025)
by: Du, Yongkang, et al.
Published: (2025)
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
by: Wu, Yutong, et al.
Published: (2024)
by: Wu, Yutong, et al.
Published: (2024)
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision?
by: Fan, Lishui, et al.
Published: (2026)
by: Fan, Lishui, et al.
Published: (2026)
Understanding and Bridging the Planner-Coder Gap: A Systematic Study on the Robustness of Multi-Agent Systems for Code Generation
by: Lyu, Zongyi, et al.
Published: (2025)
by: Lyu, Zongyi, et al.
Published: (2025)
DebugLM: Learning Traceable Training Data Provenance for LLMs
by: Mo, Wenjie Jacky, et al.
Published: (2026)
by: Mo, Wenjie Jacky, et al.
Published: (2026)
TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs
by: Shen, Qingchao, et al.
Published: (2026)
by: Shen, Qingchao, et al.
Published: (2026)
Red Teaming Program Repair Agents: When Correct Patches can Hide Vulnerabilities
by: Chen, Simin, et al.
Published: (2025)
by: Chen, Simin, et al.
Published: (2025)
MeltRTL: Multi-Expert LLMs with Inference-time Intervention for RTL Code Generation
by: Mashnoor, Nowfel, et al.
Published: (2026)
by: Mashnoor, Nowfel, et al.
Published: (2026)
AdaCoder: An Adaptive Planning and Multi-Agent Framework for Function-Level Code Generation
by: Zhu, Yueheng, et al.
Published: (2025)
by: Zhu, Yueheng, et al.
Published: (2025)
What Makes Good In-context Demonstrations for Code Intelligence Tasks with LLMs?
by: Gao, Shuzheng, et al.
Published: (2023)
by: Gao, Shuzheng, et al.
Published: (2023)
Chart2Code-MoLA: Efficient Multi-Modal Code Generation via Adaptive Expert Routing
by: Wang, Yifei, et al.
Published: (2025)
by: Wang, Yifei, et al.
Published: (2025)
ConceptCoder: Improve Code Reasoning via Concept Learning
by: Rahman, Md Mahbubur, et al.
Published: (2026)
by: Rahman, Md Mahbubur, et al.
Published: (2026)
PlayCoder: Making LLM-Generated GUI Code Playable
by: Peng, Zhiyuan, et al.
Published: (2026)
by: Peng, Zhiyuan, et al.
Published: (2026)
Red Teaming Contemporary AI Models: Insights from Spanish and Basque Perspectives
by: Romero-Arjona, Miguel, et al.
Published: (2025)
by: Romero-Arjona, Miguel, et al.
Published: (2025)
TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated Code
by: Huang, Jiangping, et al.
Published: (2026)
by: Huang, Jiangping, et al.
Published: (2026)
WybeCoder: Verified Imperative Code Generation
by: Gloeckle, Fabian, et al.
Published: (2026)
by: Gloeckle, Fabian, et al.
Published: (2026)
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
by: Ni, Yuansheng, et al.
Published: (2025)
by: Ni, Yuansheng, et al.
Published: (2025)
Towards Engineering Multi-Agent LLMs: A Protocol-Driven Approach
by: Mao, Zhenyu, et al.
Published: (2025)
by: Mao, Zhenyu, et al.
Published: (2025)
Hybrid Privacy Policy-Code Consistency Check using Knowledge Graphs and LLMs
by: Mao, Zhenyu, et al.
Published: (2025)
by: Mao, Zhenyu, et al.
Published: (2025)
MUCOCO: Automated Consistency Testing of Code LLMs
by: Chou, Chua Jin, et al.
Published: (2026)
by: Chou, Chua Jin, et al.
Published: (2026)
SparseCoder: Identifier-Aware Sparse Transformer for File-Level Code Summarization
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails
by: Wen, Xiaofei, et al.
Published: (2025)
by: Wen, Xiaofei, et al.
Published: (2025)
StructCoder: Structure-Aware Transformer for Code Generation
by: Tipirneni, Sindhu, et al.
Published: (2022)
by: Tipirneni, Sindhu, et al.
Published: (2022)
o1-Coder: an o1 Replication for Coding
by: Zhang, Yuxiang, et al.
Published: (2024)
by: Zhang, Yuxiang, et al.
Published: (2024)
BabelCoder: Agentic Code Translation with Specification Alignment
by: Rabbi, Fazle, et al.
Published: (2025)
by: Rabbi, Fazle, et al.
Published: (2025)
Seed-Coder: Let the Code Model Curate Data for Itself
by: Seed, ByteDance, et al.
Published: (2025)
by: Seed, ByteDance, et al.
Published: (2025)
Prompt Engineering or Fine-Tuning: An Empirical Assessment of LLMs for Code
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
by: Wang, Minxing, et al.
Published: (2026)
by: Wang, Minxing, et al.
Published: (2026)
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
by: Dou, Shihan, et al.
Published: (2024)
by: Dou, Shihan, et al.
Published: (2024)
SparseCoder: Advancing Source Code Analysis with Sparse Attention and Learned Token Pruning
by: Yang, Xueqi, et al.
Published: (2023)
by: Yang, Xueqi, et al.
Published: (2023)
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models
by: Yu, Hao, et al.
Published: (2023)
by: Yu, Hao, et al.
Published: (2023)
Does Teaming-Up LLMs Improve Secure Code Generation? A Comprehensive Evaluation with Multi-LLMSecCodeEval
by: Sabir, Bushra, et al.
Published: (2026)
by: Sabir, Bushra, et al.
Published: (2026)
Delving into Parameter-Efficient Fine-Tuning in Code Change Learning: An Empirical Study
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
by: He, Ping, et al.
Published: (2025)
by: He, Ping, et al.
Published: (2025)
Similar Items
-
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
by: Guo, Chengquan, et al.
Published: (2024) -
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
by: Guo, Chengquan, et al.
Published: (2025) -
RedCodeAgent: Automatic Red-teaming Agent against Diverse Code Agents
by: Guo, Chengquan, et al.
Published: (2025) -
SynthCoder: A Synthetical Strategy to Tune LLMs for Code Completion
by: Yu, Dongjun, et al.
Published: (2025) -
ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation
by: He, Minghua, et al.
Published: (2025)