ProxyWar: Dynamic Assessment of LLM Code Generation in Game Arenas
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Wenjun, Wang, Xinyu, Wu, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025)
von: Fu, Jia, et al.
Veröffentlicht: (2025)
Code Copycat Conundrum: Demystifying Repetition in LLM-based Code Generation
von: Liu, Mingwei, et al.
Veröffentlicht: (2025)
von: Liu, Mingwei, et al.
Veröffentlicht: (2025)
Uncovering LLM-Generated Code: A Zero-Shot Synthetic Code Detector via Code Rewriting
von: Ye, Tong, et al.
Veröffentlicht: (2024)
von: Ye, Tong, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for LLM-based Code Generation
von: Xu, Senrong, et al.
Veröffentlicht: (2026)
von: Xu, Senrong, et al.
Veröffentlicht: (2026)
SALLM: Security Assessment of Generated Code
von: Siddiq, Mohammed Latif, et al.
Veröffentlicht: (2023)
von: Siddiq, Mohammed Latif, et al.
Veröffentlicht: (2023)
Dynamic Stability of LLM-Generated Code
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
RustEvo^2: An Evolving Benchmark for API Evolution in LLM-based Rust Code Generation
von: Liang, Linxi, et al.
Veröffentlicht: (2025)
von: Liang, Linxi, et al.
Veröffentlicht: (2025)
Investigating The Smells of LLM Generated Code
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2025)
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2025)
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
On Unified Prompt Tuning for Request Quality Assurance in Public Code Review
von: Chen, Xinyu, et al.
Veröffentlicht: (2024)
von: Chen, Xinyu, et al.
Veröffentlicht: (2024)
Knowledge-Guided Prompt Learning for Request Quality Assurance in Public Code Review
von: Li, Lin, et al.
Veröffentlicht: (2024)
von: Li, Lin, et al.
Veröffentlicht: (2024)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
Synergizing Code Coverage and Gameplay Intent: Coverage-Aware Game Playtesting with LLM-Guided Reinforcement Learning
von: Mu, Enhong, et al.
Veröffentlicht: (2025)
von: Mu, Enhong, et al.
Veröffentlicht: (2025)
FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
von: Wu, Yue, et al.
Veröffentlicht: (2025)
von: Wu, Yue, et al.
Veröffentlicht: (2025)
A Performance Study of LLM-Generated Code on Leetcode
von: Coignion, Tristan, et al.
Veröffentlicht: (2024)
von: Coignion, Tristan, et al.
Veröffentlicht: (2024)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
CodeCircuit: Toward Inferring LLM-Generated Code Correctness via Attribution Graphs
von: He, Yicheng, et al.
Veröffentlicht: (2026)
von: He, Yicheng, et al.
Veröffentlicht: (2026)
LLM-Empowered Event-Chain Driven Code Generation for ADAS in SDV systems
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
CodeFuse-CommitEval: Towards Benchmarking LLM's Power on Commit Message and Code Change Inconsistency Detection
von: Zhang, Qingyu, et al.
Veröffentlicht: (2025)
von: Zhang, Qingyu, et al.
Veröffentlicht: (2025)
DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells
von: Cherief, Houcine Abdelkader, et al.
Veröffentlicht: (2026)
von: Cherief, Houcine Abdelkader, et al.
Veröffentlicht: (2026)
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation
von: Dente, Francesco, et al.
Veröffentlicht: (2026)
von: Dente, Francesco, et al.
Veröffentlicht: (2026)
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach
von: Sepidband, Melika, et al.
Veröffentlicht: (2025)
von: Sepidband, Melika, et al.
Veröffentlicht: (2025)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
von: Liu, Fang, et al.
Veröffentlicht: (2024)
von: Liu, Fang, et al.
Veröffentlicht: (2024)
Hallucination in LLM-Based Code Generation: An Automotive Case Study
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety
von: Uddin, S M Jamil
Veröffentlicht: (2026)
von: Uddin, S M Jamil
Veröffentlicht: (2026)
Unseen Horizons: Unveiling the Real Capability of LLM Code Generation Beyond the Familiar
von: Zhang, Yuanliang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanliang, et al.
Veröffentlicht: (2024)
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
von: Della Porta, Antonio, et al.
Veröffentlicht: (2025)
von: Della Porta, Antonio, et al.
Veröffentlicht: (2025)
Correctness isnt Efficiency: Runtime Memory Divergence in LLM-Generated Code
von: Rajput, Prateek, et al.
Veröffentlicht: (2026)
von: Rajput, Prateek, et al.
Veröffentlicht: (2026)
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
von: Akli, Amal, et al.
Veröffentlicht: (2026)
von: Akli, Amal, et al.
Veröffentlicht: (2026)
The Readability Spectrum: Patterns, Issues, and Prompt Effects in LLM-Generated Code
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
von: Cramer, Marcos, et al.
Veröffentlicht: (2025)
von: Cramer, Marcos, et al.
Veröffentlicht: (2025)
Using a Feedback Loop for LLM-based Infrastructure as Code Generation
von: Palavalli, Mayur Amarnath, et al.
Veröffentlicht: (2024)
von: Palavalli, Mayur Amarnath, et al.
Veröffentlicht: (2024)
Uncovering Intention through LLM-Driven Code Snippet Description Generation
von: Nugroho, Yusuf Sulistyo, et al.
Veröffentlicht: (2025)
von: Nugroho, Yusuf Sulistyo, et al.
Veröffentlicht: (2025)
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
von: Chakroborti, Apu Kumar, et al.
Veröffentlicht: (2025)
von: Chakroborti, Apu Kumar, et al.
Veröffentlicht: (2025)
Eliminating Hallucination-Induced Errors in LLM Code Generation with Functional Clustering
von: Ravuri, Chaitanya, et al.
Veröffentlicht: (2025)
von: Ravuri, Chaitanya, et al.
Veröffentlicht: (2025)
AI-Assisted Assessment of Coding Practices in Modern Code Review
von: Vijayvergiya, Manushree, et al.
Veröffentlicht: (2024)
von: Vijayvergiya, Manushree, et al.
Veröffentlicht: (2024)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
von: Lam, Man Ho, et al.
Veröffentlicht: (2025)
von: Lam, Man Ho, et al.
Veröffentlicht: (2025)
CodeVision: Detecting LLM-Generated Code Using 2D Token Probability Maps and Vision Models
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025) -
Code Copycat Conundrum: Demystifying Repetition in LLM-based Code Generation
von: Liu, Mingwei, et al.
Veröffentlicht: (2025) -
Uncovering LLM-Generated Code: A Zero-Shot Synthetic Code Detector via Code Rewriting
von: Ye, Tong, et al.
Veröffentlicht: (2024) -
Uncertainty Quantification for LLM-based Code Generation
von: Xu, Senrong, et al.
Veröffentlicht: (2026) -
SALLM: Security Assessment of Generated Code
von: Siddiq, Mohammed Latif, et al.
Veröffentlicht: (2023)