Evaluating AI-generated code for C++, Fortran, Go, Java, Julia, Matlab, Python, R, and Rust
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Diehl, Patrick, Nader, Noujoud, Brandt, Steve, Kaiser, Hartmut |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Can LLMs Find Bugs in Code? An Evaluation from Beginner Errors to Security Vulnerabilities in Python and C++
par: Mhatre, Akshay, et autres
Publié: (2025)
par: Mhatre, Akshay, et autres
Publié: (2025)
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks
par: Nader, Noujoud, et autres
Publié: (2025)
par: Nader, Noujoud, et autres
Publié: (2025)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
par: Diehl, Patrick, et autres
Publié: (2025)
par: Diehl, Patrick, et autres
Publié: (2025)
LLM Benchmarking with LLaMA2: Evaluating Code Development Performance Across Multiple Programming Languages
par: Diehl, Patrick, et autres
Publié: (2025)
par: Diehl, Patrick, et autres
Publié: (2025)
From Legacy Fortran to Portable Kokkos: An Autonomous Agentic AI Workflow
par: Gupta, Sparsh, et autres
Publié: (2025)
par: Gupta, Sparsh, et autres
Publié: (2025)
From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python
par: Wang, Jinhua, et autres
Publié: (2026)
par: Wang, Jinhua, et autres
Publié: (2026)
REMODEL-LLM: Transforming C code to Java using LLMs
par: Gupta, Aryan, et autres
Publié: (2025)
par: Gupta, Aryan, et autres
Publié: (2025)
EvoC2Rust: A Skeleton-guided Framework for Project-Level C-to-Rust Translation
par: Wang, Chaofan, et autres
Publié: (2025)
par: Wang, Chaofan, et autres
Publié: (2025)
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
par: May, Victor, et autres
Publié: (2025)
par: May, Victor, et autres
Publié: (2025)
Project-Level C-to-Rust Translation via Pointer Knowledge Graphs
par: Yuan, Zhiqiang, et autres
Publié: (2025)
par: Yuan, Zhiqiang, et autres
Publié: (2025)
Quality Evaluation of COBOL to Java Code Transformation
par: Froimovich, Shmulik, et autres
Publié: (2025)
par: Froimovich, Shmulik, et autres
Publié: (2025)
RustEvo^2: An Evolving Benchmark for API Evolution in LLM-based Rust Code Generation
par: Liang, Linxi, et autres
Publié: (2025)
par: Liang, Linxi, et autres
Publié: (2025)
Evaluating perturbation robustness of generative systems that use COBOL code inputs
par: Ackerman, Samuel, et autres
Publié: (2025)
par: Ackerman, Samuel, et autres
Publié: (2025)
Benchmarking the Parallel 1D Heat Equation Solver in Chapel, Charm++, C++, HPX, Go, Julia, Python, Rust, Swift, and Java
par: Diehl, Patrick, et autres
Publié: (2023)
par: Diehl, Patrick, et autres
Publié: (2023)
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
par: Almukhtar, Mohamed, et autres
Publié: (2026)
par: Almukhtar, Mohamed, et autres
Publié: (2026)
Automated Proof Generation for Rust Code via Self-Evolution
par: Chen, Tianyu, et autres
Publié: (2024)
par: Chen, Tianyu, et autres
Publié: (2024)
Adaptive Hierarchical Evaluation of LLMs and SAST tools for CWE Prediction in Python
par: Adnan, Muntasir, et autres
Publié: (2026)
par: Adnan, Muntasir, et autres
Publié: (2026)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
par: Barua, Saikat, et autres
Publié: (2024)
par: Barua, Saikat, et autres
Publié: (2024)
Developing a High-Performance Process Mining Library with Java and Python Bindings in Rust
par: Küsters, Aaron, et autres
Publié: (2024)
par: Küsters, Aaron, et autres
Publié: (2024)
Raw Pointer Rewriting with LLMs for Translating C to Safer Rust
par: Gao, Yifei, et autres
Publié: (2025)
par: Gao, Yifei, et autres
Publié: (2025)
Feedback Loops and Code Perturbations in LLM-based Software Engineering: A Case Study on a C-to-Rust Translation System
par: Weiss, Martin, et autres
Publié: (2025)
par: Weiss, Martin, et autres
Publié: (2025)
Go-Oracle: Automated Test Oracle for Go Concurrency Bugs
par: Tsimpourlas, Foivos, et autres
Publié: (2024)
par: Tsimpourlas, Foivos, et autres
Publié: (2024)
Automated Testing of COBOL to Java Transformation
par: Hans, Sandeep, et autres
Publié: (2025)
par: Hans, Sandeep, et autres
Publié: (2025)
Automated Validation of COBOL to Java Transformation
par: Kumar, Atul, et autres
Publié: (2025)
par: Kumar, Atul, et autres
Publié: (2025)
Assessing LLM code generation quality through path planning tasks
par: Chen, Wanyi, et autres
Publié: (2025)
par: Chen, Wanyi, et autres
Publié: (2025)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
par: Xing, Xing, et autres
Publié: (2025)
par: Xing, Xing, et autres
Publié: (2025)
Canonical Intermediate Representation for LLM-based optimization problem formulation and code generation
par: Lyu, Zhongyuan, et autres
Publié: (2026)
par: Lyu, Zhongyuan, et autres
Publié: (2026)
Embedding Software Intent: Lightweight Java Module Recovery
par: He, Yirui, et autres
Publié: (2025)
par: He, Yirui, et autres
Publié: (2025)
PyVeritas: On Verifying Python via LLM-Based Transpilation and Bounded Model Checking for C
par: Orvalho, Pedro, et autres
Publié: (2025)
par: Orvalho, Pedro, et autres
Publié: (2025)
Automated QoR improvement in OpenROAD with coding agents
par: Ghose, Amur, et autres
Publié: (2026)
par: Ghose, Amur, et autres
Publié: (2026)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
par: Larbi, Maya, et autres
Publié: (2025)
par: Larbi, Maya, et autres
Publié: (2025)
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
par: Misra, Diganta, et autres
Publié: (2025)
par: Misra, Diganta, et autres
Publié: (2025)
Better Python Programming for all: With the focus on Maintainability
par: Shivashankar, Karthik, et autres
Publié: (2024)
par: Shivashankar, Karthik, et autres
Publié: (2024)
Call-Chain-Aware LLM-Based Test Generation for Java Projects
par: Wang, Guancheng, et autres
Publié: (2026)
par: Wang, Guancheng, et autres
Publié: (2026)
A systematic review of generative AI usage for IT project management
par: Anghel, Ionut, et autres
Publié: (2026)
par: Anghel, Ionut, et autres
Publié: (2026)
Automating the Correctness Assessment of AI-generated Code for Security Contexts
par: Cotroneo, Domenico, et autres
Publié: (2023)
par: Cotroneo, Domenico, et autres
Publié: (2023)
ENCRUST: Encapsulated Substitution and Agentic Refinement on a Live Scaffold for Safe C-to-Rust Translation
par: Sim, Hohyun, et autres
Publié: (2026)
par: Sim, Hohyun, et autres
Publié: (2026)
GoNoGo: An Efficient LLM-based Multi-Agent System for Streamlining Automotive Software Release Decision-Making
par: Khoee, Arsham Gholamzadeh, et autres
Publié: (2024)
par: Khoee, Arsham Gholamzadeh, et autres
Publié: (2024)
LLMs for Automated Unit Test Generation and Assessment in Java: The AgoneTest Framework
par: Lops, Andrea, et autres
Publié: (2025)
par: Lops, Andrea, et autres
Publié: (2025)
GenAI-based test case generation and execution in SDV platform
par: Zyberaj, Denesa, et autres
Publié: (2025)
par: Zyberaj, Denesa, et autres
Publié: (2025)
Documents similaires
-
Can LLMs Find Bugs in Code? An Evaluation from Beginner Errors to Security Vulnerabilities in Python and C++
par: Mhatre, Akshay, et autres
Publié: (2025) -
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks
par: Nader, Noujoud, et autres
Publié: (2025) -
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
par: Diehl, Patrick, et autres
Publié: (2025) -
LLM Benchmarking with LLaMA2: Evaluating Code Development Performance Across Multiple Programming Languages
par: Diehl, Patrick, et autres
Publié: (2025) -
From Legacy Fortran to Portable Kokkos: An Autonomous Agentic AI Workflow
par: Gupta, Sparsh, et autres
Publié: (2025)