Effective LLM-Driven Code Generation with Pythoness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Levin, Kyla H., Gwilt, Kyle, Berger, Emery D., Freund, Stephen N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ChatDBG: Augmenting Debugging with Large Language Models
von: Levin, Kyla H., et al.
Veröffentlicht: (2024)
von: Levin, Kyla H., et al.
Veröffentlicht: (2024)
Dynamic Stability of LLM-Generated Code
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
von: Chen, Le, et al.
Veröffentlicht: (2025)
von: Chen, Le, et al.
Veröffentlicht: (2025)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
CoverUp: Effective High Coverage Test Generation for Python
von: Pizzorno, Juan Altmayer, et al.
Veröffentlicht: (2024)
von: Pizzorno, Juan Altmayer, et al.
Veröffentlicht: (2024)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
von: Peng, Yun, et al.
Veröffentlicht: (2024)
von: Peng, Yun, et al.
Veröffentlicht: (2024)
Benchmarking LLM Code Generation for Audio Programming with Visual Dataflow Languages
von: Zhang, William, et al.
Veröffentlicht: (2024)
von: Zhang, William, et al.
Veröffentlicht: (2024)
Is Functional Correctness Enough to Evaluate Code Language Models? Exploring Diversity of Generated Codes
von: Chon, Heejae, et al.
Veröffentlicht: (2024)
von: Chon, Heejae, et al.
Veröffentlicht: (2024)
EnvTrace: Simulation-Based Semantic Evaluation of LLM Code via Execution Trace Alignment -- Demonstrated at Synchrotron Beamlines
von: van der Vleuten, Noah, et al.
Veröffentlicht: (2025)
von: van der Vleuten, Noah, et al.
Veröffentlicht: (2025)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
Hydra: Efficient, Correct Code Generation via Checkpoint-and-Rollback Support
von: Du, Alexander, et al.
Veröffentlicht: (2026)
von: Du, Alexander, et al.
Veröffentlicht: (2026)
Assessing GPT-4-Vision's Capabilities in UML-Based Code Generation
von: Antal, Gábor, et al.
Veröffentlicht: (2024)
von: Antal, Gábor, et al.
Veröffentlicht: (2024)
Self-Improving Code Generation via Semantic Entropy and Behavioral Consensus
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
von: Zhou, Tianyang, et al.
Veröffentlicht: (2025)
von: Zhou, Tianyang, et al.
Veröffentlicht: (2025)
From Code Generation to Software Testing: AI Copilot with Context-Based RAG
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
AutoMCQ -- Automatically Generate Code Comprehension Questions using GenAI
von: Goodfellow, Martin, et al.
Veröffentlicht: (2025)
von: Goodfellow, Martin, et al.
Veröffentlicht: (2025)
What Were You Thinking? An LLM-Driven Large-Scale Study of Refactoring Motivations in Open-Source Projects
von: Robredo, Mikel, et al.
Veröffentlicht: (2025)
von: Robredo, Mikel, et al.
Veröffentlicht: (2025)
Incoherence as Oracle-less Measure of Error in LLM-Based Code Generation
von: Valentin, Thomas, et al.
Veröffentlicht: (2025)
von: Valentin, Thomas, et al.
Veröffentlicht: (2025)
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
von: Huang, Zhechong, et al.
Veröffentlicht: (2025)
von: Huang, Zhechong, et al.
Veröffentlicht: (2025)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
Getting Python Types Right with RightTyper
von: Pizzorno, Juan Altmayer, et al.
Veröffentlicht: (2025)
von: Pizzorno, Juan Altmayer, et al.
Veröffentlicht: (2025)
Reconsidering "Reconsidering Custom Memory Allocation"
von: van Kempen, Nicolas, et al.
Veröffentlicht: (2026)
von: van Kempen, Nicolas, et al.
Veröffentlicht: (2026)
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
von: Misra, Diganta, et al.
Veröffentlicht: (2025)
von: Misra, Diganta, et al.
Veröffentlicht: (2025)
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
von: Wallraven, Stephan, et al.
Veröffentlicht: (2026)
von: Wallraven, Stephan, et al.
Veröffentlicht: (2026)
Agentic Code Reasoning
von: Ugare, Shubham, et al.
Veröffentlicht: (2026)
von: Ugare, Shubham, et al.
Veröffentlicht: (2026)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
Once4All: Skeleton-Guided SMT Solver Fuzzing with LLM-Synthesized Generators
von: Sun, Maolin, et al.
Veröffentlicht: (2025)
von: Sun, Maolin, et al.
Veröffentlicht: (2025)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
von: Saieva, Anthony, et al.
Veröffentlicht: (2023)
von: Saieva, Anthony, et al.
Veröffentlicht: (2023)
Is Self-Repair a Silver Bullet for Code Generation?
von: Olausson, Theo X., et al.
Veröffentlicht: (2023)
von: Olausson, Theo X., et al.
Veröffentlicht: (2023)
PPM: Automated Generation of Diverse Programming Problems for Benchmarking Code Generation Models
von: Chen, Simin, et al.
Veröffentlicht: (2024)
von: Chen, Simin, et al.
Veröffentlicht: (2024)
Assessing Code Understanding in LLMs
von: Laneve, Cosimo, et al.
Veröffentlicht: (2025)
von: Laneve, Cosimo, et al.
Veröffentlicht: (2025)
AI-Mediated Code Comment Improvement
von: Dhakal, Maria, et al.
Veröffentlicht: (2025)
von: Dhakal, Maria, et al.
Veröffentlicht: (2025)
Ranking LLM-Generated Loop Invariants for Program Verification
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
LLMON: An LLM-native Markup Language to Leverage Structure and Semantics at the LLM Interface
von: Hind, Michael, et al.
Veröffentlicht: (2026)
von: Hind, Michael, et al.
Veröffentlicht: (2026)
AInsteinBench: Benchmarking Coding Agents on Scientific Repositories
von: Duston, Titouan, et al.
Veröffentlicht: (2025)
von: Duston, Titouan, et al.
Veröffentlicht: (2025)
AI-Assisted Fixes to Code Review Comments at Scale
von: Maddila, Chandra, et al.
Veröffentlicht: (2025)
von: Maddila, Chandra, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ChatDBG: Augmenting Debugging with Large Language Models
von: Levin, Kyla H., et al.
Veröffentlicht: (2024) -
Dynamic Stability of LLM-Generated Code
von: Rajput, Prateek, et al.
Veröffentlicht: (2025) -
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
von: Chen, Le, et al.
Veröffentlicht: (2025) -
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026) -
CoverUp: Effective High Coverage Test Generation for Python
von: Pizzorno, Juan Altmayer, et al.
Veröffentlicht: (2024)