LLM Performance for Code Generation on Noisy Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Sendyka, Radzim, Cabrera, Christian, Paleyes, Andrei, Robinson, Diana, Lawrence, Neil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Code Roulette: How Prompt Variability Affects LLM Code Generation
by: Paleyes, Andrei, et al.
Published: (2025)
by: Paleyes, Andrei, et al.
Published: (2025)
Optimising for Energy Efficiency and Performance in Machine Learning
by: Ferreira, Emile Dos Santos, et al.
Published: (2026)
by: Ferreira, Emile Dos Santos, et al.
Published: (2026)
Machine Learning Systems: A Survey from a Data-Oriented Perspective
by: Cabrera, Christian, et al.
Published: (2023)
by: Cabrera, Christian, et al.
Published: (2023)
Self-sustaining Software Systems (S4): Towards Improved Interpretability and Adaptation
by: Cabrera, Christian, et al.
Published: (2024)
by: Cabrera, Christian, et al.
Published: (2024)
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
by: Tsai, Yun-Da, et al.
Published: (2024)
by: Tsai, Yun-Da, et al.
Published: (2024)
RocketPPA: Code-Level Power, Performance, and Area Prediction via LLM and Mixture of Experts
by: Abdollahi, Armin, et al.
Published: (2025)
by: Abdollahi, Armin, et al.
Published: (2025)
Studying LLM Performance on Closed- and Open-source Data
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
Clotho: Measuring Task-Specific Pre-Generation Test Adequacy for LLM Inputs
by: Yoon, Juyeon, et al.
Published: (2025)
by: Yoon, Juyeon, et al.
Published: (2025)
Leveraging LLMs for Legacy Code Modernization: Challenges and Opportunities for LLM-Generated Documentation
by: Diggs, Colin, et al.
Published: (2024)
by: Diggs, Colin, et al.
Published: (2024)
Enhancing Code Quality with Generative AI: Boosting Developer Warning Compliance
by: Chang, Hansen, et al.
Published: (2025)
by: Chang, Hansen, et al.
Published: (2025)
Breaking Memorization Barriers in LLM Code Fine-Tuning via Information Bottleneck for Improved Generalization
by: Wang, Changsheng, et al.
Published: (2025)
by: Wang, Changsheng, et al.
Published: (2025)
A Survey on LLM-based Code Generation for Low-Resource and Domain-Specific Programming Languages
by: Joel, Sathvik, et al.
Published: (2024)
by: Joel, Sathvik, et al.
Published: (2024)
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
by: Rahman, Musfiqur, et al.
Published: (2025)
by: Rahman, Musfiqur, et al.
Published: (2025)
Code-Aware Prompting: A study of Coverage Guided Test Generation in Regression Setting using LLM
by: Ryan, Gabriel, et al.
Published: (2024)
by: Ryan, Gabriel, et al.
Published: (2024)
Renaissance of Literate Programming in the Era of LLMs: Enhancing LLM-Based Code Generation in Large-Scale Projects
by: Zhang, Wuyang, et al.
Published: (2024)
by: Zhang, Wuyang, et al.
Published: (2024)
Hybrid-Gym: Training Coding Agents to Generalize Across Tasks
by: Xie, Yiqing, et al.
Published: (2026)
by: Xie, Yiqing, et al.
Published: (2026)
SemRep: Generative Code Representation Learning with Code Transformations
by: Li, Weichen, et al.
Published: (2026)
by: Li, Weichen, et al.
Published: (2026)
Think Anywhere in Code Generation
by: Jiang, Xue, et al.
Published: (2026)
by: Jiang, Xue, et al.
Published: (2026)
Wisdom and Delusion of LLM Ensembles for Code Generation and Repair
by: Vallecillos-Ruiz, Fernando, et al.
Published: (2025)
by: Vallecillos-Ruiz, Fernando, et al.
Published: (2025)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
by: Yan, Kaiwen, et al.
Published: (2025)
by: Yan, Kaiwen, et al.
Published: (2025)
Refining Joint Text and Source Code Embeddings for Retrieval Task with Parameter-Efficient Fine-Tuning
by: Galliamov, Karim, et al.
Published: (2024)
by: Galliamov, Karim, et al.
Published: (2024)
OSS-Bench: Benchmark Generator for Coding LLMs
by: Jiang, Yuancheng, et al.
Published: (2025)
by: Jiang, Yuancheng, et al.
Published: (2025)
LLM-Based Design Pattern Detection
by: Schindler, Christian, et al.
Published: (2025)
by: Schindler, Christian, et al.
Published: (2025)
Enhancing LLM-Based Test Generation by Eliminating Covered Code
by: Xu, WeiZhe, et al.
Published: (2026)
by: Xu, WeiZhe, et al.
Published: (2026)
Semantic Voting: Execution-Grounded Consensus for LLM Code Generation
by: Jiang, Shan, et al.
Published: (2026)
by: Jiang, Shan, et al.
Published: (2026)
LogSieve: Task-Aware CI Log Reduction for Sustainable LLM-Based Analysis
by: Barnes, Marcus Emmanuel, et al.
Published: (2026)
by: Barnes, Marcus Emmanuel, et al.
Published: (2026)
StructCoder: Structure-Aware Transformer for Code Generation
by: Tipirneni, Sindhu, et al.
Published: (2022)
by: Tipirneni, Sindhu, et al.
Published: (2022)
Leveraging Reviewer Experience in Code Review Comment Generation
by: Lin, Hong Yi, et al.
Published: (2024)
by: Lin, Hong Yi, et al.
Published: (2024)
CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation
by: Peng, Jinjun, et al.
Published: (2025)
by: Peng, Jinjun, et al.
Published: (2025)
Carbon Footprint Evaluation of Code Generation through LLM as a Service
by: Vartziotis, Tina, et al.
Published: (2025)
by: Vartziotis, Tina, et al.
Published: (2025)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
by: Khant, Kyi Shin, et al.
Published: (2025)
by: Khant, Kyi Shin, et al.
Published: (2025)
Code Graph Model (CGM): A Graph-Integrated Large Language Model for Repository-Level Software Engineering Tasks
by: Tao, Hongyuan, et al.
Published: (2025)
by: Tao, Hongyuan, et al.
Published: (2025)
Leveraging Reward Models for Guiding Code Review Comment Generation
by: Sghaier, Oussama Ben, et al.
Published: (2025)
by: Sghaier, Oussama Ben, et al.
Published: (2025)
Ensuring Functional Correctness of Large Code Models with Selective Generation
by: Jeong, Jaewoo, et al.
Published: (2025)
by: Jeong, Jaewoo, et al.
Published: (2025)
Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards
by: Jolfaei, Erfan Aghadavoodi, et al.
Published: (2026)
by: Jolfaei, Erfan Aghadavoodi, et al.
Published: (2026)
FRANC: A Lightweight Framework for High-Quality Code Generation
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
Design-Specification Tiling for ICL-based CAD Code Generation
by: Du, Yali, et al.
Published: (2026)
by: Du, Yali, et al.
Published: (2026)
Context-Augmented Code Generation Using Programming Knowledge Graphs
by: Seddik, Shahd, et al.
Published: (2026)
by: Seddik, Shahd, et al.
Published: (2026)
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks
by: Siddiq, Mohammed Latif, et al.
Published: (2024)
by: Siddiq, Mohammed Latif, et al.
Published: (2024)
Citation-Grounded Code Comprehension: Preventing LLM Hallucination Through Hybrid Retrieval and Graph-Augmented Context
by: Arafat, Jahidul
Published: (2025)
by: Arafat, Jahidul
Published: (2025)
Similar Items
-
Code Roulette: How Prompt Variability Affects LLM Code Generation
by: Paleyes, Andrei, et al.
Published: (2025) -
Optimising for Energy Efficiency and Performance in Machine Learning
by: Ferreira, Emile Dos Santos, et al.
Published: (2026) -
Machine Learning Systems: A Survey from a Data-Oriented Perspective
by: Cabrera, Christian, et al.
Published: (2023) -
Self-sustaining Software Systems (S4): Towards Improved Interpretability and Adaptation
by: Cabrera, Christian, et al.
Published: (2024) -
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
by: Tsai, Yun-Da, et al.
Published: (2024)