Salvato in:
| Autori principali: | Lin, Zhenru, Yao, Yiqun, Yuan, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2403.01784 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CodeMind: Evaluating Large Language Models for Code Reasoning
di: Liu, Changshu, et al.
Pubblicazione: (2024)
di: Liu, Changshu, et al.
Pubblicazione: (2024)
FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation
di: Pham, Loc, et al.
Pubblicazione: (2026)
di: Pham, Loc, et al.
Pubblicazione: (2026)
Grammar-Based Code Representation: Is It a Worthy Pursuit for LLMs?
di: Liang, Qingyuan, et al.
Pubblicazione: (2025)
di: Liang, Qingyuan, et al.
Pubblicazione: (2025)
Assessing Code Understanding in LLMs
di: Laneve, Cosimo, et al.
Pubblicazione: (2025)
di: Laneve, Cosimo, et al.
Pubblicazione: (2025)
AutoCode: LLMs as Problem Setters for Competitive Programming
di: Zhou, Shang, et al.
Pubblicazione: (2025)
di: Zhou, Shang, et al.
Pubblicazione: (2025)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
di: Kim, Su-Hyeon, et al.
Pubblicazione: (2025)
di: Kim, Su-Hyeon, et al.
Pubblicazione: (2025)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
di: Wang, Peiding, et al.
Pubblicazione: (2025)
di: Wang, Peiding, et al.
Pubblicazione: (2025)
Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility
di: Maveli, Nickil, et al.
Pubblicazione: (2026)
di: Maveli, Nickil, et al.
Pubblicazione: (2026)
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization
di: Zhao, Yang, et al.
Pubblicazione: (2024)
di: Zhao, Yang, et al.
Pubblicazione: (2024)
VeriEquivBench: An Equivalence Score for Ground-Truth-Free Evaluation of Formally Verifiable Code
di: Zeng, Lingfei, et al.
Pubblicazione: (2025)
di: Zeng, Lingfei, et al.
Pubblicazione: (2025)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
di: Dai, Hankun, et al.
Pubblicazione: (2025)
di: Dai, Hankun, et al.
Pubblicazione: (2025)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
di: Tang, Hao, et al.
Pubblicazione: (2024)
di: Tang, Hao, et al.
Pubblicazione: (2024)
Is Functional Correctness Enough to Evaluate Code Language Models? Exploring Diversity of Generated Codes
di: Chon, Heejae, et al.
Pubblicazione: (2024)
di: Chon, Heejae, et al.
Pubblicazione: (2024)
From Prompts to Performance: Evaluating LLMs for Task-based Parallel Code Generation
di: Bantel, Linus, et al.
Pubblicazione: (2026)
di: Bantel, Linus, et al.
Pubblicazione: (2026)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
di: Fang, Sen, et al.
Pubblicazione: (2025)
di: Fang, Sen, et al.
Pubblicazione: (2025)
AutoMCQ -- Automatically Generate Code Comprehension Questions using GenAI
di: Goodfellow, Martin, et al.
Pubblicazione: (2025)
di: Goodfellow, Martin, et al.
Pubblicazione: (2025)
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
di: Chen, Le, et al.
Pubblicazione: (2025)
di: Chen, Le, et al.
Pubblicazione: (2025)
CodeChain: Towards Modular Code Generation Through Chain of Self-revisions with Representative Sub-modules
di: Le, Hung, et al.
Pubblicazione: (2023)
di: Le, Hung, et al.
Pubblicazione: (2023)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
di: Le-Cong, Thanh, et al.
Pubblicazione: (2025)
di: Le-Cong, Thanh, et al.
Pubblicazione: (2025)
Code Broker: A Multi-Agent System for Automated Code Quality Assessment
di: Attrah, Samer
Pubblicazione: (2026)
di: Attrah, Samer
Pubblicazione: (2026)
Bench4HLS: End-to-End Evaluation of LLMs in High-Level Synthesis Code Generation
di: Khan, M Zafir Sadik, et al.
Pubblicazione: (2026)
di: Khan, M Zafir Sadik, et al.
Pubblicazione: (2026)
A Survey of Neural Code Intelligence: Paradigms, Advances and Beyond
di: Sun, Qiushi, et al.
Pubblicazione: (2024)
di: Sun, Qiushi, et al.
Pubblicazione: (2024)
MCTS-SQL: Light-Weight LLMs can Master the Text-to-SQL through Monte Carlo Tree Search
di: Yuan, Shuozhi, et al.
Pubblicazione: (2025)
di: Yuan, Shuozhi, et al.
Pubblicazione: (2025)
Code Simulation Challenges for Large Language Models
di: La Malfa, Emanuele, et al.
Pubblicazione: (2024)
di: La Malfa, Emanuele, et al.
Pubblicazione: (2024)
ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization
di: Bae, Suyoung, et al.
Pubblicazione: (2026)
di: Bae, Suyoung, et al.
Pubblicazione: (2026)
CSSG: Measuring Code Similarity with Semantic Graphs
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
di: Sun, Zhensu, et al.
Pubblicazione: (2026)
di: Sun, Zhensu, et al.
Pubblicazione: (2026)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
di: Peng, Yun, et al.
Pubblicazione: (2024)
di: Peng, Yun, et al.
Pubblicazione: (2024)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
di: Shi, Yuling, et al.
Pubblicazione: (2024)
di: Shi, Yuling, et al.
Pubblicazione: (2024)
Conditioning LLMs to Generate Code-Switched Text
di: Heredia, Maite, et al.
Pubblicazione: (2025)
di: Heredia, Maite, et al.
Pubblicazione: (2025)
LongCodeBench: Evaluating Coding LLMs at 1M Context Windows
di: Rando, Stefano, et al.
Pubblicazione: (2025)
di: Rando, Stefano, et al.
Pubblicazione: (2025)
Agentic Code Reasoning
di: Ugare, Shubham, et al.
Pubblicazione: (2026)
di: Ugare, Shubham, et al.
Pubblicazione: (2026)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
di: Saieva, Anthony, et al.
Pubblicazione: (2023)
di: Saieva, Anthony, et al.
Pubblicazione: (2023)
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks
di: Dandamudi, Rohit, et al.
Pubblicazione: (2024)
di: Dandamudi, Rohit, et al.
Pubblicazione: (2024)
DOCE: Finding the Sweet Spot for Execution-Based Code Generation
di: Li, Haau-Sing, et al.
Pubblicazione: (2024)
di: Li, Haau-Sing, et al.
Pubblicazione: (2024)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
di: Lang, Nguyet-Anh H., et al.
Pubblicazione: (2026)
di: Lang, Nguyet-Anh H., et al.
Pubblicazione: (2026)
MHRC-Bench: A Multilingual Hardware Repository-Level Code Completion benchmark
di: Zou, Qingyun, et al.
Pubblicazione: (2026)
di: Zou, Qingyun, et al.
Pubblicazione: (2026)
Linguacodus: A Synergistic Framework for Transformative Code Generation in Machine Learning Pipelines
di: Trofimova, Ekaterina, et al.
Pubblicazione: (2024)
di: Trofimova, Ekaterina, et al.
Pubblicazione: (2024)
A Case Study on the Effectiveness of LLMs in Verification with Proof Assistants
di: Bayazıt, Barış, et al.
Pubblicazione: (2025)
di: Bayazıt, Barış, et al.
Pubblicazione: (2025)
VisCoder2: Building Multi-Language Visualization Coding Agents
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CodeMind: Evaluating Large Language Models for Code Reasoning
di: Liu, Changshu, et al.
Pubblicazione: (2024) -
FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation
di: Pham, Loc, et al.
Pubblicazione: (2026) -
Grammar-Based Code Representation: Is It a Worthy Pursuit for LLMs?
di: Liang, Qingyuan, et al.
Pubblicazione: (2025) -
Assessing Code Understanding in LLMs
di: Laneve, Cosimo, et al.
Pubblicazione: (2025) -
AutoCode: LLMs as Problem Setters for Competitive Programming
di: Zhou, Shang, et al.
Pubblicazione: (2025)