LLMs Lean on Priors, Not Programming Language Semantics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thimmaiah, Aditya, Zhang, Jiyang, Srinivasa, Jayanth, Li, Junyi Jessy, Gligoric, Milos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pattern-Based Peephole Optimizations with Java JIT Tests
von: Zang, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Zang, Zhiqiang, et al.
Veröffentlicht: (2024)
Object Graph Programming
von: Thimmaiah, Aditya, et al.
Veröffentlicht: (2024)
von: Thimmaiah, Aditya, et al.
Veröffentlicht: (2024)
exLong: Generating Exceptional Behavior Tests with Large Language Models
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
von: Wang, Ming, et al.
Veröffentlicht: (2024)
von: Wang, Ming, et al.
Veröffentlicht: (2024)
AutoCode: LLMs as Problem Setters for Competitive Programming
von: Zhou, Shang, et al.
Veröffentlicht: (2025)
von: Zhou, Shang, et al.
Veröffentlicht: (2025)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
Ranking LLM-Generated Loop Invariants for Program Verification
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
Benchmarking LLM Code Generation for Audio Programming with Visual Dataflow Languages
von: Zhang, William, et al.
Veröffentlicht: (2024)
von: Zhang, William, et al.
Veröffentlicht: (2024)
Is Programming by Example solved by LLMs?
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Refactoring Programs Using Large Language Models with Few-Shot Examples
von: Shirafuji, Atsushi, et al.
Veröffentlicht: (2023)
von: Shirafuji, Atsushi, et al.
Veröffentlicht: (2023)
SuperCoder: Assembly Program Superoptimization with Large Language Models
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
von: Tang, Hao, et al.
Veröffentlicht: (2024)
von: Tang, Hao, et al.
Veröffentlicht: (2024)
Can LLMs Enable Verification in Mainstream Programming?
von: Shefer, Aleksandr, et al.
Veröffentlicht: (2025)
von: Shefer, Aleksandr, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Compiler and Language Server Feedback
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
PPM: Automated Generation of Diverse Programming Problems for Benchmarking Code Generation Models
von: Chen, Simin, et al.
Veröffentlicht: (2024)
von: Chen, Simin, et al.
Veröffentlicht: (2024)
Structured Program Synthesis using LLMs: Results and Insights from the IPARC Challenge
von: Surana, Shraddha, et al.
Veröffentlicht: (2025)
von: Surana, Shraddha, et al.
Veröffentlicht: (2025)
CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language
von: Cheng, Junhang, et al.
Veröffentlicht: (2026)
von: Cheng, Junhang, et al.
Veröffentlicht: (2026)
Program Skeletons for Automated Program Translation
von: Wang, Bo, et al.
Veröffentlicht: (2025)
von: Wang, Bo, et al.
Veröffentlicht: (2025)
CodeMind: Evaluating Large Language Models for Code Reasoning
von: Liu, Changshu, et al.
Veröffentlicht: (2024)
von: Liu, Changshu, et al.
Veröffentlicht: (2024)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
VisCoder2: Building Multi-Language Visualization Coding Agents
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
Enhancing Automated Loop Invariant Generation for Complex Programs with Large Language Models
von: Liu, Ruibang, et al.
Veröffentlicht: (2024)
von: Liu, Ruibang, et al.
Veröffentlicht: (2024)
Towards Repository-Level Program Verification with Large Language Models
von: Zhong, Si Cheng, et al.
Veröffentlicht: (2025)
von: Zhong, Si Cheng, et al.
Veröffentlicht: (2025)
LLMON: An LLM-native Markup Language to Leverage Structure and Semantics at the LLM Interface
von: Hind, Michael, et al.
Veröffentlicht: (2026)
von: Hind, Michael, et al.
Veröffentlicht: (2026)
Natural Language-Oriented Programming (NLOP): Towards Democratizing Software Creation
von: Beheshti, Amin
Veröffentlicht: (2024)
von: Beheshti, Amin
Veröffentlicht: (2024)
Static Program Slicing Using Language Models With Dataflow-Aware Pretraining and Constrained Decoding
von: He, Pengfei, et al.
Veröffentlicht: (2026)
von: He, Pengfei, et al.
Veröffentlicht: (2026)
OSVBench: Benchmarking LLMs on Specification Generation Tasks for Operating System Verification
von: Li, Shangyu, et al.
Veröffentlicht: (2025)
von: Li, Shangyu, et al.
Veröffentlicht: (2025)
A Taxonomy of Prompt Defects in LLM Systems
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
Self-Improving Code Generation via Semantic Entropy and Behavioral Consensus
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
von: Dumitran, Adrian Marius, et al.
Veröffentlicht: (2024)
von: Dumitran, Adrian Marius, et al.
Veröffentlicht: (2024)
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
von: Xia, Shihao, et al.
Veröffentlicht: (2026)
von: Xia, Shihao, et al.
Veröffentlicht: (2026)
A Survey of Neural Code Intelligence: Paradigms, Advances and Beyond
von: Sun, Qiushi, et al.
Veröffentlicht: (2024)
von: Sun, Qiushi, et al.
Veröffentlicht: (2024)
InvAASTCluster: On Applying Invariant-Based Program Clustering to Introductory Programming Assignments
von: Orvalho, Pedro, et al.
Veröffentlicht: (2022)
von: Orvalho, Pedro, et al.
Veröffentlicht: (2022)
Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs
von: Zhang, Weixing, et al.
Veröffentlicht: (2025)
von: Zhang, Weixing, et al.
Veröffentlicht: (2025)
Raw Pointer Rewriting with LLMs for Translating C to Safer Rust
von: Gao, Yifei, et al.
Veröffentlicht: (2025)
von: Gao, Yifei, et al.
Veröffentlicht: (2025)
The New Compiler Stack: A Survey on the Synergy of LLMs and Compilers
von: Zhang, Shuoming, et al.
Veröffentlicht: (2026)
von: Zhang, Shuoming, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Pattern-Based Peephole Optimizations with Java JIT Tests
von: Zang, Zhiqiang, et al.
Veröffentlicht: (2024) -
Object Graph Programming
von: Thimmaiah, Aditya, et al.
Veröffentlicht: (2024) -
exLong: Generating Exceptional Behavior Tests with Large Language Models
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024) -
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
von: Wang, Ming, et al.
Veröffentlicht: (2024) -
AutoCode: LLMs as Problem Setters for Competitive Programming
von: Zhou, Shang, et al.
Veröffentlicht: (2025)