Grounding Data Science Code Generation with Input-Output Specifications
Fuente:
arXiv
Saved in:
| Main Authors: | Wen, Yeming, Yin, Pengcheng, Shi, Kensen, Michalewski, Henryk, Chaudhuri, Swarat, Polozov, Alex |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NExT: Teaching Large Language Models to Reason about Code Execution
by: Ni, Ansong, et al.
Published: (2024)
by: Ni, Ansong, et al.
Published: (2024)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
by: Thakur, Amitayush, et al.
Published: (2025)
by: Thakur, Amitayush, et al.
Published: (2025)
MIO: Multiverse Debugging in the Face of Input/Output -- Extended Version with Additional Appendices
by: Lauwaerts, Tom, et al.
Published: (2025)
by: Lauwaerts, Tom, et al.
Published: (2025)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
by: Petrukha, Ivan, et al.
Published: (2025)
by: Petrukha, Ivan, et al.
Published: (2025)
Correctness-Guaranteed Code Generation via Constrained Decoding
by: Li, Lingxiao, et al.
Published: (2025)
by: Li, Lingxiao, et al.
Published: (2025)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
by: Kadosh, Tal, et al.
Published: (2023)
by: Kadosh, Tal, et al.
Published: (2023)
SIMCOPILOT: Evaluating Large Language Models for Copilot-Style Code Generation
by: Jiang, Mingchao, et al.
Published: (2025)
by: Jiang, Mingchao, et al.
Published: (2025)
Batched Low-Rank Adaptation of Foundation Models
by: Wen, Yeming, et al.
Published: (2023)
by: Wen, Yeming, et al.
Published: (2023)
Large Language Models for Multilingual Code Intelligence: A Survey
by: Jiang, Chao, et al.
Published: (2026)
by: Jiang, Chao, et al.
Published: (2026)
Input-Gen: Guided Generation of Stateful Inputs for Testing, Tuning, and Training
by: Ivanov, Ivan R., et al.
Published: (2024)
by: Ivanov, Ivan R., et al.
Published: (2024)
Automatically Testing Functional Properties of Code Translation Models
by: Eniser, Hasan Ferit, et al.
Published: (2023)
by: Eniser, Hasan Ferit, et al.
Published: (2023)
Language Models for Code Completion: A Practical Evaluation
by: Izadi, Maliheh, et al.
Published: (2024)
by: Izadi, Maliheh, et al.
Published: (2024)
Pareto Optimal Code Generation
by: Orlanski, Gabriel, et al.
Published: (2025)
by: Orlanski, Gabriel, et al.
Published: (2025)
Finding Missed Code Size Optimizations in Compilers using LLMs
by: Italiano, Davide, et al.
Published: (2024)
by: Italiano, Davide, et al.
Published: (2024)
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing
by: Wei, Jiayi, et al.
Published: (2023)
by: Wei, Jiayi, et al.
Published: (2023)
GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2
by: Kashmira, Savini, et al.
Published: (2025)
by: Kashmira, Savini, et al.
Published: (2025)
FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback
by: Wu, Xueqing, et al.
Published: (2025)
by: Wu, Xueqing, et al.
Published: (2025)
MLCPD: A Unified Multi-Language Code Parsing Dataset with Universal AST Schema
by: Gajjar, Jugal, et al.
Published: (2025)
by: Gajjar, Jugal, et al.
Published: (2025)
CodeFuse-Query: A Data-Centric Static Code Analysis System for Large-Scale Organizations
by: Xie, Xiaoheng, et al.
Published: (2024)
by: Xie, Xiaoheng, et al.
Published: (2024)
Automated Code Editing with Search-Generate-Modify
by: Liu, Changshu, et al.
Published: (2023)
by: Liu, Changshu, et al.
Published: (2023)
Top Leaderboard Ranking = Top Coding Proficiency, Always? EvoEval: Evolving Coding Benchmarks via LLM
by: Xia, Chunqiu Steven, et al.
Published: (2024)
by: Xia, Chunqiu Steven, et al.
Published: (2024)
Neural Models for Source Code Synthesis and Completion
by: Niyogi, Mitodru
Published: (2024)
by: Niyogi, Mitodru
Published: (2024)
Python Symbolic Execution with LLM-powered Code Generation
by: Wang, Wenhan, et al.
Published: (2024)
by: Wang, Wenhan, et al.
Published: (2024)
A Multi-Perspective Architecture for Semantic Code Search
by: Haldar, Rajarshi, et al.
Published: (2020)
by: Haldar, Rajarshi, et al.
Published: (2020)
Constrained Decoding for Fill-in-the-Middle Code Language Models via Efficient Left and Right Quotienting of Context-Sensitive Grammars
by: Melcer, Daniel, et al.
Published: (2024)
by: Melcer, Daniel, et al.
Published: (2024)
RacerF: Lightweight Static Data Race Detection for C Code
by: Dacík, Tomáš, et al.
Published: (2025)
by: Dacík, Tomáš, et al.
Published: (2025)
SynCode: LLM Generation with Grammar Augmentation
by: Ugare, Shubham, et al.
Published: (2024)
by: Ugare, Shubham, et al.
Published: (2024)
CPSLint: A Domain-Specific Language Providing Data Validation and Sanitisation for Industrial Cyber-Physical Systems
by: Odyurt, Uraz, et al.
Published: (2025)
by: Odyurt, Uraz, et al.
Published: (2025)
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
by: Chen, Le, et al.
Published: (2025)
by: Chen, Le, et al.
Published: (2025)
Can Large Language Models Simulate Symbolic Execution Output Like KLEE?
by: Feng, Rong, et al.
Published: (2025)
by: Feng, Rong, et al.
Published: (2025)
Incoherence as Oracle-less Measure of Error in LLM-Based Code Generation
by: Valentin, Thomas, et al.
Published: (2025)
by: Valentin, Thomas, et al.
Published: (2025)
Strengthening Programming Comprehension in Large Language Models through Code Generation
by: Ren, Xiaoning, et al.
Published: (2025)
by: Ren, Xiaoning, et al.
Published: (2025)
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
by: Bouras, Dimitrios Stamatios, et al.
Published: (2026)
by: Bouras, Dimitrios Stamatios, et al.
Published: (2026)
QEDCartographer: Automating Formal Verification Using Reward-Free Reinforcement Learning
by: Sanchez-Stern, Alex, et al.
Published: (2024)
by: Sanchez-Stern, Alex, et al.
Published: (2024)
The CodeInverter Suite: Control-Flow and Data-Mapping Augmented Binary Decompilation with LLMs
by: Liu, Peipei, et al.
Published: (2025)
by: Liu, Peipei, et al.
Published: (2025)
A Multi-Expert Large Language Model Architecture for Verilog Code Generation
by: Nadimi, Bardia, et al.
Published: (2024)
by: Nadimi, Bardia, et al.
Published: (2024)
COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation
by: Dau, Anh T. V., et al.
Published: (2026)
by: Dau, Anh T. V., et al.
Published: (2026)
Finding Compiler Bugs through Cross-Language Code Generator and Differential Testing
by: Feng, Qiong, et al.
Published: (2025)
by: Feng, Qiong, et al.
Published: (2025)
Challenges of Multilingual Program Specification and Analysis
by: Furia, Carlo A., et al.
Published: (2024)
by: Furia, Carlo A., et al.
Published: (2024)
FormulaCode: Evaluating Agentic Optimization on Large Codebases
by: Sehgal, Atharva, et al.
Published: (2026)
by: Sehgal, Atharva, et al.
Published: (2026)
Similar Items
-
NExT: Teaching Large Language Models to Reason about Code Execution
by: Ni, Ansong, et al.
Published: (2024) -
CLEVER: A Curated Benchmark for Formally Verified Code Generation
by: Thakur, Amitayush, et al.
Published: (2025) -
MIO: Multiverse Debugging in the Face of Input/Output -- Extended Version with Additional Appendices
by: Lauwaerts, Tom, et al.
Published: (2025) -
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
by: Petrukha, Ivan, et al.
Published: (2025) -
Correctness-Guaranteed Code Generation via Constrained Decoding
by: Li, Lingxiao, et al.
Published: (2025)