AST-T5: Structure-Aware Pretraining for Code Generation and Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Gong, Linyuan, Elhoushi, Mostafa, Cheung, Alvin |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Structure-Aware Fill-in-the-Middle Pretraining for Code
par: Gong, Linyuan, et autres
Publié: (2025)
par: Gong, Linyuan, et autres
Publié: (2025)
Evaluation of LLMs on Syntax-Aware Code Fill-in-the-Middle Tasks
par: Gong, Linyuan, et autres
Publié: (2024)
par: Gong, Linyuan, et autres
Publié: (2024)
MLCPD: A Unified Multi-Language Code Parsing Dataset with Universal AST Schema
par: Gajjar, Jugal, et autres
Publié: (2025)
par: Gajjar, Jugal, et autres
Publié: (2025)
RepoQA: Evaluating Long Context Code Understanding
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
Towards Understanding What Code Language Models Learned
par: Ahmed, Toufique, et autres
Publié: (2023)
par: Ahmed, Toufique, et autres
Publié: (2023)
SelfCodeAlign: Self-Alignment for Code Generation
par: Wei, Yuxiang, et autres
Publié: (2024)
par: Wei, Yuxiang, et autres
Publié: (2024)
IntentCoding: Amplifying User Intent in Code Generation
par: Fang, Zheng, et autres
Publié: (2026)
par: Fang, Zheng, et autres
Publié: (2026)
Planning-Aware Code Infilling via Horizon-Length Prediction
par: Ding, Yifeng, et autres
Publié: (2024)
par: Ding, Yifeng, et autres
Publié: (2024)
CodeJudge: Evaluating Code Generation with Large Language Models
par: Tong, Weixi, et autres
Publié: (2024)
par: Tong, Weixi, et autres
Publié: (2024)
Evaluating Language Models for Efficient Code Generation
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
CONCUR: Benchmarking LLMs for Concurrent Code Generation
par: Huang, Jue, et autres
Publié: (2026)
par: Huang, Jue, et autres
Publié: (2026)
Uncertainty Awareness of Large Language Models Under Code Distribution Shifts: A Benchmark Study
par: Li, Yufei, et autres
Publié: (2024)
par: Li, Yufei, et autres
Publié: (2024)
Wisdom and Delusion of LLM Ensembles for Code Generation and Repair
par: Vallecillos-Ruiz, Fernando, et autres
Publié: (2025)
par: Vallecillos-Ruiz, Fernando, et autres
Publié: (2025)
GiFT: Gibbs Fine-Tuning for Code Generation
par: Li, Haochen, et autres
Publié: (2025)
par: Li, Haochen, et autres
Publié: (2025)
MetaLint: Easy-to-Hard Generalization for Code Linting
par: Naik, Atharva, et autres
Publié: (2025)
par: Naik, Atharva, et autres
Publié: (2025)
Test Code Generation for Telecom Software Systems using Two-Stage Generative Model
par: Nabeel, Mohamad, et autres
Publié: (2024)
par: Nabeel, Mohamad, et autres
Publié: (2024)
Quantifying Contamination in Evaluating Code Generation Capabilities of Language Models
par: Riddell, Martin, et autres
Publié: (2024)
par: Riddell, Martin, et autres
Publié: (2024)
Hybrid-Gym: Training Coding Agents to Generalize Across Tasks
par: Xie, Yiqing, et autres
Publié: (2026)
par: Xie, Yiqing, et autres
Publié: (2026)
Evaluating the Generalization Capabilities of Large Language Models on Code Reasoning
par: Yang, Rem, et autres
Publié: (2025)
par: Yang, Rem, et autres
Publié: (2025)
Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering
par: Ridnik, Tal, et autres
Publié: (2024)
par: Ridnik, Tal, et autres
Publié: (2024)
SceneGenAgent: Precise Industrial Scene Generation with Coding Agent
par: Xia, Xiao, et autres
Publié: (2024)
par: Xia, Xiao, et autres
Publié: (2024)
CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation
par: Peng, Jinjun, et autres
Publié: (2025)
par: Peng, Jinjun, et autres
Publié: (2025)
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation
par: Haider, Md. Asif, et autres
Publié: (2024)
par: Haider, Md. Asif, et autres
Publié: (2024)
CoCoST: Automatic Complex Code Generation with Online Searching and Correctness Testing
par: He, Xinyi, et autres
Publié: (2024)
par: He, Xinyi, et autres
Publié: (2024)
Neuron Patching: Semantic-based Neuron-level Language Model Repair for Code Generation
par: Gu, Jian, et autres
Publié: (2023)
par: Gu, Jian, et autres
Publié: (2023)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
par: Gu, Jian, et autres
Publié: (2025)
par: Gu, Jian, et autres
Publié: (2025)
Exploring Parameter-Efficient Fine-Tuning Techniques for Code Generation with Large Language Models
par: Weyssow, Martin, et autres
Publié: (2023)
par: Weyssow, Martin, et autres
Publié: (2023)
GLLM: Self-Corrective G-Code Generation using Large Language Models with User Feedback
par: Abdelaal, Mohamed, et autres
Publié: (2025)
par: Abdelaal, Mohamed, et autres
Publié: (2025)
TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models
par: Park, Chansung, et autres
Publié: (2026)
par: Park, Chansung, et autres
Publié: (2026)
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
par: Jain, Naman, et autres
Publié: (2024)
par: Jain, Naman, et autres
Publié: (2024)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
par: Mayer, Luis, et autres
Publié: (2024)
par: Mayer, Luis, et autres
Publié: (2024)
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
par: Jiang, Juyong, et autres
Publié: (2026)
par: Jiang, Juyong, et autres
Publié: (2026)
NaturalCodeBench: Examining Coding Performance Mismatch on HumanEval and Natural User Prompts
par: Zhang, Shudan, et autres
Publié: (2024)
par: Zhang, Shudan, et autres
Publié: (2024)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
par: Xia, Yunhui, et autres
Publié: (2025)
par: Xia, Yunhui, et autres
Publié: (2025)
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences
par: Weyssow, Martin, et autres
Publié: (2024)
par: Weyssow, Martin, et autres
Publié: (2024)
PurpCode: Reasoning for Safer Code Generation
par: Liu, Jiawei, et autres
Publié: (2025)
par: Liu, Jiawei, et autres
Publié: (2025)
Learning Code Preference via Synthetic Evolution
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
StackEval: Benchmarking LLMs in Coding Assistance
par: Shah, Nidhish, et autres
Publié: (2024)
par: Shah, Nidhish, et autres
Publié: (2024)
Automatic Pull Request Description Generation Using LLMs: A T5 Model Approach
par: Sakib, Md Nazmus, et autres
Publié: (2024)
par: Sakib, Md Nazmus, et autres
Publié: (2024)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
par: Fan, Lishui, et autres
Publié: (2025)
par: Fan, Lishui, et autres
Publié: (2025)
Documents similaires
-
Structure-Aware Fill-in-the-Middle Pretraining for Code
par: Gong, Linyuan, et autres
Publié: (2025) -
Evaluation of LLMs on Syntax-Aware Code Fill-in-the-Middle Tasks
par: Gong, Linyuan, et autres
Publié: (2024) -
MLCPD: A Unified Multi-Language Code Parsing Dataset with Universal AST Schema
par: Gajjar, Jugal, et autres
Publié: (2025) -
RepoQA: Evaluating Long Context Code Understanding
par: Liu, Jiawei, et autres
Publié: (2024) -
Towards Understanding What Code Language Models Learned
par: Ahmed, Toufique, et autres
Publié: (2023)