On Pretraining for Project-Level Code Completion
Fuente:
arXiv
Salvato in:
| Autori principali: | Sapronov, Maksim, Glukhov, Evgeniy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Challenge on Optimization of Context Collection for Code Completion
di: Ustalov, Dmitry, et al.
Pubblicazione: (2025)
di: Ustalov, Dmitry, et al.
Pubblicazione: (2025)
Diff-XYZ: A Benchmark for Evaluating Diff Understanding
di: Glukhov, Evgeniy, et al.
Pubblicazione: (2025)
di: Glukhov, Evgeniy, et al.
Pubblicazione: (2025)
Long Code Arena: a Set of Benchmarks for Long-Context Code Models
di: Bogomolov, Egor, et al.
Pubblicazione: (2024)
di: Bogomolov, Egor, et al.
Pubblicazione: (2024)
Context Composing for Full Line Code Completion
di: Semenkin, Anton, et al.
Pubblicazione: (2024)
di: Semenkin, Anton, et al.
Pubblicazione: (2024)
Full Line Code Completion: Bringing AI to Desktop
di: Semenkin, Anton, et al.
Pubblicazione: (2024)
di: Semenkin, Anton, et al.
Pubblicazione: (2024)
Can Coding Agents Be General Agents?
di: Ivanov, Maksim, et al.
Pubblicazione: (2026)
di: Ivanov, Maksim, et al.
Pubblicazione: (2026)
Mellum: Production-Grade in-IDE Contextual Code Completion with Multi-File Project Understanding
di: Pavlichenko, Nikita, et al.
Pubblicazione: (2025)
di: Pavlichenko, Nikita, et al.
Pubblicazione: (2025)
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
di: Bui, Tuan-Dung, et al.
Pubblicazione: (2024)
di: Bui, Tuan-Dung, et al.
Pubblicazione: (2024)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
di: Rahman, Imranur, et al.
Pubblicazione: (2025)
di: Rahman, Imranur, et al.
Pubblicazione: (2025)
Model Cascading for Code: A Cascaded Black-Box Multi-Model Framework for Cost-Efficient Code Completion with Self-Testing
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding
di: Gong, Linyuan, et al.
Pubblicazione: (2024)
di: Gong, Linyuan, et al.
Pubblicazione: (2024)
CodeFuse-13B: A Pretrained Multi-lingual Code Large Language Model
di: Di, Peng, et al.
Pubblicazione: (2023)
di: Di, Peng, et al.
Pubblicazione: (2023)
Renaissance of Literate Programming in the Era of LLMs: Enhancing LLM-Based Code Generation in Large-Scale Projects
di: Zhang, Wuyang, et al.
Pubblicazione: (2024)
di: Zhang, Wuyang, et al.
Pubblicazione: (2024)
RocketPPA: Code-Level Power, Performance, and Area Prediction via LLM and Mixture of Experts
di: Abdollahi, Armin, et al.
Pubblicazione: (2025)
di: Abdollahi, Armin, et al.
Pubblicazione: (2025)
MatchFixAgent: Language-Agnostic Autonomous Repository-Level Code Translation Validation and Repair
di: Ibrahimzada, Ali Reza, et al.
Pubblicazione: (2025)
di: Ibrahimzada, Ali Reza, et al.
Pubblicazione: (2025)
AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validation
di: Ibrahimzada, Ali Reza, et al.
Pubblicazione: (2024)
di: Ibrahimzada, Ali Reza, et al.
Pubblicazione: (2024)
Language Models for Code Completion: A Practical Evaluation
di: Izadi, Maliheh, et al.
Pubblicazione: (2024)
di: Izadi, Maliheh, et al.
Pubblicazione: (2024)
Code Graph Model (CGM): A Graph-Integrated Large Language Model for Repository-Level Software Engineering Tasks
di: Tao, Hongyuan, et al.
Pubblicazione: (2025)
di: Tao, Hongyuan, et al.
Pubblicazione: (2025)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
di: Thillen, Alex, et al.
Pubblicazione: (2026)
di: Thillen, Alex, et al.
Pubblicazione: (2026)
JetTrain: IDE-Native Machine Learning Experiments
di: Trofimov, Artem, et al.
Pubblicazione: (2024)
di: Trofimov, Artem, et al.
Pubblicazione: (2024)
CodeSAM: Source Code Representation Learning by Infusing Self-Attention with Multi-Code-View Graphs
di: Mathai, Alex, et al.
Pubblicazione: (2024)
di: Mathai, Alex, et al.
Pubblicazione: (2024)
SemRep: Generative Code Representation Learning with Code Transformations
di: Li, Weichen, et al.
Pubblicazione: (2026)
di: Li, Weichen, et al.
Pubblicazione: (2026)
On The Importance of Reasoning for Context Retrieval in Repository-Level Code Editing
di: Kovrigin, Alexander, et al.
Pubblicazione: (2024)
di: Kovrigin, Alexander, et al.
Pubblicazione: (2024)
Code Roulette: How Prompt Variability Affects LLM Code Generation
di: Paleyes, Andrei, et al.
Pubblicazione: (2025)
di: Paleyes, Andrei, et al.
Pubblicazione: (2025)
Can I Solve It? Identifying APIs Required to Complete OSS Task
di: Santos, Fabio, et al.
Pubblicazione: (2021)
di: Santos, Fabio, et al.
Pubblicazione: (2021)
Leveraging Code Cohesion Analysis to Identify Source Code Supply Chain Attacks
di: Reuben, Maor, et al.
Pubblicazione: (2025)
di: Reuben, Maor, et al.
Pubblicazione: (2025)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
di: Yan, Kaiwen, et al.
Pubblicazione: (2025)
di: Yan, Kaiwen, et al.
Pubblicazione: (2025)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
Neural Models for Source Code Synthesis and Completion
di: Niyogi, Mitodru
Pubblicazione: (2024)
di: Niyogi, Mitodru
Pubblicazione: (2024)
Think Anywhere in Code Generation
di: Jiang, Xue, et al.
Pubblicazione: (2026)
di: Jiang, Xue, et al.
Pubblicazione: (2026)
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
di: Tsai, Yun-Da, et al.
Pubblicazione: (2024)
di: Tsai, Yun-Da, et al.
Pubblicazione: (2024)
Teaching Code Refactoring Using LLMs
di: Khairnar, Anshul, et al.
Pubblicazione: (2025)
di: Khairnar, Anshul, et al.
Pubblicazione: (2025)
Towards Verified Code Reasoning by LLMs
di: Sistla, Meghana, et al.
Pubblicazione: (2025)
di: Sistla, Meghana, et al.
Pubblicazione: (2025)
Robust Learning of Diverse Code Edits
di: Aggarwal, Tushar, et al.
Pubblicazione: (2025)
di: Aggarwal, Tushar, et al.
Pubblicazione: (2025)
Calibration and Correctness of Language Models for Code
di: Spiess, Claudio, et al.
Pubblicazione: (2024)
di: Spiess, Claudio, et al.
Pubblicazione: (2024)
Bootstrapping Coding Agents: The Specification Is the Program
di: Monperrus, Martin
Pubblicazione: (2026)
di: Monperrus, Martin
Pubblicazione: (2026)
Continuous Integration Practices in Machine Learning Projects: The Practitioners` Perspective
di: Bernardo, João Helis, et al.
Pubblicazione: (2025)
di: Bernardo, João Helis, et al.
Pubblicazione: (2025)
A Flexible Cell Classification for ML Projects in Jupyter Notebooks
di: Perez, Miguel, et al.
Pubblicazione: (2024)
di: Perez, Miguel, et al.
Pubblicazione: (2024)
LLM Performance for Code Generation on Noisy Tasks
di: Sendyka, Radzim, et al.
Pubblicazione: (2025)
di: Sendyka, Radzim, et al.
Pubblicazione: (2025)
OSS-Bench: Benchmark Generator for Coding LLMs
di: Jiang, Yuancheng, et al.
Pubblicazione: (2025)
di: Jiang, Yuancheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Challenge on Optimization of Context Collection for Code Completion
di: Ustalov, Dmitry, et al.
Pubblicazione: (2025) -
Diff-XYZ: A Benchmark for Evaluating Diff Understanding
di: Glukhov, Evgeniy, et al.
Pubblicazione: (2025) -
Long Code Arena: a Set of Benchmarks for Long-Context Code Models
di: Bogomolov, Egor, et al.
Pubblicazione: (2024) -
Context Composing for Full Line Code Completion
di: Semenkin, Anton, et al.
Pubblicazione: (2024) -
Full Line Code Completion: Bringing AI to Desktop
di: Semenkin, Anton, et al.
Pubblicazione: (2024)