Scaling Laws Behind Code Understanding Model
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lin, Jiayi, Dong, Hande, Xie, Yutao, Zhang, Lei |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Improving Code Search with Hard Negative Sampling Based on Fine-tuning
par: Dong, Hande, et autres
Publié: (2023)
par: Dong, Hande, et autres
Publié: (2023)
Towards Better Code Understanding in Decoder-Only Models with Contrastive Learning
par: Lin, Jiayi, et autres
Publié: (2024)
par: Lin, Jiayi, et autres
Publié: (2024)
AP2O-Coder: Adaptively Progressive Preference Optimization for Reducing Compilation and Runtime Errors in LLM-Generated Code
par: Zhang, Jianqing, et autres
Publié: (2025)
par: Zhang, Jianqing, et autres
Publié: (2025)
AOCI: Symbolic-Semantic Indexing for Practical Repository-Scale Code Understanding with LLMs
par: Liu, Jinshi, et autres
Publié: (2026)
par: Liu, Jinshi, et autres
Publié: (2026)
Debt Behind the AI Boom: A Large-Scale Empirical Study of AI-Generated Code in the Wild
par: Liu, Yue, et autres
Publié: (2026)
par: Liu, Yue, et autres
Publié: (2026)
Understanding Code Change with Micro-Changes
par: Chen, Lei, et autres
Publié: (2024)
par: Chen, Lei, et autres
Publié: (2024)
Towards Understanding the Characteristics of Code Generation Errors Made by Large Language Models
par: Wang, Zhijie, et autres
Publié: (2024)
par: Wang, Zhijie, et autres
Publié: (2024)
LONGCODEU: Benchmarking Long-Context Language Models on Long Code Understanding
par: Li, Jia, et autres
Publié: (2025)
par: Li, Jia, et autres
Publié: (2025)
CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding
par: Shi, Yuling, et autres
Publié: (2026)
par: Shi, Yuling, et autres
Publié: (2026)
When LLMs Lag Behind: Knowledge Conflicts from Evolving APIs in Code Generation
par: Ashik, Ahmed Nusayer, et autres
Publié: (2026)
par: Ashik, Ahmed Nusayer, et autres
Publié: (2026)
Understanding Code Understandability Improvements in Code Reviews
par: Oliveira, Delano, et autres
Publié: (2024)
par: Oliveira, Delano, et autres
Publié: (2024)
Enhancing and Reporting Robustness Boundary of Neural Code Models for Intelligent Code Understanding
par: Han, Tingxu, et autres
Publié: (2026)
par: Han, Tingxu, et autres
Publié: (2026)
Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer
par: Wang, Tianfu, et autres
Publié: (2026)
par: Wang, Tianfu, et autres
Publié: (2026)
Evaluating Small-Scale Code Models for Code Clone Detection
par: Martinez-Gil, Jorge
Publié: (2025)
par: Martinez-Gil, Jorge
Publié: (2025)
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
par: Zhang, Kechi, et autres
Publié: (2024)
par: Zhang, Kechi, et autres
Publié: (2024)
"Refactoring Runaway": Understanding and Mitigating Tangled Refactorings in Coding Agents for Issue Resolution
par: Tian, Zhao, et autres
Publié: (2026)
par: Tian, Zhao, et autres
Publié: (2026)
Understanding Everything as Code: A Taxonomy and Conceptual Model
par: Wei, Haoran, et autres
Publié: (2025)
par: Wei, Haoran, et autres
Publié: (2025)
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
par: Nakamura, Ibuki, et autres
Publié: (2025)
par: Nakamura, Ibuki, et autres
Publié: (2025)
GenCode: A Generic Data Augmentation Framework for Boosting Deep Learning-Based Code Understanding
par: Dong, Zeming, et autres
Publié: (2024)
par: Dong, Zeming, et autres
Publié: (2024)
A Hybrid Approach for EMF Code Generation:Code Templates Meet Large Language Models
par: He, Xiao, et autres
Publié: (2025)
par: He, Xiao, et autres
Publié: (2025)
Unveiling Code Clones in the Eclipse IIoT Software Ecosystem
par: Li, Zengyang, et autres
Publié: (2026)
par: Li, Zengyang, et autres
Publié: (2026)
One Bug, Hundreds Behind: LLMs for Large-Scale Bug Discovery
par: Wu, Qiushi, et autres
Publié: (2025)
par: Wu, Qiushi, et autres
Publié: (2025)
CodeGlance: Understanding Code Reasoning Challenges in LLMs through Multi-Dimensional Feature Analysis
par: Wang, Yunkun, et autres
Publié: (2026)
par: Wang, Yunkun, et autres
Publié: (2026)
CodeJudge-Eval: Can Large Language Models be Good Judges in Code Understanding?
par: Zhao, Yuwei, et autres
Publié: (2024)
par: Zhao, Yuwei, et autres
Publié: (2024)
Completion by Comprehension: Guiding Code Generation with Multi-Granularity Understanding
par: Zhao, Xinkui, et autres
Publié: (2025)
par: Zhao, Xinkui, et autres
Publié: (2025)
Understanding the AI-powered Binary Code Similarity Detection
par: Fu, Lirong, et autres
Publié: (2024)
par: Fu, Lirong, et autres
Publié: (2024)
PSD2Code: Automated Front-End Code Generation from Design Files via Multimodal Large Language Models
par: Chen, Yongxi, et autres
Publié: (2025)
par: Chen, Yongxi, et autres
Publié: (2025)
Change Logging and Mining of Change Logs of Business Processes -- A Literature Review
par: Ghahderijani, Arash Yadegari, et autres
Publié: (2025)
par: Ghahderijani, Arash Yadegari, et autres
Publié: (2025)
A Lightweight Framework for Adaptive Retrieval In Code Completion With Critique Model
par: Zhang, Wenrui, et autres
Publié: (2024)
par: Zhang, Wenrui, et autres
Publié: (2024)
Understanding Emojis :) in Useful Code Review Comments
par: Ahmed, Sharif, et autres
Publié: (2024)
par: Ahmed, Sharif, et autres
Publié: (2024)
Demystifying and Assessing Code Understandability in Java Decompilation
par: Qin, Ruixin, et autres
Publié: (2024)
par: Qin, Ruixin, et autres
Publié: (2024)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
par: Wang, Kaixin, et autres
Publié: (2025)
par: Wang, Kaixin, et autres
Publié: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
par: Manh, Dung Nguyen, et autres
Publié: (2024)
par: Manh, Dung Nguyen, et autres
Publié: (2024)
CodeSSM: Towards State Space Models for Code Understanding
par: Verma, Shweta, et autres
Publié: (2025)
par: Verma, Shweta, et autres
Publié: (2025)
Dynamic Scaling of Unit Tests for Code Reward Modeling
par: Ma, Zeyao, et autres
Publié: (2025)
par: Ma, Zeyao, et autres
Publié: (2025)
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
par: He, Mengliang, et autres
Publié: (2025)
par: He, Mengliang, et autres
Publié: (2025)
Detecting Essence Code Clones via Information Theoretic Analysis
par: Zhao, Lida, et autres
Publié: (2025)
par: Zhao, Lida, et autres
Publié: (2025)
Unveiling Memorization in Code Models
par: Yang, Zhou, et autres
Publié: (2023)
par: Yang, Zhou, et autres
Publié: (2023)
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
par: Guo, Chengquan, et autres
Publié: (2025)
par: Guo, Chengquan, et autres
Publié: (2025)
Understanding Chain-of-Thought Effectiveness in Code Generation: An Empirical and Information-Theoretic Analysis
par: Jin, Naizhu, et autres
Publié: (2025)
par: Jin, Naizhu, et autres
Publié: (2025)
Documents similaires
-
Improving Code Search with Hard Negative Sampling Based on Fine-tuning
par: Dong, Hande, et autres
Publié: (2023) -
Towards Better Code Understanding in Decoder-Only Models with Contrastive Learning
par: Lin, Jiayi, et autres
Publié: (2024) -
AP2O-Coder: Adaptively Progressive Preference Optimization for Reducing Compilation and Runtime Errors in LLM-Generated Code
par: Zhang, Jianqing, et autres
Publié: (2025) -
AOCI: Symbolic-Semantic Indexing for Practical Repository-Scale Code Understanding with LLMs
par: Liu, Jinshi, et autres
Publié: (2026) -
Debt Behind the AI Boom: A Large-Scale Empirical Study of AI-Generated Code in the Wild
par: Liu, Yue, et autres
Publié: (2026)