GenCode: A Generic Data Augmentation Framework for Boosting Deep Learning-Based Code Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Dong, Zeming, Hu, Qiang, Xie, Xiaofei, Cordy, Maxime, Papadakis, Mike, Traon, Yves Le, Zhao, Jianjun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Boosting Source Code Learning with Text-Oriented Data Augmentation: An Empirical Study
di: Dong, Zeming, et al.
Pubblicazione: (2023)
di: Dong, Zeming, et al.
Pubblicazione: (2023)
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026)
di: Akli, Amal, et al.
Pubblicazione: (2026)
One Model, Many Skills: Parameter-Efficient Fine-Tuning for Multitask Code Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026)
di: Akli, Amal, et al.
Pubblicazione: (2026)
Software Fairness: An Analysis and Survey
di: Soremekun, Ezekiel, et al.
Pubblicazione: (2022)
di: Soremekun, Ezekiel, et al.
Pubblicazione: (2022)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
di: Larbi, Maya, et al.
Pubblicazione: (2025)
di: Larbi, Maya, et al.
Pubblicazione: (2025)
Learning Generalizable Multimodal Representations for Software Vulnerability Detection
di: Dong, Zeming, et al.
Pubblicazione: (2026)
di: Dong, Zeming, et al.
Pubblicazione: (2026)
Foundation Models for Autonomous Driving System: An Initial Roadmap
di: Wu, Xiongfei, et al.
Pubblicazione: (2025)
di: Wu, Xiongfei, et al.
Pubblicazione: (2025)
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
di: AKLI, Amal, et al.
Pubblicazione: (2026)
di: AKLI, Amal, et al.
Pubblicazione: (2026)
Inferring Code Correctness from Specification
di: Florian, Tambon, et al.
Pubblicazione: (2026)
di: Florian, Tambon, et al.
Pubblicazione: (2026)
An Empirical Study of the Imbalance Issue in Software Vulnerability Detection
di: Guo, Yuejun, et al.
Pubblicazione: (2026)
di: Guo, Yuejun, et al.
Pubblicazione: (2026)
Do Code Semantics Help? A Comprehensive Study on Execution Trace-Based Information for Code Large Language Models
di: Wang, Jian, et al.
Pubblicazione: (2025)
di: Wang, Jian, et al.
Pubblicazione: (2025)
Evaluation and Improvement of Fault Detection for Large Language Models
di: Hu, Qiang, et al.
Pubblicazione: (2024)
di: Hu, Qiang, et al.
Pubblicazione: (2024)
AnomalyGen: Enhancing Log-Based Anomaly Detection with Code-Guided Data Augmentation
di: Li, Xinyu, et al.
Pubblicazione: (2026)
di: Li, Xinyu, et al.
Pubblicazione: (2026)
Scaling Laws Behind Code Understanding Model
di: Lin, Jiayi, et al.
Pubblicazione: (2024)
di: Lin, Jiayi, et al.
Pubblicazione: (2024)
On the Effectiveness of Hybrid Pooling in Mixup-Based Graph Learning for Language Processing
di: Dong, Zeming, et al.
Pubblicazione: (2022)
di: Dong, Zeming, et al.
Pubblicazione: (2022)
CodeBenchGen: Creating Scalable Execution-based Code Generation Benchmarks
di: Xie, Yiqing, et al.
Pubblicazione: (2024)
di: Xie, Yiqing, et al.
Pubblicazione: (2024)
Search-Induced Issues in Web-Augmented LLM Code Generation: Detecting and Repairing Error-Inducing Pages
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
Understanding Code Understandability Improvements in Code Reviews
di: Oliveira, Delano, et al.
Pubblicazione: (2024)
di: Oliveira, Delano, et al.
Pubblicazione: (2024)
ConAIR:Consistency-Augmented Iterative Interaction Framework to Enhance the Reliability of Code Generation
di: Dong, Jinhao, et al.
Pubblicazione: (2024)
di: Dong, Jinhao, et al.
Pubblicazione: (2024)
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
di: Guo, Chengquan, et al.
Pubblicazione: (2025)
di: Guo, Chengquan, et al.
Pubblicazione: (2025)
Intention is All You Need: Refining Your Code from Your Intention
di: Guo, Qi, et al.
Pubblicazione: (2025)
di: Guo, Qi, et al.
Pubblicazione: (2025)
FT2Ra: A Fine-Tuning-Inspired Approach to Retrieval-Augmented Code Completion
di: Guo, Qi, et al.
Pubblicazione: (2024)
di: Guo, Qi, et al.
Pubblicazione: (2024)
A Systematic Literature Review on Neural Code Translation
di: Chen, Xiang, et al.
Pubblicazione: (2025)
di: Chen, Xiang, et al.
Pubblicazione: (2025)
UniCode: Augmenting Evaluation for Code Reasoning
di: Zheng, Xinyue, et al.
Pubblicazione: (2025)
di: Zheng, Xinyue, et al.
Pubblicazione: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
di: Manh, Dung Nguyen, et al.
Pubblicazione: (2024)
di: Manh, Dung Nguyen, et al.
Pubblicazione: (2024)
You Can REST Now: Automated REST API Documentation and Testing via LLM-Assisted Request Mutations
di: Decrop, Alix, et al.
Pubblicazione: (2024)
di: Decrop, Alix, et al.
Pubblicazione: (2024)
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
di: Ma, Wei, et al.
Pubblicazione: (2022)
di: Ma, Wei, et al.
Pubblicazione: (2022)
RepoGenReflex: Enhancing Repository-Level Code Completion with Verbal Reinforcement and Retrieval-Augmented Generation
di: Wang, Jicheng, et al.
Pubblicazione: (2024)
di: Wang, Jicheng, et al.
Pubblicazione: (2024)
Boosting LLMs for Mutation Generation
di: Wang, Bo, et al.
Pubblicazione: (2026)
di: Wang, Bo, et al.
Pubblicazione: (2026)
Do LLMs generate test oracles that capture the actual or the expected program behaviour?
di: Konstantinou, Michael, et al.
Pubblicazione: (2024)
di: Konstantinou, Michael, et al.
Pubblicazione: (2024)
How well LLM-based test generation techniques perform with newer LLM versions?
di: Konstantinou, Michael, et al.
Pubblicazione: (2026)
di: Konstantinou, Michael, et al.
Pubblicazione: (2026)
Bias Testing and Mitigation in LLM-based Code Generation
di: Huang, Dong, et al.
Pubblicazione: (2023)
di: Huang, Dong, et al.
Pubblicazione: (2023)
Benchmarking and Revisiting Code Generation Assessment: A Mutation-Based Approach
di: Wang, Longtian, et al.
Pubblicazione: (2025)
di: Wang, Longtian, et al.
Pubblicazione: (2025)
TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
di: Wang, Minxing, et al.
Pubblicazione: (2026)
di: Wang, Minxing, et al.
Pubblicazione: (2026)
QuanBench: Benchmarking Quantum Code Generation with Large Language Models
di: Guo, Xiaoyu, et al.
Pubblicazione: (2025)
di: Guo, Xiaoyu, et al.
Pubblicazione: (2025)
CodeRAG-Bench: Can Retrieval Augment Code Generation?
di: Wang, Zora Zhiruo, et al.
Pubblicazione: (2024)
di: Wang, Zora Zhiruo, et al.
Pubblicazione: (2024)
CodeImprove: Program Adaptation for Deep Code Models
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
Understanding Practitioners' Expectations on Clear Code Review Comments
di: Chen, Junkai, et al.
Pubblicazione: (2024)
di: Chen, Junkai, et al.
Pubblicazione: (2024)
You Augment Me: Exploring ChatGPT-based Data Augmentation for Semantic Code Search
di: Wang, Yanlin, et al.
Pubblicazione: (2024)
di: Wang, Yanlin, et al.
Pubblicazione: (2024)
A^3-CodGen: A Repository-Level Code Generation Framework for Code Reuse with Local-Aware, Global-Aware, and Third-Party-Library-Aware
di: Liao, Dianshu, et al.
Pubblicazione: (2023)
di: Liao, Dianshu, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Boosting Source Code Learning with Text-Oriented Data Augmentation: An Empirical Study
di: Dong, Zeming, et al.
Pubblicazione: (2023) -
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026) -
One Model, Many Skills: Parameter-Efficient Fine-Tuning for Multitask Code Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026) -
Software Fairness: An Analysis and Survey
di: Soremekun, Ezekiel, et al.
Pubblicazione: (2022) -
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
di: Larbi, Maya, et al.
Pubblicazione: (2025)