ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Zehan, Zhang, Jianfei, Yin, Chuantao, Ouyang, Yuanxin, Rong, Wenge |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SBAN: A Framework & Multi-Dimensional Dataset for Large Language Model Pre-Training and Software Code Mining
par: Jelodar, Hamed, et autres
Publié: (2025)
par: Jelodar, Hamed, et autres
Publié: (2025)
CPRet: A Dataset, Benchmark, and Model for Retrieval in Competitive Programming
par: Deng, Han, et autres
Publié: (2025)
par: Deng, Han, et autres
Publié: (2025)
Deep Code Search with Naming-Agnostic Contrastive Multi-View Learning
par: Feng, Jiadong, et autres
Publié: (2024)
par: Feng, Jiadong, et autres
Publié: (2024)
MGS3: A Multi-Granularity Self-Supervised Code Search Framework
par: Li, Rui, et autres
Publié: (2025)
par: Li, Rui, et autres
Publié: (2025)
Duplicate Question Retrieval and Confirmation Time Prediction in Software Communities
par: Hazra, Rima, et autres
Publié: (2023)
par: Hazra, Rima, et autres
Publié: (2023)
Code-Craft: Hierarchical Graph-Based Code Summarization for Enhanced Context Retrieval
par: Sounthiraraj, David, et autres
Publié: (2025)
par: Sounthiraraj, David, et autres
Publié: (2025)
Context-Augmented Code Generation Using Programming Knowledge Graphs
par: Saberi, Iman, et autres
Publié: (2024)
par: Saberi, Iman, et autres
Publié: (2024)
Source Code Clone Detection Using Unsupervised Similarity Measures
par: Martinez-Gil, Jorge
Publié: (2024)
par: Martinez-Gil, Jorge
Publié: (2024)
Rewriting the Code: A Simple Method for Large Language Model Augmented Code Search
par: Li, Haochen, et autres
Publié: (2024)
par: Li, Haochen, et autres
Publié: (2024)
GNN-Coder: Boosting Semantic Code Retrieval with Combined GNNs and Transformer
par: Ye, Yufan, et autres
Publié: (2025)
par: Ye, Yufan, et autres
Publié: (2025)
In-Context Learning as an Effective Estimator of Functional Correctness of LLM-Generated Code
par: Das, Susmita, et autres
Publié: (2025)
par: Das, Susmita, et autres
Publié: (2025)
Modular Layout Synthesis (MLS): Front-end Code via Structure Normalization and Constrained Generation
par: Liu, Chong, et autres
Publié: (2025)
par: Liu, Chong, et autres
Publié: (2025)
Which Programming Language and Model Work Best With LLM-as-a-Judge For Code Retrieval?
par: Roberts, Lucas, et autres
Publié: (2025)
par: Roberts, Lucas, et autres
Publié: (2025)
A Survey on Query-based API Recommendation
par: Wei, Moshi, et autres
Publié: (2023)
par: Wei, Moshi, et autres
Publié: (2023)
LLM Agents Improve Semantic Code Search
par: Jain, Sarthak, et autres
Publié: (2024)
par: Jain, Sarthak, et autres
Publié: (2024)
Can Code Evaluation Metrics Detect Code Plagiarism?
par: Ebrahim, Fahad, et autres
Publié: (2026)
par: Ebrahim, Fahad, et autres
Publié: (2026)
AI-assisted Coding with Cody: Lessons from Context Retrieval and Evaluation for Code Recommendations
par: Hartman, Jan, et autres
Publié: (2024)
par: Hartman, Jan, et autres
Publié: (2024)
DeepCodeSeek: Real-Time API Retrieval for Context-Aware Code Generation
par: Esakkiraja, Esakkivel, et autres
Publié: (2025)
par: Esakkiraja, Esakkivel, et autres
Publié: (2025)
CodeRAG: Finding Relevant and Necessary Knowledge for Retrieval-Augmented Repository-Level Code Completion
par: Zhang, Sheng, et autres
Publié: (2025)
par: Zhang, Sheng, et autres
Publié: (2025)
Selective Shot Learning for Code Explanation
par: Bhattacharya, Paheli, et autres
Publié: (2024)
par: Bhattacharya, Paheli, et autres
Publié: (2024)
LoRACode: LoRA Adapters for Code Embeddings
par: Chaturvedi, Saumya, et autres
Publié: (2025)
par: Chaturvedi, Saumya, et autres
Publié: (2025)
SweRank: Software Issue Localization with Code Ranking
par: Reddy, Revanth Gangi, et autres
Publié: (2025)
par: Reddy, Revanth Gangi, et autres
Publié: (2025)
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
par: Gurioli, Andrea, et autres
Publié: (2026)
par: Gurioli, Andrea, et autres
Publié: (2026)
Iterative Self-Training for Code Generation via Reinforced Re-Ranking
par: Sorokin, Nikita, et autres
Publié: (2025)
par: Sorokin, Nikita, et autres
Publié: (2025)
What About Emotions? Guiding Fine-Grained Emotion Extraction from Mobile App Reviews
par: Motger, Quim, et autres
Publié: (2025)
par: Motger, Quim, et autres
Publié: (2025)
Evaluating LLM-Based Mobile App Recommendations: An Empirical Study
par: Motger, Quim, et autres
Publié: (2025)
par: Motger, Quim, et autres
Publié: (2025)
Incremental Analysis of Legacy Applications Using Knowledge Graphs for Application Modernization
par: Krishnan, Saravanan, et autres
Publié: (2025)
par: Krishnan, Saravanan, et autres
Publié: (2025)
FLOWER: Flow-Oriented Entity-Relationship Tool
par: Moskalev, Dmitry
Publié: (2025)
par: Moskalev, Dmitry
Publié: (2025)
Use as Directed? A Comparison of Software Tools Intended to Check Rigor and Transparency of Published Work
par: Eckmann, Peter, et autres
Publié: (2025)
par: Eckmann, Peter, et autres
Publié: (2025)
Leveraging Graph-RAG and Prompt Engineering to Enhance LLM-Based Automated Requirement Traceability and Compliance Checks
par: Masoudifard, Arsalan, et autres
Publié: (2024)
par: Masoudifard, Arsalan, et autres
Publié: (2024)
An Empirical Study of Multi-Agent RAG for Real-World University Admissions Counseling
par: Nguyen-Duc, Anh, et autres
Publié: (2025)
par: Nguyen-Duc, Anh, et autres
Publié: (2025)
Optimizing Retrieval Augmented Generation for Object Constraint Language
par: Li, Kevin Chenhao, et autres
Publié: (2025)
par: Li, Kevin Chenhao, et autres
Publié: (2025)
From Chaos to Automation: Enabling the Use of Unstructured Data for Robotic Process Automation
par: Kurowski, Kelly, et autres
Publié: (2025)
par: Kurowski, Kelly, et autres
Publié: (2025)
Using GPT to build a Project Management assistant for Jira environments
par: Garcia-Escribano, Joel, et autres
Publié: (2025)
par: Garcia-Escribano, Joel, et autres
Publié: (2025)
Multi-View Adaptive Contrastive Learning for Information Retrieval Based Fault Localization
par: Zhou, Chunying, et autres
Publié: (2024)
par: Zhou, Chunying, et autres
Publié: (2024)
Towards Analysing Invoices and Receipts with Amazon Textract
par: Oommen, Sneha, et autres
Publié: (2025)
par: Oommen, Sneha, et autres
Publié: (2025)
Development of an Automated Web Application for Efficient Web Scraping: Design and Implementation
par: Dutta, Alok, et autres
Publié: (2025)
par: Dutta, Alok, et autres
Publié: (2025)
Balanced Knowledge Distribution among Software Development Teams -- Observations from Open-Source and Closed-Source Software Development
par: Shafiq, Saad, et autres
Publié: (2022)
par: Shafiq, Saad, et autres
Publié: (2022)
Supporting Cross-language Cross-project Bug Localization Using Pre-trained Language Models
par: Chandramohan, Mahinthan, et autres
Publié: (2024)
par: Chandramohan, Mahinthan, et autres
Publié: (2024)
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation
par: Simon, Sebastian, et autres
Publié: (2024)
par: Simon, Sebastian, et autres
Publié: (2024)
Documents similaires
-
SBAN: A Framework & Multi-Dimensional Dataset for Large Language Model Pre-Training and Software Code Mining
par: Jelodar, Hamed, et autres
Publié: (2025) -
CPRet: A Dataset, Benchmark, and Model for Retrieval in Competitive Programming
par: Deng, Han, et autres
Publié: (2025) -
Deep Code Search with Naming-Agnostic Contrastive Multi-View Learning
par: Feng, Jiadong, et autres
Publié: (2024) -
MGS3: A Multi-Granularity Self-Supervised Code Search Framework
par: Li, Rui, et autres
Publié: (2025) -
Duplicate Question Retrieval and Confirmation Time Prediction in Software Communities
par: Hazra, Rima, et autres
Publié: (2023)