OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rahman, Musfiqur, Khatoonabadi, SayedHassan, Shihab, Emad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2025)
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2025)
Automatic Detection of LLM-Generated Code: A Comparative Case Study of Contemporary Models Across Function and Class Granularities
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024)
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024)
The Impact of Environment Configurations on the Stability of AI-Enabled Systems
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024)
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024)
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
von: Latendresse, Jasmine, et al.
Veröffentlicht: (2025)
von: Latendresse, Jasmine, et al.
Veröffentlicht: (2025)
Evaluating the Use of LLMs for Documentation to Code Traceability
von: Alor, Ebube, et al.
Veröffentlicht: (2025)
von: Alor, Ebube, et al.
Veröffentlicht: (2025)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2023)
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2023)
Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering
von: Abedu, Samuel, et al.
Veröffentlicht: (2024)
von: Abedu, Samuel, et al.
Veröffentlicht: (2024)
Evaluating the Use of LLMs for Automated DOM-Level Resolution of Web Performance Issues
von: Peters, Gideon, et al.
Veröffentlicht: (2026)
von: Peters, Gideon, et al.
Veröffentlicht: (2026)
Automated File-Level Logging Generation for Machine Learning Applications using LLMs: A Case Study using GPT-4o Mini
von: Rodriguez, Mayra Sofia Ruiz, et al.
Veröffentlicht: (2025)
von: Rodriguez, Mayra Sofia Ruiz, et al.
Veröffentlicht: (2025)
Is ChatGPT a Good Software Librarian? An Exploratory Study on the Use of ChatGPT for Software Library Recommendations
von: Latendresse, Jasmine, et al.
Veröffentlicht: (2024)
von: Latendresse, Jasmine, et al.
Veröffentlicht: (2024)
On Wasted Contributions: Understanding the Dynamics of Contributor-Abandoned Pull Requests
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2021)
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2021)
Predicting the First Response Latency of Maintainers and Contributors in Pull Requests
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2023)
von: Khatoonabadi, SayedHassan, et al.
Veröffentlicht: (2023)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
von: Alor, Ebube, et al.
Veröffentlicht: (2024)
von: Alor, Ebube, et al.
Veröffentlicht: (2024)
Will It Survive? Deciphering the Fate of AI-Generated Code in Open Source
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2026)
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2026)
Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
von: Salim, Mohamad, et al.
Veröffentlicht: (2026)
von: Salim, Mohamad, et al.
Veröffentlicht: (2026)
The Impact of Large Language Models (LLMs) on Code Review Process
von: Collante, Antonio, et al.
Veröffentlicht: (2025)
von: Collante, Antonio, et al.
Veröffentlicht: (2025)
PCBSchemaGen: Constraint-Guided Schematic Design via LLM for Printed Circuit Boards (PCB)
von: Zou, Huanghaohe, et al.
Veröffentlicht: (2026)
von: Zou, Huanghaohe, et al.
Veröffentlicht: (2026)
ClassInvGen: Class Invariant Synthesis using Large Language Models
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
On the Impact of Black-box Deployment Strategies for Edge AI on Latency and Model Performance
von: Singh, Jaskirat, et al.
Veröffentlicht: (2024)
von: Singh, Jaskirat, et al.
Veröffentlicht: (2024)
Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production
von: Irugalbandara, Chandra, et al.
Veröffentlicht: (2023)
von: Irugalbandara, Chandra, et al.
Veröffentlicht: (2023)
Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents
von: Manglik, Akshay, et al.
Veröffentlicht: (2026)
von: Manglik, Akshay, et al.
Veröffentlicht: (2026)
On The Effectiveness of One-Class Support Vector Machine in Different Defect Prediction Scenarios
von: Moussa, Rebecca, et al.
Veröffentlicht: (2022)
von: Moussa, Rebecca, et al.
Veröffentlicht: (2022)
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
von: Jewitt, James, et al.
Veröffentlicht: (2026)
von: Jewitt, James, et al.
Veröffentlicht: (2026)
SCoGen: Scenario-Centric Graph-Based Synthesis of Real-World Code Problems
von: Yao, Xifeng, et al.
Veröffentlicht: (2025)
von: Yao, Xifeng, et al.
Veröffentlicht: (2025)
Deploying Geospatial Foundation Models in the Real World: Lessons from WorldCereal
von: Butsko, Christina, et al.
Veröffentlicht: (2025)
von: Butsko, Christina, et al.
Veröffentlicht: (2025)
Towards a Neural Debugger for Python
von: Beck, Maximilian, et al.
Veröffentlicht: (2026)
von: Beck, Maximilian, et al.
Veröffentlicht: (2026)
WhatsCode: Large-Scale GenAI Deployment for Developer Efficiency at WhatsApp
von: Mao, Ke, et al.
Veröffentlicht: (2025)
von: Mao, Ke, et al.
Veröffentlicht: (2025)
SWE-Universe: Scale Real-World Verifiable Environments to Millions
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
Behavioral Augmentation of UML Class Diagrams: An Empirical Study of Large Language Models for Method Generation
von: Rouabhia, Djaber, et al.
Veröffentlicht: (2025)
von: Rouabhia, Djaber, et al.
Veröffentlicht: (2025)
SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents
von: Mündler, Niels, et al.
Veröffentlicht: (2024)
von: Mündler, Niels, et al.
Veröffentlicht: (2024)
MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion
von: Feng, Yunfei, et al.
Veröffentlicht: (2026)
von: Feng, Yunfei, et al.
Veröffentlicht: (2026)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
A Cartography of Open Collaboration in Open Source AI: Mapping Practices, Motivations, and Governance in 14 Open Large Language Model Projects
von: Linåker, Johan, et al.
Veröffentlicht: (2025)
von: Linåker, Johan, et al.
Veröffentlicht: (2025)
BuildBench: Benchmarking LLM Agents on Compiling Real-World Open-Source Software
von: Zhang, Zehua, et al.
Veröffentlicht: (2025)
von: Zhang, Zehua, et al.
Veröffentlicht: (2025)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
von: Vulićević, Jelena Ilić
Veröffentlicht: (2026)
von: Vulićević, Jelena Ilić
Veröffentlicht: (2026)
CWM: An Open-Weights LLM for Research on Code Generation with World Models
von: FAIR CodeGen team, et al.
Veröffentlicht: (2025)
von: FAIR CodeGen team, et al.
Veröffentlicht: (2025)
DataGovBench: Benchmarking LLM Agents for Real-World Data Governance Workflows
von: Liu, Zhou, et al.
Veröffentlicht: (2025)
von: Liu, Zhou, et al.
Veröffentlicht: (2025)
A State-of-the-practice Release-readiness Checklist for Generative AI-based Software Products
von: Patel, Harsh, et al.
Veröffentlicht: (2024)
von: Patel, Harsh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2025) -
Automatic Detection of LLM-Generated Code: A Comparative Case Study of Contemporary Models Across Function and Class Granularities
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024) -
The Impact of Environment Configurations on the Stability of AI-Enabled Systems
von: Rahman, Musfiqur, et al.
Veröffentlicht: (2024) -
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
von: Latendresse, Jasmine, et al.
Veröffentlicht: (2025) -
Evaluating the Use of LLMs for Documentation to Code Traceability
von: Alor, Ebube, et al.
Veröffentlicht: (2025)