Towards a Neural Debugger for Python
Fuente:
arXiv
Saved in:
| Main Authors: | Beck, Maximilian, Gehring, Jonas, Kossen, Jannik, Synnaeve, Gabriel |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution
by: Gu, Alex, et al.
Published: (2024)
by: Gu, Alex, et al.
Published: (2024)
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
by: Hassid, Michael, et al.
Published: (2024)
by: Hassid, Michael, et al.
Published: (2024)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
by: Vulićević, Jelena Ilić
Published: (2026)
by: Vulićević, Jelena Ilić
Published: (2026)
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns
by: Samsonau, Sergey V.
Published: (2026)
by: Samsonau, Sergey V.
Published: (2026)
OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research
by: Rahman, Musfiqur, et al.
Published: (2025)
by: Rahman, Musfiqur, et al.
Published: (2025)
Toward a Theory of Causation for Interpreting Neural Code Models
by: Palacio, David N., et al.
Published: (2023)
by: Palacio, David N., et al.
Published: (2023)
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents
by: Zhao, Chenyu, et al.
Published: (2026)
by: Zhao, Chenyu, et al.
Published: (2026)
Towards a Classification of Open-Source ML Models and Datasets for Software Engineering
by: González, Alexandra, et al.
Published: (2024)
by: González, Alexandra, et al.
Published: (2024)
Challenges and Paths Towards AI for Software Engineering
by: Gu, Alex, et al.
Published: (2025)
by: Gu, Alex, et al.
Published: (2025)
Functional Overlap Reranking for Neural Code Generation
by: To, Hung Quoc, et al.
Published: (2023)
by: To, Hung Quoc, et al.
Published: (2023)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
by: Bouchoucha, Rached, et al.
Published: (2024)
by: Bouchoucha, Rached, et al.
Published: (2024)
evomap: A Toolbox for Dynamic Mapping in Python
by: Matthe, Maximilian
Published: (2025)
by: Matthe, Maximilian
Published: (2025)
NeuFair: Neural Network Fairness Repair with Dropout
by: Dasu, Vishnu Asutosh, et al.
Published: (2024)
by: Dasu, Vishnu Asutosh, et al.
Published: (2024)
Monitizer: Automating Design and Evaluation of Neural Network Monitors
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Toward Explaining Large Language Models in Software Engineering Tasks
by: Vitale, Antonio, et al.
Published: (2025)
by: Vitale, Antonio, et al.
Published: (2025)
Towards Robust Agentic CUDA Kernel Benchmarking, Verification, and Optimization
by: Lange, Robert Tjarko, et al.
Published: (2025)
by: Lange, Robert Tjarko, et al.
Published: (2025)
Can Requirements Engineering Support Explainable Artificial Intelligence? Towards a User-Centric Approach for Explainability Requirements
by: Umm-e-Habiba, et al.
Published: (2022)
by: Umm-e-Habiba, et al.
Published: (2022)
Talking with Verifiers: Automatic Specification Generation for Neural Network Verification
by: Elboher, Yizhak Y., et al.
Published: (2026)
by: Elboher, Yizhak Y., et al.
Published: (2026)
SoundnessBench: A Soundness Benchmark for Neural Network Verifiers
by: Zhou, Xingjian, et al.
Published: (2024)
by: Zhou, Xingjian, et al.
Published: (2024)
Graph Neural Networks based Log Anomaly Detection and Explanation
by: Li, Zhong, et al.
Published: (2023)
by: Li, Zhong, et al.
Published: (2023)
Analysing the Behaviour of Tree-Based Neural Networks in Regression Tasks
by: Samoaa, Peter, et al.
Published: (2024)
by: Samoaa, Peter, et al.
Published: (2024)
RBT4DNN: Requirements-based Testing of Neural Networks
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
by: Zhao, Zhimin, et al.
Published: (2026)
by: Zhao, Zhimin, et al.
Published: (2026)
Debugging and Runtime Analysis of Neural Networks with VLMs (A Case Study)
by: Hu, Boyue Caroline, et al.
Published: (2025)
by: Hu, Boyue Caroline, et al.
Published: (2025)
Towards More Trustworthy and Interpretable LLMs for Code through Syntax-Grounded Explanations
by: Palacio, David N., et al.
Published: (2024)
by: Palacio, David N., et al.
Published: (2024)
SimCert: Probabilistic Certification for Behavioral Similarity in Deep Neural Network Compression
by: Li, Jingyang, et al.
Published: (2026)
by: Li, Jingyang, et al.
Published: (2026)
GREPO: A Benchmark for Graph Neural Networks on Repository-Level Bug Localization
by: Wang, Juntong, et al.
Published: (2026)
by: Wang, Juntong, et al.
Published: (2026)
First Three Years of the International Verification of Neural Networks Competition (VNN-COMP)
by: Brix, Christopher, et al.
Published: (2023)
by: Brix, Christopher, et al.
Published: (2023)
Scaling Test-Time Compute for Agentic Coding
by: Kim, Joongwon, et al.
Published: (2026)
by: Kim, Joongwon, et al.
Published: (2026)
Towards Engineering Fair and Equitable Software Systems for Managing Low-Altitude Airspace Authorizations
by: Gohar, Usman, et al.
Published: (2024)
by: Gohar, Usman, et al.
Published: (2024)
Towards Transparent and Accurate Diabetes Prediction Using Machine Learning and Explainable Artificial Intelligence
by: Khokhar, Pir Bakhsh, et al.
Published: (2025)
by: Khokhar, Pir Bakhsh, et al.
Published: (2025)
Towards Open Federated Learning Platforms: Survey and Vision from Technical and Legal Perspectives
by: Duan, Moming, et al.
Published: (2023)
by: Duan, Moming, et al.
Published: (2023)
PaperDebugger: A Plugin-Based Multi-Agent System for In-Editor Academic Writing, Review, and Editing
by: Hou, Junyi, et al.
Published: (2025)
by: Hou, Junyi, et al.
Published: (2025)
AI-Driven Code Refactoring: Using Graph Neural Networks to Enhance Software Maintainability
by: Bandarupalli, Gopichand
Published: (2025)
by: Bandarupalli, Gopichand
Published: (2025)
Predicting Open Source Software Sustainability with Deep Temporal Neural Hierarchical Architectures and Explainable AI
by: Karim, S M Rakib Ul, et al.
Published: (2026)
by: Karim, S M Rakib Ul, et al.
Published: (2026)
Heterogeneous Directed Hypergraph Neural Network over abstract syntax tree (AST) for Code Classification
by: Yang, Guang, et al.
Published: (2023)
by: Yang, Guang, et al.
Published: (2023)
Call Me Maybe: Enhancing JavaScript Call Graph Construction using Graph Neural Networks
by: Bhuiyan, Masudul Hasan Masud, et al.
Published: (2025)
by: Bhuiyan, Masudul Hasan Masud, et al.
Published: (2025)
Engineering FAIR Privacy-preserving Applications that Learn Histories of Disease
by: Duarte, Ines N., et al.
Published: (2026)
by: Duarte, Ines N., et al.
Published: (2026)
DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models
by: Kumarappan, Adarsh, et al.
Published: (2026)
by: Kumarappan, Adarsh, et al.
Published: (2026)
Similar Items
-
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
by: Wei, Yuxiang, et al.
Published: (2025) -
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution
by: Gu, Alex, et al.
Published: (2024) -
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
by: Hassid, Michael, et al.
Published: (2024) -
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
by: Vulićević, Jelena Ilić
Published: (2026) -
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns
by: Samsonau, Sergey V.
Published: (2026)