Debugging code world models
Fuente:
arXiv
Guardado en:
| Autor principal: | Rahmani, Babak |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance
por: Chen, Yongchao, et al.
Publicado: (2025)
por: Chen, Yongchao, et al.
Publicado: (2025)
SymbolicAI: A framework for logic-based approaches combining generative models and solvers
por: Dinu, Marius-Constantin, et al.
Publicado: (2024)
por: Dinu, Marius-Constantin, et al.
Publicado: (2024)
SV-LIB 1.0: A Standard Exchange Format for Software-Verification Tasks
por: Beyer, Dirk, et al.
Publicado: (2025)
por: Beyer, Dirk, et al.
Publicado: (2025)
Open Source Prover in the Attic
por: Kovács, Zoltán, et al.
Publicado: (2024)
por: Kovács, Zoltán, et al.
Publicado: (2024)
ChatDBG: Augmenting Debugging with Large Language Models
por: Levin, Kyla H., et al.
Publicado: (2024)
por: Levin, Kyla H., et al.
Publicado: (2024)
Satire: Computing Rigorous Bounds for Floating-Point Rounding Error in Mixed-Precision Loop-Free Programs
por: Tirpankar, Tanmay, et al.
Publicado: (2025)
por: Tirpankar, Tanmay, et al.
Publicado: (2025)
Proceedings 18th International Workshop on Logical and Semantic Frameworks, with Applications and 10th Workshop on Horn Clauses for Verification and Synthesis
por: Kutsia, Temur, et al.
Publicado: (2024)
por: Kutsia, Temur, et al.
Publicado: (2024)
Learning to Generate Unit Tests for Automated Debugging
por: Prasad, Archiki, et al.
Publicado: (2025)
por: Prasad, Archiki, et al.
Publicado: (2025)
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
por: Orvalho, Pedro, et al.
Publicado: (2025)
por: Orvalho, Pedro, et al.
Publicado: (2025)
debug-gym: A Text-Based Environment for Interactive Debugging
por: Yuan, Xingdi, et al.
Publicado: (2025)
por: Yuan, Xingdi, et al.
Publicado: (2025)
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
por: Luo, Ziyan, et al.
Publicado: (2023)
por: Luo, Ziyan, et al.
Publicado: (2023)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
por: Shi, Yuling, et al.
Publicado: (2024)
por: Shi, Yuling, et al.
Publicado: (2024)
Antiassociative algebra in R: introducing the evitaicossa package
por: Hankinn, Robin K. S.
Publicado: (2024)
por: Hankinn, Robin K. S.
Publicado: (2024)
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
por: Hajizadeh, Samira, et al.
Publicado: (2026)
por: Hajizadeh, Samira, et al.
Publicado: (2026)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
por: Zhang, Dylan, et al.
Publicado: (2024)
por: Zhang, Dylan, et al.
Publicado: (2024)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
por: Dai, Hankun, et al.
Publicado: (2025)
por: Dai, Hankun, et al.
Publicado: (2025)
Linguacodus: A Synergistic Framework for Transformative Code Generation in Machine Learning Pipelines
por: Trofimova, Ekaterina, et al.
Publicado: (2024)
por: Trofimova, Ekaterina, et al.
Publicado: (2024)
ReGAL: Refactoring Programs to Discover Generalizable Abstractions
por: Stengel-Eskin, Elias, et al.
Publicado: (2024)
por: Stengel-Eskin, Elias, et al.
Publicado: (2024)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
por: Wei, Anjiang, et al.
Publicado: (2025)
por: Wei, Anjiang, et al.
Publicado: (2025)
Is Programming by Example solved by LLMs?
por: Li, Wen-Ding, et al.
Publicado: (2024)
por: Li, Wen-Ding, et al.
Publicado: (2024)
TokDrift: When LLM Speaks in Subwords but Code Speaks in Grammar
por: Li, Yinxi, et al.
Publicado: (2025)
por: Li, Yinxi, et al.
Publicado: (2025)
Beyond Accuracy: Introducing a Symbolic-Mechanistic Approach to Interpretable Evaluation
por: Habibi, Reza, et al.
Publicado: (2026)
por: Habibi, Reza, et al.
Publicado: (2026)
A Comparative Study of Neurosymbolic AI Approaches to Interpretable Logical Reasoning
por: Chen, Michael K.
Publicado: (2025)
por: Chen, Michael K.
Publicado: (2025)
CoMAT: Chain of Mathematically Annotated Thought Improves Mathematical Reasoning
por: Leang, Joshua Ong Jun, et al.
Publicado: (2024)
por: Leang, Joshua Ong Jun, et al.
Publicado: (2024)
Cognitive Architectures for Language Agents
por: Sumers, Theodore R., et al.
Publicado: (2023)
por: Sumers, Theodore R., et al.
Publicado: (2023)
Jet Expansions of Residual Computation
por: Chen, Yihong, et al.
Publicado: (2024)
por: Chen, Yihong, et al.
Publicado: (2024)
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
por: Zhao, Xufeng, et al.
Publicado: (2023)
por: Zhao, Xufeng, et al.
Publicado: (2023)
Inferring Equivalence Classes from Legacy Undocumented Embedded Binaries for ISO 26262-Compliant Testing
por: De Luca, Marco, et al.
Publicado: (2026)
por: De Luca, Marco, et al.
Publicado: (2026)
HIVE: Scalable Hardware-Firmware Co-Verification using Scenario-based Decomposition and Automated Hint Extraction
por: Jayasena, Aruna, et al.
Publicado: (2023)
por: Jayasena, Aruna, et al.
Publicado: (2023)
Comparing large language models and human programmers for generating programming code
por: Hou, Wenpin, et al.
Publicado: (2024)
por: Hou, Wenpin, et al.
Publicado: (2024)
Memory as Resonance: A Biomimetic Architecture for Infinite Context Memory on Ergodic Phonetic Manifolds
por: Houichime, Tarik, et al.
Publicado: (2025)
por: Houichime, Tarik, et al.
Publicado: (2025)
Neuro-Symbolic ODE Discovery with Latent Grammar Flow
por: Yu, Karin, et al.
Publicado: (2026)
por: Yu, Karin, et al.
Publicado: (2026)
code_transformed: The Influence of Large Language Models on Code
por: Xu, Yuliang, et al.
Publicado: (2025)
por: Xu, Yuliang, et al.
Publicado: (2025)
DebugBench: Evaluating Debugging Capability of Large Language Models
por: Tian, Runchu, et al.
Publicado: (2024)
por: Tian, Runchu, et al.
Publicado: (2024)
AutoGen Studio: A No-Code Developer Tool for Building and Debugging Multi-Agent Systems
por: Dibia, Victor, et al.
Publicado: (2024)
por: Dibia, Victor, et al.
Publicado: (2024)
Understanding Tool-Augmented Agents for Lean Formalization: A Factorial Analysis
por: Zhang, Ke, et al.
Publicado: (2026)
por: Zhang, Ke, et al.
Publicado: (2026)
Incoherence as Oracle-less Measure of Error in LLM-Based Code Generation
por: Valentin, Thomas, et al.
Publicado: (2025)
por: Valentin, Thomas, et al.
Publicado: (2025)
FormalSpecCpp: A Dataset of C++ Formal Specifications created using LLMs
por: Chakraborty, Madhurima, et al.
Publicado: (2025)
por: Chakraborty, Madhurima, et al.
Publicado: (2025)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
por: Hooda, Ashish, et al.
Publicado: (2024)
por: Hooda, Ashish, et al.
Publicado: (2024)
On the Effectiveness of Machine Learning-based Call Graph Pruning: An Empirical Study
por: Mir, Amir M., et al.
Publicado: (2024)
por: Mir, Amir M., et al.
Publicado: (2024)
Ejemplares similares
-
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance
por: Chen, Yongchao, et al.
Publicado: (2025) -
SymbolicAI: A framework for logic-based approaches combining generative models and solvers
por: Dinu, Marius-Constantin, et al.
Publicado: (2024) -
SV-LIB 1.0: A Standard Exchange Format for Software-Verification Tasks
por: Beyer, Dirk, et al.
Publicado: (2025) -
Open Source Prover in the Attic
por: Kovács, Zoltán, et al.
Publicado: (2024) -
ChatDBG: Augmenting Debugging with Large Language Models
por: Levin, Kyla H., et al.
Publicado: (2024)