Treating Run-time Execution History as a First-Class Citizen: Co-Versioning Run-time Behavior alongside Code
Fuente:
arXiv
Salvato in:
| Autore principale: | Kessel, Marcus |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
di: Hu, Ruida, et al.
Pubblicazione: (2025)
di: Hu, Ruida, et al.
Pubblicazione: (2025)
You Name It, I Run It: An LLM Agent to Execute Tests of Arbitrary Projects
di: Bouzenia, Islem, et al.
Pubblicazione: (2024)
di: Bouzenia, Islem, et al.
Pubblicazione: (2024)
Encoding Version History Context for Better Code Representation
di: Nguyen, Huy, et al.
Pubblicazione: (2024)
di: Nguyen, Huy, et al.
Pubblicazione: (2024)
Thinking Before Running! Efficient Code Generation with Thorough Exploration and Optimal Refinement
di: Zhang, Xiaoqing, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoqing, et al.
Pubblicazione: (2024)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
di: Berndt, Alexander, et al.
Pubblicazione: (2026)
di: Berndt, Alexander, et al.
Pubblicazione: (2026)
Towards Observation Lakehouses: Living, Interactive Archives of Software Behavior
di: Kessel, Marcus
Pubblicazione: (2025)
di: Kessel, Marcus
Pubblicazione: (2025)
Formalization of the AADL Run-Time Services with Time
di: Larson, Brian R, et al.
Pubblicazione: (2025)
di: Larson, Brian R, et al.
Pubblicazione: (2025)
PhantomRun: Auto Repair of Compilation Errors in Embedded Open Source Software
di: Fu, Han, et al.
Pubblicazione: (2026)
di: Fu, Han, et al.
Pubblicazione: (2026)
Compositionality of Systems and Partially Ordered Runs
di: Fettke, Peter, et al.
Pubblicazione: (2026)
di: Fettke, Peter, et al.
Pubblicazione: (2026)
SelfPiCo: Self-Guided Partial Code Execution with LLMs
di: Xue, Zhipeng, et al.
Pubblicazione: (2024)
di: Xue, Zhipeng, et al.
Pubblicazione: (2024)
N-Version Assessment and Enhancement of Generative AI
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
Morescient GAI for Software Engineering (Extended Version)
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
On the Flakiness of LLM-Generated Tests for Industrial and Open-Source Database Management Systems
di: Berndt, Alexander, et al.
Pubblicazione: (2026)
di: Berndt, Alexander, et al.
Pubblicazione: (2026)
CodeScore: Evaluating Code Generation by Learning Code Execution
di: Dong, Yihong, et al.
Pubblicazione: (2023)
di: Dong, Yihong, et al.
Pubblicazione: (2023)
RM -RF: Reward Model for Run-Free Unit Test Evaluation
di: Bruches, Elena, et al.
Pubblicazione: (2026)
di: Bruches, Elena, et al.
Pubblicazione: (2026)
RAT: RunAnyThing via Fully Automated Environment Configuration
di: Huang, Renhong, et al.
Pubblicazione: (2026)
di: Huang, Renhong, et al.
Pubblicazione: (2026)
Compiling Code LLMs into Lightweight Executables
di: Shi, Jieke, et al.
Pubblicazione: (2026)
di: Shi, Jieke, et al.
Pubblicazione: (2026)
CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
Do Not Treat Code as Natural Language: Implications for Repository-Level Code Generation and Beyond
di: Le-Anh, Minh, et al.
Pubblicazione: (2026)
di: Le-Anh, Minh, et al.
Pubblicazione: (2026)
Altered Histories in Version Control System Repositories: Evidence from the Trenches
di: Rapaport, Solal, et al.
Pubblicazione: (2025)
di: Rapaport, Solal, et al.
Pubblicazione: (2025)
Keeping Behavioral Programs Alive: Specifying and Executing Liveness Requirements
di: Yaacov, Tom, et al.
Pubblicazione: (2024)
di: Yaacov, Tom, et al.
Pubblicazione: (2024)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2026)
di: Gong, Zhihao, et al.
Pubblicazione: (2026)
SpaceTime Programming: Live and Omniscient Exploration of Code and Execution
di: Döderlein, Jean-Baptiste, et al.
Pubblicazione: (2026)
di: Döderlein, Jean-Baptiste, et al.
Pubblicazione: (2026)
Implementing and Executing Static Analysis Using LLVM and CodeChecker
di: Horvath, Gabor, et al.
Pubblicazione: (2024)
di: Horvath, Gabor, et al.
Pubblicazione: (2024)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2025)
di: Gong, Zhihao, et al.
Pubblicazione: (2025)
ClassEval-T: Evaluating Large Language Models in Class-Level Code Translation
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
Assessing Coherency and Consistency of Code Execution Reasoning by Large Language Models
di: Liu, Changshu, et al.
Pubblicazione: (2025)
di: Liu, Changshu, et al.
Pubblicazione: (2025)
Should I Run My Cloud Benchmark on Black Friday?
di: Henning, Sören, et al.
Pubblicazione: (2025)
di: Henning, Sören, et al.
Pubblicazione: (2025)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
di: Kessel, Marcus
Pubblicazione: (2024)
di: Kessel, Marcus
Pubblicazione: (2024)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
di: Guo, Chengquan, et al.
Pubblicazione: (2024)
di: Guo, Chengquan, et al.
Pubblicazione: (2024)
Identifying Inaccurate Descriptions in LLM-generated Code Comments via Test Execution
di: Kang, Sungmin, et al.
Pubblicazione: (2024)
di: Kang, Sungmin, et al.
Pubblicazione: (2024)
Sifting through the Chaff: On Utilizing Execution Feedback for Ranking the Generated Code Candidates
di: Sun, Zhihong, et al.
Pubblicazione: (2024)
di: Sun, Zhihong, et al.
Pubblicazione: (2024)
ChangeGuard: Validating Code Changes via Pairwise Learning-Guided Execution
di: Gröninger, Lars, et al.
Pubblicazione: (2024)
di: Gröninger, Lars, et al.
Pubblicazione: (2024)
ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation
di: He, Minghua, et al.
Pubblicazione: (2025)
di: He, Minghua, et al.
Pubblicazione: (2025)
Demystifying Errors in LLM Reasoning Traces: An Empirical Study of Code Execution Simulation
di: Abdollahi, Mohammad, et al.
Pubblicazione: (2025)
di: Abdollahi, Mohammad, et al.
Pubblicazione: (2025)
A Tool for In-depth Analysis of Code Execution Reasoning of Large Language Models
di: Liu, Changshu, et al.
Pubblicazione: (2025)
di: Liu, Changshu, et al.
Pubblicazione: (2025)
The Influence of Code Smells in Efferent Neighbors on Class Stability
di: Zhang, Zushuai, et al.
Pubblicazione: (2026)
di: Zhang, Zushuai, et al.
Pubblicazione: (2026)
The First 50 Years of Software Reliability Engineering: A History of SRE with First Person Accounts
di: Cusick, James J.
Pubblicazione: (2019)
di: Cusick, James J.
Pubblicazione: (2019)
VersiCode: Towards Version-controllable Code Generation
di: Wu, Tongtong, et al.
Pubblicazione: (2024)
di: Wu, Tongtong, et al.
Pubblicazione: (2024)
HistoryFinder: Advancing Method-Level Source Code History Generation with Accurate Oracles and Enhanced Algorithm
di: Islam, Shahidul, et al.
Pubblicazione: (2025)
di: Islam, Shahidul, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
di: Hu, Ruida, et al.
Pubblicazione: (2025) -
You Name It, I Run It: An LLM Agent to Execute Tests of Arbitrary Projects
di: Bouzenia, Islem, et al.
Pubblicazione: (2024) -
Encoding Version History Context for Better Code Representation
di: Nguyen, Huy, et al.
Pubblicazione: (2024) -
Thinking Before Running! Efficient Code Generation with Thorough Exploration and Optimal Refinement
di: Zhang, Xiaoqing, et al.
Pubblicazione: (2024) -
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
di: Berndt, Alexander, et al.
Pubblicazione: (2026)