What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Yujie, Yu, Zhuoyun, Wang, Xuehai, Zhu, Yuqi, Zhang, Ningyu, Wei, Lanning, Du, Lun, Zheng, Da, Chen, Huajun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
El Agente Gráfico: Structured Execution Graphs for Scientific Agents
by: Bai, Jiaru, et al.
Published: (2026)
by: Bai, Jiaru, et al.
Published: (2026)
From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work
by: Rosen, Josh, et al.
Published: (2026)
by: Rosen, Josh, et al.
Published: (2026)
Sherlock: Reliable and Efficient Agentic Workflow Execution
by: Ro, Yeonju, et al.
Published: (2025)
by: Ro, Yeonju, et al.
Published: (2025)
Applying Large Language Models in Knowledge Graph-based Enterprise Modeling: Challenges and Opportunities
by: Reitemeyer, Benedikt, et al.
Published: (2025)
by: Reitemeyer, Benedikt, et al.
Published: (2025)
InnoGym: Benchmarking the Innovation Potential of AI Agents
by: Zhang, Jintian, et al.
Published: (2025)
by: Zhang, Jintian, et al.
Published: (2025)
Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis
by: Qiu, Zhisong, et al.
Published: (2026)
by: Qiu, Zhisong, et al.
Published: (2026)
BootstrapAgent: Distilling Repository Setup into Reusable Agent Knowledge
by: Fu, Sihan, et al.
Published: (2026)
by: Fu, Sihan, et al.
Published: (2026)
AutoMind: Adaptive Knowledgeable Agent for Automated Data Science
by: Ou, Yixin, et al.
Published: (2025)
by: Ou, Yixin, et al.
Published: (2025)
The Specification Gap: Coordination Failure Under Partial Knowledge in Code Agents
by: Sartori, Camilo Chacón
Published: (2026)
by: Sartori, Camilo Chacón
Published: (2026)
An Executable Benchmarking Suite for Tool-Using Agents
by: Zhong, Zhiqing, et al.
Published: (2026)
by: Zhong, Zhiqing, et al.
Published: (2026)
ToolRosella: Translating Code Repositories into Standardized Tools for Scientific Agents
by: Di, Shimin, et al.
Published: (2026)
by: Di, Shimin, et al.
Published: (2026)
Governed Evolution of Agent Runtimes through Executable Operational Cognition
by: Garralda-Barrio, Mariano
Published: (2026)
by: Garralda-Barrio, Mariano
Published: (2026)
AutoFSM: A Multi-agent Framework for FSM Code Generation with IR and SystemC-Based Testing
by: Luo, Qiuming, et al.
Published: (2025)
by: Luo, Qiuming, et al.
Published: (2025)
SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers
by: Xiang, Yanzheng, et al.
Published: (2025)
by: Xiang, Yanzheng, et al.
Published: (2025)
Reflective Paper-to-Code Reproduction Enabled by Fine-Grained Verification
by: Zhou, Mingyang, et al.
Published: (2025)
by: Zhou, Mingyang, et al.
Published: (2025)
No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows
by: Raghavan, Siddeshwar, et al.
Published: (2026)
by: Raghavan, Siddeshwar, et al.
Published: (2026)
Gecko: A Simulation Environment with Stateful Feedback for Refining Agent Tool Calls
by: Zhang, Zeyu, et al.
Published: (2026)
by: Zhang, Zeyu, et al.
Published: (2026)
Externalization in LLM Agents: A Unified Review of Memory, Skills, Protocols and Harness Engineering
by: Zhou, Chenyu, et al.
Published: (2026)
by: Zhou, Chenyu, et al.
Published: (2026)
Affordance Representation and Recognition for Autonomous Agents
by: Gidey, Habtom Kahsay, et al.
Published: (2025)
by: Gidey, Habtom Kahsay, et al.
Published: (2025)
Multi-Agent Collaboration via Cross-Team Orchestration
by: Du, Zhuoyun, et al.
Published: (2024)
by: Du, Zhuoyun, et al.
Published: (2024)
A Scoresheet for Explainable AI
by: Winikoff, Michael, et al.
Published: (2025)
by: Winikoff, Michael, et al.
Published: (2025)
The Lifecycle Workbench -- A Configurable Framework for Digitized Product Maintenance Services
by: Briechle, Dominique, et al.
Published: (2025)
by: Briechle, Dominique, et al.
Published: (2025)
Bridging the Prototype-Production Gap: A Multi-Agent System for Notebooks Transformation
by: Elhashemy, Hanya, et al.
Published: (2025)
by: Elhashemy, Hanya, et al.
Published: (2025)
ABMax: A JAX-based Agent-based Modeling Framework
by: Chaturvedi, Siddharth, et al.
Published: (2025)
by: Chaturvedi, Siddharth, et al.
Published: (2025)
SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies
by: Saxena, Siddhant, et al.
Published: (2026)
by: Saxena, Siddhant, et al.
Published: (2026)
CodeCureAgent: Automatic Classification and Repair of Static Analysis Warnings
by: Joos, Pascal, et al.
Published: (2025)
by: Joos, Pascal, et al.
Published: (2025)
Cognitive Agents Powered by Large Language Models for Agile Software Project Management
by: Cinkusz, Konrad, et al.
Published: (2025)
by: Cinkusz, Konrad, et al.
Published: (2025)
REFINE: Enhancing Program Repair Agents through Context-Aware Patch Refinement
by: Pabba, Anvith, et al.
Published: (2025)
by: Pabba, Anvith, et al.
Published: (2025)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
by: Lwin, Naing Oo, et al.
Published: (2026)
by: Lwin, Naing Oo, et al.
Published: (2026)
Real-Time BDI Agents: a model and its implementation
by: Traldi, Andrea, et al.
Published: (2022)
by: Traldi, Andrea, et al.
Published: (2022)
A Generic Modelling Framework for Last-Mile Delivery Systems
by: Gürcan, Önder, et al.
Published: (2025)
by: Gürcan, Önder, et al.
Published: (2025)
Fairness in Multi-Agent Systems for Software Engineering: An SDLC-Oriented Rapid Review
by: Yang-Smith, Corey, et al.
Published: (2026)
by: Yang-Smith, Corey, et al.
Published: (2026)
Analyzing Code Injection Attacks on LLM-based Multi-Agent Systems in Software Development
by: Bowers, Brian, et al.
Published: (2025)
by: Bowers, Brian, et al.
Published: (2025)
The Hidden Bloat in Machine Learning Systems
by: Zhang, Huaifeng, et al.
Published: (2025)
by: Zhang, Huaifeng, et al.
Published: (2025)
Runtime Composition in Dynamic System of Systems: A Systematic Review of Challenges, Solutions, Tools, and Evaluation Methods
by: Ashfaq, Muhammad, et al.
Published: (2025)
by: Ashfaq, Muhammad, et al.
Published: (2025)
A Step Towards a Universal Method for Modeling and Implementing Cross-Organizational Business Processes
by: Zeisler, Gerhard, et al.
Published: (2024)
by: Zeisler, Gerhard, et al.
Published: (2024)
LongCLI-Bench: A Preliminary Benchmark and Study for Long-horizon Agentic Programming in Command-Line Interfaces
by: Feng, Yukang, et al.
Published: (2026)
by: Feng, Yukang, et al.
Published: (2026)
Sibyl-AutoResearch: Autonomous Research Needs Self-Evolving Trial-and-Error Harnesses, Not Paper Generators
by: Wang, Chengcheng, et al.
Published: (2026)
by: Wang, Chengcheng, et al.
Published: (2026)
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
by: Akshathala, Sreemaee, et al.
Published: (2025)
by: Akshathala, Sreemaee, et al.
Published: (2025)
Qualixar OS: A Universal Operating System for AI Agent Orchestration
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
Similar Items
-
El Agente Gráfico: Structured Execution Graphs for Scientific Agents
by: Bai, Jiaru, et al.
Published: (2026) -
From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work
by: Rosen, Josh, et al.
Published: (2026) -
Sherlock: Reliable and Efficient Agentic Workflow Execution
by: Ro, Yeonju, et al.
Published: (2025) -
Applying Large Language Models in Knowledge Graph-based Enterprise Modeling: Challenges and Opportunities
by: Reitemeyer, Benedikt, et al.
Published: (2025) -
InnoGym: Benchmarking the Innovation Potential of AI Agents
by: Zhang, Jintian, et al.
Published: (2025)