AgentHub: A Registry for Discoverable, Verifiable, and Reproducible AI Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pautsch, Erik, Singla, Tanmay, Kumar, Parv, Jiang, Wenxin, Peng, Huiyun, Hassanshahi, Behnaz, Läufer, Konstantin, Thiruvathukal, George K., Davis, James C. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WIP: An Engaging Undergraduate Intro to Model Checking in Software Engineering Using TLA+
von: Läufer, Konstantin, et al.
Veröffentlicht: (2024)
von: Läufer, Konstantin, et al.
Veröffentlicht: (2024)
On the Variability of Source Code in Maven Package Rebuilds
von: Dietrich, Jens, et al.
Veröffentlicht: (2026)
von: Dietrich, Jens, et al.
Veröffentlicht: (2026)
Beyond Local Code Optimization: Multi-Agent Reasoning for Software System Optimization
von: Peng, Huiyun, et al.
Veröffentlicht: (2026)
von: Peng, Huiyun, et al.
Veröffentlicht: (2026)
Unlocking Reproducibility: Automating re-Build Process for Open-Source Software
von: Hassanshahi, Behnaz, et al.
Veröffentlicht: (2025)
von: Hassanshahi, Behnaz, et al.
Veröffentlicht: (2025)
Can Large-Language Models Help us Better Understand and Teach the Development of Energy-Efficient Software?
von: Hasler, Ryan, et al.
Veröffentlicht: (2024)
von: Hasler, Ryan, et al.
Veröffentlicht: (2024)
Large Language Models for Energy-Efficient Code: Emerging Results and Future Directions
von: Peng, Huiyun, et al.
Veröffentlicht: (2024)
von: Peng, Huiyun, et al.
Veröffentlicht: (2024)
SysLLMatic: Large Language Models are Software System Optimizers
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
Improving the Reproducibility of Deep Learning Software: An Initial Investigation through a Case Study Analysis
von: Ravi, Nikita, et al.
Veröffentlicht: (2025)
von: Ravi, Nikita, et al.
Veröffentlicht: (2025)
What do we know about Hugging Face? A systematic literature review and quantitative validation of qualitative claims
von: Jones, Jason, et al.
Veröffentlicht: (2024)
von: Jones, Jason, et al.
Veröffentlicht: (2024)
Now's the Time: Computer Science Must Evolve to Emphasize Software and Systems Engineering with Artificial Intelligence (AI)
von: Sekharan, Chandra N., et al.
Veröffentlicht: (2026)
von: Sekharan, Chandra N., et al.
Veröffentlicht: (2026)
Fingerprinting AI Coding Agents on GitHub
von: Ghaleb, Taher A.
Veröffentlicht: (2026)
von: Ghaleb, Taher A.
Veröffentlicht: (2026)
Reusing Deep Learning Models: Challenges and Directions in Software Engineering
von: Davis, James C., et al.
Veröffentlicht: (2024)
von: Davis, James C., et al.
Veröffentlicht: (2024)
Levels of Binary Equivalence for the Comparison of Binaries from Alternative Builds
von: Dietrich, Jens, et al.
Veröffentlicht: (2024)
von: Dietrich, Jens, et al.
Veröffentlicht: (2024)
"I see models being a whole other thing": An Empirical Study of Pre-Trained Model Naming Conventions and A Tool for Enhancing Naming Consistency
von: Jiang, Wenxin, et al.
Veröffentlicht: (2023)
von: Jiang, Wenxin, et al.
Veröffentlicht: (2023)
How Do Agents Perform Code Optimization? An Empirical Study
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
von: Peng, Huiyun, et al.
Veröffentlicht: (2025)
AIDev: Studying AI Coding Agents on GitHub
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
Analysis of Failures and Risks in Deep Learning Model Converters: A Case Study in the ONNX Ecosystem
von: Jajal, Purvish, et al.
Veröffentlicht: (2023)
von: Jajal, Purvish, et al.
Veröffentlicht: (2023)
Towards a Benchmark for Dependency Decision-Making
von: Singla, Tanmay, et al.
Veröffentlicht: (2026)
von: Singla, Tanmay, et al.
Veröffentlicht: (2026)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
Agentic Much? Adoption of Coding Agents on GitHub
von: Robbes, Romain, et al.
Veröffentlicht: (2026)
von: Robbes, Romain, et al.
Veröffentlicht: (2026)
Learning From Software Failures: A Case Study at a National Space Research Center
von: Anandayuvaraj, Dharun, et al.
Veröffentlicht: (2025)
von: Anandayuvaraj, Dharun, et al.
Veröffentlicht: (2025)
ZTD$_{JAVA}$: Mitigating Software Supply Chain Vulnerabilities via Zero-Trust Dependencies
von: Amusuo, Paschal C., et al.
Veröffentlicht: (2023)
von: Amusuo, Paschal C., et al.
Veröffentlicht: (2023)
What Challenges Do Developers Face in AI Agent Systems? An Empirical Study on Stack Overflow & GitHub Issues
von: Asgari, Ali, et al.
Veröffentlicht: (2025)
von: Asgari, Ali, et al.
Veröffentlicht: (2025)
Towards Verifiably Safe Tool Use for LLM Agents
von: Doshi, Aarya, et al.
Veröffentlicht: (2026)
von: Doshi, Aarya, et al.
Veröffentlicht: (2026)
Process-based Indicators of Vulnerability Re-Introducing Code Changes: An Exploratory Case Study
von: Shimmi, Samiha, et al.
Veröffentlicht: (2025)
von: Shimmi, Samiha, et al.
Veröffentlicht: (2025)
GitBug-Actions: Building Reproducible Bug-Fix Benchmarks with GitHub Actions
von: Saavedra, Nuno, et al.
Veröffentlicht: (2023)
von: Saavedra, Nuno, et al.
Veröffentlicht: (2023)
SEVerA: Verified Synthesis of Self-Evolving Agents
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
von: Yan, Shuo, et al.
Veröffentlicht: (2025)
von: Yan, Shuo, et al.
Veröffentlicht: (2025)
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
von: Ehsani, Ramtin, et al.
Veröffentlicht: (2026)
von: Ehsani, Ramtin, et al.
Veröffentlicht: (2026)
PeaTMOSS: A Dataset and Initial Analysis of Pre-Trained Models in Open-Source Software
von: Jiang, Wenxin, et al.
Veröffentlicht: (2024)
von: Jiang, Wenxin, et al.
Veröffentlicht: (2024)
From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work
von: Rosen, Josh, et al.
Veröffentlicht: (2026)
von: Rosen, Josh, et al.
Veröffentlicht: (2026)
Does SWE-Bench-Verified Test Agent Ability or Model Memory?
von: Prathifkumar, Thanosan, et al.
Veröffentlicht: (2025)
von: Prathifkumar, Thanosan, et al.
Veröffentlicht: (2025)
Teamwork makes the dream work: LLMs-Based Agents for GitHub README.MD Summarization
von: Nguyen, Duc S. H., et al.
Veröffentlicht: (2025)
von: Nguyen, Duc S. H., et al.
Veröffentlicht: (2025)
DALEQ -- Explainable Equivalence for Java Bytecode
von: Dietrich, Jens, et al.
Veröffentlicht: (2025)
von: Dietrich, Jens, et al.
Veröffentlicht: (2025)
Why Johnny Adopts Identity-Based Software Signing: A Usability Case Study of Sigstore
von: Kalu, Kelechi G., et al.
Veröffentlicht: (2025)
von: Kalu, Kelechi G., et al.
Veröffentlicht: (2025)
Training Software Engineering Agents and Verifiers with SWE-Gym
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
von: Tao, Wei, et al.
Veröffentlicht: (2024)
von: Tao, Wei, et al.
Veröffentlicht: (2024)
Generative AI in Software Testing: Current Trends and Future Directions
von: Singla, Tanish, et al.
Veröffentlicht: (2026)
von: Singla, Tanish, et al.
Veröffentlicht: (2026)
Counterexample Classification against Signal Temporal Logic Specifications
von: Zhang, Zhenya, et al.
Veröffentlicht: (2026)
von: Zhang, Zhenya, et al.
Veröffentlicht: (2026)
SERA: Soft-Verified Efficient Repository Agents
von: Shen, Ethan, et al.
Veröffentlicht: (2026)
von: Shen, Ethan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
WIP: An Engaging Undergraduate Intro to Model Checking in Software Engineering Using TLA+
von: Läufer, Konstantin, et al.
Veröffentlicht: (2024) -
On the Variability of Source Code in Maven Package Rebuilds
von: Dietrich, Jens, et al.
Veröffentlicht: (2026) -
Beyond Local Code Optimization: Multi-Agent Reasoning for Software System Optimization
von: Peng, Huiyun, et al.
Veröffentlicht: (2026) -
Unlocking Reproducibility: Automating re-Build Process for Open-Source Software
von: Hassanshahi, Behnaz, et al.
Veröffentlicht: (2025) -
Can Large-Language Models Help us Better Understand and Teach the Development of Energy-Efficient Software?
von: Hasler, Ryan, et al.
Veröffentlicht: (2024)