Prometheus: Towards Long-Horizon Codebase Navigation for Repository-Level Problem Solving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pan, Yue, Chen, Zimin, Lu, Siyu, Chu, Zhaoyang, Li, Xiang, Li, Han, Feng, Yang, Goues, Claire Le, Sarro, Federica, Monperrus, Martin, Ye, He |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
PreciseBugCollector: Extensible, Executable and Precise Bug-fix Collection
von: Ye, He, et al.
Veröffentlicht: (2023)
von: Ye, He, et al.
Veröffentlicht: (2023)
TestForge: Feedback-Driven, Agentic Test Suite Generation
von: Jain, Kush, et al.
Veröffentlicht: (2025)
von: Jain, Kush, et al.
Veröffentlicht: (2025)
Supersonic: Learning to Generate Source Code Optimizations in C/C++
von: Chen, Zimin, et al.
Veröffentlicht: (2023)
von: Chen, Zimin, et al.
Veröffentlicht: (2023)
ContextBench: A Benchmark for Context Retrieval in Coding Agents
von: Li, Han, et al.
Veröffentlicht: (2026)
von: Li, Han, et al.
Veröffentlicht: (2026)
What is a "bug"? On subjectivity, epistemic power, and implications for software research
von: Widder, David Gray, et al.
Veröffentlicht: (2024)
von: Widder, David Gray, et al.
Veröffentlicht: (2024)
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation
von: Luo, Jane, et al.
Veröffentlicht: (2025)
von: Luo, Jane, et al.
Veröffentlicht: (2025)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
von: Majgaonkar, Oorja, et al.
Veröffentlicht: (2025)
von: Majgaonkar, Oorja, et al.
Veröffentlicht: (2025)
Automatic, Expressive, and Scalable Fuzzing with Stitching
von: Green, Harrison, et al.
Veröffentlicht: (2026)
von: Green, Harrison, et al.
Veröffentlicht: (2026)
FrameShift: Learning to Resize Fuzzer Inputs Without Breaking Them
von: Green, Harrison, et al.
Veröffentlicht: (2025)
von: Green, Harrison, et al.
Veröffentlicht: (2025)
NL2Repo-Bench: Towards Long-Horizon Repository Generation Evaluation of Coding Agents
von: Ding, Jingzhe, et al.
Veröffentlicht: (2025)
von: Ding, Jingzhe, et al.
Veröffentlicht: (2025)
Automated Repair of Ambiguous Problem Descriptions for LLM-Based Code Generation
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
A Scalable Benchmark for Repository-Oriented Long-Horizon Conversational Context Management
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Fast, Fine-Grained Equivalence Checking for Neural Decompilers
von: Dramko, Luke, et al.
Veröffentlicht: (2025)
von: Dramko, Luke, et al.
Veröffentlicht: (2025)
Idioms: Neural Decompilation With Joint Code and Type Definition Prediction
von: Dramko, Luke, et al.
Veröffentlicht: (2025)
von: Dramko, Luke, et al.
Veröffentlicht: (2025)
Environment-in-the-Loop: Rethinking Code Migration with LLM-based Agents
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method
von: Song, Xinshuai, et al.
Veröffentlicht: (2024)
von: Song, Xinshuai, et al.
Veröffentlicht: (2024)
Prometheus
Veröffentlicht: (2023)
Veröffentlicht: (2023)
On the Problem Characteristics of Multi-objective Pseudo-Boolean Functions in Runtime Analysis
von: Liang, Zimin, et al.
Veröffentlicht: (2025)
von: Liang, Zimin, et al.
Veröffentlicht: (2025)
AR Forcing: Towards Long-Horizon Robot Navigation World Model
von: Yang, Yifei, et al.
Veröffentlicht: (2026)
von: Yang, Yifei, et al.
Veröffentlicht: (2026)
Generative AI for Testing of Autonomous Driving Systems: A Survey
von: Song, Qunying, et al.
Veröffentlicht: (2025)
von: Song, Qunying, et al.
Veröffentlicht: (2025)
Echo: Graph-Enhanced Retrieval and Execution Feedback for Issue Reproduction Test Generation
von: Fei, Zhiwei, et al.
Veröffentlicht: (2026)
von: Fei, Zhiwei, et al.
Veröffentlicht: (2026)
Can Interest-Bearing Positions Solve the Long-Horizon Problem in Prediction Markets?
von: Maresca, Caleb
Veröffentlicht: (2026)
von: Maresca, Caleb
Veröffentlicht: (2026)
Prometheus Reimagined
von: Lin, Albert C.
Veröffentlicht: (2024)
von: Lin, Albert C.
Veröffentlicht: (2024)
Prometheus and the Liver
von: van Rosmalen, Julia, et al.
Veröffentlicht: (2022)
von: van Rosmalen, Julia, et al.
Veröffentlicht: (2022)
LongNav-R1: Horizon-Adaptive Multi-Turn RL for Long-Horizon VLA Navigation
von: Hu, Yue, et al.
Veröffentlicht: (2026)
von: Hu, Yue, et al.
Veröffentlicht: (2026)
Code Compass: A Study on the Challenges of Navigating Unfamiliar Codebases
von: Agrawal, Ekansh, et al.
Veröffentlicht: (2024)
von: Agrawal, Ekansh, et al.
Veröffentlicht: (2024)
STRIDE: Simple Type Recognition In Decompiled Executables
von: Green, Harrison, et al.
Veröffentlicht: (2024)
von: Green, Harrison, et al.
Veröffentlicht: (2024)
Security Vulnerability Detection with Multitask Self-Instructed Fine-Tuning of Large Language Models
von: Yang, Aidan Z. H., et al.
Veröffentlicht: (2024)
von: Yang, Aidan Z. H., et al.
Veröffentlicht: (2024)
Gistify! Codebase-Level Understanding via Runtime Execution
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Understanding Fairness in Software Engineering: Insights from Stack Exchange
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
Psychological Safety Framework in Pull-based Open Source Projects
von: Sesari, Emeralda, et al.
Veröffentlicht: (2025)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2025)
On The Effectiveness of One-Class Support Vector Machine in Different Defect Prediction Scenarios
von: Moussa, Rebecca, et al.
Veröffentlicht: (2022)
von: Moussa, Rebecca, et al.
Veröffentlicht: (2022)
BayesInsights: Modelling Software Delivery and Developer Experience with Bayesian Networks at Bloomberg
von: Kirbas, Serkan, et al.
Veröffentlicht: (2026)
von: Kirbas, Serkan, et al.
Veröffentlicht: (2026)
It is Giving Major Satisfaction: Why Fairness Matters for Software Practitioners
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
Remember Your Trace: Memory-Guided Long-Horizon Agentic Framework for Consistent and Hierarchical Repository-Level Code Documentation
von: Bae, Suyoung, et al.
Veröffentlicht: (2026)
von: Bae, Suyoung, et al.
Veröffentlicht: (2026)
ITER: Iterative Neural Repair for Multi-Location Patches
von: Ye, He, et al.
Veröffentlicht: (2023)
von: Ye, He, et al.
Veröffentlicht: (2023)
On Scalability of Multi-Objective Evolutionary Algorithms on Combinatorial Optimisation Problems
von: Tang, Menghao, et al.
Veröffentlicht: (2026)
von: Tang, Menghao, et al.
Veröffentlicht: (2026)
TerminalWorld: Benchmarking Agents on Real-World Terminal Tasks
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2026)
Toward Executable Repository-Level Code Generation via Environment Alignment
von: Pan, Ruwei, et al.
Veröffentlicht: (2026)
von: Pan, Ruwei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
von: Li, Xiang, et al.
Veröffentlicht: (2026) -
PreciseBugCollector: Extensible, Executable and Precise Bug-fix Collection
von: Ye, He, et al.
Veröffentlicht: (2023) -
TestForge: Feedback-Driven, Agentic Test Suite Generation
von: Jain, Kush, et al.
Veröffentlicht: (2025) -
Supersonic: Learning to Generate Source Code Optimizations in C/C++
von: Chen, Zimin, et al.
Veröffentlicht: (2023) -
ContextBench: A Benchmark for Context Retrieval in Coding Agents
von: Li, Han, et al.
Veröffentlicht: (2026)