Can Github issues be solved with Tree Of Thoughts?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | La Rosa, Ricardo, Hulse, Corey, Liu, Bangdi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Performance Review on LLM for solving leetcode problems
von: Wang, Lun, et al.
Veröffentlicht: (2025)
von: Wang, Lun, et al.
Veröffentlicht: (2025)
Analysis of Commit Signing on Github
von: Shittu, Abubakar Sadiq, et al.
Veröffentlicht: (2026)
von: Shittu, Abubakar Sadiq, et al.
Veröffentlicht: (2026)
Enhancing Interpretability in Software Change Management with Chain-of-Thought Reasoning
von: Sun, Yongqian, et al.
Veröffentlicht: (2025)
von: Sun, Yongqian, et al.
Veröffentlicht: (2025)
Can Agents Fix Agent Issues?
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
Lemur: Log Parsing with Entropy Sampling and Chain-of-Thought Merging
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
CMSA algorithm for solving the prioritized pairwise test data generation problem in software product lines
von: Ferrer, Javier, et al.
Veröffentlicht: (2024)
von: Ferrer, Javier, et al.
Veröffentlicht: (2024)
Can LLM Generate Regression Tests for Software Commits?
von: Liu, Jing, et al.
Veröffentlicht: (2025)
von: Liu, Jing, et al.
Veröffentlicht: (2025)
Quality-Driven Agentic Reasoning for LLM-Assisted Software Design: Questions-of-Thoughts (QoT) as a Time-Series Self-QA Chain
von: Liu, Yen-Ku, et al.
Veröffentlicht: (2026)
von: Liu, Yen-Ku, et al.
Veröffentlicht: (2026)
The Hitchhiker's Guide to Program Analysis, Part II: Deep Thoughts by LLMs
von: Li, Haonan, et al.
Veröffentlicht: (2025)
von: Li, Haonan, et al.
Veröffentlicht: (2025)
Can AI Models Direct Each Other? Organizational Structure as a Probe into Training Limitations
von: Liu, Rui
Veröffentlicht: (2026)
von: Liu, Rui
Veröffentlicht: (2026)
Understanding Software Engineering Agents: A Study of Thought-Action-Result Trajectories
von: Bouzenia, Islem, et al.
Veröffentlicht: (2025)
von: Bouzenia, Islem, et al.
Veröffentlicht: (2025)
Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2026)
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2026)
Adaptive Root Cause Localization for Microservice Systems with Multi-Agent Recursion-of-Thought
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2025)
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2025)
Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2026)
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2026)
From Code-Centric to Intent-Centric Software Engineering: A Reflexive Thematic Analysis of Generative AI, Agentic Systems, and Engineering Accountability
von: De La Cruz, Elyson
Veröffentlicht: (2026)
von: De La Cruz, Elyson
Veröffentlicht: (2026)
Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
Can LLMs Replace Humans During Code Chunking?
von: Glasz, Christopher, et al.
Veröffentlicht: (2025)
von: Glasz, Christopher, et al.
Veröffentlicht: (2025)
Can an LLM Detect Instances of Microservice Infrastructure Patterns?
von: Duarte, Carlos Eduardo, et al.
Veröffentlicht: (2026)
von: Duarte, Carlos Eduardo, et al.
Veröffentlicht: (2026)
Can LLMs Generate User Stories and Assess Their Quality?
von: Quattrocchi, Giovanni, et al.
Veröffentlicht: (2025)
von: Quattrocchi, Giovanni, et al.
Veröffentlicht: (2025)
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check?
von: Bansal, Srijan, et al.
Veröffentlicht: (2026)
von: Bansal, Srijan, et al.
Veröffentlicht: (2026)
Can GPT-4 Replicate Empirical Software Engineering Research?
von: Liang, Jenny T., et al.
Veröffentlicht: (2023)
von: Liang, Jenny T., et al.
Veröffentlicht: (2023)
Accuracy Can Lie: On the Impact of Surrogate Model in Configuration Tuning
von: Chen, Pengzhou, et al.
Veröffentlicht: (2025)
von: Chen, Pengzhou, et al.
Veröffentlicht: (2025)
Breaking the Myth: Can Small Models Infer Postconditions Too?
von: Zhang, Gehao, et al.
Veröffentlicht: (2025)
von: Zhang, Gehao, et al.
Veröffentlicht: (2025)
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End"
von: Sovrano, Francesco, et al.
Veröffentlicht: (2025)
von: Sovrano, Francesco, et al.
Veröffentlicht: (2025)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
ProgramBench: Can Language Models Rebuild Programs From Scratch?
von: Yang, John, et al.
Veröffentlicht: (2026)
von: Yang, John, et al.
Veröffentlicht: (2026)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
I Can Find You in Seconds! Leveraging Large Language Models for Code Authorship Attribution
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
GitHub's Copilot Code Review: Can AI Spot Security Flaws Before You Commit?
von: Amro, Amena, et al.
Veröffentlicht: (2025)
von: Amro, Amena, et al.
Veröffentlicht: (2025)
Reasoning Efficiently Through Adaptive Chain-of-Thought Compression: A Self-Optimizing Framework
von: Huang, Kerui, et al.
Veröffentlicht: (2025)
von: Huang, Kerui, et al.
Veröffentlicht: (2025)
Generating Verifiable Chain of Thoughts from Exection-Traces
von: Thakur, Shailja, et al.
Veröffentlicht: (2025)
von: Thakur, Shailja, et al.
Veröffentlicht: (2025)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
PyResBugs: A Dataset of Residual Python Bugs for Natural Language-Driven Fault Injection
von: Cotroneo, Domenico, et al.
Veröffentlicht: (2025)
von: Cotroneo, Domenico, et al.
Veröffentlicht: (2025)
Seed-CTS: Unleashing the Power of Tree Search for Superior Performance in Competitive Coding Tasks
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
$T^3$: Multi-level Tree-based Automatic Program Repair with Large Language Models
von: Liu, Quanming, et al.
Veröffentlicht: (2025)
von: Liu, Quanming, et al.
Veröffentlicht: (2025)
EZASP -- Facilitating the usage of ASP
von: Martins, Rafael, et al.
Veröffentlicht: (2026)
von: Martins, Rafael, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Performance Review on LLM for solving leetcode problems
von: Wang, Lun, et al.
Veröffentlicht: (2025) -
Analysis of Commit Signing on Github
von: Shittu, Abubakar Sadiq, et al.
Veröffentlicht: (2026) -
Enhancing Interpretability in Software Change Management with Chain-of-Thought Reasoning
von: Sun, Yongqian, et al.
Veröffentlicht: (2025) -
Can Agents Fix Agent Issues?
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025) -
Lemur: Log Parsing with Entropy Sampling and Chain-of-Thought Merging
von: Zhang, Wei, et al.
Veröffentlicht: (2024)