Tool-Augmented LLMs as a Universal Interface for IDEs
Fuente:
arXiv
Saved in:
| Main Authors: | Zharov, Yaroslav, Khudyakov, Yury, Fedotova, Evgeniia, Grigorenko, Evgeny, Bogomolov, Egor |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GitGoodBench: A Novel Benchmark For Evaluating Agentic Performance On Git
by: Lindenbauer, Tobias, et al.
Published: (2025)
by: Lindenbauer, Tobias, et al.
Published: (2025)
Dynamic Retrieval-Augmented Generation
by: Shapkin, Anton, et al.
Published: (2023)
by: Shapkin, Anton, et al.
Published: (2023)
Step Rejection Fine-Tuning: A Practical Distillation Recipe
by: Slinko, Igor, et al.
Published: (2026)
by: Slinko, Igor, et al.
Published: (2026)
EnvBench: A Benchmark for Automated Environment Setup
by: Eliseeva, Aleksandra, et al.
Published: (2025)
by: Eliseeva, Aleksandra, et al.
Published: (2025)
The Complexity Trap: Simple Observation Masking Is as Efficient as LLM Summarization for Agent Context Management
by: Lindenbauer, Tobias, et al.
Published: (2025)
by: Lindenbauer, Tobias, et al.
Published: (2025)
PIPer: On-Device Environment Setup via Online Reinforcement Learning
by: Kovrigin, Alexander, et al.
Published: (2025)
by: Kovrigin, Alexander, et al.
Published: (2025)
TreeRanker: Fast and Model-agnostic Ranking System for Code Suggestions in IDEs
by: Cipollone, Daniele, et al.
Published: (2025)
by: Cipollone, Daniele, et al.
Published: (2025)
On Problems of Implicit Context Compression for Software Engineering Agents
by: Gelvan, Kirill, et al.
Published: (2026)
by: Gelvan, Kirill, et al.
Published: (2026)
Leveraging LLMs, IDEs, and Semantic Embeddings for Automated Move Method Refactoring
by: Bellur, Abhiram, et al.
Published: (2025)
by: Bellur, Abhiram, et al.
Published: (2025)
Drawing Pandas: A Benchmark for LLMs in Generating Plotting Code
by: Galimzyanov, Timur, et al.
Published: (2024)
by: Galimzyanov, Timur, et al.
Published: (2024)
On the Integration of Spectrum-Based Fault Localization Tools into IDEs
by: Szatmári, Attila, et al.
Published: (2024)
by: Szatmári, Attila, et al.
Published: (2024)
Diff-XYZ: A Benchmark for Evaluating Diff Understanding
by: Glukhov, Evgeniy, et al.
Published: (2025)
by: Glukhov, Evgeniy, et al.
Published: (2025)
In-IDE Toolkit for Developers of AI-Based Features
by: Sokolov, Yaroslav, et al.
Published: (2026)
by: Sokolov, Yaroslav, et al.
Published: (2026)
Context Composing for Full Line Code Completion
by: Semenkin, Anton, et al.
Published: (2024)
by: Semenkin, Anton, et al.
Published: (2024)
Untangling Knots: Leveraging LLM for Error Resolution in Computational Notebooks
by: Grotov, Konstantin, et al.
Published: (2024)
by: Grotov, Konstantin, et al.
Published: (2024)
Embedding-based search in JetBrains IDEs
by: Abramov, Evgeny, et al.
Published: (2024)
by: Abramov, Evgeny, et al.
Published: (2024)
Together We Go Further: LLMs and IDE Static Analysis for Extract Method Refactoring
by: Pomian, Dorin, et al.
Published: (2024)
by: Pomian, Dorin, et al.
Published: (2024)
Bridging Education and Development: IDEs as Interactive Learning Platforms
by: Birillo, Anastasiia, et al.
Published: (2024)
by: Birillo, Anastasiia, et al.
Published: (2024)
On The Importance of Reasoning for Context Retrieval in Repository-Level Code Editing
by: Kovrigin, Alexander, et al.
Published: (2024)
by: Kovrigin, Alexander, et al.
Published: (2024)
Hidden Gems in the Rough: Computational Notebooks as an Uncharted Oasis for IDEs
by: Titov, Sergey, et al.
Published: (2024)
by: Titov, Sergey, et al.
Published: (2024)
Challenge on Optimization of Context Collection for Code Completion
by: Ustalov, Dmitry, et al.
Published: (2025)
by: Ustalov, Dmitry, et al.
Published: (2025)
AI IDEs or Autonomous Agents? Measuring the Impact of Coding Agents on Software Development
by: Agarwal, Shyam, et al.
Published: (2026)
by: Agarwal, Shyam, et al.
Published: (2026)
Wired for Reuse: Automating Context-Aware Code Adaptation in IDEs via LLM-Based Agent
by: Wang, Taiming, et al.
Published: (2025)
by: Wang, Taiming, et al.
Published: (2025)
Developer Needs and Feasible Features for AI Assistants in IDEs
by: Sergeyuk, Agnia, et al.
Published: (2024)
by: Sergeyuk, Agnia, et al.
Published: (2024)
Towards Realistic Evaluation of Commit Message Generation by Matching Online and Offline Settings
by: Tsvetkov, Petr, et al.
Published: (2024)
by: Tsvetkov, Petr, et al.
Published: (2024)
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
by: Skopin, Egor, et al.
Published: (2026)
by: Skopin, Egor, et al.
Published: (2026)
Investigating Tool-Memory Conflicts in Tool-Augmented LLMs
by: Cheng, Jiali, et al.
Published: (2026)
by: Cheng, Jiali, et al.
Published: (2026)
Trust Calibration in IDEs: Paving the Way for Widespread Adoption of AI Refactoring
by: Borg, Markus
Published: (2024)
by: Borg, Markus
Published: (2024)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
by: Shibaev, Egor, et al.
Published: (2024)
by: Shibaev, Egor, et al.
Published: (2024)
EM-Assist: Safe Automated ExtractMethod Refactoring with LLMs
by: Pomian, Dorin, et al.
Published: (2024)
by: Pomian, Dorin, et al.
Published: (2024)
Long Code Arena: a Set of Benchmarks for Long-Context Code Models
by: Bogomolov, Egor, et al.
Published: (2024)
by: Bogomolov, Egor, et al.
Published: (2024)
KOALA: a Configurable Tool for Collecting IDE Data When Solving Programming Tasks
by: Karol, Daniil, et al.
Published: (2025)
by: Karol, Daniil, et al.
Published: (2025)
Tool for Supporting Debugging and Understanding of Normative Requirements Using LLMs
by: Kleijwegt, Alex, et al.
Published: (2025)
by: Kleijwegt, Alex, et al.
Published: (2025)
Interact and React: Exploring Gender Patterns in Development and the Impact on Innovation and Robustness of a User Interface Tool
by: Brooke, Sian
Published: (2025)
by: Brooke, Sian
Published: (2025)
Finding Important Stack Frames in Large Systems
by: Khvorov, Aleksandr, et al.
Published: (2025)
by: Khvorov, Aleksandr, et al.
Published: (2025)
Runtime-Augmented LLMs for Crash Detection and Diagnosis in ML Notebooks
by: Wang, Yiran, et al.
Published: (2026)
by: Wang, Yiran, et al.
Published: (2026)
Multi-Agent Coordinated Rename Refactoring
by: Bellur, Abhiram, et al.
Published: (2026)
by: Bellur, Abhiram, et al.
Published: (2026)
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
Evolving Triple Knowledge-Augmented LLMs for Code Translation in Repository Context
by: Ou, Guangsheng, et al.
Published: (2025)
by: Ou, Guangsheng, et al.
Published: (2025)
Benchmarking Failures in Tool-Augmented Language Models
by: Treviño, Eduardo, et al.
Published: (2025)
by: Treviño, Eduardo, et al.
Published: (2025)
Similar Items
-
GitGoodBench: A Novel Benchmark For Evaluating Agentic Performance On Git
by: Lindenbauer, Tobias, et al.
Published: (2025) -
Dynamic Retrieval-Augmented Generation
by: Shapkin, Anton, et al.
Published: (2023) -
Step Rejection Fine-Tuning: A Practical Distillation Recipe
by: Slinko, Igor, et al.
Published: (2026) -
EnvBench: A Benchmark for Automated Environment Setup
by: Eliseeva, Aleksandra, et al.
Published: (2025) -
The Complexity Trap: Simple Observation Masking Is as Efficient as LLM Summarization for Agent Context Management
by: Lindenbauer, Tobias, et al.
Published: (2025)