A Study on Leveraging Search and Self-Feedback for Agent Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | K, Karthikeyan, Yuan, Michelle, Mansimov, Elman, Margatina, Katerina, Pratik, Anurag, Bonadiman, Daniele, Sunkara, Monica, Zhang, Yi, Benajiba, Yassine |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MemInsight: Autonomous Memory Augmentation for LLM Agents
von: Salama, Rana, et al.
Veröffentlicht: (2025)
von: Salama, Rana, et al.
Veröffentlicht: (2025)
Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development
von: Shen, Ming, et al.
Veröffentlicht: (2025)
von: Shen, Ming, et al.
Veröffentlicht: (2025)
Bootstrapping LLM-based Task-Oriented Dialogue Agents via Self-Talk
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
CONFETTI: Conversational Function-Calling Evaluation Through Turn-Level Interactions
von: Alkhouli, Tamer, et al.
Veröffentlicht: (2025)
von: Alkhouli, Tamer, et al.
Veröffentlicht: (2025)
Explicit Trait Inference for Multi-Agent Coordination
von: Abdurahman, Suhaib, et al.
Veröffentlicht: (2026)
von: Abdurahman, Suhaib, et al.
Veröffentlicht: (2026)
TReMu: Towards Neuro-Symbolic Temporal Reasoning for LLM-Agents with Memory in Multi-Session Dialogues
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
Towards Effective GenAI Multi-Agent Collaboration: Design and Evaluation for Enterprise Applications
von: Shu, Raphael, et al.
Veröffentlicht: (2024)
von: Shu, Raphael, et al.
Veröffentlicht: (2024)
Inference time LLM alignment in single and multidomain preference spectrum
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
Barriers to Discrete Reasoning with Transformers: A Survey Across Depth, Exactness, and Bandwidth
von: Yuan, Michelle, et al.
Veröffentlicht: (2026)
von: Yuan, Michelle, et al.
Veröffentlicht: (2026)
Self-supervised Analogical Learning using Language Models
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
General Purpose Verification for Chain of Thought Prompting
von: Vacareanu, Robert, et al.
Veröffentlicht: (2024)
von: Vacareanu, Robert, et al.
Veröffentlicht: (2024)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
von: Li, Bryan, et al.
Veröffentlicht: (2024)
von: Li, Bryan, et al.
Veröffentlicht: (2024)
Arabic Named Entity Recognition
von: Yassine Benajiba
Veröffentlicht: (2010)
von: Yassine Benajiba
Veröffentlicht: (2010)
Automated Composition of Agents: A Knapsack Approach for Agentic Component Selection
von: Yuan, Michelle, et al.
Veröffentlicht: (2025)
von: Yuan, Michelle, et al.
Veröffentlicht: (2025)
MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
FLAP: Flow-Adhering Planning with Constrained Decoding in LLMs
von: Roy, Shamik, et al.
Veröffentlicht: (2024)
von: Roy, Shamik, et al.
Veröffentlicht: (2024)
Diable: Efficient Dialogue State Tracking as Operations on Tables
von: Lesci, Pietro, et al.
Veröffentlicht: (2023)
von: Lesci, Pietro, et al.
Veröffentlicht: (2023)
The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2024)
von: Kirk, Hannah Rose, et al.
Veröffentlicht: (2024)
Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
von: Singh, Jyotika, et al.
Veröffentlicht: (2025)
von: Singh, Jyotika, et al.
Veröffentlicht: (2025)
NewsQs: Multi-Source Question Generation for the Inquiring Mind
von: Hwang, Alyssa, et al.
Veröffentlicht: (2024)
von: Hwang, Alyssa, et al.
Veröffentlicht: (2024)
In-Token Rationality Optimization: Towards Accurate and Concise LLM Reasoning via Self-Feedback
von: Zhu, Mingye, et al.
Veröffentlicht: (2025)
von: Zhu, Mingye, et al.
Veröffentlicht: (2025)
L-MARS: Legal Multi-Agent Workflow with Orchestrated Reasoning and Agentic Search
von: Wang, Ziqi, et al.
Veröffentlicht: (2025)
von: Wang, Ziqi, et al.
Veröffentlicht: (2025)
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
von: Li, Jian, et al.
Veröffentlicht: (2026)
von: Li, Jian, et al.
Veröffentlicht: (2026)
QueryAgent: A Reliable and Efficient Reasoning Framework with Environmental Feedback-based Self-Correction
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
EvolveSearch: An Iterative Self-Evolving Search Agent
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
DeAL: Decoding-time Alignment for Large Language Models
von: Huang, James Y., et al.
Veröffentlicht: (2024)
von: Huang, James Y., et al.
Veröffentlicht: (2024)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
von: Banerjee, Tanushree, et al.
Veröffentlicht: (2024)
von: Banerjee, Tanushree, et al.
Veröffentlicht: (2024)
Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent
von: Huang, Ziyang, et al.
Veröffentlicht: (2025)
von: Huang, Ziyang, et al.
Veröffentlicht: (2025)
DETOUR: An Interactive Benchmark for Dual-Agent Search and Reasoning
von: Siyan, Li, et al.
Veröffentlicht: (2026)
von: Siyan, Li, et al.
Veröffentlicht: (2026)
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
von: Zhao, Chunxu, et al.
Veröffentlicht: (2026)
von: Zhao, Chunxu, et al.
Veröffentlicht: (2026)
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM Agents
von: Jin, Bowen, et al.
Veröffentlicht: (2025)
von: Jin, Bowen, et al.
Veröffentlicht: (2025)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
von: Wan, Guangya, et al.
Veröffentlicht: (2024)
von: Wan, Guangya, et al.
Veröffentlicht: (2024)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
von: Jin, Bowen, et al.
Veröffentlicht: (2025)
von: Jin, Bowen, et al.
Veröffentlicht: (2025)
SIGHT: Reinforcement Learning with Self-Evidence and Information-Gain Diverse Branching for Search Agent
von: Zhong, Wenlin, et al.
Veröffentlicht: (2026)
von: Zhong, Wenlin, et al.
Veröffentlicht: (2026)
Science Consultant Agent
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
Enhancing Language Agent Strategic Reasoning through Self-Play in Adversarial Games
von: Zhang, Yikai, et al.
Veröffentlicht: (2025)
von: Zhang, Yikai, et al.
Veröffentlicht: (2025)
SAMULE: Self-Learning Agents Enhanced by Multi-level Reflection
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MemInsight: Autonomous Memory Augmentation for LLM Agents
von: Salama, Rana, et al.
Veröffentlicht: (2025) -
Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development
von: Shen, Ming, et al.
Veröffentlicht: (2025) -
Bootstrapping LLM-based Task-Oriented Dialogue Agents via Self-Talk
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024) -
CONFETTI: Conversational Function-Calling Evaluation Through Turn-Level Interactions
von: Alkhouli, Tamer, et al.
Veröffentlicht: (2025) -
Explicit Trait Inference for Multi-Agent Coordination
von: Abdurahman, Suhaib, et al.
Veröffentlicht: (2026)