A Study on Leveraging Search and Self-Feedback for Agent Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | K, Karthikeyan, Yuan, Michelle, Mansimov, Elman, Margatina, Katerina, Pratik, Anurag, Bonadiman, Daniele, Sunkara, Monica, Zhang, Yi, Benajiba, Yassine |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MemInsight: Autonomous Memory Augmentation for LLM Agents
por: Salama, Rana, et al.
Publicado: (2025)
por: Salama, Rana, et al.
Publicado: (2025)
Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development
por: Shen, Ming, et al.
Publicado: (2025)
por: Shen, Ming, et al.
Publicado: (2025)
Bootstrapping LLM-based Task-Oriented Dialogue Agents via Self-Talk
por: Ulmer, Dennis, et al.
Publicado: (2024)
por: Ulmer, Dennis, et al.
Publicado: (2024)
CONFETTI: Conversational Function-Calling Evaluation Through Turn-Level Interactions
por: Alkhouli, Tamer, et al.
Publicado: (2025)
por: Alkhouli, Tamer, et al.
Publicado: (2025)
Explicit Trait Inference for Multi-Agent Coordination
por: Abdurahman, Suhaib, et al.
Publicado: (2026)
por: Abdurahman, Suhaib, et al.
Publicado: (2026)
TReMu: Towards Neuro-Symbolic Temporal Reasoning for LLM-Agents with Memory in Multi-Session Dialogues
por: Ge, Yubin, et al.
Publicado: (2025)
por: Ge, Yubin, et al.
Publicado: (2025)
Towards Effective GenAI Multi-Agent Collaboration: Design and Evaluation for Enterprise Applications
por: Shu, Raphael, et al.
Publicado: (2024)
por: Shu, Raphael, et al.
Publicado: (2024)
Inference time LLM alignment in single and multidomain preference spectrum
por: Shahriar, Sadat, et al.
Publicado: (2024)
por: Shahriar, Sadat, et al.
Publicado: (2024)
Barriers to Discrete Reasoning with Transformers: A Survey Across Depth, Exactness, and Bandwidth
por: Yuan, Michelle, et al.
Publicado: (2026)
por: Yuan, Michelle, et al.
Publicado: (2026)
Self-supervised Analogical Learning using Language Models
por: Zhou, Ben, et al.
Publicado: (2025)
por: Zhou, Ben, et al.
Publicado: (2025)
General Purpose Verification for Chain of Thought Prompting
por: Vacareanu, Robert, et al.
Publicado: (2024)
por: Vacareanu, Robert, et al.
Publicado: (2024)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
por: Li, Bryan, et al.
Publicado: (2024)
por: Li, Bryan, et al.
Publicado: (2024)
Arabic Named Entity Recognition
por: Yassine Benajiba
Publicado: (2010)
por: Yassine Benajiba
Publicado: (2010)
Automated Composition of Agents: A Knapsack Approach for Agentic Component Selection
por: Yuan, Michelle, et al.
Publicado: (2025)
por: Yuan, Michelle, et al.
Publicado: (2025)
MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation
por: Singh, Jyotika, et al.
Publicado: (2026)
por: Singh, Jyotika, et al.
Publicado: (2026)
FLAP: Flow-Adhering Planning with Constrained Decoding in LLMs
por: Roy, Shamik, et al.
Publicado: (2024)
por: Roy, Shamik, et al.
Publicado: (2024)
Diable: Efficient Dialogue State Tracking as Operations on Tables
por: Lesci, Pietro, et al.
Publicado: (2023)
por: Lesci, Pietro, et al.
Publicado: (2023)
The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
por: Kirk, Hannah Rose, et al.
Publicado: (2024)
por: Kirk, Hannah Rose, et al.
Publicado: (2024)
Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty
por: Feng, Yu, et al.
Publicado: (2024)
por: Feng, Yu, et al.
Publicado: (2024)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
por: Singh, Jyotika, et al.
Publicado: (2025)
por: Singh, Jyotika, et al.
Publicado: (2025)
NewsQs: Multi-Source Question Generation for the Inquiring Mind
por: Hwang, Alyssa, et al.
Publicado: (2024)
por: Hwang, Alyssa, et al.
Publicado: (2024)
In-Token Rationality Optimization: Towards Accurate and Concise LLM Reasoning via Self-Feedback
por: Zhu, Mingye, et al.
Publicado: (2025)
por: Zhu, Mingye, et al.
Publicado: (2025)
L-MARS: Legal Multi-Agent Workflow with Orchestrated Reasoning and Agentic Search
por: Wang, Ziqi, et al.
Publicado: (2025)
por: Wang, Ziqi, et al.
Publicado: (2025)
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
por: Li, Jian, et al.
Publicado: (2026)
por: Li, Jian, et al.
Publicado: (2026)
QueryAgent: A Reliable and Efficient Reasoning Framework with Environmental Feedback-based Self-Correction
por: Huang, Xiang, et al.
Publicado: (2024)
por: Huang, Xiang, et al.
Publicado: (2024)
EvolveSearch: An Iterative Self-Evolving Search Agent
por: Zhang, Dingchu, et al.
Publicado: (2025)
por: Zhang, Dingchu, et al.
Publicado: (2025)
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
por: Zheng, Caleb, et al.
Publicado: (2026)
por: Zheng, Caleb, et al.
Publicado: (2026)
DeAL: Decoding-time Alignment for Large Language Models
por: Huang, James Y., et al.
Publicado: (2024)
por: Huang, James Y., et al.
Publicado: (2024)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
por: Gonzalez-Pumariega, Gonzalo, et al.
Publicado: (2025)
por: Gonzalez-Pumariega, Gonzalo, et al.
Publicado: (2025)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
por: Banerjee, Tanushree, et al.
Publicado: (2024)
por: Banerjee, Tanushree, et al.
Publicado: (2024)
Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent
por: Huang, Ziyang, et al.
Publicado: (2025)
por: Huang, Ziyang, et al.
Publicado: (2025)
DETOUR: An Interactive Benchmark for Dual-Agent Search and Reasoning
por: Siyan, Li, et al.
Publicado: (2026)
por: Siyan, Li, et al.
Publicado: (2026)
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
por: Zhao, Chunxu, et al.
Publicado: (2026)
por: Zhao, Chunxu, et al.
Publicado: (2026)
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM Agents
por: Jin, Bowen, et al.
Publicado: (2025)
por: Jin, Bowen, et al.
Publicado: (2025)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
por: Wan, Guangya, et al.
Publicado: (2024)
por: Wan, Guangya, et al.
Publicado: (2024)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
por: Jin, Bowen, et al.
Publicado: (2025)
por: Jin, Bowen, et al.
Publicado: (2025)
SIGHT: Reinforcement Learning with Self-Evidence and Information-Gain Diverse Branching for Search Agent
por: Zhong, Wenlin, et al.
Publicado: (2026)
por: Zhong, Wenlin, et al.
Publicado: (2026)
Science Consultant Agent
por: K, Karthikeyan, et al.
Publicado: (2025)
por: K, Karthikeyan, et al.
Publicado: (2025)
Enhancing Language Agent Strategic Reasoning through Self-Play in Adversarial Games
por: Zhang, Yikai, et al.
Publicado: (2025)
por: Zhang, Yikai, et al.
Publicado: (2025)
SAMULE: Self-Learning Agents Enhanced by Multi-level Reflection
por: Ge, Yubin, et al.
Publicado: (2025)
por: Ge, Yubin, et al.
Publicado: (2025)
Ejemplares similares
-
MemInsight: Autonomous Memory Augmentation for LLM Agents
por: Salama, Rana, et al.
Publicado: (2025) -
Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development
por: Shen, Ming, et al.
Publicado: (2025) -
Bootstrapping LLM-based Task-Oriented Dialogue Agents via Self-Talk
por: Ulmer, Dennis, et al.
Publicado: (2024) -
CONFETTI: Conversational Function-Calling Evaluation Through Turn-Level Interactions
por: Alkhouli, Tamer, et al.
Publicado: (2025) -
Explicit Trait Inference for Multi-Agent Coordination
por: Abdurahman, Suhaib, et al.
Publicado: (2026)