Saved in:
Bibliographic Details
Main Authors: Guan, Jian, Wu, Wei, Wen, Zujie, Xu, Peng, Wang, Hongning, Huang, Minlie
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2402.01469
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916452880089088
author Guan, Jian
Wu, Wei
Wen, Zujie
Xu, Peng
Wang, Hongning
Huang, Minlie
author_facet Guan, Jian
Wu, Wei
Wen, Zujie
Xu, Peng
Wang, Hongning
Huang, Minlie
contents The notable success of large language models (LLMs) has sparked an upsurge in building language agents to complete various complex tasks. We present AMOR, an agent framework based on open-source LLMs, which reasons with external knowledge bases and adapts to specific domains through human supervision to the reasoning process. AMOR builds reasoning logic over a finite state machine (FSM) that solves problems through autonomous executions and transitions over disentangled modules. This allows humans to provide direct feedback to the individual modules, and thus naturally forms process supervision. Based on this reasoning and feedback framework, we develop AMOR through two-stage fine-tuning: warm-up and adaptation. The former fine-tunes the LLM with examples automatically constructed from various public datasets, enabling AMOR to generalize across different knowledge environments, while the latter tailors AMOR to specific domains using process feedback. Extensive experiments across multiple domains demonstrate the advantage of AMOR to strong baselines, thanks to its FSM-based reasoning and process feedback mechanism. The code and data are publicly available at \url{https://github.com/JianGuanTHU/AMOR}.
format Preprint
id arxiv_https___arxiv_org_abs_2402_01469
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
Guan, Jian
Wu, Wei
Wen, Zujie
Xu, Peng
Wang, Hongning
Huang, Minlie
Computation and Language
The notable success of large language models (LLMs) has sparked an upsurge in building language agents to complete various complex tasks. We present AMOR, an agent framework based on open-source LLMs, which reasons with external knowledge bases and adapts to specific domains through human supervision to the reasoning process. AMOR builds reasoning logic over a finite state machine (FSM) that solves problems through autonomous executions and transitions over disentangled modules. This allows humans to provide direct feedback to the individual modules, and thus naturally forms process supervision. Based on this reasoning and feedback framework, we develop AMOR through two-stage fine-tuning: warm-up and adaptation. The former fine-tunes the LLM with examples automatically constructed from various public datasets, enabling AMOR to generalize across different knowledge environments, while the latter tailors AMOR to specific domains using process feedback. Extensive experiments across multiple domains demonstrate the advantage of AMOR to strong baselines, thanks to its FSM-based reasoning and process feedback mechanism. The code and data are publicly available at \url{https://github.com/JianGuanTHU/AMOR}.
title AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
topic Computation and Language
url https://arxiv.org/abs/2402.01469