Agentic Entropy-Balanced Policy Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Guanting, Bao, Licheng, Wang, Zhongyuan, Zhao, Kangzhi, Li, Xiaoxi, Jin, Jiajie, Yang, Jinghan, Mao, Hangyu, Zhang, Fuzheng, Gai, Kun, Zhou, Guorui, Zhu, Yutao, Wen, Ji-Rong, Dou, Zhicheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agentic Reinforced Policy Optimization
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
Search-o1: Agentic Search-Enhanced Large Reasoning Models
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
WebThinker: Empowering Large Reasoning Models with Deep Research Capability
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
HiRA: A Hierarchical Reasoning Framework for Decoupled Planning and Execution in Deep Search
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
DeepAgent: A General Reasoning Agent with Scalable Toolsets
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
RecThinker: An Agentic Framework for Tool-Augmented Reasoning in Recommendation
von: Zhang, Haobo, et al.
Veröffentlicht: (2026)
von: Zhang, Haobo, et al.
Veröffentlicht: (2026)
Progressive Multimodal Reasoning via Active Retrieval
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
Query-oriented Data Augmentation for Session Search
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research
von: Jin, Jiajie, et al.
Veröffentlicht: (2024)
von: Jin, Jiajie, et al.
Veröffentlicht: (2024)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
GISA: A Benchmark for General Information-Seeking Assistant
von: Zhu, Yutao, et al.
Veröffentlicht: (2026)
von: Zhu, Yutao, et al.
Veröffentlicht: (2026)
Neuro-Symbolic Query Compiler
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
From Matching to Generation: A Survey on Generative Information Retrieval
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
Agentic-R: Learning to Retrieve for Agentic Search
von: Liu, Wenhan, et al.
Veröffentlicht: (2026)
von: Liu, Wenhan, et al.
Veröffentlicht: (2026)
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
ChatShopBuddy: Towards Reliable Conversational Shopping Agents via Reinforcement Learning
von: Cheng, Yiruo, et al.
Veröffentlicht: (2026)
von: Cheng, Yiruo, et al.
Veröffentlicht: (2026)
DemoRank: Selecting Effective Demonstrations for Large Language Models in Ranking Task
von: Liu, Wenhan, et al.
Veröffentlicht: (2024)
von: Liu, Wenhan, et al.
Veröffentlicht: (2024)
Cognitive Personalized Search Integrating Large Language Models with an Efficient Memory Mechanism
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
R$^3$AG: Retriever Routing for Retrieval-Augmented Generation
von: Zhao, Tong, et al.
Veröffentlicht: (2026)
von: Zhao, Tong, et al.
Veröffentlicht: (2026)
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
LaSER: Internalizing Explicit Reasoning into Latent Space for Dense Retrieval
von: Jin, Jiajie, et al.
Veröffentlicht: (2026)
von: Jin, Jiajie, et al.
Veröffentlicht: (2026)
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems
von: Tan, Jiejun, et al.
Veröffentlicht: (2024)
von: Tan, Jiejun, et al.
Veröffentlicht: (2024)
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
von: Zhang, Chenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Chenghao, et al.
Veröffentlicht: (2025)
RetroLLM: Empowering Large Language Models to Retrieve Fine-grained Evidence within Generation
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
INTERS: Unlocking the Power of Large Language Models in Search with Instruction Tuning
von: Zhu, Yutao, et al.
Veröffentlicht: (2024)
von: Zhu, Yutao, et al.
Veröffentlicht: (2024)
Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
Session-level Normalization and Click-through Data Enhancement for Session-based Evaluation
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
An Analysis on Matching Mechanisms and Token Pruning for Late-interaction Models
von: Liu, Qi, et al.
Veröffentlicht: (2024)
von: Liu, Qi, et al.
Veröffentlicht: (2024)
A Survey of Conversational Search
von: Mo, Fengran, et al.
Veröffentlicht: (2024)
von: Mo, Fengran, et al.
Veröffentlicht: (2024)
HoME: Hierarchy of Multi-Gate Experts for Multi-Task Learning at Kuaishou
von: Wang, Xu, et al.
Veröffentlicht: (2024)
von: Wang, Xu, et al.
Veröffentlicht: (2024)
Learning Interpretable Legal Case Retrieval via Knowledge-Guided Case Reformulation
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
Laser: Governing Long-Horizon Agentic Search via Structured Protocol and Context Register
von: Wang, Shuting, et al.
Veröffentlicht: (2025)
von: Wang, Shuting, et al.
Veröffentlicht: (2025)
Leveraging LLM-Assisted Query Understanding for Live Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning
von: Tan, Jiejun, et al.
Veröffentlicht: (2026)
von: Tan, Jiejun, et al.
Veröffentlicht: (2026)
Large Language Models for Information Retrieval: A Survey
von: Zhu, Yutao, et al.
Veröffentlicht: (2023)
von: Zhu, Yutao, et al.
Veröffentlicht: (2023)
GRank: Towards Target-Aware and Streamlined Industrial Retrieval with a Generate-Rank Framework
von: Sun, Yijia, et al.
Veröffentlicht: (2025)
von: Sun, Yijia, et al.
Veröffentlicht: (2025)
GEMs: Breaking the Long-Sequence Barrier in Generative Recommendation with a Multi-Stream Decoder
von: Zhou, Yu, et al.
Veröffentlicht: (2026)
von: Zhou, Yu, et al.
Veröffentlicht: (2026)
Metacognitive Retrieval-Augmented Large Language Models
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
PROMISE: Process Reward Models Unlock Test-Time Scaling Laws in Generative Recommendations
von: Guo, Chengcheng, et al.
Veröffentlicht: (2026)
von: Guo, Chengcheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Agentic Reinforced Policy Optimization
von: Dong, Guanting, et al.
Veröffentlicht: (2025) -
Search-o1: Agentic Search-Enhanced Large Reasoning Models
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025) -
WebThinker: Empowering Large Reasoning Models with Deep Research Capability
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025) -
HiRA: A Hierarchical Reasoning Framework for Decoupled Planning and Execution in Deep Search
von: Jin, Jiajie, et al.
Veröffentlicht: (2025) -
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024)