Guardado en:
| Autores principales: | Xiao, Ziyan, Zhu, Yinghao, Peng, Liang, Yu, Lequan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.00740 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks
por: Zhu, Yinghao, et al.
Publicado: (2024)
por: Zhu, Yinghao, et al.
Publicado: (2024)
AgentHER: Hindsight Experience Replay for LLM Agent Trajectory Relabeling
por: Ding, Liang
Publicado: (2026)
por: Ding, Liang
Publicado: (2026)
Retrieval-Augmented LLM Agents: Learning to Learn from Experience
por: Ferraz, Thomas Palmeira, et al.
Publicado: (2026)
por: Ferraz, Thomas Palmeira, et al.
Publicado: (2026)
Adaptive Activation Steering: A Tuning-Free LLM Truthfulness Improvement Method for Diverse Hallucinations Categories
por: Wang, Tianlong, et al.
Publicado: (2024)
por: Wang, Tianlong, et al.
Publicado: (2024)
Dyna-Mind: Learning to Simulate from Experience for Better AI Agents
por: Yu, Xiao, et al.
Publicado: (2025)
por: Yu, Xiao, et al.
Publicado: (2025)
Echo: Learning from Experience Data via User-Driven Refinement
por: Dong, Hande, et al.
Publicado: (2026)
por: Dong, Hande, et al.
Publicado: (2026)
Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning
por: Zhang, Xichen, et al.
Publicado: (2025)
por: Zhang, Xichen, et al.
Publicado: (2025)
Weaver: Foundation Models for Creative Writing
por: Wang, Tiannan, et al.
Publicado: (2024)
por: Wang, Tiannan, et al.
Publicado: (2024)
Contextual Experience Replay for Self-Improvement of Language Agents
por: Liu, Yitao, et al.
Publicado: (2025)
por: Liu, Yitao, et al.
Publicado: (2025)
FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration
por: Jung, Dongwon, et al.
Publicado: (2025)
por: Jung, Dongwon, et al.
Publicado: (2025)
ToolBridge: An Open-Source Dataset to Equip LLMs with External Tool Capabilities
por: Jin, Zhenchao, et al.
Publicado: (2024)
por: Jin, Zhenchao, et al.
Publicado: (2024)
MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks
por: Zhu, Yinghao, et al.
Publicado: (2025)
por: Zhu, Yinghao, et al.
Publicado: (2025)
Experience Compression Spectrum: Unifying Memory, Skills, and Rules in LLM Agents
por: Zhang, Xing, et al.
Publicado: (2026)
por: Zhang, Xing, et al.
Publicado: (2026)
Positive Experience Reflection for Agents in Interactive Text Environments
por: Lippmann, Philip, et al.
Publicado: (2024)
por: Lippmann, Philip, et al.
Publicado: (2024)
Experiences Build Characters: The Linguistic Origins and Functional Impact of LLM Personality
por: Wang, Xi, et al.
Publicado: (2026)
por: Wang, Xi, et al.
Publicado: (2026)
CoWork-X: Experience-Optimized Co-Evolution for Multi-Agent Collaboration System
por: Lin, Zexin, et al.
Publicado: (2026)
por: Lin, Zexin, et al.
Publicado: (2026)
ExGRPO: Learning to Reason from Experience
por: Zhan, Runzhe, et al.
Publicado: (2025)
por: Zhan, Runzhe, et al.
Publicado: (2025)
Rethinking Text-based Protein Understanding: Retrieval or LLM?
por: Wu, Juntong, et al.
Publicado: (2025)
por: Wu, Juntong, et al.
Publicado: (2025)
From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms
por: Luo, Jinghao, et al.
Publicado: (2026)
por: Luo, Jinghao, et al.
Publicado: (2026)
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
por: Wu, Rong, et al.
Publicado: (2025)
por: Wu, Rong, et al.
Publicado: (2025)
Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving
por: Tang, Xiangru, et al.
Publicado: (2025)
por: Tang, Xiangru, et al.
Publicado: (2025)
Budget-Aware Routing for Long Clinical Text
por: Qureshi, Khizar, et al.
Publicado: (2026)
por: Qureshi, Khizar, et al.
Publicado: (2026)
DriVLMe: Enhancing LLM-based Autonomous Driving Agents with Embodied and Social Experiences
por: Huang, Yidong, et al.
Publicado: (2024)
por: Huang, Yidong, et al.
Publicado: (2024)
Efficient Preference-based Reinforcement Learning via Aligned Experience Estimation
por: Bai, Fengshuo, et al.
Publicado: (2024)
por: Bai, Fengshuo, et al.
Publicado: (2024)
Auditing medical multi-agent AI reveals risks of false consensus
por: Zhu, Yinghao, et al.
Publicado: (2025)
por: Zhu, Yinghao, et al.
Publicado: (2025)
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
por: Yang, Cheng, et al.
Publicado: (2025)
por: Yang, Cheng, et al.
Publicado: (2025)
XSkill: Continual Learning from Experience and Skills in Multimodal Agents
por: Jiang, Guanyu, et al.
Publicado: (2026)
por: Jiang, Guanyu, et al.
Publicado: (2026)
Improving Interactive Diagnostic Ability of a Large Language Model Agent Through Clinical Experience Learning
por: Sun, Zhoujian, et al.
Publicado: (2025)
por: Sun, Zhoujian, et al.
Publicado: (2025)
CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports
por: Zhang, Xiao Yu Cindy, et al.
Publicado: (2025)
por: Zhang, Xiao Yu Cindy, et al.
Publicado: (2025)
Agent Learning via Early Experience
por: Zhang, Kai, et al.
Publicado: (2025)
por: Zhang, Kai, et al.
Publicado: (2025)
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
por: Panaganti, Kishan, et al.
Publicado: (2026)
por: Panaganti, Kishan, et al.
Publicado: (2026)
Attribute Structuring Improves LLM-Based Evaluation of Clinical Text Summaries
por: Gero, Zelalem, et al.
Publicado: (2024)
por: Gero, Zelalem, et al.
Publicado: (2024)
Building Self-Evolving Agents via Experience-Driven Lifelong Learning: A Framework and Benchmark
por: Cai, Yuxuan, et al.
Publicado: (2025)
por: Cai, Yuxuan, et al.
Publicado: (2025)
HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research
por: Zhu, Yinghao, et al.
Publicado: (2025)
por: Zhu, Yinghao, et al.
Publicado: (2025)
SmartSwitch: Advancing LLM Reasoning by Overcoming Underthinking via Promoting Deeper Thought Exploration
por: Zhang, Xichen, et al.
Publicado: (2025)
por: Zhang, Xichen, et al.
Publicado: (2025)
ExpSeek: Self-Triggered Experience Seeking for Web Agents
por: Zhang, Wenyuan, et al.
Publicado: (2026)
por: Zhang, Wenyuan, et al.
Publicado: (2026)
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
por: Yu, Xiao, et al.
Publicado: (2023)
por: Yu, Xiao, et al.
Publicado: (2023)
Data-efficient Targeted Token-level Preference Optimization for LLM-based Text-to-Speech
por: Kotoge, Rikuto, et al.
Publicado: (2025)
por: Kotoge, Rikuto, et al.
Publicado: (2025)
Optimizing Long-Form Clinical Text Generation with Claim-Based Rewards
por: Jhaveri, Samyak, et al.
Publicado: (2025)
por: Jhaveri, Samyak, et al.
Publicado: (2025)
Self-Verification Dilemma: Experience-Driven Suppression of Overused Checking in LLM Reasoning
por: Long, Quanyu, et al.
Publicado: (2026)
por: Long, Quanyu, et al.
Publicado: (2026)
Ejemplares similares
-
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks
por: Zhu, Yinghao, et al.
Publicado: (2024) -
AgentHER: Hindsight Experience Replay for LLM Agent Trajectory Relabeling
por: Ding, Liang
Publicado: (2026) -
Retrieval-Augmented LLM Agents: Learning to Learn from Experience
por: Ferraz, Thomas Palmeira, et al.
Publicado: (2026) -
Adaptive Activation Steering: A Tuning-Free LLM Truthfulness Improvement Method for Diverse Hallucinations Categories
por: Wang, Tianlong, et al.
Publicado: (2024) -
Dyna-Mind: Learning to Simulate from Experience for Better AI Agents
por: Yu, Xiao, et al.
Publicado: (2025)