SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jian, Jin, Yizhang, Liu, Dongqi, Ding, Hang, Wu, Jiafu, Chen, Dongsheng, Shen, Yunhang, Qin, Yulei, Tai, Ying, Wang, Chengjie, Yuan, Xiaotong, Wang, Yabiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Search Agent with One Line of Code
von: Li, Jian, et al.
Veröffentlicht: (2026)
von: Li, Jian, et al.
Veröffentlicht: (2026)
RoleRMBench & RoleRM: Towards Reward Modeling for Profile-Based Role Play in Dialogue Systems
von: Ding, Hang, et al.
Veröffentlicht: (2025)
von: Ding, Hang, et al.
Veröffentlicht: (2025)
EvolveSearch: An Iterative Self-Evolving Search Agent
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025)
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
von: Liu, Dongqi, et al.
Veröffentlicht: (2026)
von: Liu, Dongqi, et al.
Veröffentlicht: (2026)
PVG: Progressive Vision Graph for Vision Recognition
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
PiT: Progressive Diffusion Transformer
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
A Survey on Benchmarks of Multimodal Large Language Models
von: Li, Jian, et al.
Veröffentlicht: (2024)
von: Li, Jian, et al.
Veröffentlicht: (2024)
AutoMaAS: Self-Evolving Multi-Agent Architecture Search for Large Language Models
von: Ma, Bo, et al.
Veröffentlicht: (2025)
von: Ma, Bo, et al.
Veröffentlicht: (2025)
Dr. Zero: Self-Evolving Search Agents without Training Data
von: Yue, Zhenrui, et al.
Veröffentlicht: (2026)
von: Yue, Zhenrui, et al.
Veröffentlicht: (2026)
AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents
von: Wang, Pan, et al.
Veröffentlicht: (2026)
von: Wang, Pan, et al.
Veröffentlicht: (2026)
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking
von: Zhang, Junke, et al.
Veröffentlicht: (2026)
von: Zhang, Junke, et al.
Veröffentlicht: (2026)
LLaVA-VSD: Large Language-and-Vision Assistant for Visual Spatial Description
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
von: Jin, Yizhang, et al.
Veröffentlicht: (2024)
ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
EXG: Self-Evolving Agents with Experience Graphs
von: Jin, Yuxin, et al.
Veröffentlicht: (2026)
von: Jin, Yuxin, et al.
Veröffentlicht: (2026)
YOLO-NAS-Bench: A Surrogate Benchmark with Self-Evolving Predictors for YOLO Architecture Search
von: Li, Zhe, et al.
Veröffentlicht: (2026)
von: Li, Zhe, et al.
Veröffentlicht: (2026)
Knowledge-Graph Paths as Intermediate Supervision for Self-Evolving Search Agents
von: Wu, Huyu, et al.
Veröffentlicht: (2026)
von: Wu, Huyu, et al.
Veröffentlicht: (2026)
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents
von: Liu, Jiaqi, et al.
Veröffentlicht: (2026)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2026)
SELAUR: Self Evolving LLM Agent via Uncertainty-aware Rewards
von: Zhang, Dengjia, et al.
Veröffentlicht: (2026)
von: Zhang, Dengjia, et al.
Veröffentlicht: (2026)
SmartSearch: Process Reward-Guided Query Refinement for Search Agents
von: Wen, Tongyu, et al.
Veröffentlicht: (2026)
von: Wen, Tongyu, et al.
Veröffentlicht: (2026)
AdapNet: Adaptive Noise-Based Network for Low-Quality Image Retrieval
von: Zhang, Sihe, et al.
Veröffentlicht: (2024)
von: Zhang, Sihe, et al.
Veröffentlicht: (2024)
SwiftVideo: A Unified Framework for Few-Step Video Generation through Trajectory-Distribution Alignment
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
Dynamic Mixture of Latent Memories for Self-Evolving Agents
von: Yu, Dianzhi, et al.
Veröffentlicht: (2026)
von: Yu, Dianzhi, et al.
Veröffentlicht: (2026)
SE-GA: Memory-Augmented Self-Evolution for GUI Agents
von: Jin, Shilong, et al.
Veröffentlicht: (2026)
von: Jin, Shilong, et al.
Veröffentlicht: (2026)
Beyond Outcome Reward: Decoupling Search and Answering Improves LLM Agents
von: Wang, Yiding, et al.
Veröffentlicht: (2025)
von: Wang, Yiding, et al.
Veröffentlicht: (2025)
How to Train Your Deep Research Agent? Prompt, Reward, and Policy Optimization in Search-R1
von: Xu, Yinuo, et al.
Veröffentlicht: (2026)
von: Xu, Yinuo, et al.
Veröffentlicht: (2026)
Self-Evolved Reward Learning for LLMs
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
FactorMiner: A Self-Evolving Agent with Skills and Experience Memory for Financial Alpha Discovery
von: Wang, Yanlong, et al.
Veröffentlicht: (2026)
von: Wang, Yanlong, et al.
Veröffentlicht: (2026)
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory
von: Wei, Tianxin, et al.
Veröffentlicht: (2025)
von: Wei, Tianxin, et al.
Veröffentlicht: (2025)
Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training
von: An, Sohyun, et al.
Veröffentlicht: (2026)
von: An, Sohyun, et al.
Veröffentlicht: (2026)
PARMESAN: Parameter-Free Memory Search and Transduction for Dense Prediction Tasks
von: Winter, Philip Matthias, et al.
Veröffentlicht: (2024)
von: Winter, Philip Matthias, et al.
Veröffentlicht: (2024)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
What Makes a Good Diffusion Planner for Decision Making?
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
SEDM: Scalable Self-Evolving Distributed Memory for Agents
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?
von: Fan, HuiMing, et al.
Veröffentlicht: (2026)
von: Fan, HuiMing, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Improving Search Agent with One Line of Code
von: Li, Jian, et al.
Veröffentlicht: (2026) -
RoleRMBench & RoleRM: Towards Reward Modeling for Profile-Based Role Play in Dialogue Systems
von: Ding, Hang, et al.
Veröffentlicht: (2025) -
EvolveSearch: An Iterative Self-Evolving Search Agent
von: Zhang, Dingchu, et al.
Veröffentlicht: (2025) -
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026) -
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
von: Liu, Dongqi, et al.
Veröffentlicht: (2026)