Improving Search Agent with One Line of Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jian, Chen, Dongsheng, Xu, Zhenhua, Jin, Yizhang, Wu, Jiafu, Wang, Chengjie, Yuan, Xiaotong, Wang, Yabiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
von: Li, Jian, et al.
Veröffentlicht: (2026)
von: Li, Jian, et al.
Veröffentlicht: (2026)
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026)
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
von: Liu, Dongqi, et al.
Veröffentlicht: (2026)
von: Liu, Dongqi, et al.
Veröffentlicht: (2026)
Cautious Optimizers: Improving Training with One Line of Code
von: Liang, Kaizhao, et al.
Veröffentlicht: (2024)
von: Liang, Kaizhao, et al.
Veröffentlicht: (2024)
Planning In Natural Language Improves LLM Search For Code Generation
von: Wang, Evan, et al.
Veröffentlicht: (2024)
von: Wang, Evan, et al.
Veröffentlicht: (2024)
Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries
von: Zhou, Liqi, et al.
Veröffentlicht: (2026)
von: Zhou, Liqi, et al.
Veröffentlicht: (2026)
AdapNet: Adaptive Noise-Based Network for Low-Quality Image Retrieval
von: Zhang, Sihe, et al.
Veröffentlicht: (2024)
von: Zhang, Sihe, et al.
Veröffentlicht: (2024)
SELA: Tree-Search Enhanced LLM Agents for Automated Machine Learning
von: Chi, Yizhou, et al.
Veröffentlicht: (2024)
von: Chi, Yizhou, et al.
Veröffentlicht: (2024)
Stochastic Adversarial Networks for Multi-Domain Text Classification
von: Wang, Xu, et al.
Veröffentlicht: (2024)
von: Wang, Xu, et al.
Veröffentlicht: (2024)
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
ReCode: Unify Plan and Action for Universal Granularity Control
von: Yu, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoyang, et al.
Veröffentlicht: (2025)
Improving Zero-Shot Cross-Lingual Transfer via Progressive Code-Switching
von: Li, Zhuoran, et al.
Veröffentlicht: (2024)
von: Li, Zhuoran, et al.
Veröffentlicht: (2024)
Reading Between the Lines: The One-Sided Conversation Problem
von: Ebert, Victoria, et al.
Veröffentlicht: (2025)
von: Ebert, Victoria, et al.
Veröffentlicht: (2025)
Verbal Process Supervision Elicits Better Coding Agents
von: Chen, Hao-Yuan, et al.
Veröffentlicht: (2025)
von: Chen, Hao-Yuan, et al.
Veröffentlicht: (2025)
Direct Multi-Turn Preference Optimization for Language Agents
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
Beyond Sequential Reranking: Reranker-Guided Search Improves Reasoning Intensive Retrieval
von: Xu, Haike, et al.
Veröffentlicht: (2025)
von: Xu, Haike, et al.
Veröffentlicht: (2025)
Improving Diffusion Language Model Decoding through Joint Search in Generation Order and Token Space
von: Shen, Yangyi, et al.
Veröffentlicht: (2026)
von: Shen, Yangyi, et al.
Veröffentlicht: (2026)
A Multi-Perspective Architecture for Semantic Code Search
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2020)
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2020)
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
von: Xia, Yunhui, et al.
Veröffentlicht: (2025)
von: Xia, Yunhui, et al.
Veröffentlicht: (2025)
EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning
von: Xu, Wujiang, et al.
Veröffentlicht: (2025)
von: Xu, Wujiang, et al.
Veröffentlicht: (2025)
From I/O to Code with Discovery Agent
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
Provable Knowledge Acquisition and Extraction in One-Layer Transformers
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
Patch the Distribution Mismatch: RL Rewriting Agent for Stable Off-Policy SFT
von: Wang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Wang, Jiacheng, et al.
Veröffentlicht: (2026)
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention
von: Jiang, Huiqiang, et al.
Veröffentlicht: (2024)
von: Jiang, Huiqiang, et al.
Veröffentlicht: (2024)
Self-Execution Simulation Improves Coding Models
von: Maimon, Gallil, et al.
Veröffentlicht: (2026)
von: Maimon, Gallil, et al.
Veröffentlicht: (2026)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
Demystifying and Enhancing the Efficiency of Large Language Model Based Search Agents
von: Yang, Tiannuo, et al.
Veröffentlicht: (2025)
von: Yang, Tiannuo, et al.
Veröffentlicht: (2025)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
Learning Virtual Machine Scheduling in Cloud Computing through Language Agents
von: Wu, JieHao, et al.
Veröffentlicht: (2025)
von: Wu, JieHao, et al.
Veröffentlicht: (2025)
\$OneMillion-Bench: How Far are Language Agents from Human Experts?
von: Yang, Qianyu, et al.
Veröffentlicht: (2026)
von: Yang, Qianyu, et al.
Veröffentlicht: (2026)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Enhancing Rare Codes via Probability-Biased Directed Graph Attention for Long-Tail ICD Coding
von: Chen, Tianlei, et al.
Veröffentlicht: (2025)
von: Chen, Tianlei, et al.
Veröffentlicht: (2025)
Fine-Tuning is Subgraph Search: A New Lens on Learning Dynamics
von: Li, Yueyan, et al.
Veröffentlicht: (2025)
von: Li, Yueyan, et al.
Veröffentlicht: (2025)
AutoMLGen: Navigating Fine-Grained Optimization for Coding Agents
von: Du, Shangheng, et al.
Veröffentlicht: (2025)
von: Du, Shangheng, et al.
Veröffentlicht: (2025)
Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
Rebellious Student: Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR
von: Kim, Jeonghye, et al.
Veröffentlicht: (2026)
von: Kim, Jeonghye, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
von: Li, Jian, et al.
Veröffentlicht: (2026) -
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
von: Xu, Zhenhua, et al.
Veröffentlicht: (2026) -
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
von: Liu, Dongqi, et al.
Veröffentlicht: (2026) -
Cautious Optimizers: Improving Training with One Line of Code
von: Liang, Kaizhao, et al.
Veröffentlicht: (2024) -
Planning In Natural Language Improves LLM Search For Code Generation
von: Wang, Evan, et al.
Veröffentlicht: (2024)