Improving Search Agent with One Line of Code
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Jian, Chen, Dongsheng, Xu, Zhenhua, Jin, Yizhang, Wu, Jiafu, Wang, Chengjie, Yuan, Xiaotong, Wang, Yabiao |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
par: Li, Jian, et autres
Publié: (2026)
par: Li, Jian, et autres
Publié: (2026)
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
par: Xu, Zhenhua, et autres
Publié: (2026)
par: Xu, Zhenhua, et autres
Publié: (2026)
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
par: Liu, Dongqi, et autres
Publié: (2026)
par: Liu, Dongqi, et autres
Publié: (2026)
Cautious Optimizers: Improving Training with One Line of Code
par: Liang, Kaizhao, et autres
Publié: (2024)
par: Liang, Kaizhao, et autres
Publié: (2024)
Planning In Natural Language Improves LLM Search For Code Generation
par: Wang, Evan, et autres
Publié: (2024)
par: Wang, Evan, et autres
Publié: (2024)
Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries
par: Zhou, Liqi, et autres
Publié: (2026)
par: Zhou, Liqi, et autres
Publié: (2026)
AdapNet: Adaptive Noise-Based Network for Low-Quality Image Retrieval
par: Zhang, Sihe, et autres
Publié: (2024)
par: Zhang, Sihe, et autres
Publié: (2024)
SELA: Tree-Search Enhanced LLM Agents for Automated Machine Learning
par: Chi, Yizhou, et autres
Publié: (2024)
par: Chi, Yizhou, et autres
Publié: (2024)
Stochastic Adversarial Networks for Multi-Domain Text Classification
par: Wang, Xu, et autres
Publié: (2024)
par: Wang, Xu, et autres
Publié: (2024)
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
par: Wang, Hanlin, et autres
Publié: (2025)
par: Wang, Hanlin, et autres
Publié: (2025)
ReCode: Unify Plan and Action for Universal Granularity Control
par: Yu, Zhaoyang, et autres
Publié: (2025)
par: Yu, Zhaoyang, et autres
Publié: (2025)
Improving Zero-Shot Cross-Lingual Transfer via Progressive Code-Switching
par: Li, Zhuoran, et autres
Publié: (2024)
par: Li, Zhuoran, et autres
Publié: (2024)
Reading Between the Lines: The One-Sided Conversation Problem
par: Ebert, Victoria, et autres
Publié: (2025)
par: Ebert, Victoria, et autres
Publié: (2025)
Verbal Process Supervision Elicits Better Coding Agents
par: Chen, Hao-Yuan, et autres
Publié: (2025)
par: Chen, Hao-Yuan, et autres
Publié: (2025)
Direct Multi-Turn Preference Optimization for Language Agents
par: Shi, Wentao, et autres
Publié: (2024)
par: Shi, Wentao, et autres
Publié: (2024)
Beyond Sequential Reranking: Reranker-Guided Search Improves Reasoning Intensive Retrieval
par: Xu, Haike, et autres
Publié: (2025)
par: Xu, Haike, et autres
Publié: (2025)
Improving Diffusion Language Model Decoding through Joint Search in Generation Order and Token Space
par: Shen, Yangyi, et autres
Publié: (2026)
par: Shen, Yangyi, et autres
Publié: (2026)
A Multi-Perspective Architecture for Semantic Code Search
par: Haldar, Rajarshi, et autres
Publié: (2020)
par: Haldar, Rajarshi, et autres
Publié: (2020)
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
par: Wang, Xiyao, et autres
Publié: (2024)
par: Wang, Xiyao, et autres
Publié: (2024)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
par: Xia, Yunhui, et autres
Publié: (2025)
par: Xia, Yunhui, et autres
Publié: (2025)
EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning
par: Xu, Wujiang, et autres
Publié: (2025)
par: Xu, Wujiang, et autres
Publié: (2025)
From I/O to Code with Discovery Agent
par: Dong, Yihong, et autres
Publié: (2026)
par: Dong, Yihong, et autres
Publié: (2026)
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
par: Wang, Yubo, et autres
Publié: (2025)
par: Wang, Yubo, et autres
Publié: (2025)
A Survey on Code Generation with LLM-based Agents
par: Dong, Yihong, et autres
Publié: (2025)
par: Dong, Yihong, et autres
Publié: (2025)
Provable Knowledge Acquisition and Extraction in One-Layer Transformers
par: Xu, Ruichen, et autres
Publié: (2025)
par: Xu, Ruichen, et autres
Publié: (2025)
Patch the Distribution Mismatch: RL Rewriting Agent for Stable Off-Policy SFT
par: Wang, Jiacheng, et autres
Publié: (2026)
par: Wang, Jiacheng, et autres
Publié: (2026)
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention
par: Jiang, Huiqiang, et autres
Publié: (2024)
par: Jiang, Huiqiang, et autres
Publié: (2024)
Self-Execution Simulation Improves Coding Models
par: Maimon, Gallil, et autres
Publié: (2026)
par: Maimon, Gallil, et autres
Publié: (2026)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
par: Zhang, Zeyu, et autres
Publié: (2025)
par: Zhang, Zeyu, et autres
Publié: (2025)
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents
par: Wang, Jiawei, et autres
Publié: (2025)
par: Wang, Jiawei, et autres
Publié: (2025)
Demystifying and Enhancing the Efficiency of Large Language Model Based Search Agents
par: Yang, Tiannuo, et autres
Publié: (2025)
par: Yang, Tiannuo, et autres
Publié: (2025)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
par: Wang, Hanlin, et autres
Publié: (2025)
par: Wang, Hanlin, et autres
Publié: (2025)
Learning Virtual Machine Scheduling in Cloud Computing through Language Agents
par: Wu, JieHao, et autres
Publié: (2025)
par: Wu, JieHao, et autres
Publié: (2025)
\$OneMillion-Bench: How Far are Language Agents from Human Experts?
par: Yang, Qianyu, et autres
Publié: (2026)
par: Yang, Qianyu, et autres
Publié: (2026)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
par: Wei, Anjiang, et autres
Publié: (2025)
par: Wei, Anjiang, et autres
Publié: (2025)
Enhancing Rare Codes via Probability-Biased Directed Graph Attention for Long-Tail ICD Coding
par: Chen, Tianlei, et autres
Publié: (2025)
par: Chen, Tianlei, et autres
Publié: (2025)
Fine-Tuning is Subgraph Search: A New Lens on Learning Dynamics
par: Li, Yueyan, et autres
Publié: (2025)
par: Li, Yueyan, et autres
Publié: (2025)
AutoMLGen: Navigating Fine-Grained Optimization for Coding Agents
par: Du, Shangheng, et autres
Publié: (2025)
par: Du, Shangheng, et autres
Publié: (2025)
Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
par: Liu, Youwei, et autres
Publié: (2026)
par: Liu, Youwei, et autres
Publié: (2026)
Rebellious Student: Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR
par: Kim, Jeonghye, et autres
Publié: (2026)
par: Kim, Jeonghye, et autres
Publié: (2026)
Documents similaires
-
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
par: Li, Jian, et autres
Publié: (2026) -
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
par: Xu, Zhenhua, et autres
Publié: (2026) -
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
par: Liu, Dongqi, et autres
Publié: (2026) -
Cautious Optimizers: Improving Training with One Line of Code
par: Liang, Kaizhao, et autres
Publié: (2024) -
Planning In Natural Language Improves LLM Search For Code Generation
par: Wang, Evan, et autres
Publié: (2024)