Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yifei, Yang, Xu, Yang, Xiao, Xian, Bowen, Li, Qizheng, Fang, Shikai, Li, Jingyuan, Wang, Jian, Xu, Mingrui, Liu, Weiqing, Bian, Jiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
R&D-Agent: An LLM-Agent Framework Towards Autonomous Data Science
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents
von: Li, Qizheng, et al.
Veröffentlicht: (2026)
von: Li, Qizheng, et al.
Veröffentlicht: (2026)
Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?
von: Chen, Wanyi, et al.
Veröffentlicht: (2026)
von: Chen, Wanyi, et al.
Veröffentlicht: (2026)
MarS: a Financial Market Simulation Engine Powered by Generative Foundation Model
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
R&D-Agent-Quant: A Multi-Agent Framework for Data-Centric Factors and Model Joint Optimization
von: Li, Yuante, et al.
Veröffentlicht: (2025)
von: Li, Yuante, et al.
Veröffentlicht: (2025)
Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation
von: Chen, Jiawei, et al.
Veröffentlicht: (2026)
von: Chen, Jiawei, et al.
Veröffentlicht: (2026)
MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline
von: Qiang, Rushi, et al.
Veröffentlicht: (2025)
von: Qiang, Rushi, et al.
Veröffentlicht: (2025)
Functional Complexity-adaptive Temporal Tensor Decomposition
von: Chen, Panqi, et al.
Veröffentlicht: (2025)
von: Chen, Panqi, et al.
Veröffentlicht: (2025)
Generating Full-field Evolution of Physical Dynamics from Irregular Sparse Observations
von: Chen, Panqi, et al.
Veröffentlicht: (2025)
von: Chen, Panqi, et al.
Veröffentlicht: (2025)
Less Is More: Generating Time Series with LLaMA-Style Autoregression in Simple Factorized Latent Spaces
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
Controllable Financial Market Generation with Diffusion Guided Meta Agent
von: Huang, Yu-Hao, et al.
Veröffentlicht: (2024)
von: Huang, Yu-Hao, et al.
Veröffentlicht: (2024)
BPQP: A Differentiable Convex Optimization Framework for Efficient End-to-End Learning
von: Pan, Jianming, et al.
Veröffentlicht: (2024)
von: Pan, Jianming, et al.
Veröffentlicht: (2024)
AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench
von: Toledo, Edan, et al.
Veröffentlicht: (2025)
von: Toledo, Edan, et al.
Veröffentlicht: (2025)
MLE-STAR: Machine Learning Engineering Agent via Search and Targeted Refinement
von: Nam, Jaehyun, et al.
Veröffentlicht: (2025)
von: Nam, Jaehyun, et al.
Veröffentlicht: (2025)
MLE-Dojo: Interactive Environments for Empowering LLM Agents in Machine Learning Engineering
von: Qiang, Rushi, et al.
Veröffentlicht: (2025)
von: Qiang, Rushi, et al.
Veröffentlicht: (2025)
A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning
von: Ahadian, Pegah, et al.
Veröffentlicht: (2026)
von: Ahadian, Pegah, et al.
Veröffentlicht: (2026)
Think 360°: Evaluating the Width-centric Reasoning Capability of MLLMs Beyond Depth
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
Policy Guided Tree Search for Enhanced LLM Reasoning
von: Li, Yang
Veröffentlicht: (2025)
von: Li, Yang
Veröffentlicht: (2025)
Agent-Based Modelling for Real-World Stock Markets under Behavioral Economic Principles
von: He, Tianlang, et al.
Veröffentlicht: (2023)
von: He, Tianlang, et al.
Veröffentlicht: (2023)
LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Distributed Multi-Agent Bandits Over Erdős-Rényi Random Networks
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
Towards Data-Centric Automatic R&D
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
von: Chen, Haotian, et al.
Veröffentlicht: (2024)
MARS$^2$: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation
von: Li, Pengfei, et al.
Veröffentlicht: (2026)
von: Li, Pengfei, et al.
Veröffentlicht: (2026)
rStar2-Agent: Agentic Reasoning Technical Report
von: Shang, Ning, et al.
Veröffentlicht: (2025)
von: Shang, Ning, et al.
Veröffentlicht: (2025)
GARAD-SLAM: 3D GAussian splatting for Real-time Anti Dynamic SLAM
von: Li, Mingrui, et al.
Veröffentlicht: (2025)
von: Li, Mingrui, et al.
Veröffentlicht: (2025)
Simple and Faster Algorithms for Knapsack
von: He, Qizheng, et al.
Veröffentlicht: (2023)
von: He, Qizheng, et al.
Veröffentlicht: (2023)
Locomo-Plus: Beyond-Factual Cognitive Memory Evaluation Framework for LLM Agents
von: Li, Yifei, et al.
Veröffentlicht: (2026)
von: Li, Yifei, et al.
Veröffentlicht: (2026)
SRA-MCTS: Self-driven Reasoning Augmentation with Monte Carlo Tree Search for Code Generation
von: Xu, Bin, et al.
Veröffentlicht: (2024)
von: Xu, Bin, et al.
Veröffentlicht: (2024)
SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks
von: Zhou, Yifei, et al.
Veröffentlicht: (2025)
von: Zhou, Yifei, et al.
Veröffentlicht: (2025)
Random pairing MLE for estimation of item parameters in Rasch model
von: Yang, Yuepeng, et al.
Veröffentlicht: (2024)
von: Yang, Yuepeng, et al.
Veröffentlicht: (2024)
Gradient Co-occurrence Analysis for Detecting Unsafe Prompts in Large Language Models
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
SRR-Judge: Step-Level Rating and Refinement for Enhancing Search-Integrated Reasoning in Search Agents
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
Toward Understanding Mechanistic Regulation of Body Size and Growth Control in Bivalve Mollusks
von: Ahmed Mokrani, et al.
Veröffentlicht: (2024)
von: Ahmed Mokrani, et al.
Veröffentlicht: (2024)
Beyond MLE: Investigating SEARNN for Low-Resourced Neural Machine Translation
von: Emezue, Chris
Veröffentlicht: (2024)
von: Emezue, Chris
Veröffentlicht: (2024)
MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
von: Chan, Jun Shern, et al.
Veröffentlicht: (2024)
von: Chan, Jun Shern, et al.
Veröffentlicht: (2024)
Fast MLE and MAPE-Based Device Activity Detection for Grant-Free Access via PSCA and PSCA-Net
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
Beyond Token-Level Policy Gradients for Complex Reasoning with Large Language Models
von: Xu, Mufan, et al.
Veröffentlicht: (2026)
von: Xu, Mufan, et al.
Veröffentlicht: (2026)
MG-TSD: Multi-Granularity Time Series Diffusion Models with Guided Learning Process
von: Fan, Xinyao, et al.
Veröffentlicht: (2024)
von: Fan, Xinyao, et al.
Veröffentlicht: (2024)
End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
Collaborative Evolving Strategy for Automatic Data-Centric Development
von: Yang, Xu, et al.
Veröffentlicht: (2024)
von: Yang, Xu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
R&D-Agent: An LLM-Agent Framework Towards Autonomous Data Science
von: Yang, Xu, et al.
Veröffentlicht: (2025) -
FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents
von: Li, Qizheng, et al.
Veröffentlicht: (2026) -
Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?
von: Chen, Wanyi, et al.
Veröffentlicht: (2026) -
MarS: a Financial Market Simulation Engine Powered by Generative Foundation Model
von: Li, Junjie, et al.
Veröffentlicht: (2024) -
R&D-Agent-Quant: A Multi-Agent Framework for Data-Centric Factors and Model Joint Optimization
von: Li, Yuante, et al.
Veröffentlicht: (2025)