LaMDAgent: An Autonomous Framework for Post-Training Pipeline Optimization via LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yano, Taro, Ishibashi, Yoichi, Oyamada, Masafumi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Effective Harness Engineering for Algorithm Discovery with Coding Agents
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2026)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2026)
Mining Hidden Thoughts from Texts: Evaluating Continual Pretraining with Synthetic Data for LLM Reasoning
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2025)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2025)
Can Large Language Models Invent Algorithms to Improve Themselves?: Algorithm Discovery for Recursive Self-Improvement through Reinforcement Learning
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024)
An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability
von: Yamauchi, Yusuke, et al.
Veröffentlicht: (2025)
von: Yamauchi, Yusuke, et al.
Veröffentlicht: (2025)
Can a Crow Hatch a Falcon? Lineage Matters in Predicting Large Language Model Performance
von: Tamura, Takuya, et al.
Veröffentlicht: (2025)
von: Tamura, Takuya, et al.
Veröffentlicht: (2025)
MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making
von: Kim, Yubin, et al.
Veröffentlicht: (2024)
von: Kim, Yubin, et al.
Veröffentlicht: (2024)
Jellyfish: A Large Language Model for Data Preprocessing
von: Zhang, Haochen, et al.
Veröffentlicht: (2023)
von: Zhang, Haochen, et al.
Veröffentlicht: (2023)
Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework
von: Enomoto, Masafumi, et al.
Veröffentlicht: (2026)
von: Enomoto, Masafumi, et al.
Veröffentlicht: (2026)
$M^3$ Scaling Law: Optimizing Multi-Epoch, Multi-Lingual, and Multi-Stage Training for Low-Resource Language Models
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
Fixing It in Post: A Comparative Study of LLM Post-Training Data Quality and Model Performance
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
Unified Mind Model: Reimagining Autonomous Agents in the LLM Era
von: Hu, Pengbo, et al.
Veröffentlicht: (2025)
von: Hu, Pengbo, et al.
Veröffentlicht: (2025)
Look Before You Leap: Autonomous Exploration for LLM Agents
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks
von: Schmidt, Jan-Philipp
Veröffentlicht: (2026)
von: Schmidt, Jan-Philipp
Veröffentlicht: (2026)
Synthesizing Post-Training Data for LLMs through Multi-Agent Simulation
von: Tang, Shuo, et al.
Veröffentlicht: (2024)
von: Tang, Shuo, et al.
Veröffentlicht: (2024)
AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training
von: Pan, Chengjun, et al.
Veröffentlicht: (2026)
von: Pan, Chengjun, et al.
Veröffentlicht: (2026)
Read More, Think More: Revisiting Observation Reduction for Web Agents
von: Enomoto, Masafumi, et al.
Veröffentlicht: (2026)
von: Enomoto, Masafumi, et al.
Veröffentlicht: (2026)
Training an LLM-as-a-Judge Model: Pipeline, Insights, and Practical Lessons
von: Hu, Renjun, et al.
Veröffentlicht: (2025)
von: Hu, Renjun, et al.
Veröffentlicht: (2025)
STACK: Adversarial Attacks on LLM Safeguard Pipelines
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
Context Quality Matters in Training Fusion-in-Decoder for Extractive Open-Domain Question Answering
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
KLong: Training LLM Agent for Extremely Long-horizon Tasks
von: Liu, Yue, et al.
Veröffentlicht: (2026)
von: Liu, Yue, et al.
Veröffentlicht: (2026)
TextMineX: Data, Evaluation Framework and Ontology-guided LLM Pipeline for Humanitarian Mine Action
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
English is Not All You Need: Systematically Exploring the Role of Multilinguality in LLM Post-Training
von: Dhaliwal, Mehak, et al.
Veröffentlicht: (2026)
von: Dhaliwal, Mehak, et al.
Veröffentlicht: (2026)
Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline
von: Zeng, Guancheng, et al.
Veröffentlicht: (2024)
von: Zeng, Guancheng, et al.
Veröffentlicht: (2024)
Performance Evaluation of Emotion Classification in Japanese Using RoBERTa and DeBERTa
von: Takenaka, Yoichi
Veröffentlicht: (2025)
von: Takenaka, Yoichi
Veröffentlicht: (2025)
DEPO: Dual-Efficiency Preference Optimization for LLM Agents
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
LLM-Human Pipeline for Cultural Context Grounding of Conversations
von: Pujari, Rajkumar, et al.
Veröffentlicht: (2024)
von: Pujari, Rajkumar, et al.
Veröffentlicht: (2024)
ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering
von: Liu, Zexi, et al.
Veröffentlicht: (2025)
von: Liu, Zexi, et al.
Veröffentlicht: (2025)
MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments
von: Gao, Yicheng, et al.
Veröffentlicht: (2026)
von: Gao, Yicheng, et al.
Veröffentlicht: (2026)
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
From Utterance to Vividity: Training Expressive Subtitle Translation LLM via Adaptive Local Preference Optimization
von: Cui, Chaoqun, et al.
Veröffentlicht: (2026)
von: Cui, Chaoqun, et al.
Veröffentlicht: (2026)
AI Planning Framework for LLM-Based Web Agents
von: Shahnovsky, Orit, et al.
Veröffentlicht: (2026)
von: Shahnovsky, Orit, et al.
Veröffentlicht: (2026)
AutoKaggle: A Multi-Agent Framework for Autonomous Data Science Competitions
von: Li, Ziming, et al.
Veröffentlicht: (2024)
von: Li, Ziming, et al.
Veröffentlicht: (2024)
Best-of-$\infty$ -- Asymptotic Performance of Test-Time LLM Ensembling
von: Komiyama, Junpei, et al.
Veröffentlicht: (2025)
von: Komiyama, Junpei, et al.
Veröffentlicht: (2025)
AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
von: Tang, Jiabin, et al.
Veröffentlicht: (2025)
von: Tang, Jiabin, et al.
Veröffentlicht: (2025)
BackdoorAgent: A Unified Framework for Backdoor Attacks on LLM-based Agents
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Effective Harness Engineering for Algorithm Discovery with Coding Agents
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2026) -
Mining Hidden Thoughts from Texts: Evaluating Continual Pretraining with Synthetic Data for LLM Reasoning
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2025) -
Can Large Language Models Invent Algorithms to Improve Themselves?: Algorithm Discovery for Recursive Self-Improvement through Reinforcement Learning
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024) -
Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2024) -
An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability
von: Yamauchi, Yusuke, et al.
Veröffentlicht: (2025)