TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ma, Zerun, Wang, Guoqiang, Xie, Xinchen, Chen, Yicheng, Du, He, Li, Bowen, Sun, Yanan, Liu, Wenran, Chen, Kai, Li, Yining |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning
par: Chen, Yicheng, et autres
Publié: (2026)
par: Chen, Yicheng, et autres
Publié: (2026)
MIG: Automatic Data Selection for Instruction Tuning by Maximizing Information Gain in Semantic Space
par: Chen, Yicheng, et autres
Publié: (2025)
par: Chen, Yicheng, et autres
Publié: (2025)
GTA: A Benchmark for General Tool Agents
par: Wang, Jize, et autres
Publié: (2024)
par: Wang, Jize, et autres
Publié: (2024)
RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation
par: Jiang, Tao, et autres
Publié: (2024)
par: Jiang, Tao, et autres
Publié: (2024)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
par: Zhou, Huichi, et autres
Publié: (2025)
par: Zhou, Huichi, et autres
Publié: (2025)
Auto Cherry-Picker: Learning from High-quality Generative Data Driven by Language
par: Chen, Yicheng, et autres
Publié: (2024)
par: Chen, Yicheng, et autres
Publié: (2024)
CooperLLM: Cloud-Edge-End Cooperative Federated Fine-tuning for LLMs via ZOO-based Gradient Correction
par: Sun, He, et autres
Publié: (2026)
par: Sun, He, et autres
Publié: (2026)
CompileAgent: Automated Real-World Repo-Level Compilation with Tool-Integrated LLM-based Agent System
par: Hu, Li, et autres
Publié: (2025)
par: Hu, Li, et autres
Publié: (2025)
Agent Data Protocol: Unifying Datasets for Diverse, Effective Fine-tuning of LLM Agents
par: Song, Yueqi, et autres
Publié: (2025)
par: Song, Yueqi, et autres
Publié: (2025)
CharacterShot: Controllable and Consistent 4D Character Animation
par: Gao, Junyao, et autres
Publié: (2025)
par: Gao, Junyao, et autres
Publié: (2025)
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP
par: Lin, Shuyang, et autres
Publié: (2024)
par: Lin, Shuyang, et autres
Publié: (2024)
An LLM-LVLM Driven Agent for Iterative and Fine-Grained Image Editing
par: Liang, Zihan, et autres
Publié: (2025)
par: Liang, Zihan, et autres
Publié: (2025)
Mechanism Design for LLM Fine-tuning with Multiple Reward Models
par: Sun, Haoran, et autres
Publié: (2024)
par: Sun, Haoran, et autres
Publié: (2024)
The Double-Edged Sword of Open-Ended Interaction: How LLM-Driven NPCs Affect Players' Cognitive Load and Gaming Experience
par: Hsu, Ting-Chen, et autres
Publié: (2026)
par: Hsu, Ting-Chen, et autres
Publié: (2026)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
par: Liu, Xuannan, et autres
Publié: (2026)
par: Liu, Xuannan, et autres
Publié: (2026)
Data-efficient Fine-tuning for LLM-based Recommendation
par: Lin, Xinyu, et autres
Publié: (2024)
par: Lin, Xinyu, et autres
Publié: (2024)
Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning
par: Chen, Sirui, et autres
Publié: (2026)
par: Chen, Sirui, et autres
Publié: (2026)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
par: Ren, Jie, et autres
Publié: (2025)
par: Ren, Jie, et autres
Publié: (2025)
AgentExpt: Automating AI Experiment Design with LLM-based Resource Retrieval Agent
par: Li, Yu, et autres
Publié: (2025)
par: Li, Yu, et autres
Publié: (2025)
Analyzing the Effect of Noise in LLM Fine-tuning
par: Li, Lingfang, et autres
Publié: (2026)
par: Li, Lingfang, et autres
Publié: (2026)
APPT: Boosting Automated Patch Correctness Prediction via Fine-tuning Pre-trained Models
par: Zhang, Quanjun, et autres
Publié: (2023)
par: Zhang, Quanjun, et autres
Publié: (2023)
Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization
par: Du, He, et autres
Publié: (2026)
par: Du, He, et autres
Publié: (2026)
TREX: Tokenizer Regression for Optimal Data Mixture
par: Won, Inho, et autres
Publié: (2026)
par: Won, Inho, et autres
Publié: (2026)
LoRA-PAR: A Flexible Dual-System LoRA Partitioning Approach to Efficient LLM Fine-Tuning
par: Huang, Yining, et autres
Publié: (2025)
par: Huang, Yining, et autres
Publié: (2025)
UFO: Unfair-to-Fair Evolving Mitigates Unfairness in LLM-based Recommender Systems via Self-Play Fine-tuning
par: Zhang, Jiaming, et autres
Publié: (2025)
par: Zhang, Jiaming, et autres
Publié: (2025)
Window-based Membership Inference Attacks Against Fine-tuned Large Language Models
par: Chen, Yuetian, et autres
Publié: (2026)
par: Chen, Yuetian, et autres
Publié: (2026)
Alleviating the Fear of Losing Alignment in LLM Fine-tuning
par: Yang, Kang, et autres
Publié: (2025)
par: Yang, Kang, et autres
Publié: (2025)
Safeguarding LLM Fine-tuning via Push-Pull Distributional Alignment
par: Wang, Haozhong, et autres
Publié: (2026)
par: Wang, Haozhong, et autres
Publié: (2026)
Pyramid-Driven Alignment: Pyramid Principle Guided Integration of Large Language Models and Knowledge Graphs
par: Sun, Lei, et autres
Publié: (2024)
par: Sun, Lei, et autres
Publié: (2024)
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
par: Zheng, Han, et autres
Publié: (2026)
par: Zheng, Han, et autres
Publié: (2026)
Diffusion-Sharpening: Fine-tuning Diffusion Models with Denoising Trajectory Sharpening
par: Tian, Ye, et autres
Publié: (2025)
par: Tian, Ye, et autres
Publié: (2025)
Learning-guided Prioritized Planning for Lifelong Multi-Agent Path Finding in Warehouse Automation
par: Zheng, Han, et autres
Publié: (2026)
par: Zheng, Han, et autres
Publié: (2026)
Federated Co-tuning Framework for Large and Small Language Models
par: Fan, Tao, et autres
Publié: (2024)
par: Fan, Tao, et autres
Publié: (2024)
Fine-tuned LLM-based Code Migration Framework
par: Grynets, Oleg, et autres
Publié: (2025)
par: Grynets, Oleg, et autres
Publié: (2025)
Token-level Data Selection for Safe LLM Fine-tuning
par: Li, Yanping, et autres
Publié: (2026)
par: Li, Yanping, et autres
Publié: (2026)
Harmonia: Algorithm-Hardware Co-Design for Memory- and Compute-Efficient BFP-based LLM Inference
par: Wang, Xinyu, et autres
Publié: (2026)
par: Wang, Xinyu, et autres
Publié: (2026)
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning
par: Zou, Jiaru, et autres
Publié: (2025)
par: Zou, Jiaru, et autres
Publié: (2025)
Outlier-weighed Layerwise Sampling for LLM Fine-tuning
par: Li, Pengxiang, et autres
Publié: (2024)
par: Li, Pengxiang, et autres
Publié: (2024)
PentestAgent: Incorporating LLM Agents to Automated Penetration Testing
par: Shen, Xiangmin, et autres
Publié: (2024)
par: Shen, Xiangmin, et autres
Publié: (2024)
FedLoDrop: Federated LoRA with Dropout for Generalized LLM Fine-tuning
par: Xie, Sijing, et autres
Publié: (2025)
par: Xie, Sijing, et autres
Publié: (2025)
Documents similaires
-
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning
par: Chen, Yicheng, et autres
Publié: (2026) -
MIG: Automatic Data Selection for Instruction Tuning by Maximizing Information Gain in Semantic Space
par: Chen, Yicheng, et autres
Publié: (2025) -
GTA: A Benchmark for General Tool Agents
par: Wang, Jize, et autres
Publié: (2024) -
RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation
par: Jiang, Tao, et autres
Publié: (2024) -
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
par: Zhou, Huichi, et autres
Publié: (2025)