Tree Training: Accelerating Agentic LLMs Training via Shared Prefix Reuse
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jinghui, Wang, Shaojie, Cui, Yinghan, Chen, Xuxing, Wang, Chao, Huang, Liang, Tang, Can, Zhang, Xiaojiang, Peng, Junyi, Wan, Li, Zhang, Haotian, Chen, Bin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Schedule-Level Shared-Prefix Reuse for LLM RL Training
von: Li, Pengbo, et al.
Veröffentlicht: (2026)
von: Li, Pengbo, et al.
Veröffentlicht: (2026)
SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
von: Zhang, Xiaojiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaojiang, et al.
Veröffentlicht: (2025)
Prefix Grouper: Efficient GRPO Training through Shared-Prefix Forward
von: Liu, Zikang, et al.
Veröffentlicht: (2025)
von: Liu, Zikang, et al.
Veröffentlicht: (2025)
CoDec: Prefix-Shared Decoding Kernel for LLMs
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
SeamlessFlow: A Trainer Agent Isolation RL Framework Achieving Bubble-Free Pipelines via Tag Scheduling
von: Wang, Jinghui, et al.
Veröffentlicht: (2025)
von: Wang, Jinghui, et al.
Veröffentlicht: (2025)
Accelerating Direct Preference Optimization with Prefix Sharing
von: Wang, Franklin, et al.
Veröffentlicht: (2024)
von: Wang, Franklin, et al.
Veröffentlicht: (2024)
Bifurcated Attention: Accelerating Massively Parallel Decoding with Shared Prefixes in LLMs
von: Athiwaratkun, Ben, et al.
Veröffentlicht: (2024)
von: Athiwaratkun, Ben, et al.
Veröffentlicht: (2024)
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
Accelerating Suffix Jailbreak attacks with Prefix-Shared KV-cache
von: Wang, Xinhai, et al.
Veröffentlicht: (2026)
von: Wang, Xinhai, et al.
Veröffentlicht: (2026)
How to Select Pre-Trained Code Models for Reuse? A Learning Perspective
von: Bi, Zhangqian, et al.
Veröffentlicht: (2025)
von: Bi, Zhangqian, et al.
Veröffentlicht: (2025)
From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
Marconi: Prefix Caching for the Era of Hybrid LLMs
von: Pan, Rui, et al.
Veröffentlicht: (2024)
von: Pan, Rui, et al.
Veröffentlicht: (2024)
LLM Query Scheduling with Prefix Reuse and Latency Constraints
von: Dexter, Gregory, et al.
Veröffentlicht: (2025)
von: Dexter, Gregory, et al.
Veröffentlicht: (2025)
Fully First-Order Methods for Decentralized Bilevel Optimization
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
von: Wang, Xi, et al.
Veröffentlicht: (2026)
von: Wang, Xi, et al.
Veröffentlicht: (2026)
Structure-Guided Adversarial Training of Diffusion Models
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Joint Training Across Multiple Activation Sparsity Regimes
von: Wang, Haotian
Veröffentlicht: (2026)
von: Wang, Haotian
Veröffentlicht: (2026)
TAP: A Token-Adaptive Predictor Framework for Training-Free Diffusion Acceleration
von: Zhu, Haowei, et al.
Veröffentlicht: (2026)
von: Zhu, Haowei, et al.
Veröffentlicht: (2026)
ResearchGPT: Benchmarking and Training LLMs for End-to-End Computer Science Research Workflows
von: Wang, Penghao, et al.
Veröffentlicht: (2025)
von: Wang, Penghao, et al.
Veröffentlicht: (2025)
PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
An Augmented Lagrangian Method for Training Recurrent Neural Networks
von: Wang, Yue, et al.
Veröffentlicht: (2024)
von: Wang, Yue, et al.
Veröffentlicht: (2024)
PrefixQuant: Eliminating Outliers by Prefixed Tokens for Large Language Models Quantization
von: Chen, Mengzhao, et al.
Veröffentlicht: (2024)
von: Chen, Mengzhao, et al.
Veröffentlicht: (2024)
Exploring Dynamic Properties of Backdoor Training Through Information Bottleneck
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
KAT-Coder Technical Report
von: Zhan, Zizheng, et al.
Veröffentlicht: (2025)
von: Zhan, Zizheng, et al.
Veröffentlicht: (2025)
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs
von: Lu, Keer, et al.
Veröffentlicht: (2024)
von: Lu, Keer, et al.
Veröffentlicht: (2024)
From Prefix Cache to Fusion RAG Cache: Accelerating LLM Inference in Retrieval-Augmented Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
von: Wang, Jiahao, et al.
Veröffentlicht: (2026)
FRDiff : Feature Reuse for Universal Training-free Acceleration of Diffusion Models
von: So, Junhyuk, et al.
Veröffentlicht: (2023)
von: So, Junhyuk, et al.
Veröffentlicht: (2023)
Part II: ROLL Flash -- Accelerating RLVR and Agentic Training with Asynchrony
von: Lu, Han, et al.
Veröffentlicht: (2025)
von: Lu, Han, et al.
Veröffentlicht: (2025)
Can LLMs Threaten Human Survival? Benchmarking Potential Existential Threats from LLMs via Prefix Completion
von: Cui, Yu, et al.
Veröffentlicht: (2025)
von: Cui, Yu, et al.
Veröffentlicht: (2025)
Regurgitative Training: The Value of Real Data in Training Large Language Models
von: Zhang, Jinghui, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghui, et al.
Veröffentlicht: (2024)
AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse
von: Yu, Zichao, et al.
Veröffentlicht: (2025)
von: Yu, Zichao, et al.
Veröffentlicht: (2025)
sEMG-Based Joint Angle Estimation via Hierarchical Spiking Attentional Feature Decomposition Network
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
Accelerating Compound LLM Training Workloads with Maestro
von: Yuan, Xiulong, et al.
Veröffentlicht: (2026)
von: Yuan, Xiulong, et al.
Veröffentlicht: (2026)
Training-free Diffusion Acceleration with Bottleneck Sampling
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
Chain-of-Models Pre-Training: Rethinking Training Acceleration of Vision Foundation Models
von: Fan, Jiawei, et al.
Veröffentlicht: (2026)
von: Fan, Jiawei, et al.
Veröffentlicht: (2026)
Reuse your FLOPs: Scaling RL on Hard Problems by Conditioning on Very Off-Policy Prefixes
von: Setlur, Amrith, et al.
Veröffentlicht: (2026)
von: Setlur, Amrith, et al.
Veröffentlicht: (2026)
Agentic Critical Training
von: Liu, Weize, et al.
Veröffentlicht: (2026)
von: Liu, Weize, et al.
Veröffentlicht: (2026)
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
von: Li, Yafu, et al.
Veröffentlicht: (2025)
von: Li, Yafu, et al.
Veröffentlicht: (2025)
Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache
von: Cui, Bowen, et al.
Veröffentlicht: (2026)
von: Cui, Bowen, et al.
Veröffentlicht: (2026)
LoongTrain: Efficient Training of Long-Sequence LLMs with Head-Context Parallelism
von: Gu, Diandian, et al.
Veröffentlicht: (2024)
von: Gu, Diandian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Schedule-Level Shared-Prefix Reuse for LLM RL Training
von: Li, Pengbo, et al.
Veröffentlicht: (2026) -
SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
von: Zhang, Xiaojiang, et al.
Veröffentlicht: (2025) -
Prefix Grouper: Efficient GRPO Training through Shared-Prefix Forward
von: Liu, Zikang, et al.
Veröffentlicht: (2025) -
CoDec: Prefix-Shared Decoding Kernel for LLMs
von: Wang, Zhibin, et al.
Veröffentlicht: (2025) -
SeamlessFlow: A Trainer Agent Isolation RL Framework Achieving Bubble-Free Pipelines via Tag Scheduling
von: Wang, Jinghui, et al.
Veröffentlicht: (2025)