TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Jinyang, Liao, Chonghua, Feng, Mingkuan, Zhang, Shuai, Wen, Zhengqi, Luo, Haoran, Yang, Ling, Xu, Huazhe, Tao, Jianhua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS
by: Wu, Jinyang, et al.
Published: (2024)
by: Wu, Jinyang, et al.
Published: (2024)
RadialRouter: Structured Representation for Efficient and Robust Large Language Models Routing
by: Jin, Ruihan, et al.
Published: (2025)
by: Jin, Ruihan, et al.
Published: (2025)
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
by: Feng, Mingkuan, et al.
Published: (2025)
by: Feng, Mingkuan, et al.
Published: (2025)
AStar: Boosting Multimodal Reasoning with Automated Structured Thinking
by: Wu, Jinyang, et al.
Published: (2025)
by: Wu, Jinyang, et al.
Published: (2025)
Two-Stage Regularization-Based Structured Pruning for LLMs
by: Feng, Mingkuan, et al.
Published: (2025)
by: Feng, Mingkuan, et al.
Published: (2025)
Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs
by: Jin, Ruihan, et al.
Published: (2026)
by: Jin, Ruihan, et al.
Published: (2026)
Entropy Regularizing Activation: Boosting Continuous Control, Large Language Models, and Image Classification with Activation as Entropy Constraints
by: Kang, Zilin, et al.
Published: (2025)
by: Kang, Zilin, et al.
Published: (2025)
Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
Pandora's Box or Aladdin's Lamp: A Comprehensive Analysis Revealing the Role of RAG Noise in Large Language Models
by: Wu, Jinyang, et al.
Published: (2024)
by: Wu, Jinyang, et al.
Published: (2024)
Atlas: Orchestrating Heterogeneous Models and Tools for Multi-Domain Complex Reasoning
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
by: Yang, Ling, et al.
Published: (2025)
by: Yang, Ling, et al.
Published: (2025)
SSL: Sweet Spot Learning for Differentiated Guidance in Agentic Optimization
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
SuperCorrect: Advancing Small LLM Reasoning with Thought Template Distillation and Self-Correction
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models
by: Jin, Ruihan, et al.
Published: (2024)
by: Jin, Ruihan, et al.
Published: (2024)
Protein–Ligand Structure Prediction by Template‐Guided Ensemble Docking Strategy
by: Keqiong Zhang, et al.
Published: (2025)
by: Keqiong Zhang, et al.
Published: (2025)
Improving Large Language Models Function Calling and Interpretability via Guided-Structured Templates
by: Dang, Hy, et al.
Published: (2025)
by: Dang, Hy, et al.
Published: (2025)
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
by: Xie, Tian, et al.
Published: (2025)
by: Xie, Tian, et al.
Published: (2025)
Automating Agent Hijacking via Structural Template Injection
by: Deng, Xinhao, et al.
Published: (2026)
by: Deng, Xinhao, et al.
Published: (2026)
Discovering Conceptual Knowledge with Analytic Ontology Templates for Articulated Objects
by: Sun, Jianhua, et al.
Published: (2024)
by: Sun, Jianhua, et al.
Published: (2024)
Debunk and Infer: Multimodal Fake News Detection via Diffusion-Generated Evidence and LLM Reasoning
by: Yan, Kaiying, et al.
Published: (2025)
by: Yan, Kaiying, et al.
Published: (2025)
ReACT-Drug: Reaction-Template Guided Reinforcement Learning for de novo Drug Design
by: Yadunandan, R, et al.
Published: (2025)
by: Yadunandan, R, et al.
Published: (2025)
From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning
by: Feng, Yichao, et al.
Published: (2025)
by: Feng, Yichao, et al.
Published: (2025)
KS-LLM: Knowledge Selection of Large Language Models with Evidence Document for Question Answering
by: Zheng, Xinxin, et al.
Published: (2024)
by: Zheng, Xinxin, et al.
Published: (2024)
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning
by: Zha, Kaiwen, et al.
Published: (2025)
by: Zha, Kaiwen, et al.
Published: (2025)
Dependency Parsing with the Structuralized Prompt Template
by: Kim, Keunha, et al.
Published: (2025)
by: Kim, Keunha, et al.
Published: (2025)
ReshapeIT: Reliable Shape Interaction with Implicit Template for Anatomical Structure Reconstruction
by: Zhang, Minghui, et al.
Published: (2023)
by: Zhang, Minghui, et al.
Published: (2023)
AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition
by: Lian, Zheng, et al.
Published: (2026)
by: Lian, Zheng, et al.
Published: (2026)
Looking Backward: Retrospective Backward Synthesis for Goal-Conditioned GFlowNets
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Mapping Template
by: Scrocca, Mario, et al.
Published: (2026)
by: Scrocca, Mario, et al.
Published: (2026)
TemplatePackage
by: Schmiedmayer, Paul, et al.
Published: (2025)
by: Schmiedmayer, Paul, et al.
Published: (2025)
Templating Aggregation
by: Krapivsky, P. L., et al.
Published: (2024)
by: Krapivsky, P. L., et al.
Published: (2024)
Quantum Template Resonance Optimization (QTRO) A Template-Driven Variational Strategy for Structured Convergence in the TFIM
by: Quinto, Antonio
Published: (2026)
by: Quinto, Antonio
Published: (2026)
Template‐Guided Silicon Micromotor Assembly for Enhanced Cell Manipulation
by: Yuxin Gao, et al.
Published: (2024)
by: Yuxin Gao, et al.
Published: (2024)
Template‐Guided Silicon Micromotor Assembly for Enhanced Cell Manipulation
by: Yuxin Gao, et al.
Published: (2024)
by: Yuxin Gao, et al.
Published: (2024)
Reinforcement Learning-Based Prompt Template Stealing for Text-to-Image Models
by: Zou, Xiaotian
Published: (2025)
by: Zou, Xiaotian
Published: (2025)
Comparing Template-based and Template-free Language Model Probing
by: Shaier, Sagi, et al.
Published: (2024)
by: Shaier, Sagi, et al.
Published: (2024)
TORSO: Template-Oriented Reasoning Towards General Tasks
by: Kim, Minhyuk, et al.
Published: (2025)
by: Kim, Minhyuk, et al.
Published: (2025)
Verifying Lock-free Search Structure Templates
by: Patel, Nisarg, et al.
Published: (2024)
by: Patel, Nisarg, et al.
Published: (2024)
Biomimetic Catalase‐Templated Nanoprobes for MRI‐Guided Oxygen‐Supplemented Photodynamic Therapy in Breast Cancer
by: Wen Gu, et al.
Published: (2026)
by: Wen Gu, et al.
Published: (2026)
Similar Items
-
Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS
by: Wu, Jinyang, et al.
Published: (2024) -
RadialRouter: Structured Representation for Efficient and Robust Large Language Models Routing
by: Jin, Ruihan, et al.
Published: (2025) -
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
by: Feng, Mingkuan, et al.
Published: (2025) -
AStar: Boosting Multimodal Reasoning with Automated Structured Thinking
by: Wu, Jinyang, et al.
Published: (2025) -
Two-Stage Regularization-Based Structured Pruning for LLMs
by: Feng, Mingkuan, et al.
Published: (2025)