Group Pattern Selection Optimization: Let LRMs Pick the Right Pattern for Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Hanbin, Song, Jingwei, Li, Jinpeng, Mi, Fei, Shang, Lifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Teaching Large Reasoning Models Effective Reflection
von: Wang, Hanbin, et al.
Veröffentlicht: (2026)
von: Wang, Hanbin, et al.
Veröffentlicht: (2026)
Let LRMs Break Free from Overthinking via Self-Braking Tuning
von: Zhao, Haoran, et al.
Veröffentlicht: (2025)
von: Zhao, Haoran, et al.
Veröffentlicht: (2025)
KDRL: Post-Training Reasoning LLMs via Unified Knowledge Distillation and Reinforcement Learning
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
von: Yang, Junxiao, et al.
Veröffentlicht: (2025)
von: Yang, Junxiao, et al.
Veröffentlicht: (2025)
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
von: Puerto, Haritz, et al.
Veröffentlicht: (2026)
von: Puerto, Haritz, et al.
Veröffentlicht: (2026)
GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning
von: Wang, Jingyi, et al.
Veröffentlicht: (2026)
von: Wang, Jingyi, et al.
Veröffentlicht: (2026)
Patterns Over Principles: The Fragility of Inductive Reasoning in LLMs under Noisy Observations
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
Stop Unnecessary Reflection: Training LRMs for Efficient Reasoning with Adaptive Reflection and Length Coordinated Penalty
von: Yu, Zewei, et al.
Veröffentlicht: (2026)
von: Yu, Zewei, et al.
Veröffentlicht: (2026)
Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners
von: Csaba, Botos, et al.
Veröffentlicht: (2026)
von: Csaba, Botos, et al.
Veröffentlicht: (2026)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
von: Gao, Fan, et al.
Veröffentlicht: (2025)
von: Gao, Fan, et al.
Veröffentlicht: (2025)
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
Breaking Thought Patterns: A Multi-Dimensional Reasoning Framework for LLMs
von: Tang, Xintong, et al.
Veröffentlicht: (2025)
von: Tang, Xintong, et al.
Veröffentlicht: (2025)
Random Graph Set and Evidence Pattern Reasoning Model
von: Zhan, Tianxiang, et al.
Veröffentlicht: (2024)
von: Zhan, Tianxiang, et al.
Veröffentlicht: (2024)
Socratic-PRMBench: Benchmarking Process Reward Models with Systematic Reasoning Patterns
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
von: Chen, Xingwu, et al.
Veröffentlicht: (2025)
von: Chen, Xingwu, et al.
Veröffentlicht: (2025)
Reasoning Pattern Alignment Merging for Adaptive Reasoning
von: Zhong, Zhaofeng, et al.
Veröffentlicht: (2026)
von: Zhong, Zhaofeng, et al.
Veröffentlicht: (2026)
Margin-aware Fuzzy Rough Feature Selection: Bridging Uncertainty Characterization and Pattern Classification
von: Xu, Suping, et al.
Veröffentlicht: (2025)
von: Xu, Suping, et al.
Veröffentlicht: (2025)
Self-Consistency Boosts Calibration for Math Reasoning
von: Wang, Ante, et al.
Veröffentlicht: (2024)
von: Wang, Ante, et al.
Veröffentlicht: (2024)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
von: Li, Siyi, et al.
Veröffentlicht: (2026)
von: Li, Siyi, et al.
Veröffentlicht: (2026)
Fine-Grained Self-Endorsement Improves Factuality and Reasoning
von: Wang, Ante, et al.
Veröffentlicht: (2024)
von: Wang, Ante, et al.
Veröffentlicht: (2024)
How Does the Thinking Step Influence Model Safety? An Entropy-based Safety Reminder for LRMs
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2026)
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2026)
ICPO: Intrinsic Confidence-Driven Group Relative Preference Optimization for Efficient Reinforcement Learning
von: Wang, Jinpeng, et al.
Veröffentlicht: (2025)
von: Wang, Jinpeng, et al.
Veröffentlicht: (2025)
DocPuzzle: A Process-Aware Benchmark for Evaluating Realistic Long-Context Reasoning Capabilities
von: Zhuang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhuang, Tianyi, et al.
Veröffentlicht: (2025)
Insights on Disagreement Patterns in Multimodal Safety Perception across Diverse Rater Groups
von: Rastogi, Charvi, et al.
Veröffentlicht: (2024)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2024)
Entropy Centroids as Intrinsic Rewards for Test-Time Scaling
von: Zhao, Wenshuo, et al.
Veröffentlicht: (2026)
von: Zhao, Wenshuo, et al.
Veröffentlicht: (2026)
Reasoning Pattern Matters: Learning to Reason without Human Rationales
von: Pang, Chaoxu, et al.
Veröffentlicht: (2025)
von: Pang, Chaoxu, et al.
Veröffentlicht: (2025)
Implicit Bias-Like Patterns in Reasoning Models
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
von: Panaganti, Kishan, et al.
Veröffentlicht: (2026)
von: Panaganti, Kishan, et al.
Veröffentlicht: (2026)
Veritas: Generalizable Deepfake Detection via Pattern-Aware Reasoning
von: Tan, Hao, et al.
Veröffentlicht: (2025)
von: Tan, Hao, et al.
Veröffentlicht: (2025)
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
On Evaluation of Unsupervised Feature Selection for Pattern Classification
von: Kim, Gyu-Il, et al.
Veröffentlicht: (2026)
von: Kim, Gyu-Il, et al.
Veröffentlicht: (2026)
From GPS Points to Travel Patterns: Flexible and Semantic Trajectory Generation with LLMs
von: Zhou, Silin, et al.
Veröffentlicht: (2026)
von: Zhou, Silin, et al.
Veröffentlicht: (2026)
Do Code LLMs Understand Design Patterns?
von: Pan, Zhenyu, et al.
Veröffentlicht: (2025)
von: Pan, Zhenyu, et al.
Veröffentlicht: (2025)
Your thoughts tell who you are: Characterize the reasoning patterns of LRMs
von: Chen, Yida, et al.
Veröffentlicht: (2025)
von: Chen, Yida, et al.
Veröffentlicht: (2025)
Data Management For Training Large Language Models: A Survey
von: Wang, Zige, et al.
Veröffentlicht: (2023)
von: Wang, Zige, et al.
Veröffentlicht: (2023)
PACAD-Based Structural Reasoning for AGI: From Probabilistic Pattern-Matching to Verifiable Canonical Architectures – paper 1, version 2
von: Brown, Cameron
Veröffentlicht: (2026)
von: Brown, Cameron
Veröffentlicht: (2026)
Explainable Knowledge Tracing via Probabilistic Embeddings and Pattern-based Reasoning
von: Wu, Siyu, et al.
Veröffentlicht: (2026)
von: Wu, Siyu, et al.
Veröffentlicht: (2026)
Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions
von: Spracklen, Patrick
Veröffentlicht: (2026)
von: Spracklen, Patrick
Veröffentlicht: (2026)
Are We Scaling the Right Thing? A System Perspective on Test-Time Scaling
von: Zhao, Youpeng, et al.
Veröffentlicht: (2025)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Teaching Large Reasoning Models Effective Reflection
von: Wang, Hanbin, et al.
Veröffentlicht: (2026) -
Let LRMs Break Free from Overthinking via Self-Braking Tuning
von: Zhao, Haoran, et al.
Veröffentlicht: (2025) -
KDRL: Post-Training Reasoning LLMs via Unified Knowledge Distillation and Reinforcement Learning
von: Xu, Hongling, et al.
Veröffentlicht: (2025) -
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
von: Yang, Junxiao, et al.
Veröffentlicht: (2025) -
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
von: Puerto, Haritz, et al.
Veröffentlicht: (2026)