Your Language Model May Think Too Rigidly: Achieving Reasoning Consistency with Symmetry-Enhanced Training
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Yihang, Cen, Zhepeng, Li, Miao, Han, William, Zhang, Yuyou, Liu, Emerson, Liu, Zuxin, Gan, Chuang, Zhao, Ding |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safety is Not Only About Refusal: Reasoning-Enhanced Fine-tuning for Interpretable LLM Safety
by: Zhang, Yuyou, et al.
Published: (2025)
by: Zhang, Yuyou, et al.
Published: (2025)
Feasibility Consistent Representation Learning for Safe Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
Behavior Injection: Preparing Language Models for Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2025)
by: Cen, Zhepeng, et al.
Published: (2025)
Learning from Sparse Offline Datasets via Conservative Density Estimation
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2023)
by: Yao, Yihang, et al.
Published: (2023)
Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization
by: Yao, Yihang, et al.
Published: (2026)
by: Yao, Yihang, et al.
Published: (2026)
CrashAgent: Crash Scenario Generation via Multi-modal Reasoning
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
Signal, Image, or Symbolic: Exploring the Best Input Representation for Electrocardiogram-Language Models Through a Unified Framework
by: Han, William, et al.
Published: (2025)
by: Han, William, et al.
Published: (2025)
OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2024)
by: Yao, Yihang, et al.
Published: (2024)
Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels
by: Cen, Zhepeng, et al.
Published: (2025)
by: Cen, Zhepeng, et al.
Published: (2025)
Bridging the Training-Inference Gap in LLMs by Leveraging Self-Generated Tokens
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning
by: Gan, Siyuan, et al.
Published: (2026)
by: Gan, Siyuan, et al.
Published: (2026)
Exceptional Enhancement of Optical Anisotropy Achieved via the Strategy of Combining Rigid Groups with High Symmetry and π‐Conjugated Organic Groups in Hybrid Fluorides
by: Ru‐Ling Tang, et al.
Published: (2025)
by: Ru‐Ling Tang, et al.
Published: (2025)
We May Have Come Too Far, Too Fast
SpinBench: Perspective and Rotation as a Lens on Spatial Reasoning in VLMs
by: Zhang, Yuyou, et al.
Published: (2025)
by: Zhang, Yuyou, et al.
Published: (2025)
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers
by: Barron, Joshua, et al.
Published: (2025)
by: Barron, Joshua, et al.
Published: (2025)
Bipedalism for Quadrupedal Robots: Versatile Loco-Manipulation through Risk-Adaptive Reinforcement Learning
by: Zhang, Yuyou, et al.
Published: (2025)
by: Zhang, Yuyou, et al.
Published: (2025)
Taming the Long-Tail: Efficient Reasoning RL Training with Adaptive Drafter
by: Hu, Qinghao, et al.
Published: (2025)
by: Hu, Qinghao, et al.
Published: (2025)
Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning
by: Gan, Zeyu, et al.
Published: (2025)
by: Gan, Zeyu, et al.
Published: (2025)
QuietPaw: Learning Quadrupedal Locomotion with Versatile Noise Preference Alignment
by: Zhang, Yuyou, et al.
Published: (2025)
by: Zhang, Yuyou, et al.
Published: (2025)
Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning
by: Zhang, Xue, et al.
Published: (2025)
by: Zhang, Xue, et al.
Published: (2025)
Steering LLM Thinking with Budget Guidance
by: Li, Junyan, et al.
Published: (2025)
by: Li, Junyan, et al.
Published: (2025)
Reverse Thinking Enhances Missing Information Detection in Large Language Models
by: Liu, Yuxin, et al.
Published: (2025)
by: Liu, Yuxin, et al.
Published: (2025)
Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs
by: Tan, Hexiang, et al.
Published: (2025)
by: Tan, Hexiang, et al.
Published: (2025)
Achieving binary weight and activation for LLMs using Post-Training Quantization
by: Song, Siqing, et al.
Published: (2025)
by: Song, Siqing, et al.
Published: (2025)
Wishful Thinking is Risky Thinking
by: Burgh, Jarrod, et al.
Published: (2023)
by: Burgh, Jarrod, et al.
Published: (2023)
Tailored Primitive Initialization is the Secret Key to Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2025)
by: Yao, Yihang, et al.
Published: (2025)
ELF: A Family of Encoder-Free ECG-Language Models
by: Han, William, et al.
Published: (2026)
by: Han, William, et al.
Published: (2026)
A Remeshing Method via Adaptive Multiple Original-Facet-Clipping and Centroidal Voronoi Tessellation
by: Fei, Yue, et al.
Published: (2025)
by: Fei, Yue, et al.
Published: (2025)
As We May Think, Information Systems Do Not.
by: Paisley, William J.
Published: (1968)
by: Paisley, William J.
Published: (1968)
Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training
by: Zhong, Qihuang, et al.
Published: (2026)
by: Zhong, Qihuang, et al.
Published: (2026)
Thinking as Compression: Your Reasoning Model is Secretly a Context Compressor
by: Ma, Guoxin, et al.
Published: (2026)
by: Ma, Guoxin, et al.
Published: (2026)
Think How Your Teammates Think: Active Inference Can Benefit Decentralized Execution
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
Breaking Symmetries Leads to Diverse Quadrupedal Gaits
by: Ding, Jiayu, et al.
Published: (2023)
by: Ding, Jiayu, et al.
Published: (2023)
SQLFixAgent: Towards Semantic-Accurate Text-to-SQL Parsing via Consistency-Enhanced Multi-Agent Collaboration
by: Cen, Jipeng, et al.
Published: (2024)
by: Cen, Jipeng, et al.
Published: (2024)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
by: Yang, Dayu, et al.
Published: (2025)
by: Yang, Dayu, et al.
Published: (2025)
It's Never Too Late to Promote Your Library Services
by: Labiak, Beth
Published: (2007)
by: Labiak, Beth
Published: (2007)
Your Reasoning Benchmark May Not Test Reasoning: Revealing Perception Bottleneck in Abstract Reasoning Benchmarks
by: Wang, Xinhe, et al.
Published: (2025)
by: Wang, Xinhe, et al.
Published: (2025)
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning
by: Zhong, Han, et al.
Published: (2025)
by: Zhong, Han, et al.
Published: (2025)
MM-THEBench: Do Reasoning MLLMs Think Reasonably?
by: Huang, Zhidian, et al.
Published: (2026)
by: Huang, Zhidian, et al.
Published: (2026)
Similar Items
-
Safety is Not Only About Refusal: Reasoning-Enhanced Fine-tuning for Interpretable LLM Safety
by: Zhang, Yuyou, et al.
Published: (2025) -
Feasibility Consistent Representation Learning for Safe Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2024) -
Behavior Injection: Preparing Language Models for Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2025) -
Learning from Sparse Offline Datasets via Conservative Density Estimation
by: Cen, Zhepeng, et al.
Published: (2024) -
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2023)