Correct Reasoning Paths Visit Shared Decision Pivots
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Dongkyu, Zhang, Amy B. Z., Fehri, Bilel, Wang, Sheng, Chunara, Rumi, Cai, Hengrui, Song, Rui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Clinical Models Change Treatment Decisions?
by: Cho, Dongkyu, et al.
Published: (2026)
by: Cho, Dongkyu, et al.
Published: (2026)
Forget Forgetting: Continual Learning in a World of Abundant Memory
by: Cho, Dongkyu, et al.
Published: (2025)
by: Cho, Dongkyu, et al.
Published: (2025)
Dealing with the Evil Twins: Improving Random Augmentation by Addressing Catastrophic Forgetting of Diverse Augmentations
by: Cho, Dongkyu, et al.
Published: (2025)
by: Cho, Dongkyu, et al.
Published: (2025)
Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration
by: Cho, Dongkyu, et al.
Published: (2025)
by: Cho, Dongkyu, et al.
Published: (2025)
Tree of Concepts: Interpretable Continual Learners in Non-Stationary Clinical Domains
by: Cho, Dongkyu, et al.
Published: (2026)
by: Cho, Dongkyu, et al.
Published: (2026)
Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning
by: Cai, Hengrui, et al.
Published: (2023)
by: Cai, Hengrui, et al.
Published: (2023)
Reasoning Can Be Restored by Correcting a Few Decision Tokens
by: Shen, Changshuo, et al.
Published: (2026)
by: Shen, Changshuo, et al.
Published: (2026)
Supervised Fine-Tuning versus Reinforcement Learning: A Study of Post-Training Methods for Large Language Models
by: Jiang, Haitao, et al.
Published: (2026)
by: Jiang, Haitao, et al.
Published: (2026)
PABU: Progress-Aware Belief Update for Efficient LLM Agents
by: Jiang, Haitao, et al.
Published: (2026)
by: Jiang, Haitao, et al.
Published: (2026)
Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge
by: Zhang, Wenbo, et al.
Published: (2026)
by: Zhang, Wenbo, et al.
Published: (2026)
Configurable Fairness: Direct Optimization of Parity Metrics via Vision-Language Models
by: Zhang, Miao, et al.
Published: (2024)
by: Zhang, Miao, et al.
Published: (2024)
Mitigating Urban-Rural Disparities in Contrastive Representation Learning with Satellite Imagery
by: Zhang, Miao, et al.
Published: (2022)
by: Zhang, Miao, et al.
Published: (2022)
Text Rationalization for Robust Causal Effect Estimation
by: Zhang, Lijinghua, et al.
Published: (2025)
by: Zhang, Lijinghua, et al.
Published: (2025)
Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay
by: Marek, Martin, et al.
Published: (2026)
by: Marek, Martin, et al.
Published: (2026)
REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment
by: Ye, Kai, et al.
Published: (2026)
by: Ye, Kai, et al.
Published: (2026)
MPRM: A Markov Path-based Rule Miner for Efficient and Interpretable Knowledge Graph Reasoning
by: Li, Mingyang, et al.
Published: (2025)
by: Li, Mingyang, et al.
Published: (2025)
Disparate Effect Of Missing Mediators On Transportability of Causal Effects
by: Mhasawade, Vishwali, et al.
Published: (2024)
by: Mhasawade, Vishwali, et al.
Published: (2024)
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models
by: Qian, Zhe, et al.
Published: (2026)
by: Qian, Zhe, et al.
Published: (2026)
Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation
by: Zhang, Wenbo, et al.
Published: (2025)
by: Zhang, Wenbo, et al.
Published: (2025)
Impact on Public Health Decision Making by Utilizing Big Data Without Domain Knowledge
by: Zhang, Miao, et al.
Published: (2024)
by: Zhang, Miao, et al.
Published: (2024)
Process or Result? Manipulated Ending Tokens Can Mislead Reasoning LLMs to Ignore the Correct Reasoning Steps
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Conservative Distributional Reinforcement Learning with Safety Constraints
by: Zhang, Hengrui, et al.
Published: (2022)
by: Zhang, Hengrui, et al.
Published: (2022)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
by: Wan, Guangya, et al.
Published: (2024)
by: Wan, Guangya, et al.
Published: (2024)
Cumulative Path-Level Semantic Reasoning for Inductive Knowledge Graph Completion
by: Wang, Jiapu, et al.
Published: (2026)
by: Wang, Jiapu, et al.
Published: (2026)
PathReasoning: A multimodal reasoning agent for query-based ROI navigation on whole-slide images
by: Zhang, Kunpeng, et al.
Published: (2025)
by: Zhang, Kunpeng, et al.
Published: (2025)
Driving with Regulation: Trustworthy and Interpretable Decision-Making for Autonomous Driving with Retrieval-Augmented Reasoning
by: Cai, Tianhui, et al.
Published: (2024)
by: Cai, Tianhui, et al.
Published: (2024)
Conformal Diffusion Models for Individual Treatment Effect Estimation and Inference
by: Cai, Hengrui, et al.
Published: (2024)
by: Cai, Hengrui, et al.
Published: (2024)
KnowPath: Knowledge-enhanced Reasoning via LLM-generated Inference Paths over Knowledge Graphs
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories
by: Nian, Junjie, et al.
Published: (2026)
by: Nian, Junjie, et al.
Published: (2026)
An Identifiable Cost-Aware Causal Decision-Making Framework Using Counterfactual Reasoning
by: Cai, Ruichu, et al.
Published: (2025)
by: Cai, Ruichu, et al.
Published: (2025)
CivRealm: A Learning and Reasoning Odyssey in Civilization for Decision-Making Agents
by: Qi, Siyuan, et al.
Published: (2024)
by: Qi, Siyuan, et al.
Published: (2024)
PILOC: A Pheromone Inverse Guidance Mechanism and Local-Communication Framework for Dynamic Target Search of Multi-Agent in Unknown Environments
by: Liu, Hengrui, et al.
Published: (2025)
by: Liu, Hengrui, et al.
Published: (2025)
Breaking the Safety-Capability Tradeoff: Reinforcement Learning with Verifiable Rewards Maintains Safety Guardrails in LLMs
by: Cho, Dongkyu Derek, et al.
Published: (2025)
by: Cho, Dongkyu Derek, et al.
Published: (2025)
Conditional Synthesis of 3D Molecules with Time Correction Sampler
by: Jung, Hojung, et al.
Published: (2024)
by: Jung, Hojung, et al.
Published: (2024)
PhiloBERTA: A Transformer-Based Cross-Lingual Analysis of Greek and Latin Lexicons
by: Allbert, Rumi, et al.
Published: (2025)
by: Allbert, Rumi, et al.
Published: (2025)
Explainable Session-based Recommendation via Path Reasoning
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
by: Zhang, Kongcheng, et al.
Published: (2025)
by: Zhang, Kongcheng, et al.
Published: (2025)
Reflective Confidence: Correcting Reasoning Flaws via Online Self-Correction
by: Zeng, Qinglin, et al.
Published: (2025)
by: Zeng, Qinglin, et al.
Published: (2025)
PivotRL: High Accuracy Agentic Post-Training at Low Compute Cost
by: Yi, Junkeun, et al.
Published: (2026)
by: Yi, Junkeun, et al.
Published: (2026)
CADSmith: Multi-Agent CAD Generation with Programmatic Geometric Validation
by: Barkley, Jesse, et al.
Published: (2026)
by: Barkley, Jesse, et al.
Published: (2026)
Similar Items
-
Do Clinical Models Change Treatment Decisions?
by: Cho, Dongkyu, et al.
Published: (2026) -
Forget Forgetting: Continual Learning in a World of Abundant Memory
by: Cho, Dongkyu, et al.
Published: (2025) -
Dealing with the Evil Twins: Improving Random Augmentation by Addressing Catastrophic Forgetting of Diverse Augmentations
by: Cho, Dongkyu, et al.
Published: (2025) -
Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration
by: Cho, Dongkyu, et al.
Published: (2025) -
Tree of Concepts: Interpretable Continual Learners in Non-Stationary Clinical Domains
by: Cho, Dongkyu, et al.
Published: (2026)