OPE: Overcoming Information Saturation in Parallel Thinking via Outline-Guided Path Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Qi, Wang, Jianing, Kong, Deyang, Xi, Xiangyu, Zhang, Jianfei, Lu, Yi, Wang, Jingang, Wang, Wei, Zhang, Shikun, Ye, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Autoformalizer with Tool Feedback
by: Guo, Qi, et al.
Published: (2025)
by: Guo, Qi, et al.
Published: (2025)
Rethinking the Sampling Criteria in Reinforcement Learning for LLM Reasoning: A Competence-Difficulty Alignment Perspective
by: Kong, Deyang, et al.
Published: (2025)
by: Kong, Deyang, et al.
Published: (2025)
SampleMix: A Sample-wise Pre-training Data Mixing Strategey by Coordinating Data Quality and Diversity
by: Xi, Xiangyu, et al.
Published: (2025)
by: Xi, Xiangyu, et al.
Published: (2025)
Advancing Block Diffusion Language Models for Test-Time Scaling
by: Lu, Yi, et al.
Published: (2026)
by: Lu, Yi, et al.
Published: (2026)
HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness
by: Wang, Jianing, et al.
Published: (2026)
by: Wang, Jianing, et al.
Published: (2026)
Crossing symmetry of OPE statistics
by: Wang, Diandian
Published: (2025)
by: Wang, Diandian
Published: (2025)
Instruction Data Selection via Answer Divergence
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Learning to Reason Across Parallel Samples for LLM Reasoning
by: Qi, Jianing, et al.
Published: (2025)
by: Qi, Jianing, et al.
Published: (2025)
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
by: Wang, Jianing, et al.
Published: (2026)
by: Wang, Jianing, et al.
Published: (2026)
SG-Bench: Evaluating LLM Safety Generalization Across Diverse Tasks and Prompt Types
by: Mou, Yutao, et al.
Published: (2024)
by: Mou, Yutao, et al.
Published: (2024)
Data Selection for Multi-turn Dialogue Instruction Tuning
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Goal-Guided Efficient Exploration via Large Language Model in Reinforcement Learning
by: Qi, Yajie, et al.
Published: (2025)
by: Qi, Yajie, et al.
Published: (2025)
Analytic Four-Point Lightlike Form Factors and OPE of Null-Wrapped Polygons
by: Guo, Yuanhong, et al.
Published: (2022)
by: Guo, Yuanhong, et al.
Published: (2022)
Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
CoderUJB: An Executable and Unified Java Benchmark for Practical Programming Scenarios
by: Zeng, Zhengran, et al.
Published: (2024)
by: Zeng, Zhengran, et al.
Published: (2024)
On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency
by: Wang, Yiming, et al.
Published: (2026)
by: Wang, Yiming, et al.
Published: (2026)
Internal stimulation source near‐field small target electrical impedance tomography methodology (ISEIT) in pulmonary interventional surgery
by: Wei Zhang, et al.
Published: (2026)
by: Wei Zhang, et al.
Published: (2026)
The OPE Approach to Renormalization: Operator Mixing
by: Zhang, Jinpeng, et al.
Published: (2026)
by: Zhang, Jinpeng, et al.
Published: (2026)
Reinforcement Learning for Tool-Integrated Interleaved Thinking towards Cross-Domain Generalization
by: Chen, Zhengyu, et al.
Published: (2025)
by: Chen, Zhengyu, et al.
Published: (2025)
Don't Get Lost in the Trees: Streamlining LLM Reasoning by Overcoming Tree Search Exploration Pitfalls
by: Wang, Ante, et al.
Published: (2025)
by: Wang, Ante, et al.
Published: (2025)
Scalable and Certifiable Graph Unlearning: Overcoming the Approximation Error Barrier
by: Yi, Lu, et al.
Published: (2024)
by: Yi, Lu, et al.
Published: (2024)
Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-Future
by: Wang, Yidong, et al.
Published: (2025)
by: Wang, Yidong, et al.
Published: (2025)
Virasoro OPE Blocks, Causal Diamonds, and Higher-Dimensional CFT
by: Haehl, Felix M., et al.
Published: (2025)
by: Haehl, Felix M., et al.
Published: (2025)
Advancing Precise Outline-Conditioned Text Generation with Task Duality and Explicit Outline Control
by: Li, Yunzhe, et al.
Published: (2023)
by: Li, Yunzhe, et al.
Published: (2023)
Seed&Steer: Guiding Large Language Models with Compilable Prefix and Branch Signals for Unit Test Generation
by: Zhou, Shuaiyu, et al.
Published: (2025)
by: Zhou, Shuaiyu, et al.
Published: (2025)
What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity
by: Li, Haoxi, et al.
Published: (2026)
by: Li, Haoxi, et al.
Published: (2026)
Mitigating Spurious Correlations with Causal Logit Perturbation
by: Zhou, Xiaoling, et al.
Published: (2025)
by: Zhou, Xiaoling, et al.
Published: (2025)
SaRO: Enhancing LLM Safety through Reasoning-based Alignment
by: Mou, Yutao, et al.
Published: (2025)
by: Mou, Yutao, et al.
Published: (2025)
Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework
by: Jia, Hongrui, et al.
Published: (2025)
by: Jia, Hongrui, et al.
Published: (2025)
ZeroFlow: Overcoming Catastrophic Forgetting is Easier than You Think
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
The Fox-Wolfram Moment of Jet Production in Relativistic Heavy Ion Collisions
by: Kong, Wei-Xi, et al.
Published: (2024)
by: Kong, Wei-Xi, et al.
Published: (2024)
On the sharp critical mass threshold for the 3D Patlak-Keller-Segel-Navier-Stokes system via Couette flow
by: Cui, Shikun, et al.
Published: (2025)
by: Cui, Shikun, et al.
Published: (2025)
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
by: Goli, Hossein, et al.
Published: (2025)
by: Goli, Hossein, et al.
Published: (2025)
EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration
by: Wu, Jianfei, et al.
Published: (2026)
by: Wu, Jianfei, et al.
Published: (2026)
UP3: Unsupervised Predictive Path Planner for Mobile Robots in Unknown Environment
by: Jianing Luo, et al.
Published: (2025)
by: Jianing Luo, et al.
Published: (2025)
SAGE: Spuriousness-Aware Guided Prompt Exploration for Mitigating Multimodal Bias
by: Ye, Wenqian, et al.
Published: (2025)
by: Ye, Wenqian, et al.
Published: (2025)
ResiHP: Taming LLM Training Failures with Dynamic Hybrid Parallelism
by: Ma, Tenghui, et al.
Published: (2026)
by: Ma, Tenghui, et al.
Published: (2026)
Outline or Solid? The Role of Icon Style on User's Perception
by: Zhangfan Shen, et al.
Published: (2025)
by: Zhangfan Shen, et al.
Published: (2025)
Parallel-R1: Towards Parallel Thinking via Reinforcement Learning
by: Zheng, Tong, et al.
Published: (2025)
by: Zheng, Tong, et al.
Published: (2025)
Similar Items
-
Autoformalizer with Tool Feedback
by: Guo, Qi, et al.
Published: (2025) -
Rethinking the Sampling Criteria in Reinforcement Learning for LLM Reasoning: A Competence-Difficulty Alignment Perspective
by: Kong, Deyang, et al.
Published: (2025) -
SampleMix: A Sample-wise Pre-training Data Mixing Strategey by Coordinating Data Quality and Diversity
by: Xi, Xiangyu, et al.
Published: (2025) -
Advancing Block Diffusion Language Models for Test-Time Scaling
by: Lu, Yi, et al.
Published: (2026) -
HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness
by: Wang, Jianing, et al.
Published: (2026)