OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Shang, Chai, Wenhao, Liu, Kaiyuan, Mao, Huanzhi, Mang, Qiuyang, Shang, Jingbo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoCode: LLMs as Problem Setters for Competitive Programming
by: Zhou, Shang, et al.
Published: (2025)
by: Zhou, Shang, et al.
Published: (2025)
Neural Bradley-Terry Rating: Quantifying Properties from Comparisons
by: Fujii, Satoru
Published: (2023)
by: Fujii, Satoru
Published: (2023)
Efficient Portfolio Selection through Preference Aggregation with Quicksort and the Bradley--Terry Model
by: Ge, Yurun, et al.
Published: (2025)
by: Ge, Yurun, et al.
Published: (2025)
Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives
by: Sun, Hao, et al.
Published: (2024)
by: Sun, Hao, et al.
Published: (2024)
Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live
by: Li, Hanchen, et al.
Published: (2025)
by: Li, Hanchen, et al.
Published: (2025)
FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale
by: He, Runyuan, et al.
Published: (2026)
by: He, Runyuan, et al.
Published: (2026)
Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment
by: Zhang, Yifan, et al.
Published: (2024)
by: Zhang, Yifan, et al.
Published: (2024)
Toward Student-Oriented Teacher Network Training For Knowledge Distillation
by: Dong, Chengyu, et al.
Published: (2022)
by: Dong, Chengyu, et al.
Published: (2022)
Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning
by: Zhu, Yekun, et al.
Published: (2025)
by: Zhu, Yekun, et al.
Published: (2025)
ParallelMuse: Agentic Parallel Thinking for Deep Information Seeking
by: Li, Baixuan, et al.
Published: (2025)
by: Li, Baixuan, et al.
Published: (2025)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
Efficient Test-Time Scaling via Temporal Reasoning Aggregation
by: Li, Jiakun, et al.
Published: (2026)
by: Li, Jiakun, et al.
Published: (2026)
READ: Improving Relation Extraction from an ADversarial Perspective
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Assertion-Conditioned Compliance: A Provenance-Aware Vulnerability in Multi-Turn Tool-Calling Agents
by: Waqas, Daud, et al.
Published: (2025)
by: Waqas, Daud, et al.
Published: (2025)
Parallel Thinking, Sequential Answering: Bridging NAR and AR for Efficient Reasoning
by: Ai, Qihang, et al.
Published: (2025)
by: Ai, Qihang, et al.
Published: (2025)
Think Deep, Think Fast: Investigating Efficiency of Verifier-free Inference-time-scaling Methods
by: Wang, Junlin, et al.
Published: (2025)
by: Wang, Junlin, et al.
Published: (2025)
When is the consistent prediction likely to be a correct prediction?
by: Nguyen, Alex, et al.
Published: (2024)
by: Nguyen, Alex, et al.
Published: (2024)
See and Think: Embodied Agent in Virtual Environment
by: Zhao, Zhonghan, et al.
Published: (2023)
by: Zhao, Zhonghan, et al.
Published: (2023)
Enhancing Spatial Reasoning through Visual and Textual Thinking
by: Liang, Xun, et al.
Published: (2025)
by: Liang, Xun, et al.
Published: (2025)
Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
A Tale of LLMs and Induced Small Proxies: Scalable Agents for Knowledge Mining
by: Zhang, Sipeng, et al.
Published: (2025)
by: Zhang, Sipeng, et al.
Published: (2025)
Vector-ICL: In-context Learning with Continuous Vector Representations
by: Zhuang, Yufan, et al.
Published: (2024)
by: Zhuang, Yufan, et al.
Published: (2024)
Text Generation Beyond Discrete Token Sampling
by: Zhuang, Yufan, et al.
Published: (2025)
by: Zhuang, Yufan, et al.
Published: (2025)
ReasonIF: Large Reasoning Models Fail to Follow Instructions During Reasoning
by: Kwon, Yongchan, et al.
Published: (2025)
by: Kwon, Yongchan, et al.
Published: (2025)
Don't Think Longer, Think Wisely: Optimizing Thinking Dynamics for Large Reasoning Models
by: An, Sohyun, et al.
Published: (2025)
by: An, Sohyun, et al.
Published: (2025)
PoTable: Towards Systematic Thinking via Plan-then-Execute Stage Reasoning on Tables
by: Mao, Qingyang, et al.
Published: (2024)
by: Mao, Qingyang, et al.
Published: (2024)
Explainable Session-based Recommendation via Path Reasoning
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Scaling Retrieval-Augmented Reasoning with Parallel Search and Explicit Merging
by: Liu, Jiabei, et al.
Published: (2026)
by: Liu, Jiabei, et al.
Published: (2026)
Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs
by: Feng, Guangyu, et al.
Published: (2026)
by: Feng, Guangyu, et al.
Published: (2026)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
by: Zheng, Zihan, et al.
Published: (2025)
by: Zheng, Zihan, et al.
Published: (2025)
Explainable Chain-of-Thought Reasoning: An Empirical Analysis on State-Aware Reasoning Dynamics
by: Yu, Sheldon, et al.
Published: (2025)
by: Yu, Sheldon, et al.
Published: (2025)
Parallel Belief Revision via Order Aggregation
by: Chandler, Jake, et al.
Published: (2025)
by: Chandler, Jake, et al.
Published: (2025)
Parallel Belief Contraction via Order Aggregation
by: Chandler, Jake, et al.
Published: (2025)
by: Chandler, Jake, et al.
Published: (2025)
KAG-Thinker: Interactive Thinking and Deep Reasoning in LLMs via Knowledge-Augmented Generation
by: Zhang, Dalong, et al.
Published: (2025)
by: Zhang, Dalong, et al.
Published: (2025)
Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning
by: Ko, Dayoon, et al.
Published: (2025)
by: Ko, Dayoon, et al.
Published: (2025)
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
by: Li, Guanghao, et al.
Published: (2025)
by: Li, Guanghao, et al.
Published: (2025)
Graph Neural Aggregation-diffusion with Metastability
by: Cui, Kaiyuan, et al.
Published: (2024)
by: Cui, Kaiyuan, et al.
Published: (2024)
Learning a Decision Tree Algorithm with Transformers
by: Zhuang, Yufan, et al.
Published: (2024)
by: Zhuang, Yufan, et al.
Published: (2024)
DeepThink3D: Enhancing Large Language Models with Programmatic Reasoning in Complex 3D Situated Reasoning Tasks
by: Song, Jiayi, et al.
Published: (2025)
by: Song, Jiayi, et al.
Published: (2025)
Finish First, Perfect Later: Test-Time Token-Level Cross-Validation for Diffusion Large Language Models
by: Tian, Runchu, et al.
Published: (2025)
by: Tian, Runchu, et al.
Published: (2025)
Similar Items
-
AutoCode: LLMs as Problem Setters for Competitive Programming
by: Zhou, Shang, et al.
Published: (2025) -
Neural Bradley-Terry Rating: Quantifying Properties from Comparisons
by: Fujii, Satoru
Published: (2023) -
Efficient Portfolio Selection through Preference Aggregation with Quicksort and the Bradley--Terry Model
by: Ge, Yurun, et al.
Published: (2025) -
Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives
by: Sun, Hao, et al.
Published: (2024) -
Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live
by: Li, Hanchen, et al.
Published: (2025)