Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Ngoc-Hieu, Shojaee, Parshin, Nguyen, Phuc Minh, Zhang, Nan, Reddy, Chandan K, Doan, Khoa D, Zhang, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
Towards Scientific Discovery with Generative AI: Progress, Opportunities, and Challenges
von: Reddy, Chandan K, et al.
Veröffentlicht: (2024)
von: Reddy, Chandan K, et al.
Veröffentlicht: (2024)
LLM-FE: Automated Feature Engineering for Tabular Data with LLMs as Evolutionary Optimizers
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
SURFACEBENCH: A Geometry-Aware Benchmark for Symbolic Surface Discovery
von: Kabra, Sanchit, et al.
Veröffentlicht: (2025)
von: Kabra, Sanchit, et al.
Veröffentlicht: (2025)
SNIP: Bridging Mathematical Symbolic and Numeric Realms with Unified Pre-training
von: Meidani, Kazem, et al.
Veröffentlicht: (2023)
von: Meidani, Kazem, et al.
Veröffentlicht: (2023)
Mitigating Reward Over-optimization in Direct Alignment Algorithms with Importance Sampling
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
Discovering Heuristics with Large Language Models (LLMs) for Mixed-Integer Programs: Single-Machine Scheduling
von: Çetinkaya, İbrahim Oğuz, et al.
Veröffentlicht: (2025)
von: Çetinkaya, İbrahim Oğuz, et al.
Veröffentlicht: (2025)
LLM-SR: Scientific Equation Discovery via Programming with Large Language Models
von: Shojaee, Parshin, et al.
Veröffentlicht: (2024)
von: Shojaee, Parshin, et al.
Veröffentlicht: (2024)
The Reasoning Boundary Paradox: How Reinforcement Learning Constrains Language Models
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
Cold-start Recommendation by Personalized Embedding Region Elicitation
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2024)
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2024)
Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories
von: Beigi, Mohammad, et al.
Veröffentlicht: (2025)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2025)
Pauli nonlocality and the nucleon effective mass
von: Khoa, Dao T., et al.
Veröffentlicht: (2024)
von: Khoa, Dao T., et al.
Veröffentlicht: (2024)
When Reasoning Meets Compression: Understanding the Effects of LLMs Compression on Large Reasoning Models
von: Zhang, Nan, et al.
Veröffentlicht: (2025)
von: Zhang, Nan, et al.
Veröffentlicht: (2025)
Robust Aggregation for Federated Sequential Recommendation with Sparse and Poisoned Data
von: Nguyen, Minh Hieu
Veröffentlicht: (2026)
von: Nguyen, Minh Hieu
Veröffentlicht: (2026)
[Paper 5] The Information Ignition Principle: Why the Universe Inflates
von: Nguyen, Khoa
Veröffentlicht: (2026)
von: Nguyen, Khoa
Veröffentlicht: (2026)
Nuclear Rainbow of Core-Symmetric Systems
von: Phuc, Nguyen Tri Toan, et al.
Veröffentlicht: (2026)
von: Phuc, Nguyen Tri Toan, et al.
Veröffentlicht: (2026)
Nuclear rainbow of the symmetric nucleus-nucleus system: Interchange of the nearside and farside scattering
von: Phuc, Nguyen Tri Toan, et al.
Veröffentlicht: (2024)
von: Phuc, Nguyen Tri Toan, et al.
Veröffentlicht: (2024)
Evaluating the Prognostic Accuracy of New Scores for In‐Hospital Outcomes in Cirrhotic Patients With Esophageal Variceal Bleeding
von: Khoa Phuoc Nguyen, et al.
Veröffentlicht: (2026)
von: Khoa Phuoc Nguyen, et al.
Veröffentlicht: (2026)
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
Wicked Oddities: Selectively Poisoning for Effective Clean-Label Backdoor Attacks
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
von: Nguyen, Quang H., et al.
Veröffentlicht: (2024)
Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context
von: Alizadeh, Keivan, et al.
Veröffentlicht: (2026)
von: Alizadeh, Keivan, et al.
Veröffentlicht: (2026)
EventCap
von: Nguyen, Phuc-Tan, et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc-Tan, et al.
Veröffentlicht: (2024)
Retrospective Feature Estimation for Continual Learning
von: Nguyen, Nghia D., et al.
Veröffentlicht: (2024)
von: Nguyen, Nghia D., et al.
Veröffentlicht: (2024)
Environmental taxes and the economy: New evidence of the shadow economy
von: Canh Phuc Nguyen, et al.
Veröffentlicht: (2026)
von: Canh Phuc Nguyen, et al.
Veröffentlicht: (2026)
Metric constructions and fixed point theorems in product spaces
von: Hieu, Doan Huu, et al.
Veröffentlicht: (2026)
von: Hieu, Doan Huu, et al.
Veröffentlicht: (2026)
Economic Policy Uncertainty and Gambling Preference: Evidence From the Asia‐Pacific Stock Markets
von: Khoa Dang Duong, et al.
Veröffentlicht: (2026)
von: Khoa Dang Duong, et al.
Veröffentlicht: (2026)
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
The Effect of Fiscal Decentralization on Tax Compliance Time: Evidence From the World Bank Doing Business Project
von: Nguyen Doan
Veröffentlicht: (2025)
von: Nguyen Doan
Veröffentlicht: (2025)
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
von: Nguyen, Hieu Minh "Jord"
Veröffentlicht: (2025)
von: Nguyen, Hieu Minh "Jord"
Veröffentlicht: (2025)
SAMSA: Efficient Transformer for Many Data Modalities
von: Lenhat, Minh, et al.
Veröffentlicht: (2024)
von: Lenhat, Minh, et al.
Veröffentlicht: (2024)
Risk of Osteoporosis and Degraded Trabecular Bone Score in Rheumatoid Arthritis Patients
von: Huong Nguyen, et al.
Veröffentlicht: (2025)
von: Huong Nguyen, et al.
Veröffentlicht: (2025)
A von Neumann-Jordan Constant of Non-Normable Metrics
von: Hieu, Doan Huu, et al.
Veröffentlicht: (2026)
von: Hieu, Doan Huu, et al.
Veröffentlicht: (2026)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
Amortized Optimal Transport from Sliced Potentials
von: Truong, Minh-Phuc, et al.
Veröffentlicht: (2026)
von: Truong, Minh-Phuc, et al.
Veröffentlicht: (2026)
Venomancer: Towards Imperceptible and Target-on-Demand Backdoor Attacks in Federated Learning
von: Nguyen, Son, et al.
Veröffentlicht: (2024)
von: Nguyen, Son, et al.
Veröffentlicht: (2024)
Improving Vietnamese Legal Document Retrieval using Synthetic Data
von: Tien, Son Pham, et al.
Veröffentlicht: (2024)
von: Tien, Son Pham, et al.
Veröffentlicht: (2024)
Are you SURE? Enhancing Multimodal Pretraining with Missing Modalities through Uncertainty Estimation
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
Folding model approach to the elastic $p+^{12,13}$C scattering at low energies and radiative capture $^{12,13}$C$(p,γ)$ reactions
von: Anh, Nguyen Le, et al.
Veröffentlicht: (2020)
von: Anh, Nguyen Le, et al.
Veröffentlicht: (2020)
Split Learning without Local Weight Sharing to Enhance Client-side Data Privacy
von: Pham, Ngoc Duy, et al.
Veröffentlicht: (2022)
von: Pham, Ngoc Duy, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025) -
Towards Scientific Discovery with Generative AI: Progress, Opportunities, and Challenges
von: Reddy, Chandan K, et al.
Veröffentlicht: (2024) -
LLM-FE: Automated Feature Engineering for Tabular Data with LLMs as Evolutionary Optimizers
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025) -
SURFACEBENCH: A Geometry-Aware Benchmark for Symbolic Surface Discovery
von: Kabra, Sanchit, et al.
Veröffentlicht: (2025) -
SNIP: Bridging Mathematical Symbolic and Numeric Realms with Unified Pre-training
von: Meidani, Kazem, et al.
Veröffentlicht: (2023)