Saved in:
| Main Authors: | Fan, Shicheng, Hao, Haochang, Min, Dehai, Liu, Weihao, Yu, Philip S., Cheng, Lu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.29648 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
by: Liu, Weihao, et al.
Published: (2026)
by: Liu, Weihao, et al.
Published: (2026)
QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation
by: Min, Dehai, et al.
Published: (2025)
by: Min, Dehai, et al.
Published: (2025)
Let's Verify Math Questions Step by Step
by: Shen, Chengyu, et al.
Published: (2025)
by: Shen, Chengyu, et al.
Published: (2025)
ComposeRAG: A Modular and Composable RAG for Corpus-Grounded Multi-Hop Question Answering
by: Wu, Ruofan, et al.
Published: (2025)
by: Wu, Ruofan, et al.
Published: (2025)
On Early Detection of Hallucinations in Factual Question Answering
by: Snyder, Ben, et al.
Published: (2023)
by: Snyder, Ben, et al.
Published: (2023)
EpiQAL: Benchmarking Large Language Models in Epidemiological Question Answering and Reasoning
by: Wei, Mingyang, et al.
Published: (2026)
by: Wei, Mingyang, et al.
Published: (2026)
Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content
by: Bhalerao, Parth, et al.
Published: (2026)
by: Bhalerao, Parth, et al.
Published: (2026)
UQA: Corpus for Urdu Question Answering
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
AlphaMath Almost Zero: Process Supervision without Process
by: Chen, Guoxin, et al.
Published: (2024)
by: Chen, Guoxin, et al.
Published: (2024)
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
by: Jeong, Minbyul, et al.
Published: (2024)
by: Jeong, Minbyul, et al.
Published: (2024)
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data
by: Min, Dehai, et al.
Published: (2024)
by: Min, Dehai, et al.
Published: (2024)
Context Filtering with Reward Modeling in Question Answering
by: Kim, Sangryul, et al.
Published: (2024)
by: Kim, Sangryul, et al.
Published: (2024)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
by: Pronesti, Massimiliano, et al.
Published: (2026)
by: Pronesti, Massimiliano, et al.
Published: (2026)
Fine-tuning Large Language Models for Improving Factuality in Legal Question Answering
by: Hu, Yinghao, et al.
Published: (2025)
by: Hu, Yinghao, et al.
Published: (2025)
Structured List-Grounded Question Answering
by: Sung, Mujeen, et al.
Published: (2024)
by: Sung, Mujeen, et al.
Published: (2024)
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
by: Qian, Deniz, et al.
Published: (2026)
by: Qian, Deniz, et al.
Published: (2026)
Pre-training, Fine-tuning and Re-ranking: A Three-Stage Framework for Legal Question Answering
by: Ni, Shiwen, et al.
Published: (2024)
by: Ni, Shiwen, et al.
Published: (2024)
P2S: Probabilistic Process Supervision for General-Domain Reasoning Question Answering
by: Zhong, Wenlin, et al.
Published: (2026)
by: Zhong, Wenlin, et al.
Published: (2026)
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering
by: Liang, Sichu, et al.
Published: (2025)
by: Liang, Sichu, et al.
Published: (2025)
Who's Asking? Evaluating LLM Robustness to Inquiry Personas in Factual Question Answering
by: Akpinar, Nil-Jana, et al.
Published: (2025)
by: Akpinar, Nil-Jana, et al.
Published: (2025)
Lessons from Training Grounded LLMs with Verifiable Rewards
by: Sim, Shang Hong, et al.
Published: (2025)
by: Sim, Shang Hong, et al.
Published: (2025)
DPRM: A Dual Implicit Process Reward Model in Multi-Hop Question Answering
by: Wang, Xinyi, et al.
Published: (2025)
by: Wang, Xinyi, et al.
Published: (2025)
AdaptR1: Reinforcement Learning Based Adaptive Interleaved Thinking in Multi-hop Question Answering
by: Wang, Yuxin, et al.
Published: (2026)
by: Wang, Yuxin, et al.
Published: (2026)
FacLens: Transferable Probe for Foreseeing Non-Factuality in Fact-Seeking Question Answering of Large Language Models
by: Wang, Yanling, et al.
Published: (2024)
by: Wang, Yanling, et al.
Published: (2024)
AutoPSV: Automated Process-Supervised Verifier
by: Lu, Jianqiao, et al.
Published: (2024)
by: Lu, Jianqiao, et al.
Published: (2024)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
by: Fan, Lishui, et al.
Published: (2025)
by: Fan, Lishui, et al.
Published: (2025)
When Language Shapes Thought: Cross-Lingual Transfer of Factual Knowledge in Question Answering
by: Kang, Eojin, et al.
Published: (2025)
by: Kang, Eojin, et al.
Published: (2025)
Chain of Condition: Construct, Verify and Solve Conditions for Conditional Question Answering
by: Lin, Jiuheng, et al.
Published: (2024)
by: Lin, Jiuheng, et al.
Published: (2024)
Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains
by: Su, Yi, et al.
Published: (2025)
by: Su, Yi, et al.
Published: (2025)
Weakly Supervised Gaussian Contrastive Grounding with Large Multimodal Models for Video Question Answering
by: Wang, Haibo, et al.
Published: (2024)
by: Wang, Haibo, et al.
Published: (2024)
Reward-SQL: Boosting Text-to-SQL via Stepwise Reasoning and Process-Supervised Rewards
by: Zhang, Yuxin, et al.
Published: (2025)
by: Zhang, Yuxin, et al.
Published: (2025)
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search
by: Wu, Fang, et al.
Published: (2025)
by: Wu, Fang, et al.
Published: (2025)
LiCQA : A Lightweight Complex Question Answering System
by: Saha, Sourav, et al.
Published: (2026)
by: Saha, Sourav, et al.
Published: (2026)
Reasoning by Commented Code for Table Question Answering
by: Pyo, Seho, et al.
Published: (2026)
by: Pyo, Seho, et al.
Published: (2026)
SPAGHETTI: Open-Domain Question Answering from Heterogeneous Data Sources with Retrieval and Semantic Parsing
by: Zhang, Heidi C., et al.
Published: (2024)
by: Zhang, Heidi C., et al.
Published: (2024)
FoRAG: Factuality-optimized Retrieval Augmented Generation for Web-enhanced Long-form Question Answering
by: Cai, Tianchi, et al.
Published: (2024)
by: Cai, Tianchi, et al.
Published: (2024)
Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in Large Language Models
by: Kim, Soyeon, et al.
Published: (2025)
by: Kim, Soyeon, et al.
Published: (2025)
SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems
by: Hao, Haochang, et al.
Published: (2026)
by: Hao, Haochang, et al.
Published: (2026)
Exploring Generative Process Reward Modeling for Semi-Structured Data: A Case Study of Table Question Answering
by: Tang, Lei, et al.
Published: (2025)
by: Tang, Lei, et al.
Published: (2025)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
by: Wang, Zengzhi, et al.
Published: (2023)
by: Wang, Zengzhi, et al.
Published: (2023)
Similar Items
-
Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
by: Liu, Weihao, et al.
Published: (2026) -
QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation
by: Min, Dehai, et al.
Published: (2025) -
Let's Verify Math Questions Step by Step
by: Shen, Chengyu, et al.
Published: (2025) -
ComposeRAG: A Modular and Composable RAG for Corpus-Grounded Multi-Hop Question Answering
by: Wu, Ruofan, et al.
Published: (2025) -
On Early Detection of Hallucinations in Factual Question Answering
by: Snyder, Ben, et al.
Published: (2023)