CoSineVerifier: Tool-Augmented Answer Verification for Computation-Oriented Scientific Questions
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Ruixiang, An, Zhenwei, Wen, Yuntao, Le, Ran, Jia, Yiming, Yang, Chen, Chen, Zongchao, Chen, Lisi, Gao, Shen, Shang, Shuo, Song, Yang, Zhang, Tao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PACE: Prefix-Protected and Difficulty-Aware Compression for Efficient Reasoning
by: Feng, Ruixiang, et al.
Published: (2026)
by: Feng, Ruixiang, et al.
Published: (2026)
ToolMind Technical Report: A Large-Scale, Reasoning-Enhanced Tool-Use Dataset
by: Yang, Chen, et al.
Published: (2025)
by: Yang, Chen, et al.
Published: (2025)
Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation
by: Peng, Guangyue, et al.
Published: (2026)
by: Peng, Guangyue, et al.
Published: (2026)
Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts
by: Yang, Chen, et al.
Published: (2026)
by: Yang, Chen, et al.
Published: (2026)
CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis
by: Feng, Ruixiang, et al.
Published: (2025)
by: Feng, Ruixiang, et al.
Published: (2025)
Verif.ai: Towards an Open-Source Scientific Generative Question-Answering System with Referenced and Verifiable Answers
by: Košprdić, Miloš, et al.
Published: (2024)
by: Košprdić, Miloš, et al.
Published: (2024)
Nanbeige4-3B Technical Report: Exploring the Frontier of Small Language Models
by: Yang, Chen, et al.
Published: (2025)
by: Yang, Chen, et al.
Published: (2025)
From GPS Points to Travel Patterns: Flexible and Semantic Trajectory Generation with LLMs
by: Zhou, Silin, et al.
Published: (2026)
by: Zhou, Silin, et al.
Published: (2026)
Scientific QA System with Verifiable Answers
by: Ljajić, Adela, et al.
Published: (2024)
by: Ljajić, Adela, et al.
Published: (2024)
TOOL4POI: A Tool-Augmented LLM Framework for Next POI Recommendation
by: Wang, Dongsheng, et al.
Published: (2025)
by: Wang, Dongsheng, et al.
Published: (2025)
Region-Point Joint Representation for Effective Trajectory Similarity Learning
by: Long, Hao, et al.
Published: (2025)
by: Long, Hao, et al.
Published: (2025)
What Affects the Stability of Tool Learning? An Empirical Study on the Robustness of Tool Learning Frameworks
by: Huang, Chengrui, et al.
Published: (2024)
by: Huang, Chengrui, et al.
Published: (2024)
Knowledge-Augmented Question Error Correction for Chinese Question Answer System with QuestionRAG
by: Qiu, Longpeng, et al.
Published: (2025)
by: Qiu, Longpeng, et al.
Published: (2025)
ReCoQA: A Benchmark for Tool-Augmented and Multi-Step Reasoning in Real Estate Question and Answering
by: Zhang, Yindong, et al.
Published: (2026)
by: Zhang, Yindong, et al.
Published: (2026)
xVerify: Efficient Answer Verifier for Reasoning Model Evaluations
by: Chen, Ding, et al.
Published: (2025)
by: Chen, Ding, et al.
Published: (2025)
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
by: Li, Yuyi, et al.
Published: (2025)
by: Li, Yuyi, et al.
Published: (2025)
DALK: Dynamic Co-Augmentation of LLMs and KG to answer Alzheimer's Disease Questions with Scientific Literature
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
RED: Effective Trajectory Representation Learning with Comprehensive Information
by: Zhou, Silin, et al.
Published: (2024)
by: Zhou, Silin, et al.
Published: (2024)
Efficient Model-Agnostic Continual Learning for Next POI Recommendation
by: Wang, Chenhao, et al.
Published: (2025)
by: Wang, Chenhao, et al.
Published: (2025)
Grid and Road Expressions Are Complementary for Trajectory Representation Learning
by: Zhou, Silin, et al.
Published: (2024)
by: Zhou, Silin, et al.
Published: (2024)
Computer MCQ Questions And Answers
by: Easy Quizzz
Published: (2026)
by: Easy Quizzz
Published: (2026)
Blurred Encoding for Trajectory Representation Learning
by: Zhou, Silin, et al.
Published: (2025)
by: Zhou, Silin, et al.
Published: (2025)
EEE-QA: Exploring Effective and Efficient Question-Answer Representations
by: Hu, Zhanghao, et al.
Published: (2024)
by: Hu, Zhanghao, et al.
Published: (2024)
AutoVerifier: An Agentic Automated Verification Framework Using Large Language Models
by: Du, Yuntao, et al.
Published: (2026)
by: Du, Yuntao, et al.
Published: (2026)
Incentivizing LLMs to Self-Verify Their Answers
by: Zhang, Fuxiang, et al.
Published: (2025)
by: Zhang, Fuxiang, et al.
Published: (2025)
Intuitive Axial Augmentation Using Polar-Sine-Based Piecewise Distortion for Medical Slice-Wise Segmentation
by: Zhang, Yiqin, et al.
Published: (2024)
by: Zhang, Yiqin, et al.
Published: (2024)
To Answer or to Refuse? Investigating the Effect of Refusal to Answer Privacy‐Invasive Question on Applicants' Perceived Hireability
by: Wanlu Li, et al.
Published: (2024)
by: Wanlu Li, et al.
Published: (2024)
Rehearsing Answers to Probable Questions with Perspective-Taking
by: Shih, Yung-Yu, et al.
Published: (2024)
by: Shih, Yung-Yu, et al.
Published: (2024)
Can We Verify Step by Step for Incorrect Answer Detection?
by: Xu, Xin, et al.
Published: (2024)
by: Xu, Xin, et al.
Published: (2024)
SCI-Verifier: Scientific Verifier with Thinking
by: Zheng, Shenghe, et al.
Published: (2025)
by: Zheng, Shenghe, et al.
Published: (2025)
DRE: Generating Recommendation Explanations by Aligning Large Language Models at Data-level
by: Gao, Shen, et al.
Published: (2024)
by: Gao, Shen, et al.
Published: (2024)
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
by: Chen, Xupeng, et al.
Published: (2026)
by: Chen, Xupeng, et al.
Published: (2026)
Shrinking the Generation-Verification Gap with Weak Verifiers
by: Saad-Falcon, Jon, et al.
Published: (2025)
by: Saad-Falcon, Jon, et al.
Published: (2025)
Prophet: Prompting Large Language Models with Complementary Answer Heuristics for Knowledge-based Visual Question Answering
by: Yu, Zhou, et al.
Published: (2023)
by: Yu, Zhou, et al.
Published: (2023)
Sine-transform-based fast solvers for Riesz fractional nonlinear Schrödinger equations with attractive nonlinearities
by: Chen, Chao, et al.
Published: (2024)
by: Chen, Chao, et al.
Published: (2024)
PaperArena: An Evaluation Benchmark for Tool-Augmented Agentic Reasoning on Scientific Literature
by: Wang, Daoyu, et al.
Published: (2025)
by: Wang, Daoyu, et al.
Published: (2025)
Answer Retrieval in Legal Community Question Answering
by: Askari, Arian, et al.
Published: (2024)
by: Askari, Arian, et al.
Published: (2024)
SWE-Master: Unleashing the Potential of Software Engineering Agents via Post-Training
by: Song, Huatong, et al.
Published: (2026)
by: Song, Huatong, et al.
Published: (2026)
Fast inverse lithography based on a model-driven block stacking convolutional neural network
by: Chen, Ruixiang, et al.
Published: (2024)
by: Chen, Ruixiang, et al.
Published: (2024)
RPDR: A Round-trip Prediction-Based Data Augmentation Framework for Long-Tail Question Answering
by: Zhang, Yiming, et al.
Published: (2026)
by: Zhang, Yiming, et al.
Published: (2026)
Similar Items
-
PACE: Prefix-Protected and Difficulty-Aware Compression for Efficient Reasoning
by: Feng, Ruixiang, et al.
Published: (2026) -
ToolMind Technical Report: A Large-Scale, Reasoning-Enhanced Tool-Use Dataset
by: Yang, Chen, et al.
Published: (2025) -
Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation
by: Peng, Guangyue, et al.
Published: (2026) -
Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts
by: Yang, Chen, et al.
Published: (2026) -
CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis
by: Feng, Ruixiang, et al.
Published: (2025)