WarrantScore: Modeling Warrants between Claims and Evidence for Substantiation Evaluation in Peer Reviews
Fuente:
arXiv
Saved in:
| Main Authors: | Mori, Kiyotada, Tanaka, Shohei, Hirasawa, Tosho, Kozuno, Tadashi, Yoshino, Koichiro, Ushiku, Yoshitaka |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
by: Xu, Yuzheng, et al.
Published: (2026)
by: Xu, Yuzheng, et al.
Published: (2026)
SciPostGen: Bridging the Gap between Scientific Papers and Poster Layouts
by: Inadumi, Shun, et al.
Published: (2025)
by: Inadumi, Shun, et al.
Published: (2025)
MK2 at PBIG Competition: A Prompt Generation Solution
by: Xu, Yuzheng, et al.
Published: (2025)
by: Xu, Yuzheng, et al.
Published: (2025)
SBS Figures: Pre-training Figure QA from Stage-by-Stage Synthesized Images
by: Shinoda, Risa, et al.
Published: (2024)
by: Shinoda, Risa, et al.
Published: (2024)
Dialogue Response Prefetching Based on Semantic Similarity and Prediction Confidence of Language Model
by: Mori, Kiyotada, et al.
Published: (2025)
by: Mori, Kiyotada, et al.
Published: (2025)
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
by: Saito, Kuniaki, et al.
Published: (2026)
by: Saito, Kuniaki, et al.
Published: (2026)
HalDec-Bench: Benchmarking Hallucination Detector in Image Captioning
by: Saito, Kuniaki, et al.
Published: (2025)
by: Saito, Kuniaki, et al.
Published: (2025)
What Do Humans Hear When Interacting? Experiments on Selective Listening for Evaluating ASR of Spoken Dialogue Systems
by: Mori, Kiyotada, et al.
Published: (2025)
by: Mori, Kiyotada, et al.
Published: (2025)
COM Kitchens: An Unedited Overhead-view Video Dataset as a Vision-Language Benchmark
by: Maeda, Koki, et al.
Published: (2024)
by: Maeda, Koki, et al.
Published: (2024)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
by: Kawano, Seiya, et al.
Published: (2024)
by: Kawano, Seiya, et al.
Published: (2024)
Pruning Multilingual Large Language Models for Multilingual Inference
by: Kim, Hwichan, et al.
Published: (2024)
by: Kim, Hwichan, et al.
Published: (2024)
Assessing the Capabilities of LLMs in Humor:A Multi-dimensional Analysis of Oogiri Generation and Evaluation
by: Sakabe, Ritsu, et al.
Published: (2025)
by: Sakabe, Ritsu, et al.
Published: (2025)
Royal Warrant
by: Parfait, Shaun
Published: (2026)
by: Parfait, Shaun
Published: (2026)
Self Iterative Label Refinement via Robust Unlabeled Learning
by: Asano, Hikaru, et al.
Published: (2025)
by: Asano, Hikaru, et al.
Published: (2025)
Informational Content of Warrant Trading Prior to Interim Monthly‐Revenue Report: Evidence From the Taiwan Warrant Market
by: Che‐Chia Chang, et al.
Published: (2025)
by: Che‐Chia Chang, et al.
Published: (2025)
Where-to-Unmask: Ground-Truth-Guided Unmasking Order Learning for Masked Diffusion Language Models
by: Asano, Hikaru, et al.
Published: (2026)
by: Asano, Hikaru, et al.
Published: (2026)
SciPostLayoutTree: A Dataset for Structural Analysis of Scientific Posters
by: Tanaka, Shohei, et al.
Published: (2025)
by: Tanaka, Shohei, et al.
Published: (2025)
SciPostLayout: A Dataset for Layout Analysis and Layout Generation of Scientific Posters
by: Tanaka, Shohei, et al.
Published: (2024)
by: Tanaka, Shohei, et al.
Published: (2024)
Vision-Language Interpreter for Robot Task Planning
by: Shirai, Keisuke, et al.
Published: (2023)
by: Shirai, Keisuke, et al.
Published: (2023)
Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG
by: Qian, Pin, et al.
Published: (2026)
by: Qian, Pin, et al.
Published: (2026)
Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures
by: Inadumi, Shun, et al.
Published: (2025)
by: Inadumi, Shun, et al.
Published: (2025)
Peerispect: Claim Verification in Scientific Peer Reviews
by: Ghorbanpour, Ali, et al.
Published: (2026)
by: Ghorbanpour, Ali, et al.
Published: (2026)
Is Allergy Evaluation Warranted in Patients With Otitis Media With Effusion?
by: Sophie G. Shay, et al.
Published: (2026)
by: Sophie G. Shay, et al.
Published: (2026)
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
by: Sato, Takuma, et al.
Published: (2025)
by: Sato, Takuma, et al.
Published: (2025)
Evaluating the Capability of Video Question Generation for Expert Knowledge Elicitation
by: Zhang, Huaying, et al.
Published: (2025)
by: Zhang, Huaying, et al.
Published: (2025)
Recipe Generation from Unsegmented Cooking Videos
by: Nishimura, Taichi, et al.
Published: (2022)
by: Nishimura, Taichi, et al.
Published: (2022)
ReviewScore: Misinformed Peer Review Detection with Large Language Models
by: Ryu, Hyun, et al.
Published: (2025)
by: Ryu, Hyun, et al.
Published: (2025)
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models
by: Yoshitake, Michiko, et al.
Published: (2024)
by: Yoshitake, Michiko, et al.
Published: (2024)
The Account of Warrants in Bermejo-Luque’s Giving Reasons
by: ROBERT C. PINTO
Published: (2011)
by: ROBERT C. PINTO
Published: (2011)
Where is the answer? Investigating Positional Bias in Language Model Knowledge Extraction
by: Saito, Kuniaki, et al.
Published: (2024)
by: Saito, Kuniaki, et al.
Published: (2024)
Decoupling Scores and Text: The Politeness Principle in Peer Review
by: Wen, Yingxuan
Published: (2026)
by: Wen, Yingxuan
Published: (2026)
RIGOURATE: Quantifying Scientific Exaggeration with Evidence-Aligned Claim Evaluation
by: James, Joseph, et al.
Published: (2026)
by: James, Joseph, et al.
Published: (2026)
PeerPrism: Peer Evaluation Expertise vs Review-writing AI
by: Sadeghian, Soroush, et al.
Published: (2026)
by: Sadeghian, Soroush, et al.
Published: (2026)
PatentScore: Multi-dimensional Evaluation of LLM-Generated Patent Claims
by: Yoo, Yongmin, et al.
Published: (2025)
by: Yoo, Yongmin, et al.
Published: (2025)
Domain Analysis, Literary Warrant, and Consensus: The Case of Fiction Studies.
by: Beghtol, Clare
Published: (1995)
by: Beghtol, Clare
Published: (1995)
Rapport-Driven Virtual Agent: Rapport Building Dialogue Strategy for Improving User Experience at First Meeting
by: Baihaqi, Muhammad Yeza, et al.
Published: (2024)
by: Baihaqi, Muhammad Yeza, et al.
Published: (2024)
RAVE: Retrieval and Scoring Aware Verifiable Claim Detection
by: Li, Yufeng, et al.
Published: (2025)
by: Li, Yufeng, et al.
Published: (2025)
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
by: Inadumi, Shun, et al.
Published: (2024)
by: Inadumi, Shun, et al.
Published: (2024)
MaterialFigBENCH: benchmark dataset with figures for evaluating college-level materials science problem-solving abilities of multimodal large language models
by: Yoshitake, Michiko, et al.
Published: (2026)
by: Yoshitake, Michiko, et al.
Published: (2026)
Proactive User Information Acquisition via Chats on User-Favored Topics
by: Sato, Shiki, et al.
Published: (2025)
by: Sato, Shiki, et al.
Published: (2025)
Similar Items
-
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
by: Xu, Yuzheng, et al.
Published: (2026) -
SciPostGen: Bridging the Gap between Scientific Papers and Poster Layouts
by: Inadumi, Shun, et al.
Published: (2025) -
MK2 at PBIG Competition: A Prompt Generation Solution
by: Xu, Yuzheng, et al.
Published: (2025) -
SBS Figures: Pre-training Figure QA from Stage-by-Stage Synthesized Images
by: Shinoda, Risa, et al.
Published: (2024) -
Dialogue Response Prefetching Based on Semantic Similarity and Prediction Confidence of Language Model
by: Mori, Kiyotada, et al.
Published: (2025)