VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Jasmine, Dantsev, Danylo, Sun, Muyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FGTR: Fine-Grained Multi-Table Retrieval via Hierarchical LLM Reasoning
von: Sun, Chaojie, et al.
Veröffentlicht: (2026)
von: Sun, Chaojie, et al.
Veröffentlicht: (2026)
AgentPRM: Process Reward Models for LLM Agents via Step-Wise Promise and Progress
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
Judging with Personality and Confidence: A Study on Personality-Conditioned LLM Relevance Assessment
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
Task-Adaptive Embedding Refinement via Test-time LLM Guidance
von: Gera, Ariel, et al.
Veröffentlicht: (2026)
von: Gera, Ariel, et al.
Veröffentlicht: (2026)
Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion
von: Dai, Wei, et al.
Veröffentlicht: (2024)
von: Dai, Wei, et al.
Veröffentlicht: (2024)
Navigating Ideation Space: Decomposed Conceptual Representations for Positioning Scientific Ideas
von: Shen, Yuexi, et al.
Veröffentlicht: (2026)
von: Shen, Yuexi, et al.
Veröffentlicht: (2026)
Drowning in Documents: Consequences of Scaling Reranker Inference
von: Jacob, Mathew, et al.
Veröffentlicht: (2024)
von: Jacob, Mathew, et al.
Veröffentlicht: (2024)
Atomic Information Flow: A Network Flow Model for Tool Attributions in RAG Systems
von: Gao, James, et al.
Veröffentlicht: (2026)
von: Gao, James, et al.
Veröffentlicht: (2026)
Graph-based Confidence Calibration for Large Language Models
von: Li, Yukun, et al.
Veröffentlicht: (2024)
von: Li, Yukun, et al.
Veröffentlicht: (2024)
Diagnosing LLM Reranker Behavior Under Fixed Evidence Pools
von: Arat, Baris, et al.
Veröffentlicht: (2026)
von: Arat, Baris, et al.
Veröffentlicht: (2026)
Scaling Up LLM Reviews for Google Ads Content Moderation
von: Qiao, Wei, et al.
Veröffentlicht: (2024)
von: Qiao, Wei, et al.
Veröffentlicht: (2024)
Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Reranking
von: LeVine, Will, et al.
Veröffentlicht: (2025)
von: LeVine, Will, et al.
Veröffentlicht: (2025)
LawLLM: Law Large Language Model for the US Legal System
von: Shu, Dong, et al.
Veröffentlicht: (2024)
von: Shu, Dong, et al.
Veröffentlicht: (2024)
MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification
von: MiroMind Team, et al.
Veröffentlicht: (2026)
von: MiroMind Team, et al.
Veröffentlicht: (2026)
Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
von: Su, Zhengyang, et al.
Veröffentlicht: (2026)
von: Su, Zhengyang, et al.
Veröffentlicht: (2026)
Transforming User Defined Criteria into Explainable Indicators with an Integrated LLM AHP System
von: Bang, Geonwoo, et al.
Veröffentlicht: (2025)
von: Bang, Geonwoo, et al.
Veröffentlicht: (2025)
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
von: Zhang, Yunfan, et al.
Veröffentlicht: (2026)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2026)
Evaluation of LLM-based Strategies for the Extraction of Food Product Information from Online Shops
von: Brosch, Christoph, et al.
Veröffentlicht: (2025)
von: Brosch, Christoph, et al.
Veröffentlicht: (2025)
FIRE: Fact-checking with Iterative Retrieval and Verification
von: Xie, Zhuohan, et al.
Veröffentlicht: (2024)
von: Xie, Zhuohan, et al.
Veröffentlicht: (2024)
SemCSE: Semantic Contrastive Sentence Embeddings Using LLM-Generated Summaries For Scientific Abstracts
von: Brinner, Marc, et al.
Veröffentlicht: (2025)
von: Brinner, Marc, et al.
Veröffentlicht: (2025)
PASC: Pipeline-Aware Conformal Prediction with Joint Coverage Guarantees for Multi-Stage NLP and LLM Pipelines
von: Kotte, Varun
Veröffentlicht: (2026)
von: Kotte, Varun
Veröffentlicht: (2026)
How do Large Language Models Understand Relevance? A Mechanistic Interpretability Perspective
von: Liu, Qi, et al.
Veröffentlicht: (2025)
von: Liu, Qi, et al.
Veröffentlicht: (2025)
LLM vs. Lawyers: Identifying a Subset of Summary Judgments in a Large UK Case Law Dataset
von: Izzidien, Ahmed, et al.
Veröffentlicht: (2024)
von: Izzidien, Ahmed, et al.
Veröffentlicht: (2024)
Description-Based Text Similarity
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2023)
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2023)
On the Theoretical Limitations of Embedding-Based Retrieval
von: Weller, Orion, et al.
Veröffentlicht: (2025)
von: Weller, Orion, et al.
Veröffentlicht: (2025)
MixLM: High-Throughput and Effective LLM Ranking via Text-Embedding Mix-Interaction
von: Li, Guoyao, et al.
Veröffentlicht: (2025)
von: Li, Guoyao, et al.
Veröffentlicht: (2025)
Integrating Large Language Models with Graphical Session-Based Recommendation
von: Guo, Naicheng, et al.
Veröffentlicht: (2024)
von: Guo, Naicheng, et al.
Veröffentlicht: (2024)
A comparison of latent semantic analysis and correspondence analysis of document-term matrices
von: Qi, Qianqian, et al.
Veröffentlicht: (2021)
von: Qi, Qianqian, et al.
Veröffentlicht: (2021)
Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2026)
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2026)
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
Large Language Model Can Be a Foundation for Hidden Rationale-Based Retrieval
von: Ji, Luo, et al.
Veröffentlicht: (2024)
von: Ji, Luo, et al.
Veröffentlicht: (2024)
GraphER: An Efficient Graph-Based Enrichment and Reranking Method for Retrieval-Augmented Generation
von: Miao, Ruizhong, et al.
Veröffentlicht: (2026)
von: Miao, Ruizhong, et al.
Veröffentlicht: (2026)
ORBIT: Preserving Foundational Language Capabilities in GenRetrieval via Origin-Regulated Merging
von: Verma, Neha, et al.
Veröffentlicht: (2026)
von: Verma, Neha, et al.
Veröffentlicht: (2026)
ELMO: Efficiency via Low-precision and Peak Memory Optimization in Large Output Spaces
von: Zhang, Jinbin, et al.
Veröffentlicht: (2025)
von: Zhang, Jinbin, et al.
Veröffentlicht: (2025)
Optimizing Small Transformer-Based Language Models for Multi-Label Sentiment Analysis in Short Texts
von: Neumann, Julius, et al.
Veröffentlicht: (2025)
von: Neumann, Julius, et al.
Veröffentlicht: (2025)
CURE:Circuit-Aware Unlearning for LLM-based Recommendation
von: Chen, Ziheng, et al.
Veröffentlicht: (2026)
von: Chen, Ziheng, et al.
Veröffentlicht: (2026)
Scaling Test-Time Inference with Policy-Optimized, Dynamic Retrieval-Augmented Generation via KV Caching and Decoding
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2025)
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2025)
Membership Inference Attacks on LLM-based Recommender Systems
von: He, Jiajie, et al.
Veröffentlicht: (2025)
von: He, Jiajie, et al.
Veröffentlicht: (2025)
Simple Is Effective: The Roles of Graphs and Large Language Models in Knowledge-Graph-Based Retrieval-Augmented Generation
von: Li, Mufei, et al.
Veröffentlicht: (2024)
von: Li, Mufei, et al.
Veröffentlicht: (2024)
An Integrated Data Processing Framework for Pretraining Foundation Models
von: Sun, Yiding, et al.
Veröffentlicht: (2024)
von: Sun, Yiding, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FGTR: Fine-Grained Multi-Table Retrieval via Hierarchical LLM Reasoning
von: Sun, Chaojie, et al.
Veröffentlicht: (2026) -
AgentPRM: Process Reward Models for LLM Agents via Step-Wise Promise and Progress
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025) -
Judging with Personality and Confidence: A Study on Personality-Conditioned LLM Relevance Assessment
von: Chen, Nuo, et al.
Veröffentlicht: (2026) -
Task-Adaptive Embedding Refinement via Test-time LLM Guidance
von: Gera, Ariel, et al.
Veröffentlicht: (2026) -
Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion
von: Dai, Wei, et al.
Veröffentlicht: (2024)