Proper Scoring Rules for Agentic Uncertainty Quantification
Fuente:
arXiv
Saved in:
| Main Authors: | Raghu, Suresh, Pandey, Satwik, Pandey, Shashwat |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SELFDOUBT: Uncertainty Quantification for Reasoning LLMs via the Hedge-to-Verify Ratio
by: Pandey, Satwik, et al.
Published: (2026)
by: Pandey, Satwik, et al.
Published: (2026)
Don't Blink: Evidence Collapse during Multimodal Reasoning
by: Raghu, Suresh, et al.
Published: (2026)
by: Raghu, Suresh, et al.
Published: (2026)
ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules
by: Landsgesell, Jonas, et al.
Published: (2026)
by: Landsgesell, Jonas, et al.
Published: (2026)
Agentic Uncertainty Quantification
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
Teaching Models To Survive: Proper Scoring Rule and Stochastic Optimization with Competing Risks
by: Alberge, Julie, et al.
Published: (2024)
by: Alberge, Julie, et al.
Published: (2024)
Survival Models: Proper Scoring Rule and Stochastic Optimization with Competing Risks
by: Alberge, Julie, et al.
Published: (2024)
by: Alberge, Julie, et al.
Published: (2024)
Evaluating Agentic AI in the Wild: Failure Modes, Drift Patterns, and a Production Evaluation Framework
by: Pandey, Mukund
Published: (2026)
by: Pandey, Mukund
Published: (2026)
Uncertainty Quantification for Regression using Proper Scoring Rules
by: Fishkov, Alexander, et al.
Published: (2025)
by: Fishkov, Alexander, et al.
Published: (2025)
Distributional Regression with Tabular Foundation Models: Evaluating Probabilistic Predictions via Proper Scoring Rules
by: Landsgesell, Jonas, et al.
Published: (2026)
by: Landsgesell, Jonas, et al.
Published: (2026)
Uncertainty Quantification with Proper Scoring Rules: Adjusting Measures to Prediction Tasks
by: Hofman, Paul, et al.
Published: (2025)
by: Hofman, Paul, et al.
Published: (2025)
David vs. Goliath: Can Small Models Win Big with Agentic AI in Hardware Design?
by: Shankar, Shashwat, et al.
Published: (2025)
by: Shankar, Shashwat, et al.
Published: (2025)
Robust Layerwise Scaling Rules by Proper Weight Decay Tuning
by: Fan, Zhiyuan, et al.
Published: (2025)
by: Fan, Zhiyuan, et al.
Published: (2025)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
by: Sun, Kangkang, et al.
Published: (2026)
by: Sun, Kangkang, et al.
Published: (2026)
Abductive and Contrastive Explanations for Scoring Rules in Voting
by: Contet, Clément, et al.
Published: (2024)
by: Contet, Clément, et al.
Published: (2024)
Agentic Uncertainty Reveals Agentic Overconfidence
by: Kaddour, Jean, et al.
Published: (2026)
by: Kaddour, Jean, et al.
Published: (2026)
PowerChain: A Verifiable Agentic AI System for Automating Distribution Grid Analyses
by: Badmus, Emmanuel O., et al.
Published: (2025)
by: Badmus, Emmanuel O., et al.
Published: (2025)
Uncertainty Quantification in the Tsetlin Machine
by: Helin, Runar, et al.
Published: (2025)
by: Helin, Runar, et al.
Published: (2025)
Uncertainty Quantification in SVM prediction
by: Anand, Pritam
Published: (2025)
by: Anand, Pritam
Published: (2025)
Zero-Direction Probing: A Linear-Algebraic Framework for Deep Analysis of Large-Language-Model Drift
by: Pandey, Amit
Published: (2025)
by: Pandey, Amit
Published: (2025)
TRACE: Capability-Targeted Agentic Training
by: Kang, Hangoo, et al.
Published: (2026)
by: Kang, Hangoo, et al.
Published: (2026)
Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems
by: Babu, Rahul Suresh, et al.
Published: (2026)
by: Babu, Rahul Suresh, et al.
Published: (2026)
AgenticRAG: Agentic Retrieval for Enterprise Knowledge Bases
by: Suresh, Susheel, et al.
Published: (2026)
by: Suresh, Susheel, et al.
Published: (2026)
UbiQTree: Uncertainty Quantification in XAI with Tree Ensembles
by: Dubey, Akshat, et al.
Published: (2025)
by: Dubey, Akshat, et al.
Published: (2025)
Aligned Textual Scoring Rules
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
Credal Ensemble Distillation for Uncertainty Quantification
by: Wang, Kaizheng, et al.
Published: (2025)
by: Wang, Kaizheng, et al.
Published: (2025)
Fair Uncertainty Quantification for Depression Prediction
by: Li, Yonghong, et al.
Published: (2025)
by: Li, Yonghong, et al.
Published: (2025)
Torch-Uncertainty: A Deep Learning Framework for Uncertainty Quantification
by: Lafage, Adrien, et al.
Published: (2025)
by: Lafage, Adrien, et al.
Published: (2025)
MAQA: Evaluating Uncertainty Quantification in LLMs Regarding Data Uncertainty
by: Yang, Yongjin, et al.
Published: (2024)
by: Yang, Yongjin, et al.
Published: (2024)
On Training Survival Models with Scoring Rules
by: Kopper, Philipp, et al.
Published: (2024)
by: Kopper, Philipp, et al.
Published: (2024)
Complementing Self-Consistency with Cross-Model Disagreement for Uncertainty Quantification
by: Hamidieh, Kimia, et al.
Published: (2026)
by: Hamidieh, Kimia, et al.
Published: (2026)
Uncertainty Quantification in LLM Agents: Foundations, Emerging Challenges, and Opportunities
by: Oh, Changdae, et al.
Published: (2026)
by: Oh, Changdae, et al.
Published: (2026)
ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues
by: Khandelwal, Dinesh, et al.
Published: (2026)
by: Khandelwal, Dinesh, et al.
Published: (2026)
Recoverability Has a Law: The ERR Measure for Tool-Augmented Agents
by: Vuddanti, Sri Vatsa, et al.
Published: (2026)
by: Vuddanti, Sri Vatsa, et al.
Published: (2026)
STELLAR: Structure-guided LLM Assertion Retrieval and Generation for Formal Verification
by: Rajabi, Saeid, et al.
Published: (2025)
by: Rajabi, Saeid, et al.
Published: (2025)
Uncertainty Quantification for LLM-based Code Generation
by: Xu, Senrong, et al.
Published: (2026)
by: Xu, Senrong, et al.
Published: (2026)
A Rate-Distortion View of Uncertainty Quantification
by: Apostolopoulou, Ifigeneia, et al.
Published: (2024)
by: Apostolopoulou, Ifigeneia, et al.
Published: (2024)
Uncertainty Quantification via Stable Distribution Propagation
by: Petersen, Felix, et al.
Published: (2024)
by: Petersen, Felix, et al.
Published: (2024)
GNN's Uncertainty Quantification using Self-Distillation
by: Daneshvar, Hirad, et al.
Published: (2025)
by: Daneshvar, Hirad, et al.
Published: (2025)
Tackling Fake Forgetting through Uncertainty Quantification
by: Shi, Yingdan, et al.
Published: (2025)
by: Shi, Yingdan, et al.
Published: (2025)
Towards Uncertainty Quantification in Generative Model Learning
by: Morales, Giorgio, et al.
Published: (2025)
by: Morales, Giorgio, et al.
Published: (2025)
Similar Items
-
SELFDOUBT: Uncertainty Quantification for Reasoning LLMs via the Hedge-to-Verify Ratio
by: Pandey, Satwik, et al.
Published: (2026) -
Don't Blink: Evidence Collapse during Multimodal Reasoning
by: Raghu, Suresh, et al.
Published: (2026) -
ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules
by: Landsgesell, Jonas, et al.
Published: (2026) -
Agentic Uncertainty Quantification
by: Zhang, Jiaxin, et al.
Published: (2026) -
Teaching Models To Survive: Proper Scoring Rule and Stochastic Optimization with Competing Risks
by: Alberge, Julie, et al.
Published: (2024)