Improving Predictor Reliability with Selective Recalibration
Fuente:
arXiv
Saved in:
| Main Authors: | Zollo, Thomas P., Deng, Zhun, Snell, Jake C., Pitassi, Toniann, Zemel, Richard |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
by: Zollo, Thomas P., et al.
Published: (2023)
by: Zollo, Thomas P., et al.
Published: (2023)
Distribution-Free Statistical Dispersion Control for Societal Applications
by: Deng, Zhun, et al.
Published: (2023)
by: Deng, Zhun, et al.
Published: (2023)
Test-Time Warmup for Multimodal Large Language Models
by: Rajaneesh, Nikita, et al.
Published: (2025)
by: Rajaneesh, Nikita, et al.
Published: (2025)
Poly-attention: a general scheme for higher-order self-attention
by: Chakrabarti, Sayak, et al.
Published: (2026)
by: Chakrabarti, Sayak, et al.
Published: (2026)
Every Bit Counts: A Theoretical Study of Precision-Expressivity Tradeoffs in Quantized Transformers
by: Chakrabarti, Sayak, et al.
Published: (2026)
by: Chakrabarti, Sayak, et al.
Published: (2026)
Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions
by: Ding, Ruomeng, et al.
Published: (2026)
by: Ding, Ruomeng, et al.
Published: (2026)
Adaptive Elicitation of Latent Information Using Natural Language
by: Wang, Jimmy, et al.
Published: (2025)
by: Wang, Jimmy, et al.
Published: (2025)
Towards Effective Discrimination Testing for Generative AI
by: Zollo, Thomas P., et al.
Published: (2024)
by: Zollo, Thomas P., et al.
Published: (2024)
Conformal Prediction as Bayesian Quadrature
by: Snell, Jake C., et al.
Published: (2025)
by: Snell, Jake C., et al.
Published: (2025)
Confidence Calibration in Vision-Language-Action Models
by: Zollo, Thomas P, et al.
Published: (2025)
by: Zollo, Thomas P, et al.
Published: (2025)
QuEst: Enhancing Estimates of Quantile-Based Distributional Measures Using Model Predictions
by: Deng, Zhun, et al.
Published: (2025)
by: Deng, Zhun, et al.
Published: (2025)
Replay Can Provably Increase Forgetting
by: Mahdaviyeh, Yasaman, et al.
Published: (2025)
by: Mahdaviyeh, Yasaman, et al.
Published: (2025)
Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation
by: Zollo, Thomas, et al.
Published: (2026)
by: Zollo, Thomas, et al.
Published: (2026)
Tell Me What To Learn: Generalizing Neural Memory to be Controllable in Natural Language
by: Bennett, Max S., et al.
Published: (2026)
by: Bennett, Max S., et al.
Published: (2026)
Meta-Learning at Scale for Large Language Models via Low-Rank Amortized Bayesian Meta-Learning
by: Zhang, Liyi, et al.
Published: (2025)
by: Zhang, Liyi, et al.
Published: (2025)
Level Up: Defining and Exploiting Transitional Problems for Curriculum Learning
by: Tang, Zhenwei, et al.
Published: (2026)
by: Tang, Zhenwei, et al.
Published: (2026)
Differential privacy from axioms
by: Blanc, Guy, et al.
Published: (2025)
by: Blanc, Guy, et al.
Published: (2025)
Learning Human-Aligned Representations with Contrastive Learning and Generative Similarity
by: Marjieh, Raja, et al.
Published: (2024)
by: Marjieh, Raja, et al.
Published: (2024)
Few-Shot Recalibration of Language Models
by: Li, Xiang Lisa, et al.
Published: (2024)
by: Li, Xiang Lisa, et al.
Published: (2024)
Guiding LLM Decision-Making with Fairness Reward Models
by: Hall, Zara, et al.
Published: (2025)
by: Hall, Zara, et al.
Published: (2025)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
by: Liu, Zhangyi, et al.
Published: (2026)
by: Liu, Zhangyi, et al.
Published: (2026)
Adaptive MSD-Splitting: Enhancing C4.5 and Random Forests for Skewed Continuous Attributes
by: Lee, Jake
Published: (2026)
by: Lee, Jake
Published: (2026)
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
by: Zhong, Huiying, et al.
Published: (2024)
by: Zhong, Huiying, et al.
Published: (2024)
Towards a Novel Perspective on Adversarial Examples Driven by Frequency
by: Zhang, Zhun, et al.
Published: (2024)
by: Zhang, Zhun, et al.
Published: (2024)
Sample Margin-Aware Recalibration of Temperature Scaling
by: Guo, Haolan, et al.
Published: (2025)
by: Guo, Haolan, et al.
Published: (2025)
POEM: Explore Unexplored Reliable Samples to Enhance Test-Time Adaptation
by: Yi, Chang'an, et al.
Published: (2025)
by: Yi, Chang'an, et al.
Published: (2025)
Beyond Greedy Exits: Improved Early Exit Decisions for Risk Control and Reliability
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
Learning Universal Predictors
by: Grau-Moya, Jordi, et al.
Published: (2024)
by: Grau-Moya, Jordi, et al.
Published: (2024)
A Reinforcement Learning-Based Task Mapping Method to Improve the Reliability of Clustered Manycores
by: Hossein-Khani, Fatemeh, et al.
Published: (2024)
by: Hossein-Khani, Fatemeh, et al.
Published: (2024)
Improving Prediction Certainty Estimation for Reliable Early Exiting via Null Space Projection
by: He, Jianing, et al.
Published: (2025)
by: He, Jianing, et al.
Published: (2025)
$f$-Trajectory Balance: A Loss Family for Tuning GFlowNets, Generative Models, and LLMs with Off- and On-Policy Data
by: Fawkes, Jake, et al.
Published: (2026)
by: Fawkes, Jake, et al.
Published: (2026)
Why Uncertainty Calibration Matters for Reliable Perturbation-based Explanations
by: Decker, Thomas, et al.
Published: (2025)
by: Decker, Thomas, et al.
Published: (2025)
Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity
by: Grant, Satchel, et al.
Published: (2026)
by: Grant, Satchel, et al.
Published: (2026)
Embedding Reliability Verification Constraints into Generation Expansion Planning
by: Liu, Peng, et al.
Published: (2025)
by: Liu, Peng, et al.
Published: (2025)
Less is More: Improving LLM Alignment via Preference Data Selection
by: Deng, Xun, et al.
Published: (2025)
by: Deng, Xun, et al.
Published: (2025)
Improving the Computational Efficiency and Explainability of GeoAggregator
by: Deng, Rui, et al.
Published: (2025)
by: Deng, Rui, et al.
Published: (2025)
Improve Knowledge Distillation via Label Revision and Data Selection
by: Lan, Weichao, et al.
Published: (2024)
by: Lan, Weichao, et al.
Published: (2024)
Stable Attention Response for Reliable Precipitation Nowcasting
by: Wen, Penghui, et al.
Published: (2026)
by: Wen, Penghui, et al.
Published: (2026)
Reliable Explanations or Random Noise? A Reliability Metric for XAI
by: Sengupta, Poushali, et al.
Published: (2026)
by: Sengupta, Poushali, et al.
Published: (2026)
Improving Multimodal Learning Balance and Sufficiency through Data Remixing
by: Ma, Xiaoyu, et al.
Published: (2025)
by: Ma, Xiaoyu, et al.
Published: (2025)
Similar Items
-
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
by: Zollo, Thomas P., et al.
Published: (2023) -
Distribution-Free Statistical Dispersion Control for Societal Applications
by: Deng, Zhun, et al.
Published: (2023) -
Test-Time Warmup for Multimodal Large Language Models
by: Rajaneesh, Nikita, et al.
Published: (2025) -
Poly-attention: a general scheme for higher-order self-attention
by: Chakrabarti, Sayak, et al.
Published: (2026) -
Every Bit Counts: A Theoretical Study of Precision-Expressivity Tradeoffs in Quantized Transformers
by: Chakrabarti, Sayak, et al.
Published: (2026)