Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search
Fuente:
arXiv
Salvato in:
| Autori principali: | Flores, Gerardo A., Deshpande, Yash, Brea, Jannis R., Wilson, Ashia C. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Proper Scoring Rule for Virtual Staining
di: Tonks, Samuel, et al.
Pubblicazione: (2026)
di: Tonks, Samuel, et al.
Pubblicazione: (2026)
Aligning Evaluation with Clinical Priorities: Calibration, Label Shift, and Error Costs
di: Flores, Gerardo A., et al.
Pubblicazione: (2025)
di: Flores, Gerardo A., et al.
Pubblicazione: (2025)
Evaluating Posterior Probabilities: Decision Theory, Proper Scoring Rules, and Calibration
di: Ferrer, Luciana, et al.
Pubblicazione: (2024)
di: Ferrer, Luciana, et al.
Pubblicazione: (2024)
A Consequentialist Critique of Binary Classification Evaluation: Theory, Practice, and Tools
di: Flores, Gerardo, et al.
Pubblicazione: (2025)
di: Flores, Gerardo, et al.
Pubblicazione: (2025)
Quantifying Aleatoric and Epistemic Uncertainty with Proper Scoring Rules
di: Hofman, Paul, et al.
Pubblicazione: (2024)
di: Hofman, Paul, et al.
Pubblicazione: (2024)
Uncertainty Quantification for Regression using Proper Scoring Rules
di: Fishkov, Alexander, et al.
Pubblicazione: (2025)
di: Fishkov, Alexander, et al.
Pubblicazione: (2025)
Language Generation with Strictly Proper Scoring Rules
di: Shao, Chenze, et al.
Pubblicazione: (2024)
di: Shao, Chenze, et al.
Pubblicazione: (2024)
Adaptive Backtracking Line Search
di: Cavalcanti, Joao V., et al.
Pubblicazione: (2024)
di: Cavalcanti, Joao V., et al.
Pubblicazione: (2024)
Position: AI Evaluations Should be Grounded on a Theory of Capability
di: Jo, Nathanael, et al.
Pubblicazione: (2025)
di: Jo, Nathanael, et al.
Pubblicazione: (2025)
Uncertainty Quantification with Proper Scoring Rules: Adjusting Measures to Prediction Tasks
di: Hofman, Paul, et al.
Pubblicazione: (2025)
di: Hofman, Paul, et al.
Pubblicazione: (2025)
The Approximate Fisher Influence Function: Faster Estimation of Data Influence in Statistical Models
di: Lev, Omri, et al.
Pubblicazione: (2024)
di: Lev, Omri, et al.
Pubblicazione: (2024)
Distributional Regression with Tabular Foundation Models: Evaluating Probabilistic Predictions via Proper Scoring Rules
di: Landsgesell, Jonas, et al.
Pubblicazione: (2026)
di: Landsgesell, Jonas, et al.
Pubblicazione: (2026)
Survival Models: Proper Scoring Rule and Stochastic Optimization with Competing Risks
di: Alberge, Julie, et al.
Pubblicazione: (2024)
di: Alberge, Julie, et al.
Pubblicazione: (2024)
Conditional Forecasts and Proper Scoring Rules for Reliable and Accurate Performative Predictions
di: Boeken, Philip, et al.
Pubblicazione: (2025)
di: Boeken, Philip, et al.
Pubblicazione: (2025)
Semivalue-based data valuation is arbitrary and gameable
di: Diehl, Hannah, et al.
Pubblicazione: (2025)
di: Diehl, Hannah, et al.
Pubblicazione: (2025)
Efficient and accurate steering of Large Language Models through attention-guided feature learning
di: Davarmanesh, Parmida, et al.
Pubblicazione: (2026)
di: Davarmanesh, Parmida, et al.
Pubblicazione: (2026)
Mean-field underdamped Langevin dynamics and its spacetime discretization
di: Fu, Qiang, et al.
Pubblicazione: (2023)
di: Fu, Qiang, et al.
Pubblicazione: (2023)
Calibrated Regression Against An Adversary Without Regret
di: Deshpande, Shachi, et al.
Pubblicazione: (2023)
di: Deshpande, Shachi, et al.
Pubblicazione: (2023)
Adaptive Kernel Selection for Stein Variational Gradient Descent
di: Melcher, Moritz, et al.
Pubblicazione: (2025)
di: Melcher, Moritz, et al.
Pubblicazione: (2025)
From Cross-Validation to SURE: Asymptotic Risk of Tuned Regularized Estimators
di: Adusumilli, Karun, et al.
Pubblicazione: (2026)
di: Adusumilli, Karun, et al.
Pubblicazione: (2026)
High-accuracy sampling from constrained spaces with the Metropolis-adjusted Preconditioned Langevin Algorithm
di: Srinivasan, Vishwak, et al.
Pubblicazione: (2024)
di: Srinivasan, Vishwak, et al.
Pubblicazione: (2024)
Improved Regret and Contextual Linear Extension for Pandora's Box and Prophet Inequality
di: Liu, Junyan, et al.
Pubblicazione: (2025)
di: Liu, Junyan, et al.
Pubblicazione: (2025)
The Fallacy of Minimizing Cumulative Regret in the Sequential Task Setting
di: Xu, Ziping, et al.
Pubblicazione: (2024)
di: Xu, Ziping, et al.
Pubblicazione: (2024)
Layered Unlearning for Adversarial Relearning
di: Qian, Timothy, et al.
Pubblicazione: (2025)
di: Qian, Timothy, et al.
Pubblicazione: (2025)
Fast sampling from constrained spaces using the Metropolis-adjusted Mirror Langevin algorithm
di: Srinivasan, Vishwak, et al.
Pubblicazione: (2023)
di: Srinivasan, Vishwak, et al.
Pubblicazione: (2023)
Better Uncertainty Calibration via Proper Scores for Classification and Beyond
di: Gruber, Sebastian G., et al.
Pubblicazione: (2022)
di: Gruber, Sebastian G., et al.
Pubblicazione: (2022)
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
di: Cömer, Can, et al.
Pubblicazione: (2025)
di: Cömer, Can, et al.
Pubblicazione: (2025)
UCD: Unlearning in LLMs via Contrastive Decoding
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2025)
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2025)
Stopping Bayesian Optimization with Probabilistic Regret Bounds
di: Wilson, James T.
Pubblicazione: (2024)
di: Wilson, James T.
Pubblicazione: (2024)
Contextual Pandora's Box
di: Atsidakou, Alexia, et al.
Pubblicazione: (2022)
di: Atsidakou, Alexia, et al.
Pubblicazione: (2022)
Forecast Evaluation and the Relationship of Regret and Calibration
di: Derr, Rabanus, et al.
Pubblicazione: (2024)
di: Derr, Rabanus, et al.
Pubblicazione: (2024)
A Novel Framework for Uncertainty Quantification via Proper Scores for Classification and Beyond
di: Gruber, Sebastian G.
Pubblicazione: (2025)
di: Gruber, Sebastian G.
Pubblicazione: (2025)
Robust Layerwise Scaling Rules by Proper Weight Decay Tuning
di: Fan, Zhiyuan, et al.
Pubblicazione: (2025)
di: Fan, Zhiyuan, et al.
Pubblicazione: (2025)
Scoring Rules and Calibration for Imprecise Probabilities
di: Fröhlich, Christian, et al.
Pubblicazione: (2024)
di: Fröhlich, Christian, et al.
Pubblicazione: (2024)
The Fast Mixing Mechanism for Differential Privacy
di: Lev, Omri, et al.
Pubblicazione: (2026)
di: Lev, Omri, et al.
Pubblicazione: (2026)
The Gaussian Mixing Mechanism: Renyi Differential Privacy via Gaussian Sketches
di: Lev, Omri, et al.
Pubblicazione: (2025)
di: Lev, Omri, et al.
Pubblicazione: (2025)
Distributional Diffusion Models with Scoring Rules
di: De Bortoli, Valentin, et al.
Pubblicazione: (2025)
di: De Bortoli, Valentin, et al.
Pubblicazione: (2025)
Calibrated and Conformal Propensity Scores for Causal Effect Estimation
di: Deshpande, Shachi, et al.
Pubblicazione: (2023)
di: Deshpande, Shachi, et al.
Pubblicazione: (2023)
Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2026)
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2026)
Is One Score Enough? Rethinking the Evaluation of Sequentially Evolving LLM Memory
di: Dong, Songwei, et al.
Pubblicazione: (2026)
di: Dong, Songwei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Proper Scoring Rule for Virtual Staining
di: Tonks, Samuel, et al.
Pubblicazione: (2026) -
Aligning Evaluation with Clinical Priorities: Calibration, Label Shift, and Error Costs
di: Flores, Gerardo A., et al.
Pubblicazione: (2025) -
Evaluating Posterior Probabilities: Decision Theory, Proper Scoring Rules, and Calibration
di: Ferrer, Luciana, et al.
Pubblicazione: (2024) -
A Consequentialist Critique of Binary Classification Evaluation: Theory, Practice, and Tools
di: Flores, Gerardo, et al.
Pubblicazione: (2025) -
Quantifying Aleatoric and Epistemic Uncertainty with Proper Scoring Rules
di: Hofman, Paul, et al.
Pubblicazione: (2024)