When does a predictor know its own loss?
Fuente:
arXiv
Guardado en:
| Autores principales: | Gollakota, Aravind, Gopalan, Parikshit, Karan, Aayush, Peale, Charlotte, Wieder, Udi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Flexible Routing via Uncertainty Decomposition
por: Peale, Charlotte, et al.
Publicado: (2026)
por: Peale, Charlotte, et al.
Publicado: (2026)
Provable Uncertainty Decomposition via Higher-Order Calibration
por: Ahdritz, Gustaf, et al.
Publicado: (2024)
por: Ahdritz, Gustaf, et al.
Publicado: (2024)
When unlearning is free: leveraging low influence points to reduce computational costs
por: Kleiman, Anat, et al.
Publicado: (2025)
por: Kleiman, Anat, et al.
Publicado: (2025)
Calibration through the Lens of Indistinguishability
por: Gopalan, Parikshit, et al.
Publicado: (2025)
por: Gopalan, Parikshit, et al.
Publicado: (2025)
Representative Language Generation
por: Peale, Charlotte, et al.
Publicado: (2025)
por: Peale, Charlotte, et al.
Publicado: (2025)
Multigroup Robustness
por: Hu, Lunjia, et al.
Publicado: (2024)
por: Hu, Lunjia, et al.
Publicado: (2024)
Taking a Moment for Distributional Robustness
por: Hastings, Jabari, et al.
Publicado: (2024)
por: Hastings, Jabari, et al.
Publicado: (2024)
Mistake-Bounded Language Generation
por: Kleinberg, Jon, et al.
Publicado: (2026)
por: Kleinberg, Jon, et al.
Publicado: (2026)
Trace Length is a Simple Uncertainty Signal in Reasoning Models
por: Devic, Siddartha, et al.
Publicado: (2025)
por: Devic, Siddartha, et al.
Publicado: (2025)
Blink of an eye: a simple theory for feature localization in generative models
por: Li, Marvin, et al.
Publicado: (2025)
por: Li, Marvin, et al.
Publicado: (2025)
Reasoning with Sampling: Your Base Model is Smarter Than You Think
por: Karan, Aayush, et al.
Publicado: (2025)
por: Karan, Aayush, et al.
Publicado: (2025)
How Global Calibration Strengthens Multiaccuracy
por: Casacuberta, Sílvia, et al.
Publicado: (2025)
por: Casacuberta, Sílvia, et al.
Publicado: (2025)
Swap Agnostic Learning, or Characterizing Omniprediction via Multicalibration
por: Gopalan, Parikshit, et al.
Publicado: (2023)
por: Gopalan, Parikshit, et al.
Publicado: (2023)
Efficient Calibration for Decision Making
por: Gopalan, Parikshit, et al.
Publicado: (2025)
por: Gopalan, Parikshit, et al.
Publicado: (2025)
The Importance of Being Smoothly Calibrated
por: Gopalan, Parikshit, et al.
Publicado: (2026)
por: Gopalan, Parikshit, et al.
Publicado: (2026)
On Computationally Efficient Multi-Class Calibration
por: Gopalan, Parikshit, et al.
Publicado: (2024)
por: Gopalan, Parikshit, et al.
Publicado: (2024)
When Scaling Fails: Network and Fabric Effects on Distributed GPU Training Performance
por: Gopalan, Dinesh, et al.
Publicado: (2026)
por: Gopalan, Dinesh, et al.
Publicado: (2026)
When does Subagging Work?
por: Revelas, Christos, et al.
Publicado: (2024)
por: Revelas, Christos, et al.
Publicado: (2024)
ReGuidance: A Simple Diffusion Wrapper for Boosting Sample Quality on Hard Inverse Problems
por: Karan, Aayush, et al.
Publicado: (2025)
por: Karan, Aayush, et al.
Publicado: (2025)
Omnipredictors for Regression and the Approximate Rank of Convex Functions
por: Gopalan, Parikshit, et al.
Publicado: (2024)
por: Gopalan, Parikshit, et al.
Publicado: (2024)
When does a bridge become an aeroplane?
por: Dardeno, Tina A., et al.
Publicado: (2024)
por: Dardeno, Tina A., et al.
Publicado: (2024)
Unrolled denoising networks provably learn optimal Bayesian inference
por: Karan, Aayush, et al.
Publicado: (2024)
por: Karan, Aayush, et al.
Publicado: (2024)
Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles
por: Wettenstein, Ron, et al.
Publicado: (2026)
por: Wettenstein, Ron, et al.
Publicado: (2026)
WOODELF-HD: Efficient Background SHAP for High-Depth Decision Trees
por: Wettenstein, Ron, et al.
Publicado: (2026)
por: Wettenstein, Ron, et al.
Publicado: (2026)
Context-Free Synthetic Data Mitigates Forgetting
por: Bansal, Parikshit, et al.
Publicado: (2025)
por: Bansal, Parikshit, et al.
Publicado: (2025)
TF-MLPNet: Tiny Real-Time Neural Speech Separation
por: Itani, Malek, et al.
Publicado: (2025)
por: Itani, Malek, et al.
Publicado: (2025)
When does Chain-of-Thought Help: A Markovian Perspective
por: Wang, Zihan, et al.
Publicado: (2026)
por: Wang, Zihan, et al.
Publicado: (2026)
Assessing the Impact of Upselling in Online Fantasy Sports
por: Chaudhary, Aayush
Publicado: (2024)
por: Chaudhary, Aayush
Publicado: (2024)
Simulation-based inference has its own Dodelson-Schneider effect (but it knows that it does)
por: Homer, Jed, et al.
Publicado: (2024)
por: Homer, Jed, et al.
Publicado: (2024)
Finding Clustering Algorithms in the Transformer Architecture
por: Clarkson, Kenneth L., et al.
Publicado: (2025)
por: Clarkson, Kenneth L., et al.
Publicado: (2025)
How does the optimizer implicitly bias the model merging loss landscape?
por: Zhang, Chenxiang, et al.
Publicado: (2025)
por: Zhang, Chenxiang, et al.
Publicado: (2025)
Enabling Approximate Joint Sampling in Diffusion LMs
por: Bansal, Parikshit, et al.
Publicado: (2025)
por: Bansal, Parikshit, et al.
Publicado: (2025)
When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning
por: Voelcker, Claas, et al.
Publicado: (2024)
por: Voelcker, Claas, et al.
Publicado: (2024)
Learning to Route LLMs with Confidence Tokens
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
Decoupled-Value Attention for Prior-Data Fitted Networks: GP Inference for Physical Equations
por: Sharma, Kaustubh, et al.
Publicado: (2025)
por: Sharma, Kaustubh, et al.
Publicado: (2025)
Understanding Self-Supervised Learning via Gaussian Mixture Models
por: Bansal, Parikshit, et al.
Publicado: (2024)
por: Bansal, Parikshit, et al.
Publicado: (2024)
Fine-grained Soundscape Control for Augmented Hearing
por: Oh, Seunghyun, et al.
Publicado: (2026)
por: Oh, Seunghyun, et al.
Publicado: (2026)
ODD: Overlap-aware Estimation of Model Performance under Distribution Shift
por: Mishra, Aayush, et al.
Publicado: (2025)
por: Mishra, Aayush, et al.
Publicado: (2025)
On the Stability of Iterative Retraining of Generative Models on their own Data
por: Bertrand, Quentin, et al.
Publicado: (2023)
por: Bertrand, Quentin, et al.
Publicado: (2023)
PharmacoMatch: Efficient 3D Pharmacophore Screening via Neural Subgraph Matching
por: Rose, Daniel, et al.
Publicado: (2024)
por: Rose, Daniel, et al.
Publicado: (2024)
Ejemplares similares
-
Flexible Routing via Uncertainty Decomposition
por: Peale, Charlotte, et al.
Publicado: (2026) -
Provable Uncertainty Decomposition via Higher-Order Calibration
por: Ahdritz, Gustaf, et al.
Publicado: (2024) -
When unlearning is free: leveraging low influence points to reduce computational costs
por: Kleiman, Anat, et al.
Publicado: (2025) -
Calibration through the Lens of Indistinguishability
por: Gopalan, Parikshit, et al.
Publicado: (2025) -
Representative Language Generation
por: Peale, Charlotte, et al.
Publicado: (2025)