Understanding and mitigating difficulties in posterior predictive evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Agrawal, Abhinav, Domke, Justin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Disentangling impact of capacity, objective, batchsize, estimators, and step-size on flow VI
by: Agrawal, Abhinav, et al.
Published: (2024)
by: Agrawal, Abhinav, et al.
Published: (2024)
Large Language Bayes
by: Domke, Justin
Published: (2025)
by: Domke, Justin
Published: (2025)
Model-Informed Flows for Bayesian Inference
by: Ko, Joohwan, et al.
Published: (2025)
by: Ko, Joohwan, et al.
Published: (2025)
Amortized Factor Inference Networks for Posterior Inference
by: Ko, Joohwan, et al.
Published: (2026)
by: Ko, Joohwan, et al.
Published: (2026)
Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects Models
by: Lai, Jinlin, et al.
Published: (2024)
by: Lai, Jinlin, et al.
Published: (2024)
Joint control variate for faster black-box variational inference
by: Wang, Xi, et al.
Published: (2022)
by: Wang, Xi, et al.
Published: (2022)
Simulation-based stacking
by: Yao, Yuling, et al.
Published: (2023)
by: Yao, Yuling, et al.
Published: (2023)
No evaluation without fair representation : Impact of label and selection bias on the evaluation, performance and mitigation of classification models
by: Legast, Magali, et al.
Published: (2026)
by: Legast, Magali, et al.
Published: (2026)
Predictive variational inference: Learn the predictively optimal posterior distribution
by: Lai, Jinlin, et al.
Published: (2024)
by: Lai, Jinlin, et al.
Published: (2024)
Prequential posteriors
by: Sinha-Roy, Shreya, et al.
Published: (2025)
by: Sinha-Roy, Shreya, et al.
Published: (2025)
Optimistic Q-learning for average reward and episodic reinforcement learning
by: Agrawal, Priyank, et al.
Published: (2024)
by: Agrawal, Priyank, et al.
Published: (2024)
Motor Vehicle Accident Severity Prediction Via Machine Learning
by: Chandolu, Abhinav
Published: (2025)
by: Chandolu, Abhinav
Published: (2025)
CRAUM-Net: Contextual Recursive Attention with Uncertainty Modeling for Salient Object Detection
by: Sagar, Abhinav
Published: (2020)
by: Sagar, Abhinav
Published: (2020)
Analysis of Long Range Dependency Understanding in State Space Models
by: Ravikumar, Srividya, et al.
Published: (2026)
by: Ravikumar, Srividya, et al.
Published: (2026)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
by: Joshi, Abhinav, et al.
Published: (2025)
by: Joshi, Abhinav, et al.
Published: (2025)
Spoken Language Understanding on Unseen Tasks With In-Context Learning
by: Agrawal, Neeraj, et al.
Published: (2025)
by: Agrawal, Neeraj, et al.
Published: (2025)
Achilles' Heel of Mamba: Essential difficulties of the Mamba architecture demonstrated by synthetic data
by: Chen, Tianyi, et al.
Published: (2025)
by: Chen, Tianyi, et al.
Published: (2025)
Towards Understanding Self-play for LLM Reasoning
by: Chae, Justin Yang, et al.
Published: (2025)
by: Chae, Justin Yang, et al.
Published: (2025)
Fairness of Deep Ensembles: On the interplay between per-group task difficulty and under-representation
by: Claucich, Estanislao, et al.
Published: (2025)
by: Claucich, Estanislao, et al.
Published: (2025)
Q-learning with Posterior Sampling
by: Agrawal, Priyank, et al.
Published: (2025)
by: Agrawal, Priyank, et al.
Published: (2025)
Scalable Bayesian Learning with posteriors
by: Duffield, Samuel, et al.
Published: (2024)
by: Duffield, Samuel, et al.
Published: (2024)
Quality analysis and evaluation prediction of RAG retrieval based on machine learning algorithms
by: Zhang, Ruoxin, et al.
Published: (2025)
by: Zhang, Ruoxin, et al.
Published: (2025)
Lazy vs hasty: linearization in deep networks impacts learning schedule based on example difficulty
by: George, Thomas, et al.
Published: (2022)
by: George, Thomas, et al.
Published: (2022)
Understanding Task Transfer in Vision-Language Models
by: Sachdeva, Bhuvan, et al.
Published: (2025)
by: Sachdeva, Bhuvan, et al.
Published: (2025)
Matching aggregate posteriors in the variational autoencoder
by: Saha, Surojit, et al.
Published: (2023)
by: Saha, Surojit, et al.
Published: (2023)
Learning to Understand: Identifying Interactions via the Möbius Transform
by: Kang, Justin S., et al.
Published: (2024)
by: Kang, Justin S., et al.
Published: (2024)
Understanding challenges to the interpretation of disaggregated evaluations of algorithmic fairness
by: Pfohl, Stephen R., et al.
Published: (2025)
by: Pfohl, Stephen R., et al.
Published: (2025)
Machine Learning Hamiltonian Dynamical Systems with Sparse and Noisy Data
by: Thapar, Vedanta, et al.
Published: (2026)
by: Thapar, Vedanta, et al.
Published: (2026)
New allometric models for the USA create a step-change in forest carbon estimation, modeling, and mapping
by: Johnson, Lucas K., et al.
Published: (2024)
by: Johnson, Lucas K., et al.
Published: (2024)
BiRating -- Iterative averaging on a bipartite graph of Beat Saber scores, player skills, and map difficulties
by: Casanova, Juan
Published: (2025)
by: Casanova, Juan
Published: (2025)
Performance evaluation of predictive AI models to support medical decisions: Overview and guidance
by: Van Calster, Ben, et al.
Published: (2024)
by: Van Calster, Ben, et al.
Published: (2024)
Variation in prediction accuracy due to randomness in data division and fair evaluation using interval estimation
by: Goto, Isao
Published: (2024)
by: Goto, Isao
Published: (2024)
BiMi Sheets: Infosheets for bias mitigation methods
by: Defrance, MaryBeth, et al.
Published: (2025)
by: Defrance, MaryBeth, et al.
Published: (2025)
Do regularization methods for shortcut mitigation work as intended?
by: Hong, Haoyang, et al.
Published: (2025)
by: Hong, Haoyang, et al.
Published: (2025)
Review of deep learning models for crypto price prediction: implementation and evaluation
by: Wu, Jingyang, et al.
Published: (2024)
by: Wu, Jingyang, et al.
Published: (2024)
Gems: Group Emotion Profiling Through Multimodal Situational Understanding
by: Kataria, Anubhav, et al.
Published: (2025)
by: Kataria, Anubhav, et al.
Published: (2025)
Estimating problem difficulty without ground truth using Large Language Model comparisons
by: Ballon, Marthe, et al.
Published: (2025)
by: Ballon, Marthe, et al.
Published: (2025)
Rethinking the generalization of drug target affinity prediction algorithms via similarity aware evaluation
by: Zhang, Chenbin, et al.
Published: (2025)
by: Zhang, Chenbin, et al.
Published: (2025)
Physics-informed neural network for predicting fatigue life of unirradiated and irradiated austenitic and ferritic/martensitic steels under reactor-relevant conditions
by: Kori, Dhiraj S, et al.
Published: (2025)
by: Kori, Dhiraj S, et al.
Published: (2025)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
by: Fernandes, Patrick, et al.
Published: (2025)
by: Fernandes, Patrick, et al.
Published: (2025)
Similar Items
-
Disentangling impact of capacity, objective, batchsize, estimators, and step-size on flow VI
by: Agrawal, Abhinav, et al.
Published: (2024) -
Large Language Bayes
by: Domke, Justin
Published: (2025) -
Model-Informed Flows for Bayesian Inference
by: Ko, Joohwan, et al.
Published: (2025) -
Amortized Factor Inference Networks for Posterior Inference
by: Ko, Joohwan, et al.
Published: (2026) -
Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects Models
by: Lai, Jinlin, et al.
Published: (2024)