Understanding and mitigating difficulties in posterior predictive evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Agrawal, Abhinav, Domke, Justin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Disentangling impact of capacity, objective, batchsize, estimators, and step-size on flow VI
por: Agrawal, Abhinav, et al.
Publicado: (2024)
por: Agrawal, Abhinav, et al.
Publicado: (2024)
Large Language Bayes
por: Domke, Justin
Publicado: (2025)
por: Domke, Justin
Publicado: (2025)
Model-Informed Flows for Bayesian Inference
por: Ko, Joohwan, et al.
Publicado: (2025)
por: Ko, Joohwan, et al.
Publicado: (2025)
Amortized Factor Inference Networks for Posterior Inference
por: Ko, Joohwan, et al.
Publicado: (2026)
por: Ko, Joohwan, et al.
Publicado: (2026)
Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects Models
por: Lai, Jinlin, et al.
Publicado: (2024)
por: Lai, Jinlin, et al.
Publicado: (2024)
Joint control variate for faster black-box variational inference
por: Wang, Xi, et al.
Publicado: (2022)
por: Wang, Xi, et al.
Publicado: (2022)
Simulation-based stacking
por: Yao, Yuling, et al.
Publicado: (2023)
por: Yao, Yuling, et al.
Publicado: (2023)
No evaluation without fair representation : Impact of label and selection bias on the evaluation, performance and mitigation of classification models
por: Legast, Magali, et al.
Publicado: (2026)
por: Legast, Magali, et al.
Publicado: (2026)
Predictive variational inference: Learn the predictively optimal posterior distribution
por: Lai, Jinlin, et al.
Publicado: (2024)
por: Lai, Jinlin, et al.
Publicado: (2024)
Prequential posteriors
por: Sinha-Roy, Shreya, et al.
Publicado: (2025)
por: Sinha-Roy, Shreya, et al.
Publicado: (2025)
Optimistic Q-learning for average reward and episodic reinforcement learning
por: Agrawal, Priyank, et al.
Publicado: (2024)
por: Agrawal, Priyank, et al.
Publicado: (2024)
Motor Vehicle Accident Severity Prediction Via Machine Learning
por: Chandolu, Abhinav
Publicado: (2025)
por: Chandolu, Abhinav
Publicado: (2025)
CRAUM-Net: Contextual Recursive Attention with Uncertainty Modeling for Salient Object Detection
por: Sagar, Abhinav
Publicado: (2020)
por: Sagar, Abhinav
Publicado: (2020)
Analysis of Long Range Dependency Understanding in State Space Models
por: Ravikumar, Srividya, et al.
Publicado: (2026)
por: Ravikumar, Srividya, et al.
Publicado: (2026)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
por: Joshi, Abhinav, et al.
Publicado: (2025)
por: Joshi, Abhinav, et al.
Publicado: (2025)
Spoken Language Understanding on Unseen Tasks With In-Context Learning
por: Agrawal, Neeraj, et al.
Publicado: (2025)
por: Agrawal, Neeraj, et al.
Publicado: (2025)
Achilles' Heel of Mamba: Essential difficulties of the Mamba architecture demonstrated by synthetic data
por: Chen, Tianyi, et al.
Publicado: (2025)
por: Chen, Tianyi, et al.
Publicado: (2025)
Towards Understanding Self-play for LLM Reasoning
por: Chae, Justin Yang, et al.
Publicado: (2025)
por: Chae, Justin Yang, et al.
Publicado: (2025)
Fairness of Deep Ensembles: On the interplay between per-group task difficulty and under-representation
por: Claucich, Estanislao, et al.
Publicado: (2025)
por: Claucich, Estanislao, et al.
Publicado: (2025)
Q-learning with Posterior Sampling
por: Agrawal, Priyank, et al.
Publicado: (2025)
por: Agrawal, Priyank, et al.
Publicado: (2025)
Scalable Bayesian Learning with posteriors
por: Duffield, Samuel, et al.
Publicado: (2024)
por: Duffield, Samuel, et al.
Publicado: (2024)
Quality analysis and evaluation prediction of RAG retrieval based on machine learning algorithms
por: Zhang, Ruoxin, et al.
Publicado: (2025)
por: Zhang, Ruoxin, et al.
Publicado: (2025)
Lazy vs hasty: linearization in deep networks impacts learning schedule based on example difficulty
por: George, Thomas, et al.
Publicado: (2022)
por: George, Thomas, et al.
Publicado: (2022)
Understanding Task Transfer in Vision-Language Models
por: Sachdeva, Bhuvan, et al.
Publicado: (2025)
por: Sachdeva, Bhuvan, et al.
Publicado: (2025)
Matching aggregate posteriors in the variational autoencoder
por: Saha, Surojit, et al.
Publicado: (2023)
por: Saha, Surojit, et al.
Publicado: (2023)
Learning to Understand: Identifying Interactions via the Möbius Transform
por: Kang, Justin S., et al.
Publicado: (2024)
por: Kang, Justin S., et al.
Publicado: (2024)
Understanding challenges to the interpretation of disaggregated evaluations of algorithmic fairness
por: Pfohl, Stephen R., et al.
Publicado: (2025)
por: Pfohl, Stephen R., et al.
Publicado: (2025)
Machine Learning Hamiltonian Dynamical Systems with Sparse and Noisy Data
por: Thapar, Vedanta, et al.
Publicado: (2026)
por: Thapar, Vedanta, et al.
Publicado: (2026)
New allometric models for the USA create a step-change in forest carbon estimation, modeling, and mapping
por: Johnson, Lucas K., et al.
Publicado: (2024)
por: Johnson, Lucas K., et al.
Publicado: (2024)
BiRating -- Iterative averaging on a bipartite graph of Beat Saber scores, player skills, and map difficulties
por: Casanova, Juan
Publicado: (2025)
por: Casanova, Juan
Publicado: (2025)
Performance evaluation of predictive AI models to support medical decisions: Overview and guidance
por: Van Calster, Ben, et al.
Publicado: (2024)
por: Van Calster, Ben, et al.
Publicado: (2024)
Variation in prediction accuracy due to randomness in data division and fair evaluation using interval estimation
por: Goto, Isao
Publicado: (2024)
por: Goto, Isao
Publicado: (2024)
BiMi Sheets: Infosheets for bias mitigation methods
por: Defrance, MaryBeth, et al.
Publicado: (2025)
por: Defrance, MaryBeth, et al.
Publicado: (2025)
Do regularization methods for shortcut mitigation work as intended?
por: Hong, Haoyang, et al.
Publicado: (2025)
por: Hong, Haoyang, et al.
Publicado: (2025)
Review of deep learning models for crypto price prediction: implementation and evaluation
por: Wu, Jingyang, et al.
Publicado: (2024)
por: Wu, Jingyang, et al.
Publicado: (2024)
Gems: Group Emotion Profiling Through Multimodal Situational Understanding
por: Kataria, Anubhav, et al.
Publicado: (2025)
por: Kataria, Anubhav, et al.
Publicado: (2025)
Estimating problem difficulty without ground truth using Large Language Model comparisons
por: Ballon, Marthe, et al.
Publicado: (2025)
por: Ballon, Marthe, et al.
Publicado: (2025)
Rethinking the generalization of drug target affinity prediction algorithms via similarity aware evaluation
por: Zhang, Chenbin, et al.
Publicado: (2025)
por: Zhang, Chenbin, et al.
Publicado: (2025)
Physics-informed neural network for predicting fatigue life of unirradiated and irradiated austenitic and ferritic/martensitic steels under reactor-relevant conditions
por: Kori, Dhiraj S, et al.
Publicado: (2025)
por: Kori, Dhiraj S, et al.
Publicado: (2025)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
por: Fernandes, Patrick, et al.
Publicado: (2025)
por: Fernandes, Patrick, et al.
Publicado: (2025)
Ejemplares similares
-
Disentangling impact of capacity, objective, batchsize, estimators, and step-size on flow VI
por: Agrawal, Abhinav, et al.
Publicado: (2024) -
Large Language Bayes
por: Domke, Justin
Publicado: (2025) -
Model-Informed Flows for Bayesian Inference
por: Ko, Joohwan, et al.
Publicado: (2025) -
Amortized Factor Inference Networks for Posterior Inference
por: Ko, Joohwan, et al.
Publicado: (2026) -
Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects Models
por: Lai, Jinlin, et al.
Publicado: (2024)