Evaluating Model Performance Under Worst-case Subpopulations
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Mike, Mittal, Daksh, Namkoong, Hongseok, Xia, Shangzhou |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data-Driven Stochastic Modeling Using Autoregressive Sequence Models: Translating Event Tables to Queueing Dynamics
by: Mittal, Daksh, et al.
Published: (2025)
by: Mittal, Daksh, et al.
Published: (2025)
A Planning Framework for Adaptive Labeling
by: Mittal, Daksh, et al.
Published: (2025)
by: Mittal, Daksh, et al.
Published: (2025)
Architectural and Inferential Inductive Biases For Exchangeable Sequence Modeling
by: Mittal, Daksh, et al.
Published: (2025)
by: Mittal, Daksh, et al.
Published: (2025)
Empirical Likelihood for Nonsmooth Functionals
by: Namkoong, Hongseok
Published: (2026)
by: Namkoong, Hongseok
Published: (2026)
Exchangeable Sequence Models Quantify Uncertainty Over Latent Concepts
by: Ye, Naimeng, et al.
Published: (2024)
by: Ye, Naimeng, et al.
Published: (2024)
Learning With Multi-Group Guarantees For Clusterable Subpopulations
by: Dai, Jessica, et al.
Published: (2024)
by: Dai, Jessica, et al.
Published: (2024)
A Sensitivity Approach to Causal Inference Under Limited Overlap
by: Ma, Yuanzhe, et al.
Published: (2025)
by: Ma, Yuanzhe, et al.
Published: (2025)
Minimax Optimal Estimation of Stability Under Distribution Shift
by: Namkoong, Hongseok, et al.
Published: (2022)
by: Namkoong, Hongseok, et al.
Published: (2022)
SynthTools: A Framework for Scaling Synthetic Tools for Agent Development
by: Castellani, Tommaso, et al.
Published: (2025)
by: Castellani, Tommaso, et al.
Published: (2025)
The Value of Prediction in Identifying the Worst-Off
by: Fischer-Abaigar, Unai, et al.
Published: (2025)
by: Fischer-Abaigar, Unai, et al.
Published: (2025)
Assessing Generalization for Subpopulation Representative Modeling via In-Context Learning
by: Simmons, Gabriel, et al.
Published: (2024)
by: Simmons, Gabriel, et al.
Published: (2024)
Revisiting (Un)Fairness in Recourse by Minimizing Worst-Case Social Burden
by: Barrainkua, Ainhize, et al.
Published: (2025)
by: Barrainkua, Ainhize, et al.
Published: (2025)
Adaptive Elicitation of Latent Information Using Natural Language
by: Wang, Jimmy, et al.
Published: (2025)
by: Wang, Jimmy, et al.
Published: (2025)
A Broader View of Thompson Sampling
by: Qu, Yanlin, et al.
Published: (2025)
by: Qu, Yanlin, et al.
Published: (2025)
Design and Scheduling of an AI-based Queueing System
by: Lee, Jiung, et al.
Published: (2024)
by: Lee, Jiung, et al.
Published: (2024)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
by: Namkoong, Hongseok, et al.
Published: (2020)
by: Namkoong, Hongseok, et al.
Published: (2020)
LLM Generated Persona is a Promise with a Catch
by: Li, Ang, et al.
Published: (2025)
by: Li, Ang, et al.
Published: (2025)
Predictive Performance Comparison of Decision Policies Under Confounding
by: Guerdan, Luke, et al.
Published: (2024)
by: Guerdan, Luke, et al.
Published: (2024)
Differentiable Discrete Event Simulation for Queuing Network Control
by: Che, Ethan, et al.
Published: (2024)
by: Che, Ethan, et al.
Published: (2024)
Rethinking Distribution Shifts: Empirical Analysis and Inductive Modeling for Tabular Data
by: Wang, Tianyu, et al.
Published: (2023)
by: Wang, Tianyu, et al.
Published: (2023)
Benchmarking In-context Experiential Learning Through Repeated Product Recommendations
by: Yang, Gilbert, et al.
Published: (2025)
by: Yang, Gilbert, et al.
Published: (2025)
Evaluation of Machine Learning Models in Student Academic Performance Prediction
by: Sandeepa, A. G. R., et al.
Published: (2025)
by: Sandeepa, A. G. R., et al.
Published: (2025)
The New Frontier of Cybersecurity: Emerging Threats and Innovations
by: Dave, Daksh, et al.
Published: (2023)
by: Dave, Daksh, et al.
Published: (2023)
GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks
by: Patwardhan, Tejal, et al.
Published: (2025)
by: Patwardhan, Tejal, et al.
Published: (2025)
Evaluating Algorithmic Bias in Models for Predicting Academic Performance of Filipino Students
by: Švábenský, Valdemar, et al.
Published: (2024)
by: Švábenský, Valdemar, et al.
Published: (2024)
Forecasting India's Demographic Transition Under Fertility Policy Scenarios Using hybrid LSTM-PINN Model
by: Khanra, Subarna, et al.
Published: (2025)
by: Khanra, Subarna, et al.
Published: (2025)
Agent Based Simulators for Epidemic Modelling: Simulating Larger Models Using Smaller Ones
by: Mittal, Daksh, et al.
Published: (2022)
by: Mittal, Daksh, et al.
Published: (2022)
Optimal Decision Making Under Strategic Behavior
by: Tsirtsis, Stratis, et al.
Published: (2019)
by: Tsirtsis, Stratis, et al.
Published: (2019)
A Bayesian Spatial Model to Correct Under-Reporting in Urban Crowdsourcing
by: Agostini, Gabriel, et al.
Published: (2023)
by: Agostini, Gabriel, et al.
Published: (2023)
Health Insurance Coverage Rule Interpretation Corpus: Law, Policy, and Medical Guidance for Health Insurance Coverage Understanding
by: Gartner, Mike
Published: (2025)
by: Gartner, Mike
Published: (2025)
Bias Amplification Enhances Minority Group Performance
by: Li, Gaotang, et al.
Published: (2023)
by: Li, Gaotang, et al.
Published: (2023)
The Impact of Differential Feature Under-reporting on Algorithmic Fairness
by: Akpinar, Nil-Jana, et al.
Published: (2024)
by: Akpinar, Nil-Jana, et al.
Published: (2024)
Optimization-Driven Adaptive Experimentation
by: Che, Ethan, et al.
Published: (2024)
by: Che, Ethan, et al.
Published: (2024)
AExGym: Benchmarks and Environments for Adaptive Experimentation
by: Wang, Jimmy, et al.
Published: (2024)
by: Wang, Jimmy, et al.
Published: (2024)
C-Learner: Constrained Learning for Causal Inference
by: Cai, Tiffany Tianhui, et al.
Published: (2024)
by: Cai, Tiffany Tianhui, et al.
Published: (2024)
Towards Modeling Learner Performance with Large Language Models
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Learning Progression-Guided AI Evaluation of Scientific Models To Support Diverse Multi-Modal Understanding in NGSS Classroom
by: Kaldaras, Leonora, et al.
Published: (2025)
by: Kaldaras, Leonora, et al.
Published: (2025)
Mobility-GCN: a human mobility-based graph convolutional network for tracking and analyzing the spatial dynamics of the synthetic opioid crisis in the USA, 2013-2020
by: Xia, Zhiyue, et al.
Published: (2024)
by: Xia, Zhiyue, et al.
Published: (2024)
Evaluating Fairness in Self-supervised and Supervised Models for Sequential Data
by: Yfantidou, Sofia, et al.
Published: (2024)
by: Yfantidou, Sofia, et al.
Published: (2024)
Counterfactual Fairness Evaluation of Machine Learning Models on Educational Datasets
by: Kim, Woojin, et al.
Published: (2025)
by: Kim, Woojin, et al.
Published: (2025)
Similar Items
-
Data-Driven Stochastic Modeling Using Autoregressive Sequence Models: Translating Event Tables to Queueing Dynamics
by: Mittal, Daksh, et al.
Published: (2025) -
A Planning Framework for Adaptive Labeling
by: Mittal, Daksh, et al.
Published: (2025) -
Architectural and Inferential Inductive Biases For Exchangeable Sequence Modeling
by: Mittal, Daksh, et al.
Published: (2025) -
Empirical Likelihood for Nonsmooth Functionals
by: Namkoong, Hongseok
Published: (2026) -
Exchangeable Sequence Models Quantify Uncertainty Over Latent Concepts
by: Ye, Naimeng, et al.
Published: (2024)