Epidemiology of Large Language Models: A Benchmark for Observational Distribution Knowledge
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Plecko, Drago, Okanovic, Patrik, Havaldar, Shreyas, Hoefler, Torsten, Bareinboim, Elias |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Confounder Detection via Treatment Intent: A New Observational Study Design
par: Plecko, Drago, et autres
Publié: (2026)
par: Plecko, Drago, et autres
Publié: (2026)
Fairness-Accuracy Trade-Offs: A Causal Perspective
par: Plecko, Drago, et autres
Publié: (2024)
par: Plecko, Drago, et autres
Publié: (2024)
Mind the Gap: A Causal Perspective on Bias Amplification in Prediction & Decision-Making
par: Plecko, Drago, et autres
Publié: (2024)
par: Plecko, Drago, et autres
Publié: (2024)
Causal Algorithmic Recourse: Foundations and Methods
par: Plecko, Drago, et autres
Publié: (2026)
par: Plecko, Drago, et autres
Publié: (2026)
Causal Bias Detection in Generative Artificial Intelligence
par: Plecko, Drago
Publié: (2026)
par: Plecko, Drago
Publié: (2026)
Causal Fairness for Survival Analysis
par: Plecko, Drago
Publié: (2026)
par: Plecko, Drago
Publié: (2026)
Interaction Testing in Variation Analysis
par: Plecko, Drago
Publié: (2024)
par: Plecko, Drago
Publié: (2024)
EntryPrune: Neural Network Feature Selection using First Impressions
par: Zimmer, Felix, et autres
Publié: (2024)
par: Zimmer, Felix, et autres
Publié: (2024)
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
par: Wu, Boqian, et autres
Publié: (2026)
par: Wu, Boqian, et autres
Publié: (2026)
Active Model Selection for Large Language Models
par: Durmazkeser, Yavuz, et autres
Publié: (2025)
par: Durmazkeser, Yavuz, et autres
Publié: (2025)
Large Language Model Selection with Limited Annotations
par: Durmazkeser, Yavuz, et autres
Publié: (2026)
par: Durmazkeser, Yavuz, et autres
Publié: (2026)
An Algorithmic Approach for Causal Health Equity: A Look at Race Differentials in Intensive Care Unit (ICU) Outcomes
par: Plecko, Drago, et autres
Publié: (2025)
par: Plecko, Drago, et autres
Publié: (2025)
Neural Causal Abstractions
par: Xia, Kevin, et autres
Publié: (2024)
par: Xia, Kevin, et autres
Publié: (2024)
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
par: Xiao, Qiao, et autres
Publié: (2026)
par: Xiao, Qiao, et autres
Publié: (2026)
Less Greedy Equivalence Search
par: Ejaz, Adiba, et autres
Publié: (2025)
par: Ejaz, Adiba, et autres
Publié: (2025)
From Black-box to Causal-box: Towards Building More Interpretable Models
par: Hwang, Inwoo, et autres
Publié: (2025)
par: Hwang, Inwoo, et autres
Publié: (2025)
Automatic Reward Shaping from Confounded Offline Data
par: Li, Mingxuan, et autres
Publié: (2025)
par: Li, Mingxuan, et autres
Publié: (2025)
Partial Transportability for Domain Generalization
par: Jalaldoust, Kasra, et autres
Publié: (2025)
par: Jalaldoust, Kasra, et autres
Publié: (2025)
Causally Aligned Curriculum Learning
par: Li, Mingxuan, et autres
Publié: (2025)
par: Li, Mingxuan, et autres
Publié: (2025)
Fairness under Covariate Shift: Improving Fairness-Accuracy tradeoff with few Unlabeled Test Samples
par: Havaldar, Shreyas, et autres
Publié: (2023)
par: Havaldar, Shreyas, et autres
Publié: (2023)
Learning from Label Proportions: Bootstrapping Supervised Learners via Belief Propagation
par: Havaldar, Shreyas, et autres
Publié: (2023)
par: Havaldar, Shreyas, et autres
Publié: (2023)
Confounding Robust Continuous Control via Automatic Reward Shaping
par: Juliani, Mateo, et autres
Publié: (2026)
par: Juliani, Mateo, et autres
Publié: (2026)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
par: Li, Mingxuan, et autres
Publié: (2026)
par: Li, Mingxuan, et autres
Publié: (2026)
Testing Causal Models with Hidden Variables in Polynomial Delay via Conditional Independencies
par: Jeong, Hyunchai, et autres
Publié: (2024)
par: Jeong, Hyunchai, et autres
Publié: (2024)
Counterfactual Realizability
par: Raghavan, Arvind, et autres
Publié: (2025)
par: Raghavan, Arvind, et autres
Publié: (2025)
Causal Identification from Counterfactual Data: Completeness and Bounding Results
par: Raghavan, Arvind, et autres
Publié: (2026)
par: Raghavan, Arvind, et autres
Publié: (2026)
Partial Identification Approach to Counterfactual Fairness Assessment
par: Rho, Saeyoung, et autres
Publié: (2025)
par: Rho, Saeyoung, et autres
Publié: (2025)
All models are wrong, some are useful: Model Selection with Limited Labels
par: Okanovic, Patrik, et autres
Publié: (2024)
par: Okanovic, Patrik, et autres
Publié: (2024)
Chameleon: a Heterogeneous and Disaggregated Accelerator System for Retrieval-Augmented Language Models
par: Jiang, Wenqi, et autres
Publié: (2023)
par: Jiang, Wenqi, et autres
Publié: (2023)
Graph of Thoughts: Solving Elaborate Problems with Large Language Models
par: Besta, Maciej, et autres
Publié: (2023)
par: Besta, Maciej, et autres
Publié: (2023)
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models
par: Jia, Yifan, et autres
Publié: (2025)
par: Jia, Yifan, et autres
Publié: (2025)
STARLING: Self-supervised Training of Text-based Reinforcement Learning Agent with Large Language Models
par: Basavatia, Shreyas, et autres
Publié: (2024)
par: Basavatia, Shreyas, et autres
Publié: (2024)
Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge Prompting
par: Luo, Miaosen, et autres
Publié: (2025)
par: Luo, Miaosen, et autres
Publié: (2025)
RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models
par: Jin, Zhuoran, et autres
Publié: (2024)
par: Jin, Zhuoran, et autres
Publié: (2024)
AECBench: A Hierarchical Benchmark for Knowledge Evaluation of Large Language Models in the AEC Field
par: Liang, Chen, et autres
Publié: (2025)
par: Liang, Chen, et autres
Publié: (2025)
Model-Distributed Inference for Large Language Models at the Edge
par: Macario, Davide, et autres
Publié: (2025)
par: Macario, Davide, et autres
Publié: (2025)
Distributed Interpretability and Control for Large Language Models
par: Desai, Dev Arpan, et autres
Publié: (2026)
par: Desai, Dev Arpan, et autres
Publié: (2026)
Medical Interpretability and Knowledge Maps of Large Language Models
par: Marinescu, Razvan, et autres
Publié: (2025)
par: Marinescu, Razvan, et autres
Publié: (2025)
Latent Knowledge Scalpel: Precise and Massive Knowledge Editing for Large Language Models
par: Liu, Xin, et autres
Publié: (2025)
par: Liu, Xin, et autres
Publié: (2025)
SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning
par: Limozin, Alexis, et autres
Publié: (2026)
par: Limozin, Alexis, et autres
Publié: (2026)
Documents similaires
-
Confounder Detection via Treatment Intent: A New Observational Study Design
par: Plecko, Drago, et autres
Publié: (2026) -
Fairness-Accuracy Trade-Offs: A Causal Perspective
par: Plecko, Drago, et autres
Publié: (2024) -
Mind the Gap: A Causal Perspective on Bias Amplification in Prediction & Decision-Making
par: Plecko, Drago, et autres
Publié: (2024) -
Causal Algorithmic Recourse: Foundations and Methods
par: Plecko, Drago, et autres
Publié: (2026) -
Causal Bias Detection in Generative Artificial Intelligence
par: Plecko, Drago
Publié: (2026)