Causal methods for LLM development and evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Frauen, Dennis, Brockschmidt, Marie, Hess, Konstantin, Ma, Haorui, Ma, Yuchen, Maarouf, Abdurahman, Schröder, Maresa, Schweisthal, Jonas, Wang, Yuxin, Deviyani, Athiya, Parbhoo, Sonali, Krishnan, Rahul G., Feuerriegel, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Conformal Prediction for Causal Effects of Continuous Treatments
von: Schröder, Maresa, et al.
Veröffentlicht: (2024)
von: Schröder, Maresa, et al.
Veröffentlicht: (2024)
Constructing Confidence Intervals for Average Treatment Effects from Multiple Datasets
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
von: Wang, Yuxin, et al.
Veröffentlicht: (2024)
Learning Representations of Instruments for Partial Identification of Treatment Effects
von: Schweisthal, Jonas, et al.
Veröffentlicht: (2024)
von: Schweisthal, Jonas, et al.
Veröffentlicht: (2024)
Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks
von: Javurek, Emil, et al.
Veröffentlicht: (2026)
von: Javurek, Emil, et al.
Veröffentlicht: (2026)
Assessing the robustness of heterogeneous treatment effects in survival analysis under informative censoring
von: Wang, Yuxin, et al.
Veröffentlicht: (2025)
von: Wang, Yuxin, et al.
Veröffentlicht: (2025)
SurvDiff: A Diffusion Model for Generating Synthetic Data in Survival Analysis
von: Brockschmidt, Marie, et al.
Veröffentlicht: (2025)
von: Brockschmidt, Marie, et al.
Veröffentlicht: (2025)
Orthogonal Survival Learners for Estimating Heterogeneous Treatment Effects from Time-to-Event Data
von: Frauen, Dennis, et al.
Veröffentlicht: (2025)
von: Frauen, Dennis, et al.
Veröffentlicht: (2025)
Adaptive Experimentation for Censored Survival Outcomes
von: Wang, Yuxin, et al.
Veröffentlicht: (2026)
von: Wang, Yuxin, et al.
Veröffentlicht: (2026)
Causal Fairness under Unobserved Confounding: A Neural Sensitivity Framework
von: Schröder, Maresa, et al.
Veröffentlicht: (2023)
von: Schröder, Maresa, et al.
Veröffentlicht: (2023)
Nonparametric LLM Evaluation from Preference Data
von: Frauen, Dennis, et al.
Veröffentlicht: (2026)
von: Frauen, Dennis, et al.
Veröffentlicht: (2026)
LLM-Driven Treatment Effect Estimation Under Inference Time Text Confounding
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
Orthogonal Representation Learning for Estimating Causal Quantities
von: Melnychuk, Valentyn, et al.
Veröffentlicht: (2025)
von: Melnychuk, Valentyn, et al.
Veröffentlicht: (2025)
An Orthogonal Learner for Individualized Outcomes in Markov Decision Processes
von: Javurek, Emil, et al.
Veröffentlicht: (2025)
von: Javurek, Emil, et al.
Veröffentlicht: (2025)
Contextual Metric Meta-Evaluation by Measuring Local Metric Accuracy
von: Deviyani, Athiya, et al.
Veröffentlicht: (2025)
von: Deviyani, Athiya, et al.
Veröffentlicht: (2025)
DeepBlip: Estimating Conditional Average Treatment Effects Over Time
von: Ma, Haorui, et al.
Veröffentlicht: (2025)
von: Ma, Haorui, et al.
Veröffentlicht: (2025)
Targeted Synthetic Control Method
von: Wang, Yuxin, et al.
Veröffentlicht: (2026)
von: Wang, Yuxin, et al.
Veröffentlicht: (2026)
Overlap-Adaptive Regularization for Conditional Average Treatment Effect Estimation
von: Melnychuk, Valentyn, et al.
Veröffentlicht: (2025)
von: Melnychuk, Valentyn, et al.
Veröffentlicht: (2025)
Model-agnostic meta-learners for estimating heterogeneous treatment effects over time
von: Frauen, Dennis, et al.
Veröffentlicht: (2024)
von: Frauen, Dennis, et al.
Veröffentlicht: (2024)
Predicting Startup Success Using Large Language Models: A Novel In-Context Learning Approach
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2026)
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2026)
The Virality of Hate Speech on Social Media
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2022)
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2022)
A Fused Large Language Model for Predicting Startup Success
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2024)
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2024)
Analyzing User Characteristics of Hate Speech Spreaders on Social Media
von: Geissler, Dominique, et al.
Veröffentlicht: (2023)
von: Geissler, Dominique, et al.
Veröffentlicht: (2023)
Generative AI may backfire for counterspeech
von: Bär, Dominik, et al.
Veröffentlicht: (2024)
von: Bär, Dominik, et al.
Veröffentlicht: (2024)
ConfoundingSHAP: Quantifying confounding strength in causal inference
von: Brockschmidt, Marie, et al.
Veröffentlicht: (2026)
von: Brockschmidt, Marie, et al.
Veröffentlicht: (2026)
A Diffusion-Based Method for Learning the Multi-Outcome Distribution of Medical Treatments
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
DiffPO: A causal diffusion model for learning distributions of potential outcomes
von: Ma, Yuchen, et al.
Veröffentlicht: (2024)
von: Ma, Yuchen, et al.
Veröffentlicht: (2024)
Orthogonal Learner for Estimating Heterogeneous Long-Term Treatment Effects
von: Ma, Haorui, et al.
Veröffentlicht: (2026)
von: Ma, Haorui, et al.
Veröffentlicht: (2026)
Foundation Models for Causal Inference via Prior-Data Fitted Networks
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
von: Ma, Yuchen, et al.
Veröffentlicht: (2025)
Meta-Learners for Partially-Identified Treatment Effects Across Multiple Environments
von: Schweisthal, Jonas, et al.
Veröffentlicht: (2024)
von: Schweisthal, Jonas, et al.
Veröffentlicht: (2024)
Causal machine learning for predicting treatment outcomes
von: Feuerriegel, Stefan, et al.
Veröffentlicht: (2024)
von: Feuerriegel, Stefan, et al.
Veröffentlicht: (2024)
Efficient and Sharp Off-Policy Learning under Unobserved Confounding
von: Hess, Konstantin, et al.
Veröffentlicht: (2025)
von: Hess, Konstantin, et al.
Veröffentlicht: (2025)
Debiased neural operators for estimating functionals
von: Hess, Konstantin, et al.
Veröffentlicht: (2026)
von: Hess, Konstantin, et al.
Veröffentlicht: (2026)
Bayesian Neural Controlled Differential Equations for Treatment Effect Estimation
von: Hess, Konstantin, et al.
Veröffentlicht: (2023)
von: Hess, Konstantin, et al.
Veröffentlicht: (2023)
IGC-Net for conditional average potential outcome estimation over time
von: Hess, Konstantin, et al.
Veröffentlicht: (2024)
von: Hess, Konstantin, et al.
Veröffentlicht: (2024)
HQP: A Human-Annotated Dataset for Detecting Online Propaganda
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2023)
von: Maarouf, Abdurahman, et al.
Veröffentlicht: (2023)
Beyond Means: A Dynamic Framework for Predicting Customer Satisfaction
von: Naumzik, Christof, et al.
Veröffentlicht: (2025)
von: Naumzik, Christof, et al.
Veröffentlicht: (2025)
Causal Inference on Networks under Misspecified Exposure Mappings: A Partial Identification Framework
von: Schröder, Maresa, et al.
Veröffentlicht: (2026)
von: Schröder, Maresa, et al.
Veröffentlicht: (2026)
Generalized Bayes for Causal Inference
von: Javurek, Emil, et al.
Veröffentlicht: (2026)
von: Javurek, Emil, et al.
Veröffentlicht: (2026)
Treatment Effect Estimation for Optimal Decision-Making
von: Frauen, Dennis, et al.
Veröffentlicht: (2025)
von: Frauen, Dennis, et al.
Veröffentlicht: (2025)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Conformal Prediction for Causal Effects of Continuous Treatments
von: Schröder, Maresa, et al.
Veröffentlicht: (2024) -
Constructing Confidence Intervals for Average Treatment Effects from Multiple Datasets
von: Wang, Yuxin, et al.
Veröffentlicht: (2024) -
Learning Representations of Instruments for Partial Identification of Treatment Effects
von: Schweisthal, Jonas, et al.
Veröffentlicht: (2024) -
Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks
von: Javurek, Emil, et al.
Veröffentlicht: (2026) -
Assessing the robustness of heterogeneous treatment effects in survival analysis under informative censoring
von: Wang, Yuxin, et al.
Veröffentlicht: (2025)