The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Victoria, Yun, Taedong, Matarić, Maja, Canny, John, Gretton, Arthur, D'Amour, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Choosing a Proxy Metric from Past Experiments
von: Tripuraneni, Nilesh, et al.
Veröffentlicht: (2023)
von: Tripuraneni, Nilesh, et al.
Veröffentlicht: (2023)
Copula-based Sensitivity Analysis for Multi-Treatment Causal Inference with Unobserved Confounding
von: Zheng, Jiajing, et al.
Veröffentlicht: (2021)
von: Zheng, Jiajing, et al.
Veröffentlicht: (2021)
Interventional Processes for Causal Uncertainty Quantification
von: Dance, Hugh, et al.
Veröffentlicht: (2024)
von: Dance, Hugh, et al.
Veröffentlicht: (2024)
Deconfounding Scores and Representation Learning for Causal Effect Estimation with Weak Overlap
von: Clivio, Oscar, et al.
Veröffentlicht: (2026)
von: Clivio, Oscar, et al.
Veröffentlicht: (2026)
The Leaderboard Illusion
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
Sleepless Nights, Sugary Days: Creating Synthetic Users with Health Conditions for Realistic Coaching Agent Interactions
von: Yun, Taedong, et al.
Veröffentlicht: (2025)
von: Yun, Taedong, et al.
Veröffentlicht: (2025)
Kernel Treatment Effects with Adaptively Collected Data
von: Zenati, Houssam, et al.
Veröffentlicht: (2025)
von: Zenati, Houssam, et al.
Veröffentlicht: (2025)
Sequential Kernel Embedding for Mediated and Time-Varying Dose Response Curves
von: Singh, Rahul, et al.
Veröffentlicht: (2021)
von: Singh, Rahul, et al.
Veröffentlicht: (2021)
Efficient Inference after Directionally Stable Adaptive Experiments
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
Conversational Planning for Personal Plans
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2025)
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2025)
Composite Goodness-of-fit Tests with Kernels
von: Key, Oscar, et al.
Veröffentlicht: (2021)
von: Key, Oscar, et al.
Veröffentlicht: (2021)
Optimizing Language Models for Human Preferences is a Causal Inference Problem
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
Omitted Variable Bias in Language Models Under Distribution Shift
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
Demystifying Spectral Feature Learning for Instrumental Variable Regression
von: Meunier, Dimitri, et al.
Veröffentlicht: (2025)
von: Meunier, Dimitri, et al.
Veröffentlicht: (2025)
The Hardness of Validating Observational Studies with Experimental Data
von: Fawkes, Jake, et al.
Veröffentlicht: (2025)
von: Fawkes, Jake, et al.
Veröffentlicht: (2025)
On the Hardness of Conditional Independence Testing In Practice
von: He, Zheng, et al.
Veröffentlicht: (2025)
von: He, Zheng, et al.
Veröffentlicht: (2025)
Benchmarking Observational Studies with Experimental Data under Right-Censoring
von: Demirel, Ilker, et al.
Veröffentlicht: (2024)
von: Demirel, Ilker, et al.
Veröffentlicht: (2024)
LIDS: LLM Summary Inference Under the Layered Lens
von: Park, Dylan, et al.
Veröffentlicht: (2026)
von: Park, Dylan, et al.
Veröffentlicht: (2026)
Mind the Graph When Balancing Data for Fairness or Robustness
von: Schrouff, Jessica, et al.
Veröffentlicht: (2024)
von: Schrouff, Jessica, et al.
Veröffentlicht: (2024)
Evaluating Interventional Reasoning Capabilities of Large Language Models
von: Kasetty, Tejas, et al.
Veröffentlicht: (2024)
von: Kasetty, Tejas, et al.
Veröffentlicht: (2024)
Agents Thinking Fast and Slow: A Talker-Reasoner Architecture
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2024)
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2024)
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
von: Dobariya, Om, et al.
Veröffentlicht: (2025)
von: Dobariya, Om, et al.
Veröffentlicht: (2025)
Probabilistic Factorial Experimental Design for Combinatorial Interventions
von: Shyamal, Divya, et al.
Veröffentlicht: (2025)
von: Shyamal, Divya, et al.
Veröffentlicht: (2025)
Theoretical guarantees on the best-of-n alignment policy
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
Expected Reward Prediction, with Applications to Model Routing
von: Hasanaliyev, Kenan, et al.
Veröffentlicht: (2026)
von: Hasanaliyev, Kenan, et al.
Veröffentlicht: (2026)
Proxy Methods for Domain Adaptation
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
Feature Matching Intervention: Leveraging Observational Data for Causal Representation Learning
von: Li, Haoze, et al.
Veröffentlicht: (2025)
von: Li, Haoze, et al.
Veröffentlicht: (2025)
MPO: An Efficient Post-Processing Framework for Mixing Diverse Preference Alignment
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
Trust Your $\nabla$: Gradient-based Intervention Targeting for Causal Discovery
von: Olko, Mateusz, et al.
Veröffentlicht: (2022)
von: Olko, Mateusz, et al.
Veröffentlicht: (2022)
Simulation-Based Inference for Adaptive Experiments
von: Cho, Brian M, et al.
Veröffentlicht: (2025)
von: Cho, Brian M, et al.
Veröffentlicht: (2025)
ALCM: Autonomous LLM-Augmented Causal Discovery Framework
von: Khatibi, Elahe, et al.
Veröffentlicht: (2024)
von: Khatibi, Elahe, et al.
Veröffentlicht: (2024)
Cross-Validated Causal Inference: a Modern Method to Combine Experimental and Observational Data
von: Yang, Xuelin, et al.
Veröffentlicht: (2025)
von: Yang, Xuelin, et al.
Veröffentlicht: (2025)
A Design-based Solution for Causal Inference with Text: Can a Language Model Be Too Large?
von: Tierney, Graham, et al.
Veröffentlicht: (2025)
von: Tierney, Graham, et al.
Veröffentlicht: (2025)
Isolated Causal Effects of Natural Language
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
The Illusion of Learning from Observational Data: An Empirical Bayes Perspective
von: Wu, Bohan, et al.
Veröffentlicht: (2026)
von: Wu, Bohan, et al.
Veröffentlicht: (2026)
Simulation-Based Sensitivity Analysis in Optimal Treatment Regimes and Causal Decomposition with Individualized Interventions
von: Park, Soojin, et al.
Veröffentlicht: (2025)
von: Park, Soojin, et al.
Veröffentlicht: (2025)
Bias in Language Models: Beyond Trick Tests and Toward RUTEd Evaluation
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
Uncovering Bias Mechanisms in Observational Studies
von: Demirel, Ilker, et al.
Veröffentlicht: (2025)
von: Demirel, Ilker, et al.
Veröffentlicht: (2025)
A Double Machine Learning Approach to Combining Experimental and Observational Data
von: Parikh, Harsh, et al.
Veröffentlicht: (2023)
von: Parikh, Harsh, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Choosing a Proxy Metric from Past Experiments
von: Tripuraneni, Nilesh, et al.
Veröffentlicht: (2023) -
Copula-based Sensitivity Analysis for Multi-Treatment Causal Inference with Unobserved Confounding
von: Zheng, Jiajing, et al.
Veröffentlicht: (2021) -
Interventional Processes for Causal Uncertainty Quantification
von: Dance, Hugh, et al.
Veröffentlicht: (2024) -
Deconfounding Scores and Representation Learning for Causal Effect Estimation with Weak Overlap
von: Clivio, Oscar, et al.
Veröffentlicht: (2026) -
The Leaderboard Illusion
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)