Time After Time: Deep-Q Effect Estimation for Interventions on When and What to do
Fuente:
arXiv
Saved in:
| Main Authors: | Wald, Yoav, Goldstein, Mark, Efroni, Yonathan, van Amsterdam, Wouter A. C., Ranganath, Rajesh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gradient Free Deep Reinforcement Learning With TabPFN
by: Schiff, David, et al.
Published: (2025)
by: Schiff, David, et al.
Published: (2025)
Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale
by: Roth, Amit, et al.
Published: (2026)
by: Roth, Amit, et al.
Published: (2026)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited Modalities
by: Saporta, Adriel, et al.
Published: (2024)
by: Saporta, Adriel, et al.
Published: (2024)
When accurate prediction models yield harmful self-fulfilling prophecies
by: van Amsterdam, Wouter A. C., et al.
Published: (2023)
by: van Amsterdam, Wouter A. C., et al.
Published: (2023)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
by: Wu, Lili, et al.
Published: (2024)
by: Wu, Lili, et al.
Published: (2024)
What's the score? Automated Denoising Score Matching for Nonlinear Diffusions
by: Singhal, Raghav, et al.
Published: (2024)
by: Singhal, Raghav, et al.
Published: (2024)
Explanations that reveal all through the definition of encoding
by: Puli, Aahlad, et al.
Published: (2024)
by: Puli, Aahlad, et al.
Published: (2024)
Active Slice Discovery in Large Language Models
by: Zhang, Minhui, et al.
Published: (2025)
by: Zhang, Minhui, et al.
Published: (2025)
Estimating Tail Risks in Language Model Output Distributions
by: Angell, Rico, et al.
Published: (2026)
by: Angell, Rico, et al.
Published: (2026)
Open-Set Domain Adaptation Under Background Distribution Shift: Challenges and A Provably Efficient Solution
by: Chaudhari, Shravan, et al.
Published: (2025)
by: Chaudhari, Shravan, et al.
Published: (2025)
Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers
by: Ranganath, Aditya
Published: (2026)
by: Ranganath, Aditya
Published: (2026)
Black Box Causal Inference: Effect Estimation via Meta Prediction
by: Bynum, Lucius E. J., et al.
Published: (2025)
by: Bynum, Lucius E. J., et al.
Published: (2025)
Imbalanced Gradients in RL Post-Training of Multi-Task LLMs
by: Wu, Runzhe, et al.
Published: (2025)
by: Wu, Runzhe, et al.
Published: (2025)
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions
by: Shaik, Thanveer, et al.
Published: (2023)
by: Shaik, Thanveer, et al.
Published: (2023)
A causal viewpoint on prediction model performance under changes in case-mix: discrimination and calibration respond differently for prognosis and diagnosis predictions
by: van Amsterdam, Wouter A. C.
Published: (2024)
by: van Amsterdam, Wouter A. C.
Published: (2024)
Nuisances via Negativa: Adjusting for Spurious Correlations via Data Augmentation
by: Puli, Aahlad, et al.
Published: (2022)
by: Puli, Aahlad, et al.
Published: (2022)
Reasoning Stabilization Point: A Training-Time Signal for Stable Evidence and Shortcut Reliance
by: Dhayalkar, Sahil Rajesh
Published: (2026)
by: Dhayalkar, Sahil Rajesh
Published: (2026)
From algorithms to action: improving patient care requires causality
by: van Amsterdam, Wouter A. C., et al.
Published: (2022)
by: van Amsterdam, Wouter A. C., et al.
Published: (2022)
What-If Explanations Over Time: Counterfactuals for Time Series Classification
by: Schlegel, Udo, et al.
Published: (2026)
by: Schlegel, Udo, et al.
Published: (2026)
Deep Double Q-learning
by: Nagarajan, Prabhat, et al.
Published: (2025)
by: Nagarajan, Prabhat, et al.
Published: (2025)
Deep Variational Contrastive Learning for Joint Risk Stratification and Time-to-Event Estimation
by: Erbil, Pinar, et al.
Published: (2026)
by: Erbil, Pinar, et al.
Published: (2026)
Simple Optimizers for Convex Aligned Multi-Objective Optimization
by: Kretzu, Ben, et al.
Published: (2025)
by: Kretzu, Ben, et al.
Published: (2025)
Weighted Risk Invariance: Domain Generalization under Invariant Feature Shift
by: Wong, Gina, et al.
Published: (2024)
by: Wong, Gina, et al.
Published: (2024)
Intervention-Based Time Series Causal Discovery via Simulator-Generated Interventional Distributions
by: Okita, Tsuyoshi
Published: (2026)
by: Okita, Tsuyoshi
Published: (2026)
StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
by: Ranganath, Suraj, et al.
Published: (2026)
by: Ranganath, Suraj, et al.
Published: (2026)
KV Cache Quantization for Self-Forcing Video Generation: A 33-Method Empirical Study
by: Ranganath, Suraj, et al.
Published: (2026)
by: Ranganath, Suraj, et al.
Published: (2026)
Test-Time Learning of Causal Structure from Interventional Data
by: Chen, Wei, et al.
Published: (2026)
by: Chen, Wei, et al.
Published: (2026)
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning
by: van der Vaart, Pascal R., et al.
Published: (2025)
by: van der Vaart, Pascal R., et al.
Published: (2025)
Q-function Decomposition with Intervention Semantics with Factored Action Spaces
by: Lee, Junkyu, et al.
Published: (2025)
by: Lee, Junkyu, et al.
Published: (2025)
When LLM Meets Time Series: Can LLMs Perform Multi-Step Time Series Reasoning and Inference
by: Ye, Wen, et al.
Published: (2025)
by: Ye, Wen, et al.
Published: (2025)
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
by: Jayawardhana, Mayuka, et al.
Published: (2026)
by: Jayawardhana, Mayuka, et al.
Published: (2026)
Estimating Direct and Indirect Causal Effects of Spatiotemporal Interventions in Presence of Spatial Interference
by: Ali, Sahara, et al.
Published: (2024)
by: Ali, Sahara, et al.
Published: (2024)
PcLast: Discovering Plannable Continuous Latent States
by: Koul, Anurag, et al.
Published: (2023)
by: Koul, Anurag, et al.
Published: (2023)
Interactive Double Deep Q-network: Integrating Human Interventions and Evaluative Predictions in Reinforcement Learning of Autonomous Driving
by: Sygkounas, Alkis, et al.
Published: (2025)
by: Sygkounas, Alkis, et al.
Published: (2025)
New-Onset Diabetes Assessment Using Artificial Intelligence-Enhanced Electrocardiography
by: Zhang, Hao, et al.
Published: (2022)
by: Zhang, Hao, et al.
Published: (2022)
When Will It Fail?: Anomaly to Prompt for Forecasting Future Anomalies in Time Series
by: Park, Min-Yeong, et al.
Published: (2025)
by: Park, Min-Yeong, et al.
Published: (2025)
The Forecast After the Forecast: A Post-Processing Shift in Time Series
by: Liang, Daojun, et al.
Published: (2026)
by: Liang, Daojun, et al.
Published: (2026)
Finding the DeepDream for Time Series: Activation Maximization for Univariate Time Series
by: Schlegel, Udo, et al.
Published: (2024)
by: Schlegel, Udo, et al.
Published: (2024)
Deep Learning for Time Series Anomaly Detection: A Survey
by: Darban, Zahra Zamanzadeh, et al.
Published: (2022)
by: Darban, Zahra Zamanzadeh, et al.
Published: (2022)
Similar Items
-
Gradient Free Deep Reinforcement Learning With TabPFN
by: Schiff, David, et al.
Published: (2025) -
Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale
by: Roth, Amit, et al.
Published: (2026) -
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024) -
Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited Modalities
by: Saporta, Adriel, et al.
Published: (2024) -
When accurate prediction models yield harmful self-fulfilling prophecies
by: van Amsterdam, Wouter A. C., et al.
Published: (2023)