Generalized Fitted Q-Iteration with Clustered Data
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Liyuan, Wang, Jitao, Wu, Zhenke, Shi, Chengchun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Doubly Inhomogeneous Reinforcement Learning
by: Hu, Liyuan, et al.
Published: (2022)
by: Hu, Liyuan, et al.
Published: (2022)
PyCFRL: A Python library for counterfactually fair offline reinforcement learning via sequential data preprocessing
by: Zhang, Jianhan, et al.
Published: (2025)
by: Zhang, Jianhan, et al.
Published: (2025)
Counterfactually Fair Reinforcement Learning via Sequential Data Preprocessing
by: Wang, Jitao, et al.
Published: (2025)
by: Wang, Jitao, et al.
Published: (2025)
Testing Stationarity and Change Point Detection in Reinforcement Learning
by: Li, Mengbing, et al.
Published: (2022)
by: Li, Mengbing, et al.
Published: (2022)
Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning
by: Manda, Kausthubh, et al.
Published: (2025)
by: Manda, Kausthubh, et al.
Published: (2025)
Statistical Inference in Reinforcement Learning: A Selective Survey
by: Shi, Chengchun
Published: (2025)
by: Shi, Chengchun
Published: (2025)
A Measure-Theoretic Finite-Sample Theory for Adaptive-Data Fitted Q-Iteration
by: Haussmann, Manuel, et al.
Published: (2026)
by: Haussmann, Manuel, et al.
Published: (2026)
Finite-Time Bounds for Average-Reward Fitted Q-Iteration
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Generating Synthetic Electronic Health Record Data: a Methodological Scoping Review with Benchmarking on Phenotype Data and Open-Source Software
by: Chen, Xingran, et al.
Published: (2024)
by: Chen, Xingran, et al.
Published: (2024)
Multivariate Dynamic Mediation Analysis under a Reinforcement Learning Framework
by: Luo, Lan, et al.
Published: (2023)
by: Luo, Lan, et al.
Published: (2023)
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
From Authors to Reviewers: Leveraging Rankings to Improve Peer Review
by: Wang, Weichen, et al.
Published: (2025)
by: Wang, Weichen, et al.
Published: (2025)
Counterfactually Safe Reinforcement Learning
by: Li, Jingyi, et al.
Published: (2026)
by: Li, Jingyi, et al.
Published: (2026)
Robust Fitted-Q-Evaluation and Iteration under Sequentially Exogenous Unobserved Confounders
by: Bruns-Smith, David, et al.
Published: (2023)
by: Bruns-Smith, David, et al.
Published: (2023)
Pessimistic Causal Reinforcement Learning with Mediators for Confounded Offline Data
by: Wang, Danyang, et al.
Published: (2024)
by: Wang, Danyang, et al.
Published: (2024)
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024)
by: Liu, Y., et al.
Published: (2024)
Off-policy Evaluation with Deeply-abstracted States
by: Hao, Meiling, et al.
Published: (2024)
by: Hao, Meiling, et al.
Published: (2024)
A Unified Framework for Inference with General Missingness Patterns and Machine Learning Imputation
by: Chen, Xingran, et al.
Published: (2025)
by: Chen, Xingran, et al.
Published: (2025)
Reinforcement Learning from Human Feedback: A Statistical Perspective
by: Liu, Pangpang, et al.
Published: (2026)
by: Liu, Pangpang, et al.
Published: (2026)
Dual Active Learning for Reinforcement Learning from Human Feedback
by: Liu, Pangpang, et al.
Published: (2024)
by: Liu, Pangpang, et al.
Published: (2024)
Double Fairness Policy Learning: Integrating Action Fairness and Outcome Fairness in Decision-making
by: Bian, Zeyu, et al.
Published: (2026)
by: Bian, Zeyu, et al.
Published: (2026)
TabClustPFN: A Prior-Fitted Network for Tabular Data Clustering
by: Zhao, Tianqi, et al.
Published: (2026)
by: Zhao, Tianqi, et al.
Published: (2026)
Detecting LLM-Generated Text with Performance Guarantees
by: Zhou, Hongyi, et al.
Published: (2026)
by: Zhou, Hongyi, et al.
Published: (2026)
One Prompt Fits All: Universal Graph Adaptation for Pretrained Models
by: Huang, Yongqi, et al.
Published: (2025)
by: Huang, Yongqi, et al.
Published: (2025)
Off-policy Evaluation in Doubly Inhomogeneous Environments
by: Bian, Zeyu, et al.
Published: (2023)
by: Bian, Zeyu, et al.
Published: (2023)
Exponential Convergence Guarantees for Iterative Markovian Fitting
by: Silveri, Marta Gentiloni, et al.
Published: (2025)
by: Silveri, Marta Gentiloni, et al.
Published: (2025)
Spatially Robust Inference with Predicted and Missing at Random Labels
by: Salerno, Stephen, et al.
Published: (2026)
by: Salerno, Stephen, et al.
Published: (2026)
Clustering by Attention: Leveraging Prior Fitted Transformers for Data Partitioning
by: Shokry, Ahmed, et al.
Published: (2025)
by: Shokry, Ahmed, et al.
Published: (2025)
Exponential convergence rate for Iterative Markovian Fitting
by: Sokolov, Kirill, et al.
Published: (2025)
by: Sokolov, Kirill, et al.
Published: (2025)
Diffusion & Adversarial Schrödinger Bridges via Iterative Proportional Markovian Fitting
by: Kholkin, Sergei, et al.
Published: (2024)
by: Kholkin, Sergei, et al.
Published: (2024)
MENGLAN: Multiscale Enhanced Nonparametric Gas Analyzer with Lightweight Architecture and Networks
by: Duan, Zhenke, et al.
Published: (2025)
by: Duan, Zhenke, et al.
Published: (2025)
Q-Learning with Clustered-SMART (cSMART) Data: Examining Moderators in the Construction of Clustered Adaptive Interventions
by: Song, Yao, et al.
Published: (2025)
by: Song, Yao, et al.
Published: (2025)
A Two-armed Bandit Framework for A/B Testing
by: Wang, Jinjuan, et al.
Published: (2025)
by: Wang, Jinjuan, et al.
Published: (2025)
Combining Experimental and Historical Data for Policy Evaluation
by: Li, Ting, et al.
Published: (2024)
by: Li, Ting, et al.
Published: (2024)
Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems
by: Hu, Liyuan
Published: (2025)
by: Hu, Liyuan
Published: (2025)
Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Perturbation is All You Need for Extrapolating Language Models
by: Cen, Zetai, et al.
Published: (2026)
by: Cen, Zetai, et al.
Published: (2026)
Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments
by: Wen, Qianglin, et al.
Published: (2024)
by: Wen, Qianglin, et al.
Published: (2024)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026)
by: Jeong, Narim, et al.
Published: (2026)
Designing Time Series Experiments in A/B Testing with Transformer Reinforcement Learning
by: Wu, Xiangkun, et al.
Published: (2026)
by: Wu, Xiangkun, et al.
Published: (2026)
Similar Items
-
Doubly Inhomogeneous Reinforcement Learning
by: Hu, Liyuan, et al.
Published: (2022) -
PyCFRL: A Python library for counterfactually fair offline reinforcement learning via sequential data preprocessing
by: Zhang, Jianhan, et al.
Published: (2025) -
Counterfactually Fair Reinforcement Learning via Sequential Data Preprocessing
by: Wang, Jitao, et al.
Published: (2025) -
Testing Stationarity and Change Point Detection in Reinforcement Learning
by: Li, Mengbing, et al.
Published: (2022) -
Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning
by: Manda, Kausthubh, et al.
Published: (2025)