Guardado en:
| Autores principales: | Ghosh, Susobhan, Guo, Yongyi, Hung, Pei-Yao, Coughlin, Lara, Bonar, Erin, Nahum-Shani, Inbal, Walton, Maureen, Murphy, Susan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2408.15076 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
"It felt more real": Investigating the User Experience of the MiWaves Personalizing JITAI Pilot Study
por: Ghosh, Susobhan, et al.
Publicado: (2025)
por: Ghosh, Susobhan, et al.
Publicado: (2025)
reBandit: Random Effects based Online RL algorithm for Reducing Cannabis Use
por: Ghosh, Susobhan, et al.
Publicado: (2024)
por: Ghosh, Susobhan, et al.
Publicado: (2024)
Effective Monitoring of Online Decision-Making Algorithms in Digital Intervention Implementation
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Reinforcement Learning on Dyads to Enhance Medication Adherence
por: Xu, Ziping, et al.
Publicado: (2025)
por: Xu, Ziping, et al.
Publicado: (2025)
A Deployed Online Reinforcement Learning Algorithm In An Oral Health Clinical Trial
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Oralytics Reinforcement Learning Algorithm
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Dyadic Reinforcement Learning
por: Li, Shuangning, et al.
Publicado: (2023)
por: Li, Shuangning, et al.
Publicado: (2023)
Constructing Evidence-Based Tailoring Variables for Adaptive Interventions
por: Dziak, John J., et al.
Publicado: (2025)
por: Dziak, John J., et al.
Publicado: (2025)
Evaluating time-varying treatment effects in hybrid SMART-MRT designs
por: Li, Mengbing, et al.
Publicado: (2026)
por: Li, Mengbing, et al.
Publicado: (2026)
Optimal Adaptive SMART Designs with Binary Outcomes
por: Ghosh, Rik, et al.
Publicado: (2023)
por: Ghosh, Rik, et al.
Publicado: (2023)
Reproducible workflow for online AI in digital health
por: Ghosh, Susobhan, et al.
Publicado: (2025)
por: Ghosh, Susobhan, et al.
Publicado: (2025)
Online learning in bandits with predicted context
por: Guo, Yongyi, et al.
Publicado: (2023)
por: Guo, Yongyi, et al.
Publicado: (2023)
Data integration methods for micro-randomized trials
por: Huch, Easton, et al.
Publicado: (2024)
por: Huch, Easton, et al.
Publicado: (2024)
Reinforcement Learning: An Overview
por: Murphy, Kevin
Publicado: (2024)
por: Murphy, Kevin
Publicado: (2024)
Statistical Inference for Misspecified Contextual Bandits
por: Guo, Yongyi, et al.
Publicado: (2025)
por: Guo, Yongyi, et al.
Publicado: (2025)
Heuristics for Partially Observable Stochastic Contingent Planning
por: Shani, Guy
Publicado: (2024)
por: Shani, Guy
Publicado: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
por: Nofshin, Eura, et al.
Publicado: (2024)
por: Nofshin, Eura, et al.
Publicado: (2024)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
por: Biswas, Arpita, et al.
Publicado: (2023)
por: Biswas, Arpita, et al.
Publicado: (2023)
Statistical Reinforcement Learning in the Real World: A Survey of Challenges and Future Directions
por: Gazi, Asim H., et al.
Publicado: (2026)
por: Gazi, Asim H., et al.
Publicado: (2026)
RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm
por: Yang, Yongyi, et al.
Publicado: (2025)
por: Yang, Yongyi, et al.
Publicado: (2025)
Shared Control with Black Box Agents using Oracle Queries
por: Avraham, Inbal, et al.
Publicado: (2024)
por: Avraham, Inbal, et al.
Publicado: (2024)
Missing Data Multiple Imputation for Tabular Q-Learning in Online RL
por: Chasalow, Kyla, et al.
Publicado: (2025)
por: Chasalow, Kyla, et al.
Publicado: (2025)
Paladin-mini: A Compact and Efficient Grounding Model Excelling in Real-World Scenarios
por: Ivry, Dror, et al.
Publicado: (2025)
por: Ivry, Dror, et al.
Publicado: (2025)
A latent variable approach to jointly modeling longitudinal and cumulative event data using a weighted two‐stage method
por: Madeline R. Abbott, et al.
Publicado: (2024)
por: Madeline R. Abbott, et al.
Publicado: (2024)
When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models
por: Zheng, Haoran, et al.
Publicado: (2026)
por: Zheng, Haoran, et al.
Publicado: (2026)
Towards Optimizing Human-Centric Objectives in AI-Assisted Decision-Making With Offline Reinforcement Learning
por: Buçinca, Zana, et al.
Publicado: (2024)
por: Buçinca, Zana, et al.
Publicado: (2024)
Sentinel: SOTA model to protect against prompt injections
por: Ivry, Dror, et al.
Publicado: (2025)
por: Ivry, Dror, et al.
Publicado: (2025)
A Continuous-Time Dynamic Factor Model for Intensive Longitudinal Data Arising from Mobile Health Studies
por: Abbott, Madeline R., et al.
Publicado: (2023)
por: Abbott, Madeline R., et al.
Publicado: (2023)
Respiratory Manifestations in Pediatric Patients With Interleukin 6 Receptor Deficiency
por: Noga Arwas, et al.
Publicado: (2025)
por: Noga Arwas, et al.
Publicado: (2025)
Constructing Non-Markovian Decision Process via History Aggregator
por: Wang, Yongyi, et al.
Publicado: (2025)
por: Wang, Yongyi, et al.
Publicado: (2025)
Towards Fault Tolerance in Multi-Agent Reinforcement Learning
por: Shi, Yuchen, et al.
Publicado: (2024)
por: Shi, Yuchen, et al.
Publicado: (2024)
Aligning Human and Machine Attention for Enhanced Supervised Learning
por: Chriqui, Avihay, et al.
Publicado: (2025)
por: Chriqui, Avihay, et al.
Publicado: (2025)
MiMIC: Multi-Modal Indian Earnings Calls Dataset to Predict Stock Prices
por: Ghosh, Sohom, et al.
Publicado: (2025)
por: Ghosh, Sohom, et al.
Publicado: (2025)
Hybrid Cross-domain Robust Reinforcement Learning
por: Van, Linh Le Pham, et al.
Publicado: (2025)
por: Van, Linh Le Pham, et al.
Publicado: (2025)
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
por: Shen, Jingyan, et al.
Publicado: (2025)
por: Shen, Jingyan, et al.
Publicado: (2025)
A New Error Temporal Difference Algorithm for Deep Reinforcement Learning in Microgrid Optimization
por: Yao, Fulong, et al.
Publicado: (2025)
por: Yao, Fulong, et al.
Publicado: (2025)
mHC-lite: You Don't Need 20 Sinkhorn-Knopp Iterations
por: Yang, Yongyi, et al.
Publicado: (2026)
por: Yang, Yongyi, et al.
Publicado: (2026)
The Interpretation Gap in Text-to-Music Generation Models
por: Zang, Yongyi, et al.
Publicado: (2024)
por: Zang, Yongyi, et al.
Publicado: (2024)
Identification and Localization of Cometary Activity in Solar System Objects with Machine Learning
por: Bolin, Bryce T., et al.
Publicado: (2024)
por: Bolin, Bryce T., et al.
Publicado: (2024)
Ejemplares similares
-
"It felt more real": Investigating the User Experience of the MiWaves Personalizing JITAI Pilot Study
por: Ghosh, Susobhan, et al.
Publicado: (2025) -
reBandit: Random Effects based Online RL algorithm for Reducing Cannabis Use
por: Ghosh, Susobhan, et al.
Publicado: (2024) -
Effective Monitoring of Online Decision-Making Algorithms in Digital Intervention Implementation
por: Trella, Anna L., et al.
Publicado: (2024) -
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
por: Trella, Anna L., et al.
Publicado: (2024) -
Reinforcement Learning on Dyads to Enhance Medication Adherence
por: Xu, Ziping, et al.
Publicado: (2025)