Robust Reinforcement Learning for Shifting Dynamics During Deployment
Fuente:
Zenodo
Saved in:
| Main Authors: | Stanton, Samuel, Fakoor, Rasool, Mueller, Jonas, Wilson, Andrew Gordon, Smola, Alex |
|---|---|
| Format: | Recurso digital |
| Published: |
Zenodo
2021
|
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time-Varying Propensity Score to Bridge the Gap between the Past and Present
by: Fakoor, Rasool, et al.
Published: (2022)
by: Fakoor, Rasool, et al.
Published: (2022)
Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training
by: Fakoor, Rasool, et al.
Published: (2026)
by: Fakoor, Rasool, et al.
Published: (2026)
Budgeting Counterfactual for Offline RL
by: Liu, Yao, et al.
Published: (2023)
by: Liu, Yao, et al.
Published: (2023)
AlphaRouter: Quantum Circuit Routing with Reinforcement Learning and Tree Search
by: Tang, Wei, et al.
Published: (2024)
by: Tang, Wei, et al.
Published: (2024)
The Pokémon Theorem and other Fairness Impossibility Results
by: Smola, Daniel Matsui, et al.
Published: (2026)
by: Smola, Daniel Matsui, et al.
Published: (2026)
Learning the Target Network in Function Space
by: Asadi, Kavosh, et al.
Published: (2024)
by: Asadi, Kavosh, et al.
Published: (2024)
Offline Learning and Forgetting for Reasoning with Large Language Models
by: Ni, Tianwei, et al.
Published: (2025)
by: Ni, Tianwei, et al.
Published: (2025)
Humanoid robot path planning with fuzzy Markov decision processes
by: Mahdi Fakoor
Published: (2016)
by: Mahdi Fakoor
Published: (2016)
Mind the GAP: Improving Robustness to Subpopulation Shifts with Group-Aware Priors
by: Rudner, Tim G. J., et al.
Published: (2024)
by: Rudner, Tim G. J., et al.
Published: (2024)
TAIL: Task-specific Adapters for Imitation Learning with Large Pretrained Models
by: Liu, Zuxin, et al.
Published: (2023)
by: Liu, Zuxin, et al.
Published: (2023)
Romantic Relationships and Psychological Well‐Being During the Transition to College
by: Ava Trimble, et al.
Published: (2025)
by: Ava Trimble, et al.
Published: (2025)
Deep Learning is Not So Mysterious or Different
by: Wilson, Andrew Gordon
Published: (2025)
by: Wilson, Andrew Gordon
Published: (2025)
From Demonstrations to Rewards: Alignment Without Explicit Human Preferences
by: Zeng, Siliang, et al.
Published: (2025)
by: Zeng, Siliang, et al.
Published: (2025)
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
by: Zhang, Jesse, et al.
Published: (2024)
by: Zhang, Jesse, et al.
Published: (2024)
Submodular Benchmark Selection
by: Smola, Alexander
Published: (2026)
by: Smola, Alexander
Published: (2026)
Formen und Funktionen der Intertextualitaet im Prosawerk von Anton Čechov
by: Smola, Klavdia
Published: (2019)
by: Smola, Klavdia
Published: (2019)
Buscando a Nemo=Finding Nemo [VHS - video] / / Andrew Stanton
by: Stanton Andrew
by: Stanton Andrew
Automated Data Curation for Robust Language Model Fine-Tuning
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
Dual-Robust Cross-Domain Offline Reinforcement Learning Against Dynamics Shifts
by: Qiao, Zhongjian, et al.
Published: (2025)
by: Qiao, Zhongjian, et al.
Published: (2025)
Bridging the Training-Inference Gap in LLMs by Leveraging Self-Generated Tokens
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
by: Yang, Ke, et al.
Published: (2024)
by: Yang, Ke, et al.
Published: (2024)
Support-Conditioned Flow Matching Is Kernel Smoothing
by: Smola, Daniel Matsui
Published: (2026)
by: Smola, Daniel Matsui
Published: (2026)
Acción e institución en el pensamiento político de Hannah Arendt: lecturas de Sobre la Revolución
by: Julia Gabriela Smola
Published: (2017)
by: Julia Gabriela Smola
Published: (2017)
Controllable Prompt Tuning For Balancing Group Distributional Robustness
by: Phan, Hoang, et al.
Published: (2024)
by: Phan, Hoang, et al.
Published: (2024)
Decoding the Gender Gap: Addressing Gender Stereotypes and Psychological Barriers to Empower Women in Technology
by: Harehdasht, Zahra Fakoor, et al.
Published: (2025)
by: Harehdasht, Zahra Fakoor, et al.
Published: (2025)
ProactBench: Beyond What The User Asked For
by: Harfi, Sepehr, et al.
Published: (2026)
by: Harfi, Sepehr, et al.
Published: (2026)
Learning Shifts of Clinicians Who Become Clinician‐Coaches: An Exploratory Qualitative Study of Emergency Physicians
by: Andrew Rixon, et al.
Published: (2025)
by: Andrew Rixon, et al.
Published: (2025)
Microservice Deployment in Space Computing Power Networks via Robust Reinforcement Learning
by: Yu, Zhiyong, et al.
Published: (2025)
by: Yu, Zhiyong, et al.
Published: (2025)
Shifting Work Patterns with Generative AI
by: Dillon, Eleanor Wiske, et al.
Published: (2025)
by: Dillon, Eleanor Wiske, et al.
Published: (2025)
ProxFly: Robust Control for Close Proximity Quadcopter Flight via Residual Reinforcement Learning
by: Zhang, Ruiqi, et al.
Published: (2024)
by: Zhang, Ruiqi, et al.
Published: (2024)
Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning
by: Oh, Donggeon David, et al.
Published: (2026)
by: Oh, Donggeon David, et al.
Published: (2026)
The significance of feeling needed and useful to family and friends for psychological well‐being during adolescence
by: Andrew J. Fuligni, et al.
Published: (2024)
by: Andrew J. Fuligni, et al.
Published: (2024)
L3Ms -- Lagrange Large Language Models
by: Dhillon, Guneet S., et al.
Published: (2024)
by: Dhillon, Guneet S., et al.
Published: (2024)
Mapping Farmed Landscapes from Remote Sensing
by: Conserva, Michelangelo, et al.
Published: (2025)
by: Conserva, Michelangelo, et al.
Published: (2025)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction
by: Durkin, Alex, et al.
Published: (2025)
by: Durkin, Alex, et al.
Published: (2025)
ActiveLab: Active Learning with Re-Labeling by Multiple Annotators
by: Goh, Hui Wen, et al.
Published: (2023)
by: Goh, Hui Wen, et al.
Published: (2023)
Brain-Inspired Continual Learning-Robust Feature Distillation and Re-Consolidation for Class Incremental Learning
by: Khan, Hikmat, et al.
Published: (2024)
by: Khan, Hikmat, et al.
Published: (2024)
Interactive Analysis of Static, Dynamic, and Crystalline SDTrimSP Simulations: Application to Nitrogen Ion Implantation into Vanadium
by: Lebeda, Miroslav, et al.
Published: (2026)
by: Lebeda, Miroslav, et al.
Published: (2026)
Model-Based Data-Efficient and Robust Reinforcement Learning
by: Svedlund, Ludvig, et al.
Published: (2026)
by: Svedlund, Ludvig, et al.
Published: (2026)
Similar Items
-
Time-Varying Propensity Score to Bridge the Gap between the Past and Present
by: Fakoor, Rasool, et al.
Published: (2022) -
Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training
by: Fakoor, Rasool, et al.
Published: (2026) -
Budgeting Counterfactual for Offline RL
by: Liu, Yao, et al.
Published: (2023) -
AlphaRouter: Quantum Circuit Routing with Reinforcement Learning and Tree Search
by: Tang, Wei, et al.
Published: (2024) -
The Pokémon Theorem and other Fairness Impossibility Results
by: Smola, Daniel Matsui, et al.
Published: (2026)