Scalable Option Learning in High-Throughput Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Henaff, Mikael, Fujimoto, Scott, Matthews, Michael, Rabbat, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Behaviour Spaces
by: Matthews, Michael Tryfan, et al.
Published: (2026)
by: Matthews, Michael Tryfan, et al.
Published: (2026)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026)
by: Matthews, Michael, et al.
Published: (2026)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
Polar Sparsity: High Throughput Batched LLM Inferencing with Scalable Contextual Sparsity
by: Shrestha, Susav, et al.
Published: (2025)
by: Shrestha, Susav, et al.
Published: (2025)
Positive Unlabeled Contrastive Learning
by: Acharya, Anish, et al.
Published: (2022)
by: Acharya, Anish, et al.
Published: (2022)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
by: Zheng, Qinqing, et al.
Published: (2024)
by: Zheng, Qinqing, et al.
Published: (2024)
OptionZero: Planning with Learned Options
by: Huang, Po-Wei, et al.
Published: (2025)
by: Huang, Po-Wei, et al.
Published: (2025)
SOAP-RL: Sequential Option Advantage Propagation for Reinforcement Learning in POMDP Environments
by: Ishida, Shu, et al.
Published: (2024)
by: Ishida, Shu, et al.
Published: (2024)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
High-Throughput SAT Sampling
by: Ardakani, Arash, et al.
Published: (2025)
by: Ardakani, Arash, et al.
Published: (2025)
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
by: Su, DiJia, et al.
Published: (2024)
by: Su, DiJia, et al.
Published: (2024)
Data curation via joint example selection further accelerates multimodal learning
by: Evans, Talfan, et al.
Published: (2024)
by: Evans, Talfan, et al.
Published: (2024)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
Imitation Learning from Observation through Optimal Transport
by: Chang, Wei-Di, et al.
Published: (2023)
by: Chang, Wei-Di, et al.
Published: (2023)
Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
by: Bonnet, Clément, et al.
Published: (2023)
by: Bonnet, Clément, et al.
Published: (2023)
LLMPi: Optimizing LLMs for High-Throughput on Raspberry Pi
by: Ardakani, Mahsa, et al.
Published: (2025)
by: Ardakani, Mahsa, et al.
Published: (2025)
Drift Q-Learning
by: Houssaini, Anas, et al.
Published: (2026)
by: Houssaini, Anas, et al.
Published: (2026)
DiffuMamba: High-Throughput Diffusion LMs with Mamba Backbone
by: Singh, Vaibhav, et al.
Published: (2025)
by: Singh, Vaibhav, et al.
Published: (2025)
A Survey of Imitation Learning Methods, Environments and Metrics
by: Gavenski, Nathan, et al.
Published: (2024)
by: Gavenski, Nathan, et al.
Published: (2024)
Understanding Contrastive Representation Learning from Positive Unlabeled (PU) Data
by: Acharya, Anish, et al.
Published: (2024)
by: Acharya, Anish, et al.
Published: (2024)
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
by: Luo, Baiting, et al.
Published: (2025)
by: Luo, Baiting, et al.
Published: (2025)
Learning Abstract World Model for Value-preserving Planning with Options
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2024)
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2024)
Joint Learning of Hierarchical Neural Options and Abstract World Model
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
Boosting deep Reinforcement Learning using pretraining with Logical Options
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
A Review of Physics-based Machine Learning in Civil Engineering
by: Vadyala, Shashank Reddy, et al.
Published: (2021)
by: Vadyala, Shashank Reddy, et al.
Published: (2021)
Enabling High Data Throughput Reinforcement Learning on GPUs: A Domain Agnostic Framework for Data-Driven Scientific Research
by: Lan, Tian, et al.
Published: (2024)
by: Lan, Tian, et al.
Published: (2024)
Gaussian Embeddings: How JEPAs Secretly Learn Your Data Density
by: Balestriero, Randall, et al.
Published: (2025)
by: Balestriero, Randall, et al.
Published: (2025)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
Diversity-Enriched Option-Critic
by: Kamat, Anand, et al.
Published: (2020)
by: Kamat, Anand, et al.
Published: (2020)
Unveiling Options with Neural Decomposition
by: Alikhasi, Mahdi, et al.
Published: (2024)
by: Alikhasi, Mahdi, et al.
Published: (2024)
Scalable Utility-Aware Multiclass Calibration
by: Hegazy, Mahmoud, et al.
Published: (2025)
by: Hegazy, Mahmoud, et al.
Published: (2025)
Learning to Drive Safely with Hybrid Options
by: De Cooman, Bram, et al.
Published: (2025)
by: De Cooman, Bram, et al.
Published: (2025)
Explorative Imitation Learning: A Path Signature Approach for Continuous Environments
by: Gavenski, Nathan, et al.
Published: (2024)
by: Gavenski, Nathan, et al.
Published: (2024)
Coding-Enforced Resilient and Secure Aggregation for Hierarchical Federated Learning
by: Weng, Shudi, et al.
Published: (2026)
by: Weng, Shudi, et al.
Published: (2026)
Preference-based Reinforcement Learning beyond Pairwise Comparisons: Benefits of Multiple Options
by: Lee, Joongkyu, et al.
Published: (2025)
by: Lee, Joongkyu, et al.
Published: (2025)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025)
by: Ahn, Hongjoon, et al.
Published: (2025)
TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning
by: Lee, Hayeong, et al.
Published: (2026)
by: Lee, Hayeong, et al.
Published: (2026)
eagle: early approximated gradient based learning rate estimator
by: Fujimoto, Takumi, et al.
Published: (2025)
by: Fujimoto, Takumi, et al.
Published: (2025)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
by: Romio, Gabriel, et al.
Published: (2026)
by: Romio, Gabriel, et al.
Published: (2026)
Similar Items
-
Hierarchical Behaviour Spaces
by: Matthews, Michael Tryfan, et al.
Published: (2026) -
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025) -
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026) -
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024) -
Polar Sparsity: High Throughput Batched LLM Inferencing with Scalable Contextual Sparsity
by: Shrestha, Susav, et al.
Published: (2025)