Syllabus: Portable Curricula for Reinforcement Learning Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Sullivan, Ryan, Pégoud, Ryan, Rehman, Ameen Ur, Yang, Xinchen, Huang, Junyun, Verma, Aayush, Mitra, Nistha, Dickerson, John P. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Thermodynamics of Reinforcement Learning Curricula
by: Adamczyk, Jacob, et al.
Published: (2026)
by: Adamczyk, Jacob, et al.
Published: (2026)
Au-M-ol: A Unified Model for Medical Audio and Language Understanding
by: Liu, Meizhu, et al.
Published: (2026)
by: Liu, Meizhu, et al.
Published: (2026)
Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
Controlling Behavioral Diversity in Multi-Agent Reinforcement Learning
by: Bettini, Matteo, et al.
Published: (2024)
by: Bettini, Matteo, et al.
Published: (2024)
Massively Multiagent Minigames for Training Generalist Agents
by: Choe, Kyoung Whan, et al.
Published: (2024)
by: Choe, Kyoung Whan, et al.
Published: (2024)
Dual-Encoder Transformer-Based Multimodal Learning for Ischemic Stroke Lesion Segmentation Using Diffusion MRI
by: Usman, Muhammad, et al.
Published: (2025)
by: Usman, Muhammad, et al.
Published: (2025)
PIQL: Projective Implicit Q-Learning with Support Constraint for Offline Reinforcement Learning
by: Han, Xinchen, et al.
Published: (2025)
by: Han, Xinchen, et al.
Published: (2025)
KisanQRS: A Deep Learning-based Automated Query-Response System for Agricultural Decision-Making
by: Rehman, Mohammad Zia Ur, et al.
Published: (2024)
by: Rehman, Mohammad Zia Ur, et al.
Published: (2024)
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
by: Domico, Kyle, et al.
Published: (2025)
by: Domico, Kyle, et al.
Published: (2025)
Evolutionary System Prompt Learning for Reinforcement Learning in LLMs
by: Zhang, Lunjun, et al.
Published: (2026)
by: Zhang, Lunjun, et al.
Published: (2026)
AI-Driven Clinical Decision Support System for Enhanced Diabetes Diagnosis and Management
by: Rehman, Mujeeb Ur, et al.
Published: (2026)
by: Rehman, Mujeeb Ur, et al.
Published: (2026)
ASTIF: Adaptive Semantic-Temporal Integration for Cryptocurrency Price Forecasting
by: Rehman, Hafiz Saif Ur, et al.
Published: (2025)
by: Rehman, Hafiz Saif Ur, et al.
Published: (2025)
Skill-R1: Agent Skill Evolution via Reinforcement Learning
by: Vishe, Yash, et al.
Published: (2026)
by: Vishe, Yash, et al.
Published: (2026)
Automatic Constraint Policy Optimization based on Continuous Constraint Interpolation Framework for Offline Reinforcement Learning
by: Han, Xinchen, et al.
Published: (2026)
by: Han, Xinchen, et al.
Published: (2026)
RLtools: A Fast, Portable Deep Reinforcement Learning Library for Continuous Control
by: Eschmann, Jonas, et al.
Published: (2023)
by: Eschmann, Jonas, et al.
Published: (2023)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
by: Humayoo, Mahammad, et al.
Published: (2018)
by: Humayoo, Mahammad, et al.
Published: (2018)
Offscript: Automated Auditing of Instruction Adherence in LLMs
by: Clark, Nicholas, et al.
Published: (2025)
by: Clark, Nicholas, et al.
Published: (2025)
Recurrent Reinforcement Learning with Memoroids
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
Adaptive Robust Estimator for Multi-Agent Reinforcement Learning
by: Li, Zhongyi, et al.
Published: (2026)
by: Li, Zhongyi, et al.
Published: (2026)
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning
by: Chen, Yicheng, et al.
Published: (2026)
by: Chen, Yicheng, et al.
Published: (2026)
A Multimodal Framework for Depression Detection during Covid-19 via Harvesting Social Media: A Novel Dataset and Method
by: Anshul, Ashutosh, et al.
Published: (2025)
by: Anshul, Ashutosh, et al.
Published: (2025)
Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning
by: Li, Xiaozhe, et al.
Published: (2026)
by: Li, Xiaozhe, et al.
Published: (2026)
A Survey on GUI Agents with Foundation Models Enhanced by Reinforcement Learning
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
Anatomy-Guided Representation Learning Using a Transformer-Based Network for Thyroid Nodule Segmentation in Ultrasound Images
by: Farooq, Muhammad Umar, et al.
Published: (2025)
by: Farooq, Muhammad Umar, et al.
Published: (2025)
Aligning Medical Conversational AI through Online Reinforcement Learning with Information-Theoretic Rewards
by: Verma, Tanvi, et al.
Published: (2026)
by: Verma, Tanvi, et al.
Published: (2026)
Learning Game-Playing Agents with Generative Code Optimization
by: Kuang, Zhiyi, et al.
Published: (2025)
by: Kuang, Zhiyi, et al.
Published: (2025)
Robust and Efficient Communication in Multi-Agent Reinforcement Learning
by: Liu, Zejiao, et al.
Published: (2025)
by: Liu, Zejiao, et al.
Published: (2025)
Curricula for Learning Robust Policies with Factored State Representations in Changing Environments
by: Panayiotou, Panayiotis, et al.
Published: (2024)
by: Panayiotou, Panayiotis, et al.
Published: (2024)
Machine Learning and Optimization Techniques for Solving Inverse Kinematics in a 7-DOF Robotic Arm
by: Adediran, Enoch, et al.
Published: (2024)
by: Adediran, Enoch, et al.
Published: (2024)
Reinforcement Learning for Machine Learning Engineering Agents
by: Yang, Sherry, et al.
Published: (2025)
by: Yang, Sherry, et al.
Published: (2025)
Multicopy Reinforcement Learning Agents
by: Wolfe, Alicia P., et al.
Published: (2023)
by: Wolfe, Alicia P., et al.
Published: (2023)
TeachBench: A Syllabus-Grounded Framework for Evaluating Teaching Ability in Large Language Models
by: Li, Zheng, et al.
Published: (2026)
by: Li, Zheng, et al.
Published: (2026)
Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning
by: Wu, Mingze, et al.
Published: (2026)
by: Wu, Mingze, et al.
Published: (2026)
Automating the Refinement of Reinforcement Learning Specifications
by: Ambadkar, Tanmay, et al.
Published: (2025)
by: Ambadkar, Tanmay, et al.
Published: (2025)
PAC-Bayesian Reinforcement Learning Trains Generalizable Policies
by: Zitouni, Abdelkrim, et al.
Published: (2025)
by: Zitouni, Abdelkrim, et al.
Published: (2025)
Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents
by: Tan, Reuben, et al.
Published: (2025)
by: Tan, Reuben, et al.
Published: (2025)
Distinct Computations Emerge From Compositional Curricula in In-Context Learning
by: Lee, Jin Hwa, et al.
Published: (2025)
by: Lee, Jin Hwa, et al.
Published: (2025)
Flexible Blood Glucose Control: Offline Reinforcement Learning from Human Feedback
by: Emerson, Harry, et al.
Published: (2025)
by: Emerson, Harry, et al.
Published: (2025)
Accelerating Robotic Reinforcement Learning with Agent Guidance
by: Chen, Haojun, et al.
Published: (2026)
by: Chen, Haojun, et al.
Published: (2026)
Similar Items
-
Thermodynamics of Reinforcement Learning Curricula
by: Adamczyk, Jacob, et al.
Published: (2026) -
Au-M-ol: A Unified Model for Medical Audio and Language Understanding
by: Liu, Meizhu, et al.
Published: (2026) -
Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning
by: Hübotter, Jonas, et al.
Published: (2025) -
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025) -
Controlling Behavioral Diversity in Multi-Agent Reinforcement Learning
by: Bettini, Matteo, et al.
Published: (2024)