Learning Versatile Skills with Curriculum Masking
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Yao, Xie, Zhihui, Lin, Zichuan, Ye, Deheng, Li, Shuai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PIPCFR: Pseudo-outcome Imputation with Post-treatment Variables for Individual Treatment Effect Estimation
by: Lin, Zichuan, et al.
Published: (2025)
by: Lin, Zichuan, et al.
Published: (2025)
Debiased Model-based Representations for Sample-efficient Continuous Control
by: Lyu, Jiafei, et al.
Published: (2026)
by: Lyu, Jiafei, et al.
Published: (2026)
EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control
by: Yang, Kai, et al.
Published: (2025)
by: Yang, Kai, et al.
Published: (2025)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
by: Lin, Zichuan, et al.
Published: (2025)
by: Lin, Zichuan, et al.
Published: (2025)
Revisiting Discrete Soft Actor-Critic
by: Zhou, Haibin, et al.
Published: (2022)
by: Zhou, Haibin, et al.
Published: (2022)
HISR: Hindsight Information Modulated Segmental Process Rewards For Multi-turn Agentic Reinforcement Learning
by: Lu, Zhicong, et al.
Published: (2026)
by: Lu, Zhicong, et al.
Published: (2026)
The Lifecycle Principle: Stabilizing Dynamic Neural Networks with State Memory
by: Yang, Zichuan
Published: (2025)
by: Yang, Zichuan
Published: (2025)
Learning Probabilities of Causation with Mask-Augmented Data
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Neuro-symbolic Action Masking for Deep Reinforcement Learning
by: Han, Shuai, et al.
Published: (2026)
by: Han, Shuai, et al.
Published: (2026)
Automating Curriculum Learning for Reinforcement Learning using a Skill-Based Bayesian Network
by: Hsiao, Vincent, et al.
Published: (2025)
by: Hsiao, Vincent, et al.
Published: (2025)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
by: He, Longxiang, et al.
Published: (2025)
by: He, Longxiang, et al.
Published: (2025)
Curriculum Abductive Learning
by: Hu, Wen-Chao, et al.
Published: (2025)
by: Hu, Wen-Chao, et al.
Published: (2025)
Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO
by: Yu, Bowen, et al.
Published: (2026)
by: Yu, Bowen, et al.
Published: (2026)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
by: Sun, Shengjie, et al.
Published: (2025)
by: Sun, Shengjie, et al.
Published: (2025)
Reaching Consensus in Cooperative Multi-Agent Reinforcement Learning with Goal Imagination
by: Wang, Liangzhou, et al.
Published: (2024)
by: Wang, Liangzhou, et al.
Published: (2024)
Skill-Critic: Refining Learned Skills for Hierarchical Reinforcement Learning
by: Hao, Ce, et al.
Published: (2023)
by: Hao, Ce, et al.
Published: (2023)
UI-Voyager: A Self-Evolving GUI Agent Learning via Failed Experience
by: Lin, Zichuan, et al.
Published: (2026)
by: Lin, Zichuan, et al.
Published: (2026)
CL-MAE: Curriculum-Learned Masked Autoencoders
by: Madan, Neelu, et al.
Published: (2023)
by: Madan, Neelu, et al.
Published: (2023)
CBM: Curriculum by Masking
by: Jarca, Andrei, et al.
Published: (2024)
by: Jarca, Andrei, et al.
Published: (2024)
A Versatile Graph Learning Approach through LLM-based Agent
by: Wei, Lanning, et al.
Published: (2023)
by: Wei, Lanning, et al.
Published: (2023)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2023)
by: Yao, Yihang, et al.
Published: (2023)
Causally Aligned Curriculum Learning
by: Li, Mingxuan, et al.
Published: (2025)
by: Li, Mingxuan, et al.
Published: (2025)
Efficient Mitigation of Bus Bunching through Setter-Based Curriculum Learning
by: Shah, Avidan, et al.
Published: (2024)
by: Shah, Avidan, et al.
Published: (2024)
Level Up: Defining and Exploiting Transitional Problems for Curriculum Learning
by: Tang, Zhenwei, et al.
Published: (2026)
by: Tang, Zhenwei, et al.
Published: (2026)
More Agents Is All You Need
by: Li, Junyou, et al.
Published: (2024)
by: Li, Junyou, et al.
Published: (2024)
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
by: Li, Keyu, et al.
Published: (2021)
by: Li, Keyu, et al.
Published: (2021)
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
by: Xie, Johnathan, et al.
Published: (2024)
by: Xie, Johnathan, et al.
Published: (2024)
Resource Efficient Sleep Staging via Multi-Level Masking and Prompt Learning
by: Ai, Lejun, et al.
Published: (2025)
by: Ai, Lejun, et al.
Published: (2025)
Masked Diffusion Modeling for Anomaly Detection
by: Zhang, Lixing, et al.
Published: (2026)
by: Zhang, Lixing, et al.
Published: (2026)
Curriculum Learning for Safety Alignment
by: Kumar, Sandeep, et al.
Published: (2026)
by: Kumar, Sandeep, et al.
Published: (2026)
Task-Informed Anti-Curriculum by Masking Improves Downstream Performance on Text
by: Jarca, Andrei, et al.
Published: (2025)
by: Jarca, Andrei, et al.
Published: (2025)
CORD: Generalizable Cooperation via Role Diversity
by: Matsuyama, Kanefumi, et al.
Published: (2025)
by: Matsuyama, Kanefumi, et al.
Published: (2025)
Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data
by: Zhao, Shuai, et al.
Published: (2025)
by: Zhao, Shuai, et al.
Published: (2025)
Diversified Scaling Inference in Time Series Foundation Models
by: Hua, Ruijin, et al.
Published: (2026)
by: Hua, Ruijin, et al.
Published: (2026)
CURO: Curriculum Learning for Relative Overgeneralization
by: Shi, Lin, et al.
Published: (2022)
by: Shi, Lin, et al.
Published: (2022)
Learning Heterogeneous Performance-Fairness Trade-offs in Federated Learning
by: Ye, Rongguang, et al.
Published: (2025)
by: Ye, Rongguang, et al.
Published: (2025)
HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation
by: Zhao, He, et al.
Published: (2026)
by: Zhao, He, et al.
Published: (2026)
Skill-R1: Agent Skill Evolution via Reinforcement Learning
by: Vishe, Yash, et al.
Published: (2026)
by: Vishe, Yash, et al.
Published: (2026)
AudioMosaic: Contrastive Masked Audio Representation Learning
by: Huang, Hanxun, et al.
Published: (2026)
by: Huang, Hanxun, et al.
Published: (2026)
Enabling Time-series Foundation Model for Building Energy Forecasting via Contrastive Curriculum Learning
by: Liang, Rui, et al.
Published: (2024)
by: Liang, Rui, et al.
Published: (2024)
Similar Items
-
PIPCFR: Pseudo-outcome Imputation with Post-treatment Variables for Individual Treatment Effect Estimation
by: Lin, Zichuan, et al.
Published: (2025) -
Debiased Model-based Representations for Sample-efficient Continuous Control
by: Lyu, Jiafei, et al.
Published: (2026) -
EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control
by: Yang, Kai, et al.
Published: (2025) -
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
by: Lin, Zichuan, et al.
Published: (2025) -
Revisiting Discrete Soft Actor-Critic
by: Zhou, Haibin, et al.
Published: (2022)