Efficient Skill Discovery via Regret-Aware Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, He, Zhou, Ming, Zhai, Shaopeng, Sun, Ying, Xiong, Hui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024)
by: Rutherford, Alexander, et al.
Published: (2024)
Regret-Based Federated Causal Discovery with Unknown Interventions
by: Baldo, Federico, et al.
Published: (2025)
by: Baldo, Federico, et al.
Published: (2025)
Skill Weaving: Efficient LLM Improvement via Modular Skillpacks
by: Li, Zhuo, et al.
Published: (2026)
by: Li, Zhuo, et al.
Published: (2026)
Reference Grounded Skill Discovery
by: Rho, Seungeun, et al.
Published: (2025)
by: Rho, Seungeun, et al.
Published: (2025)
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024)
by: Zhao, Xufeng, et al.
Published: (2024)
Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling
by: Ye, Liqin, et al.
Published: (2026)
by: Ye, Liqin, et al.
Published: (2026)
SkillOrchestra: Learning to Route Agents via Skill Transfer
by: Wang, Jiayu, et al.
Published: (2026)
by: Wang, Jiayu, et al.
Published: (2026)
Provably Efficient Exploration in Reward Machines with Low Regret
by: Bourel, Hippolyte, et al.
Published: (2024)
by: Bourel, Hippolyte, et al.
Published: (2024)
Variational Offline Multi-agent Skill Discovery
by: Chen, Jiayu, et al.
Published: (2024)
by: Chen, Jiayu, et al.
Published: (2024)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
by: He, Jiafan, et al.
Published: (2025)
by: He, Jiafan, et al.
Published: (2025)
ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL
by: He, Zelin, et al.
Published: (2026)
by: He, Zelin, et al.
Published: (2026)
Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents
by: Zhou, Yifei, et al.
Published: (2024)
by: Zhou, Yifei, et al.
Published: (2024)
Focused Skill Discovery: Learning to Control Specific State Variables while Minimizing Side Effects
by: Carr, Jonathan Colaço, et al.
Published: (2025)
by: Carr, Jonathan Colaço, et al.
Published: (2025)
UCPO: Uncertainty-Aware Policy Optimization
by: Zeng, Xianzhou, et al.
Published: (2026)
by: Zeng, Xianzhou, et al.
Published: (2026)
PROWL: Prioritized Regret-Driven Optimization for World Model Learning
by: Güzel, Ahmet H., et al.
Published: (2026)
by: Güzel, Ahmet H., et al.
Published: (2026)
Regret-Guided Search Control for Efficient Learning in AlphaZero
by: Tsai, Yun-Jui, et al.
Published: (2026)
by: Tsai, Yun-Jui, et al.
Published: (2026)
DRPO: Efficient Reasoning via Decoupled Reward Policy Optimization
by: Li, Gang, et al.
Published: (2025)
by: Li, Gang, et al.
Published: (2025)
Leveraging Human Feedback for Semantically-Relevant Skill Discovery
by: Hussonnois, Maxence, et al.
Published: (2026)
by: Hussonnois, Maxence, et al.
Published: (2026)
POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles
by: Menet, Nicolas, et al.
Published: (2026)
by: Menet, Nicolas, et al.
Published: (2026)
Efficient Differentiable Causal Discovery via Reliable Super-Structure Learning
by: Ma, Pingchuan, et al.
Published: (2026)
by: Ma, Pingchuan, et al.
Published: (2026)
A Regret Perspective on Online Multiple Testing
by: Hao, Qingyang, et al.
Published: (2026)
by: Hao, Qingyang, et al.
Published: (2026)
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation
by: Yang, Haochen, et al.
Published: (2026)
by: Yang, Haochen, et al.
Published: (2026)
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
by: Wang, Xun, et al.
Published: (2025)
by: Wang, Xun, et al.
Published: (2025)
SUSD: Structured Unsupervised Skill Discovery through State Factorization
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2026)
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2026)
Towards Generalizable PDE Dynamics Forecasting via Physics-Guided Invariant Learning
by: Li, Siyang, et al.
Published: (2025)
by: Li, Siyang, et al.
Published: (2025)
Adversarial Environment Design via Regret-Guided Diffusion Models
by: Chung, Hojun, et al.
Published: (2024)
by: Chung, Hojun, et al.
Published: (2024)
Reasoning without Regret
by: Chitra, Tarun
Published: (2025)
by: Chitra, Tarun
Published: (2025)
GoldenStart: Q-Guided Priors and Entropy Control for Distilling Flow Policies
by: Zhang, He, et al.
Published: (2026)
by: Zhang, He, et al.
Published: (2026)
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback
by: Yang, Yongjin, et al.
Published: (2025)
by: Yang, Yongjin, et al.
Published: (2025)
Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?
by: Fan, Chenrui, et al.
Published: (2025)
by: Fan, Chenrui, et al.
Published: (2025)
OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension
by: Zhang, Zhiyuan, et al.
Published: (2026)
by: Zhang, Zhiyuan, et al.
Published: (2026)
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
by: Rho, Seungeun, et al.
Published: (2025)
by: Rho, Seungeun, et al.
Published: (2025)
Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
by: Yu, Yan, et al.
Published: (2025)
by: Yu, Yan, et al.
Published: (2025)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Language Guided Skill Discovery
by: Rho, Seungeun, et al.
Published: (2024)
by: Rho, Seungeun, et al.
Published: (2024)
Efficient Discovery of Approximate Causal Abstractions via Neural Mechanism Sparsification
by: Asiaee, Amir
Published: (2026)
by: Asiaee, Amir
Published: (2026)
CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching
by: Wu, Xingjian, et al.
Published: (2024)
by: Wu, Xingjian, et al.
Published: (2024)
MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?
by: Che, Xinyu, et al.
Published: (2026)
by: Che, Xinyu, et al.
Published: (2026)
Benign Overfitting in Adversarial Training for Vision Transformers
by: Zhang, Jiaming, et al.
Published: (2026)
by: Zhang, Jiaming, et al.
Published: (2026)
Differentiable Constraint-Based Causal Discovery
by: Zhou, Jincheng, et al.
Published: (2025)
by: Zhou, Jincheng, et al.
Published: (2025)
Similar Items
-
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024) -
Regret-Based Federated Causal Discovery with Unknown Interventions
by: Baldo, Federico, et al.
Published: (2025) -
Skill Weaving: Efficient LLM Improvement via Modular Skillpacks
by: Li, Zhuo, et al.
Published: (2026) -
Reference Grounded Skill Discovery
by: Rho, Seungeun, et al.
Published: (2025) -
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024)