Guiding Skill Discovery with Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhao, Moerland, Thomas M., Preuss, Mike, Plaat, Aske, François-Lavet, Vincent, Hu, Edward S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reset-free Reinforcement Learning with World Models
by: Yang, Zhao, et al.
Published: (2024)
by: Yang, Zhao, et al.
Published: (2024)
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024)
by: Plaat, Aske
Published: (2024)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)
by: Ponse, Koen, et al.
Published: (2025)
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)
by: Spoor, Lindsay, et al.
Published: (2025)
Reinforcement Learning for Sustainable Energy: A Survey
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Hadamax Encoding: Elevating Performance in Model-Free Atari
by: Kooi, Jacob E., et al.
Published: (2025)
by: Kooi, Jacob E., et al.
Published: (2025)
Novel RL approach for efficient Elevator Group Control Systems
by: Vaartjes, Nathan, et al.
Published: (2025)
by: Vaartjes, Nathan, et al.
Published: (2025)
Agentic Large Language Models, a survey
by: Plaat, Aske, et al.
Published: (2025)
by: Plaat, Aske, et al.
Published: (2025)
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
by: Zaciragic, Tarik, et al.
Published: (2025)
by: Zaciragic, Tarik, et al.
Published: (2025)
On the Effect of Regularization in Policy Mirror Descent
by: Kleuker, Jan Felix, et al.
Published: (2025)
by: Kleuker, Jan Felix, et al.
Published: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Baba is LLM: Reasoning in a Game with Dynamic Rules
by: van Wetten, Fien, et al.
Published: (2025)
by: van Wetten, Fien, et al.
Published: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
CausalPlayground: Addressing Data-Generation Requirements in Cutting-Edge Causality Research
by: Sauter, Andreas W M, et al.
Published: (2024)
by: Sauter, Andreas W M, et al.
Published: (2024)
Slot Structured World Models
by: Collu, Jonathan, et al.
Published: (2024)
by: Collu, Jonathan, et al.
Published: (2024)
Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability
by: Kim, Taewoon, et al.
Published: (2026)
by: Kim, Taewoon, et al.
Published: (2026)
Temporal Knowledge-Graph Memory in a Partially Observable Environment
by: Kim, Taewoon, et al.
Published: (2024)
by: Kim, Taewoon, et al.
Published: (2024)
ACTIVA: Amortized Causal Effect Estimation via Transformer-based Variational Autoencoder
by: Sauter, Andreas, et al.
Published: (2025)
by: Sauter, Andreas, et al.
Published: (2025)
CoComposer: LLM Multi-agent Collaborative Music Composition
by: Xing, Peiwen, et al.
Published: (2025)
by: Xing, Peiwen, et al.
Published: (2025)
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025)
by: Wong, Annie, et al.
Published: (2025)
Disentangled (Un)Controllable Features
by: Kooi, Jacob E., et al.
Published: (2022)
by: Kooi, Jacob E., et al.
Published: (2022)
State Design Matters: How Representations Shape Dynamic Reasoning in Large Language Models
by: Wong, Annie, et al.
Published: (2026)
by: Wong, Annie, et al.
Published: (2026)
EduGym: An Environment and Notebook Suite for Reinforcement Learning Education
by: Moerland, Thomas M., et al.
Published: (2023)
by: Moerland, Thomas M., et al.
Published: (2023)
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
by: Wong, Annie, et al.
Published: (2024)
by: Wong, Annie, et al.
Published: (2024)
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024)
by: Plaat, Aske, et al.
Published: (2024)
A Machine With Human-Like Memory Systems
by: Kim, Taewoon, et al.
Published: (2022)
by: Kim, Taewoon, et al.
Published: (2022)
A Machine with Short-Term, Episodic, and Semantic Memory Systems
by: Kim, Taewoon, et al.
Published: (2022)
by: Kim, Taewoon, et al.
Published: (2022)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
EconoJax: A Fast & Scalable Economic Simulation in Jax
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Assessing Reproducibility in Evolutionary Computation: A Case Study using Human- and LLM-based Assessment
by: Da Ros, Francesca, et al.
Published: (2026)
by: Da Ros, Francesca, et al.
Published: (2026)
A Unified Framework for Zero-Shot Reinforcement Learning
by: Di Ventura, Jacopo, et al.
Published: (2025)
by: Di Ventura, Jacopo, et al.
Published: (2025)
Reinforcement learning for Quantum Tiq-Taq-Toe
by: Dinu, Catalin-Viorel, et al.
Published: (2024)
by: Dinu, Catalin-Viorel, et al.
Published: (2024)
A Benchmark Study of Deep Reinforcement Learning Algorithms for the Container Stowage Planning Problem
by: Huang, Yunqi, et al.
Published: (2025)
by: Huang, Yunqi, et al.
Published: (2025)
Language Guided Skill Discovery
by: Rho, Seungeun, et al.
Published: (2024)
by: Rho, Seungeun, et al.
Published: (2024)
SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents
by: Zhang, Ziao, et al.
Published: (2026)
by: Zhang, Ziao, et al.
Published: (2026)
A Hybrid Intelligence Method for Argument Mining
by: van der Meer, Michiel, et al.
Published: (2024)
by: van der Meer, Michiel, et al.
Published: (2024)
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024)
by: Zhao, Xufeng, et al.
Published: (2024)
Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents
by: Zhou, Yifei, et al.
Published: (2024)
by: Zhou, Yifei, et al.
Published: (2024)
Dynamic Expert-Guided Model Averaging for Causal Discovery
by: Tench, Adrick, et al.
Published: (2026)
by: Tench, Adrick, et al.
Published: (2026)
Similar Items
-
Reset-free Reinforcement Learning with World Models
by: Yang, Zhao, et al.
Published: (2024) -
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024) -
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024) -
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025) -
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)