MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces
Fuente:
arXiv
Salvato in:
| Autori principali: | Gaven, Loris, Carta, Thomas, Romac, Clément, Colas, Cédric, Lamprier, Sylvain, Sigaud, Olivier, Oudeyer, Pierre-Yves |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
di: Gaven, Loris, et al.
Pubblicazione: (2024)
di: Gaven, Loris, et al.
Pubblicazione: (2024)
HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents
di: Carta, Thomas, et al.
Pubblicazione: (2025)
di: Carta, Thomas, et al.
Pubblicazione: (2025)
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
di: Carta, Thomas, et al.
Pubblicazione: (2023)
di: Carta, Thomas, et al.
Pubblicazione: (2023)
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2024)
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2024)
WorldLLM: Improving LLMs' world modeling using curiosity-driven theory-making
di: Levy, Guillaume, et al.
Pubblicazione: (2025)
di: Levy, Guillaume, et al.
Pubblicazione: (2025)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
di: Colas, Cédric, et al.
Pubblicazione: (2020)
di: Colas, Cédric, et al.
Pubblicazione: (2020)
CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning
di: Colas, Cédric, et al.
Pubblicazione: (2018)
di: Colas, Cédric, et al.
Pubblicazione: (2018)
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
di: Castanet, Nicolas, et al.
Pubblicazione: (2025)
di: Castanet, Nicolas, et al.
Pubblicazione: (2025)
Self-Improving Language Models for Evolutionary Program Synthesis: A Case Study on ARC-AGI
di: Pourcel, Julien, et al.
Pubblicazione: (2025)
di: Pourcel, Julien, et al.
Pubblicazione: (2025)
A Definition of Open-Ended Learning Problems for Goal-Conditioned Agents
di: Sigaud, Olivier, et al.
Pubblicazione: (2023)
di: Sigaud, Olivier, et al.
Pubblicazione: (2023)
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology
di: Tonga, Junior Cedric, et al.
Pubblicazione: (2024)
di: Tonga, Junior Cedric, et al.
Pubblicazione: (2024)
A tale of two goals: leveraging sequentiality in multi-goal scenarios
di: Serris, Olivier, et al.
Pubblicazione: (2025)
di: Serris, Olivier, et al.
Pubblicazione: (2025)
Expedition & Expansion: Leveraging Semantic Representations for Goal-Directed Exploration in Continuous Cellular Automata
di: Khajehabdollahi, Sina, et al.
Pubblicazione: (2025)
di: Khajehabdollahi, Sina, et al.
Pubblicazione: (2025)
ACES: Generating Diverse Programming Puzzles with with Autotelic Generative Models
di: Pourcel, Julien, et al.
Pubblicazione: (2023)
di: Pourcel, Julien, et al.
Pubblicazione: (2023)
Curiosity and Metacognition: Towards a Unified Framework for Learning and Education in the Age of AI
di: Desvaux, Chloé, et al.
Pubblicazione: (2026)
di: Desvaux, Chloé, et al.
Pubblicazione: (2026)
When LLMs Play the Telephone Game: Cultural Attractors as Conceptual Tools to Evaluate LLMs in Multi-turn Settings
di: Perez, Jérémy, et al.
Pubblicazione: (2024)
di: Perez, Jérémy, et al.
Pubblicazione: (2024)
Improved Performances and Motivation in Intelligent Tutoring Systems: Combining Machine Learning and Learner Choice
di: Clément, Benjamin, et al.
Pubblicazione: (2024)
di: Clément, Benjamin, et al.
Pubblicazione: (2024)
PRISM: Perception Reasoning Interleaved for Sequential Decision Making
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2026)
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2026)
PhyloLM : Inferring the Phylogeny of Large Language Models and Predicting their Performances in Benchmarks
di: Yax, Nicolas, et al.
Pubblicazione: (2024)
di: Yax, Nicolas, et al.
Pubblicazione: (2024)
LogProber: Disentangling confidence from contamination in LLM responses
di: Yax, Nicolas, et al.
Pubblicazione: (2024)
di: Yax, Nicolas, et al.
Pubblicazione: (2024)
Collective Innovation in Groups of Large Language Models
di: Nisioti, Eleni, et al.
Pubblicazione: (2024)
di: Nisioti, Eleni, et al.
Pubblicazione: (2024)
Exploring Flow-Lenia Universes with a Curiosity-driven AI Scientist: Discovering Diverse Ecosystem Dynamics
di: Michel, Thomas, et al.
Pubblicazione: (2025)
di: Michel, Thomas, et al.
Pubblicazione: (2025)
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
di: Petitbois, Mathieu, et al.
Pubblicazione: (2026)
di: Petitbois, Mathieu, et al.
Pubblicazione: (2026)
Training Table Question Answering via SQL Query Decomposition
di: Mouravieff, Raphaël, et al.
Pubblicazione: (2024)
di: Mouravieff, Raphaël, et al.
Pubblicazione: (2024)
Structural Deep Encoding for Table Question Answering
di: Mouravieff, Raphaël, et al.
Pubblicazione: (2025)
di: Mouravieff, Raphaël, et al.
Pubblicazione: (2025)
Black-Box Combinatorial Optimization with Order-Invariant Reinforcement Learning
di: Goudet, Olivier, et al.
Pubblicazione: (2025)
di: Goudet, Olivier, et al.
Pubblicazione: (2025)
Stein Variational Black-Box Combinatorial Optimization
di: Landais, Thomas, et al.
Pubblicazione: (2026)
di: Landais, Thomas, et al.
Pubblicazione: (2026)
Alone but flowing: The effects of autotelic personality and extraversion on solitary flow
di: Dwight C. K. Tse, et al.
Pubblicazione: (2024)
di: Dwight C. K. Tse, et al.
Pubblicazione: (2024)
Nature and Nature's God: A Philosophical and Scientific Defense of Aquinas's Unmoved Mover Argument. By DanielShields. Washington, D.C.: Catholic University of America Press, 2023. Pp. 328. $75.00.
di: Gaven Kerr
Pubblicazione: (2024)
di: Gaven Kerr
Pubblicazione: (2024)
Deinterleaving of Discrete Renewal Process Mixtures with Application to Electronic Support Measures
di: Pinsolle, Jean, et al.
Pubblicazione: (2024)
di: Pinsolle, Jean, et al.
Pubblicazione: (2024)
Discovering Sensorimotor Agency in Cellular Automata using Diversity Search
di: Hamon, Gautier, et al.
Pubblicazione: (2024)
di: Hamon, Gautier, et al.
Pubblicazione: (2024)
DESCRIPTION AND KEY TO THE ZOEAL STAGES OF THE CAMPYLONOTIDAE (DECAPODA, CARIDEA) FROM THE MAGELLAN REGION
di: Thatje, S., Bacardit, R., Romero, M., Tapella, F., Lovrich, G
Pubblicazione: (2001)
di: Thatje, S., Bacardit, R., Romero, M., Tapella, F., Lovrich, G
Pubblicazione: (2001)
Concrete one complex dimensional moduli spaces of hyperbolic manifolds and orbifolds
di: Elzenaar, Alex, et al.
Pubblicazione: (2022)
di: Elzenaar, Alex, et al.
Pubblicazione: (2022)
Reward-Preserving Attacks For Robust Reinforcement Learning
di: Schott, Lucas, et al.
Pubblicazione: (2026)
di: Schott, Lucas, et al.
Pubblicazione: (2026)
Offline Learning of Controllable Diverse Behaviors
di: Petitbois, Mathieu, et al.
Pubblicazione: (2025)
di: Petitbois, Mathieu, et al.
Pubblicazione: (2025)
Progress Ratio Embeddings: An Impatience Signal for Robust Length Control in Neural Text Generation
di: Botcazou, Ivanhoé, et al.
Pubblicazione: (2025)
di: Botcazou, Ivanhoé, et al.
Pubblicazione: (2025)
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms
di: Asri, Zakariae El, et al.
Pubblicazione: (2025)
di: Asri, Zakariae El, et al.
Pubblicazione: (2025)
An agent design with goal reaching guarantees for enhancement of learning
di: Osinenko, Pavel, et al.
Pubblicazione: (2024)
di: Osinenko, Pavel, et al.
Pubblicazione: (2024)
Flow-Lenia: Emergent evolutionary dynamics in mass conservative continuous cellular automata
di: Plantec, Erwan, et al.
Pubblicazione: (2025)
di: Plantec, Erwan, et al.
Pubblicazione: (2025)
Iterative On-Policy Refinement of Hierarchical Diffusion Policies for Language-Conditioned Manipulation
di: Grislain, Clemence, et al.
Pubblicazione: (2026)
di: Grislain, Clemence, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
di: Gaven, Loris, et al.
Pubblicazione: (2024) -
HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents
di: Carta, Thomas, et al.
Pubblicazione: (2025) -
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
di: Carta, Thomas, et al.
Pubblicazione: (2023) -
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
di: Aissi, Mohamed Salim, et al.
Pubblicazione: (2024) -
WorldLLM: Improving LLMs' world modeling using curiosity-driven theory-making
di: Levy, Guillaume, et al.
Pubblicazione: (2025)