A tale of two goals: leveraging sequentiality in multi-goal scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Serris, Olivier, Doncieux, Stéphane, Sigaud, Olivier |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Single-Reset Divide & Conquer Imitation Learning
von: Chenu, Alexandre, et al.
Veröffentlicht: (2024)
von: Chenu, Alexandre, et al.
Veröffentlicht: (2024)
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
Predicting Pedestrian Crossing Behavior in Germany and Japan: Insights into Model Transferability
von: Zhang, Chi, et al.
Veröffentlicht: (2024)
von: Zhang, Chi, et al.
Veröffentlicht: (2024)
Predicting and Analyzing Pedestrian Crossing Behavior at Unsignalized Crossings
von: Zhang, Chi, et al.
Veröffentlicht: (2024)
von: Zhang, Chi, et al.
Veröffentlicht: (2024)
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
von: Castanet, Nicolas, et al.
Veröffentlicht: (2025)
von: Castanet, Nicolas, et al.
Veröffentlicht: (2025)
PAC-MCTS: Bias-Aware Pruning for Robust LLM-Guided Search and Planning
von: Qian, Tianhao
Veröffentlicht: (2026)
von: Qian, Tianhao
Veröffentlicht: (2026)
Cross or Wait? Predicting Pedestrian Interaction Outcomes at Unsignalized Crossings
von: Zhang, Chi, et al.
Veröffentlicht: (2023)
von: Zhang, Chi, et al.
Veröffentlicht: (2023)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
von: Park, Junseok, et al.
Veröffentlicht: (2025)
von: Park, Junseok, et al.
Veröffentlicht: (2025)
Spatial-ViLT: Enhancing Visual Spatial Reasoning through Multi-Task Learning
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
von: Asri, Zakariae El, et al.
Veröffentlicht: (2024)
von: Asri, Zakariae El, et al.
Veröffentlicht: (2024)
Enhancing Robustness in Language-Driven Robotics: A Modular Approach to Failure Reduction
von: Garrabé, Émiland, et al.
Veröffentlicht: (2024)
von: Garrabé, Émiland, et al.
Veröffentlicht: (2024)
Reliability Quantification of Deep Reinforcement Learning-based Control
von: Yoshioka, Hitoshi, et al.
Veröffentlicht: (2023)
von: Yoshioka, Hitoshi, et al.
Veröffentlicht: (2023)
Automatic Depression Assessment using Machine Learning: A Comprehensive Survey
von: Song, Siyang, et al.
Veröffentlicht: (2025)
von: Song, Siyang, et al.
Veröffentlicht: (2025)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
von: Colas, Cédric, et al.
Veröffentlicht: (2020)
von: Colas, Cédric, et al.
Veröffentlicht: (2020)
Stealth edits to large language models
von: Sutton, Oliver J., et al.
Veröffentlicht: (2024)
von: Sutton, Oliver J., et al.
Veröffentlicht: (2024)
Learning Rate Engineering: From Coarse Single Parameter to Layered Evolution
von: Yao, Ming-Hong, et al.
Veröffentlicht: (2026)
von: Yao, Ming-Hong, et al.
Veröffentlicht: (2026)
Few-shot Quality-Diversity Optimization
von: Salehi, Achkan, et al.
Veröffentlicht: (2021)
von: Salehi, Achkan, et al.
Veröffentlicht: (2021)
Boosted Distributional Reinforcement Learning: Analysis and Healthcare Applications
von: Chen, Zequn, et al.
Veröffentlicht: (2026)
von: Chen, Zequn, et al.
Veröffentlicht: (2026)
Habits and goals in synergy: a variational Bayesian framework for behavior
von: Han, Dongqi, et al.
Veröffentlicht: (2023)
von: Han, Dongqi, et al.
Veröffentlicht: (2023)
Active Inference with a Self-Prior in the Mirror-Mark Task
von: Kim, Dongmin, et al.
Veröffentlicht: (2026)
von: Kim, Dongmin, et al.
Veröffentlicht: (2026)
MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces
von: Gaven, Loris, et al.
Veröffentlicht: (2025)
von: Gaven, Loris, et al.
Veröffentlicht: (2025)
Altered Thoughts, Altered Actions: Probing Chain-of-Thought Vulnerabilities in VLA Robotic Manipulation
von: Trinh, Tuan Duong, et al.
Veröffentlicht: (2026)
von: Trinh, Tuan Duong, et al.
Veröffentlicht: (2026)
Recent Advances of Deep Robotic Affordance Learning: A Reinforcement Learning Perspective
von: Yang, Xintong, et al.
Veröffentlicht: (2023)
von: Yang, Xintong, et al.
Veröffentlicht: (2023)
Progressive Power Homotopy for Non-convex Optimization
von: Xu, Chen
Veröffentlicht: (2026)
von: Xu, Chen
Veröffentlicht: (2026)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
A general language model for peptide function identification
von: Zhai, Jixiu, et al.
Veröffentlicht: (2025)
von: Zhai, Jixiu, et al.
Veröffentlicht: (2025)
Teacher Motion Priors: Enhancing Robot Locomotion over Challenging Terrain
von: Jin, Fangcheng, et al.
Veröffentlicht: (2025)
von: Jin, Fangcheng, et al.
Veröffentlicht: (2025)
Heuristic Search for Path Finding with Refuelling
von: Zhao, Shizhe, et al.
Veröffentlicht: (2023)
von: Zhao, Shizhe, et al.
Veröffentlicht: (2023)
Automated Feature Selection for Inverse Reinforcement Learning
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
RLPP: A Residual Method for Zero-Shot Real-World Autonomous Racing on Scaled Platforms
von: Ghignone, Edoardo, et al.
Veröffentlicht: (2025)
von: Ghignone, Edoardo, et al.
Veröffentlicht: (2025)
Graph-Instructed Neural Networks for Sparse Grid-Based Discontinuity Detectors
von: Della Santa, Francesco, et al.
Veröffentlicht: (2024)
von: Della Santa, Francesco, et al.
Veröffentlicht: (2024)
Reacting like Humans: Incorporating Intrinsic Human Behaviors into NAO through Sound-Based Reactions to Fearful and Shocking Events for Enhanced Sociability
von: Ghadami, Ali, et al.
Veröffentlicht: (2023)
von: Ghadami, Ali, et al.
Veröffentlicht: (2023)
Optimizing MoE Routers: Design, Implementation, and Evaluation in Transformer Models
von: Harvey, Daniel Fidel, et al.
Veröffentlicht: (2025)
von: Harvey, Daniel Fidel, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
MPRU: Modular Projection-Redistribution Unlearning as Output Filter for Classification Pipelines
von: Peng, Minyi, et al.
Veröffentlicht: (2025)
von: Peng, Minyi, et al.
Veröffentlicht: (2025)
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
von: Gao, Weiguo, et al.
Veröffentlicht: (2025)
von: Gao, Weiguo, et al.
Veröffentlicht: (2025)
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
von: Tyukin, Ivan Y., et al.
Veröffentlicht: (2024)
von: Tyukin, Ivan Y., et al.
Veröffentlicht: (2024)
VIPER: Visual Perception and Explainable Reasoning for Sequential Decision-Making
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2025)
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2025)
Tape: A Cellular Automata Benchmark for Evaluating Rule-Shift Generalization in Reinforcement Learning
von: Pan, Enze
Veröffentlicht: (2026)
von: Pan, Enze
Veröffentlicht: (2026)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Single-Reset Divide & Conquer Imitation Learning
von: Chenu, Alexandre, et al.
Veröffentlicht: (2024) -
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
von: Cao, Chenyang, et al.
Veröffentlicht: (2024) -
Predicting Pedestrian Crossing Behavior in Germany and Japan: Insights into Model Transferability
von: Zhang, Chi, et al.
Veröffentlicht: (2024) -
Predicting and Analyzing Pedestrian Crossing Behavior at Unsignalized Crossings
von: Zhang, Chi, et al.
Veröffentlicht: (2024) -
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
von: Castanet, Nicolas, et al.
Veröffentlicht: (2025)