Rethinking Mutual Information for Language Conditioned Skill Discovery on Imitation Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Ju, Zhaoxun, Yang, Chao, Wang, Hongbo, Qiao, Yu, Sun, Fuchun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Atomic-Probe Governance for Skill Updates in Compositional Robot Policies
por: Qin, Xue, et al.
Publicado: (2026)
por: Qin, Xue, et al.
Publicado: (2026)
Simulation-Based Counterfactual Causal Discovery on Real World Driver Behaviour
por: Howard, Rhys, et al.
Publicado: (2023)
por: Howard, Rhys, et al.
Publicado: (2023)
URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model
por: Li, Zhe, et al.
Publicado: (2025)
por: Li, Zhe, et al.
Publicado: (2025)
Learning more with the same effort: how randomization improves the robustness of a robotic deep reinforcement learning agent
por: Güitta-López, Lucía, et al.
Publicado: (2025)
por: Güitta-López, Lucía, et al.
Publicado: (2025)
Predictive Traffic Rule Compliance using Reinforcement Learning
por: Huang, Yanliang, et al.
Publicado: (2025)
por: Huang, Yanliang, et al.
Publicado: (2025)
Conservative Bias in Multi-Teacher Learning: Why Agents Prefer Low-Reward Advisors
por: Mesto, Maher, et al.
Publicado: (2025)
por: Mesto, Maher, et al.
Publicado: (2025)
ASkDAgger: Active Skill-level Data Aggregation for Interactive Imitation Learning
por: Luijkx, Jelle, et al.
Publicado: (2025)
por: Luijkx, Jelle, et al.
Publicado: (2025)
Evaluating Temporal Observation-Based Causal Discovery Techniques Applied to Road Driver Behaviour
por: Howard, Rhys, et al.
Publicado: (2023)
por: Howard, Rhys, et al.
Publicado: (2023)
SaiVLA-0: Cerebrum--Pons--Cerebellum Tripartite Architecture for Compute-Aware Vision-Language-Action
por: Shi, Xiang, et al.
Publicado: (2026)
por: Shi, Xiang, et al.
Publicado: (2026)
Energy-Efficient Quadruped Locomotion with Compliant Feet
por: Pal, Pramod, et al.
Publicado: (2026)
por: Pal, Pramod, et al.
Publicado: (2026)
Robo-CSK-Organizer: Commonsense Knowledge to Organize Detected Objects for Multipurpose Robots
por: Hidalgo, Rafael, et al.
Publicado: (2024)
por: Hidalgo, Rafael, et al.
Publicado: (2024)
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
por: Daba, Mohammed, et al.
Publicado: (2025)
por: Daba, Mohammed, et al.
Publicado: (2025)
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning
por: Zhao, Yang, et al.
Publicado: (2024)
por: Zhao, Yang, et al.
Publicado: (2024)
Sim-to-reality adaptation for Deep Reinforcement Learning applied to an underwater docking application
por: Chaarani, Alaaeddine, et al.
Publicado: (2026)
por: Chaarani, Alaaeddine, et al.
Publicado: (2026)
Deep Reinforcement Learning for Adverse Garage Scenario Generation
por: Li, Kai
Publicado: (2024)
por: Li, Kai
Publicado: (2024)
MTDrive: Multi-turn Interactive Reinforcement Learning for Autonomous Driving
por: Li, Xidong, et al.
Publicado: (2026)
por: Li, Xidong, et al.
Publicado: (2026)
Towards a Robust Soft Baby Robot With Rich Interaction Ability for Advanced Machine Learning Algorithms
por: Alhakami, Mohannad, et al.
Publicado: (2024)
por: Alhakami, Mohannad, et al.
Publicado: (2024)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
por: Zeng, Jack, et al.
Publicado: (2025)
por: Zeng, Jack, et al.
Publicado: (2025)
CUPID: Curating Data your Robot Loves with Influence Functions
por: Agia, Christopher, et al.
Publicado: (2025)
por: Agia, Christopher, et al.
Publicado: (2025)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
por: Aida, Adriana, et al.
Publicado: (2026)
por: Aida, Adriana, et al.
Publicado: (2026)
Improving Value Estimation Critically Enhances Vanilla Policy Gradient
por: Wang, Tao, et al.
Publicado: (2025)
por: Wang, Tao, et al.
Publicado: (2025)
Extrinsicaly Rewarded Soft Q Imitation Learning with Discriminator
por: Furuyama, Ryoma, et al.
Publicado: (2024)
por: Furuyama, Ryoma, et al.
Publicado: (2024)
A Framework for Neurosymbolic Robot Action Planning using Large Language Models
por: Capitanelli, Alessio, et al.
Publicado: (2023)
por: Capitanelli, Alessio, et al.
Publicado: (2023)
Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis
por: Kirk, James R., et al.
Publicado: (2023)
por: Kirk, James R., et al.
Publicado: (2023)
The Shortcomings of Force-from-Motion in Robot Learning
por: Aljalbout, Elie, et al.
Publicado: (2024)
por: Aljalbout, Elie, et al.
Publicado: (2024)
SPACeR: Self-Play Anchoring with Centralized Reference Models
por: Chang, Wei-Jer, et al.
Publicado: (2025)
por: Chang, Wei-Jer, et al.
Publicado: (2025)
When Does Adaptive Guidance Help? Belief-Aware Privileged Distillation for Autonomous Driving Under Partial Observability
por: Haklidir, Mehmet
Publicado: (2026)
por: Haklidir, Mehmet
Publicado: (2026)
MEReQ: Max-Ent Residual-Q Inverse RL for Sample-Efficient Alignment from Intervention
por: Chen, Yuxin, et al.
Publicado: (2024)
por: Chen, Yuxin, et al.
Publicado: (2024)
Simulation-Driven Railway Delay Prediction: An Imitation Learning Approach
por: Elliker, Clément, et al.
Publicado: (2025)
por: Elliker, Clément, et al.
Publicado: (2025)
Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning
por: Wang, Zizhao, et al.
Publicado: (2024)
por: Wang, Zizhao, et al.
Publicado: (2024)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
por: Palma, Alessio, et al.
Publicado: (2026)
por: Palma, Alessio, et al.
Publicado: (2026)
RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing
por: Ai, Bo, et al.
Publicado: (2024)
por: Ai, Bo, et al.
Publicado: (2024)
D-Shape: Demonstration-Shaped Reinforcement Learning via Goal Conditioning
por: Wang, Caroline, et al.
Publicado: (2022)
por: Wang, Caroline, et al.
Publicado: (2022)
LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios
por: Yao, Huaiyuan, et al.
Publicado: (2025)
por: Yao, Huaiyuan, et al.
Publicado: (2025)
Learning to Walk in Costume: Adversarial Motion Priors for Aesthetically Constrained Humanoids
por: Alvarez, Arturo Flores, et al.
Publicado: (2025)
por: Alvarez, Arturo Flores, et al.
Publicado: (2025)
RoboGrind: Intuitive and Interactive Surface Treatment with Industrial Robots
por: Alt, Benjamin, et al.
Publicado: (2024)
por: Alt, Benjamin, et al.
Publicado: (2024)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
por: Liao, Jianxing, et al.
Publicado: (2025)
por: Liao, Jianxing, et al.
Publicado: (2025)
Evaluating Bivariate Causal Statements Based on Mutual Compatibility
por: Jahn, Erik, et al.
Publicado: (2026)
por: Jahn, Erik, et al.
Publicado: (2026)
DSSE: a drone swarm search environment
por: Castanares, Manuel, et al.
Publicado: (2023)
por: Castanares, Manuel, et al.
Publicado: (2023)
Imitation learning for sim-to-real transfer of robotic cutting policies based on residual Gaussian process disturbance force model
por: Hathaway, Jamie, et al.
Publicado: (2023)
por: Hathaway, Jamie, et al.
Publicado: (2023)
Ejemplares similares
-
Atomic-Probe Governance for Skill Updates in Compositional Robot Policies
por: Qin, Xue, et al.
Publicado: (2026) -
Simulation-Based Counterfactual Causal Discovery on Real World Driver Behaviour
por: Howard, Rhys, et al.
Publicado: (2023) -
URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model
por: Li, Zhe, et al.
Publicado: (2025) -
Learning more with the same effort: how randomization improves the robustness of a robotic deep reinforcement learning agent
por: Güitta-López, Lucía, et al.
Publicado: (2025) -
Predictive Traffic Rule Compliance using Reinforcement Learning
por: Huang, Yanliang, et al.
Publicado: (2025)