Salvato in:
| Autori principali: | Liu, Shanqi, Cao, Junjie, Chen, Wenzhou, Wen, Licheng, Liu, Yong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2020
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2011.02671 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Safe and Efficient Online Convex Optimization with Linear Budget Constraints and Partial Feedback
di: Liu, Shanqi, et al.
Pubblicazione: (2024)
di: Liu, Shanqi, et al.
Pubblicazione: (2024)
True Knowledge Comes from Practice: Aligning LLMs with Embodied Environments via Reinforcement Learning
di: Tan, Weihao, et al.
Pubblicazione: (2024)
di: Tan, Weihao, et al.
Pubblicazione: (2024)
Imitation Learning from Observation with Automatic Discount Scheduling
di: Liu, Yuyang, et al.
Pubblicazione: (2023)
di: Liu, Yuyang, et al.
Pubblicazione: (2023)
Gaussian process surrogate with physical law-corrected prior for multi-coupled PDEs defined on irregular geometry
di: Tang, Pucheng, et al.
Pubblicazione: (2025)
di: Tang, Pucheng, et al.
Pubblicazione: (2025)
Diffusion Imitation from Observation
di: Huang, Bo-Ruei, et al.
Pubblicazione: (2024)
di: Huang, Bo-Ruei, et al.
Pubblicazione: (2024)
PrivacyCD: Hierarchical Unlearning for Protecting Student Privacy in Cognitive Diagnosis
di: Hou, Mingliang, et al.
Pubblicazione: (2025)
di: Hou, Mingliang, et al.
Pubblicazione: (2025)
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
di: He, Lixuan, et al.
Pubblicazione: (2025)
di: He, Lixuan, et al.
Pubblicazione: (2025)
Imitation Learning from Observation through Optimal Transport
di: Chang, Wei-Di, et al.
Pubblicazione: (2023)
di: Chang, Wei-Di, et al.
Pubblicazione: (2023)
Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance
di: Li, Jihang, et al.
Pubblicazione: (2026)
di: Li, Jihang, et al.
Pubblicazione: (2026)
Offline Imitation Learning with Variational Counterfactual Reasoning
di: He, Bowei, et al.
Pubblicazione: (2023)
di: He, Bowei, et al.
Pubblicazione: (2023)
Imitation Learning from Observations: An Autoregressive Mixture of Experts Approach
di: Wang, Renzi, et al.
Pubblicazione: (2024)
di: Wang, Renzi, et al.
Pubblicazione: (2024)
EPR-GAIL: An EPR-Enhanced Hierarchical Imitation Learning Framework to Simulate Complex User Consumption Behaviors
di: Feng, Tao, et al.
Pubblicazione: (2025)
di: Feng, Tao, et al.
Pubblicazione: (2025)
Align Your Intents: Offline Imitation Learning via Optimal Transport
di: Bobrin, Maksim, et al.
Pubblicazione: (2024)
di: Bobrin, Maksim, et al.
Pubblicazione: (2024)
Improving Generalization Ability of Robotic Imitation Learning by Resolving Causal Confusion in Observations
di: Chen, Yifei, et al.
Pubblicazione: (2025)
di: Chen, Yifei, et al.
Pubblicazione: (2025)
Imitation from Observations with Trajectory-Level Generative Embeddings
di: Qu, Yongtao, et al.
Pubblicazione: (2026)
di: Qu, Yongtao, et al.
Pubblicazione: (2026)
Beyond Observations: Reconstruction Error-Guided Irregularly Sampled Time Series Representation Learning
di: Liu, Jiexi, et al.
Pubblicazione: (2025)
di: Liu, Jiexi, et al.
Pubblicazione: (2025)
Boolean Satisfiability via Imitation Learning
di: Zhang, Zewei, et al.
Pubblicazione: (2025)
di: Zhang, Zewei, et al.
Pubblicazione: (2025)
Adversarial Imitation Learning from Visual Observations using Latent Information
di: Giammarino, Vittorio, et al.
Pubblicazione: (2023)
di: Giammarino, Vittorio, et al.
Pubblicazione: (2023)
Learning the Hierarchical Organization in Brain Network for Brain Disorder Diagnosis
di: Tang, Jingfeng, et al.
Pubblicazione: (2026)
di: Tang, Jingfeng, et al.
Pubblicazione: (2026)
Interpretable Imitation Learning with Dynamic Causal Relations
di: Zhao, Tianxiang, et al.
Pubblicazione: (2023)
di: Zhao, Tianxiang, et al.
Pubblicazione: (2023)
Hierarchical Imitation Learning of Team Behavior from Heterogeneous Demonstrations
di: Seo, Sangwon, et al.
Pubblicazione: (2025)
di: Seo, Sangwon, et al.
Pubblicazione: (2025)
CHAM-net: A Contrastive Hierarchical Adaptive Meta-network for Robust Global Methane Flux Prediction
di: Dong, Rongchao, et al.
Pubblicazione: (2026)
di: Dong, Rongchao, et al.
Pubblicazione: (2026)
Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement
di: Bloesch, Michael, et al.
Pubblicazione: (2025)
di: Bloesch, Michael, et al.
Pubblicazione: (2025)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
di: Sikchi, Harshit, et al.
Pubblicazione: (2024)
di: Sikchi, Harshit, et al.
Pubblicazione: (2024)
Hierarchical Conditional Multi-Task Learning for Streamflow Modeling
di: Xu, Shaoming, et al.
Pubblicazione: (2024)
di: Xu, Shaoming, et al.
Pubblicazione: (2024)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
di: Wu, Runzhe, et al.
Pubblicazione: (2024)
di: Wu, Runzhe, et al.
Pubblicazione: (2024)
Overcoming Knowledge Barriers: Online Imitation Learning from Visual Observation with Pretrained World Models
di: Zhang, Xingyuan, et al.
Pubblicazione: (2024)
di: Zhang, Xingyuan, et al.
Pubblicazione: (2024)
SPECI: Skill Prompts based Hierarchical Continual Imitation Learning for Robot Manipulation
di: Xu, Jingkai, et al.
Pubblicazione: (2025)
di: Xu, Jingkai, et al.
Pubblicazione: (2025)
Non-Adversarial Imitation Learning Provably Free of Compounding Errors: The Role of Bellman Constraints
di: Xu, Tian, et al.
Pubblicazione: (2026)
di: Xu, Tian, et al.
Pubblicazione: (2026)
Sequential Federated Learning in Hierarchical Architecture on Non-IID Datasets
di: Yan, Xingrun, et al.
Pubblicazione: (2024)
di: Yan, Xingrun, et al.
Pubblicazione: (2024)
Adversarial Imitation Learning via Boosting
di: Chang, Jonathan D., et al.
Pubblicazione: (2024)
di: Chang, Jonathan D., et al.
Pubblicazione: (2024)
Feasibility-aware Imitation Learning from Observations through a Hand-mounted Demonstration Interface
di: Takahashi, Kei, et al.
Pubblicazione: (2025)
di: Takahashi, Kei, et al.
Pubblicazione: (2025)
Blending Imitation and Reinforcement Learning for Robust Policy Improvement
di: Liu, Xuefeng, et al.
Pubblicazione: (2023)
di: Liu, Xuefeng, et al.
Pubblicazione: (2023)
AdaFlow: Imitation Learning with Variance-Adaptive Flow-Based Policies
di: Hu, Xixi, et al.
Pubblicazione: (2024)
di: Hu, Xixi, et al.
Pubblicazione: (2024)
Online Prediction with Limited Selectivity
di: Liu, Licheng, et al.
Pubblicazione: (2025)
di: Liu, Licheng, et al.
Pubblicazione: (2025)
An Imitative Reinforcement Learning Framework for Pursuit-Lock-Launch Missions
di: Li, Siyuan, et al.
Pubblicazione: (2024)
di: Li, Siyuan, et al.
Pubblicazione: (2024)
MEGA-DAgger: Imitation Learning with Multiple Imperfect Experts
di: Sun, Xiatao, et al.
Pubblicazione: (2023)
di: Sun, Xiatao, et al.
Pubblicazione: (2023)
Imitation Learning from Purified Demonstrations
di: Wang, Yunke, et al.
Pubblicazione: (2023)
di: Wang, Yunke, et al.
Pubblicazione: (2023)
Hierarchical Learning-based Graph Partition for Large-scale Vehicle Routing Problems
di: Pan, Yuxin, et al.
Pubblicazione: (2025)
di: Pan, Yuxin, et al.
Pubblicazione: (2025)
GUNDAM: Aligning Large Language Models with Graph Understanding
di: Ouyang, Sheng, et al.
Pubblicazione: (2024)
di: Ouyang, Sheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Safe and Efficient Online Convex Optimization with Linear Budget Constraints and Partial Feedback
di: Liu, Shanqi, et al.
Pubblicazione: (2024) -
True Knowledge Comes from Practice: Aligning LLMs with Embodied Environments via Reinforcement Learning
di: Tan, Weihao, et al.
Pubblicazione: (2024) -
Imitation Learning from Observation with Automatic Discount Scheduling
di: Liu, Yuyang, et al.
Pubblicazione: (2023) -
Gaussian process surrogate with physical law-corrected prior for multi-coupled PDEs defined on irregular geometry
di: Tang, Pucheng, et al.
Pubblicazione: (2025) -
Diffusion Imitation from Observation
di: Huang, Bo-Ruei, et al.
Pubblicazione: (2024)