Deconfounding Imitation Learning with Variational Inference
Fuente:
arXiv
Salvato in:
| Autori principali: | Vuorio, Risto, de Haan, Pim, Brehmer, Johann, Ackermann, Hanno, Dijkman, Daniel, Cohen, Taco |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Euclidean, Projective, Conformal: Choosing a Geometric Algebra for Equivariant Transformers
di: de Haan, Pim, et al.
Pubblicazione: (2023)
di: de Haan, Pim, et al.
Pubblicazione: (2023)
Does equivariance matter at scale?
di: Brehmer, Johann, et al.
Pubblicazione: (2024)
di: Brehmer, Johann, et al.
Pubblicazione: (2024)
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
di: Grislain, Clémence, et al.
Pubblicazione: (2024)
di: Grislain, Clémence, et al.
Pubblicazione: (2024)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
di: Mazzaglia, Pietro, et al.
Pubblicazione: (2024)
di: Mazzaglia, Pietro, et al.
Pubblicazione: (2024)
A Bayesian Solution To The Imitation Gap
di: Vuorio, Risto, et al.
Pubblicazione: (2024)
di: Vuorio, Risto, et al.
Pubblicazione: (2024)
Action-Constrained Imitation Learning
di: Yeh, Chia-Han, et al.
Pubblicazione: (2025)
di: Yeh, Chia-Han, et al.
Pubblicazione: (2025)
Lorentz-Equivariant Geometric Algebra Transformers for High-Energy Physics
di: Spinner, Jonas, et al.
Pubblicazione: (2024)
di: Spinner, Jonas, et al.
Pubblicazione: (2024)
Spatial Deconfounder: Interference-Aware Deconfounding for Spatial Causal Inference
di: Khot, Ayush, et al.
Pubblicazione: (2025)
di: Khot, Ayush, et al.
Pubblicazione: (2025)
A Lorentz-Equivariant Transformer for All of the LHC
di: Brehmer, Johann, et al.
Pubblicazione: (2024)
di: Brehmer, Johann, et al.
Pubblicazione: (2024)
SplAgger: Split Aggregation for Meta-Reinforcement Learning
di: Beck, Jacob, et al.
Pubblicazione: (2024)
di: Beck, Jacob, et al.
Pubblicazione: (2024)
BLIPs: Bayesian Learned Interatomic Potentials
di: Coscia, Dario, et al.
Pubblicazione: (2025)
di: Coscia, Dario, et al.
Pubblicazione: (2025)
A Tutorial on Meta-Reinforcement Learning
di: Beck, Jacob, et al.
Pubblicazione: (2023)
di: Beck, Jacob, et al.
Pubblicazione: (2023)
Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control
di: Xiong, Zheng, et al.
Pubblicazione: (2024)
di: Xiong, Zheng, et al.
Pubblicazione: (2024)
Deconfounded Time Series Forecasting: A Causal Inference Approach
di: Gao, Wentao, et al.
Pubblicazione: (2024)
di: Gao, Wentao, et al.
Pubblicazione: (2024)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
di: Mazzaglia, Pietro, et al.
Pubblicazione: (2023)
di: Mazzaglia, Pietro, et al.
Pubblicazione: (2023)
From Code to Action: Hierarchical Learning of Diffusion-VLM Policies
di: Peschl, Markus, et al.
Pubblicazione: (2025)
di: Peschl, Markus, et al.
Pubblicazione: (2025)
Noether's razor: Learning Conserved Quantities
di: van der Ouderaa, Tycho F. A., et al.
Pubblicazione: (2024)
di: van der Ouderaa, Tycho F. A., et al.
Pubblicazione: (2024)
When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift
di: Sheng, Zhecheng, et al.
Pubblicazione: (2026)
di: Sheng, Zhecheng, et al.
Pubblicazione: (2026)
Two-way Deconfounder for Off-policy Evaluation in Causal Reinforcement Learning
di: Yu, Shuguang, et al.
Pubblicazione: (2024)
di: Yu, Shuguang, et al.
Pubblicazione: (2024)
Offline Imitation Learning with Variational Counterfactual Reasoning
di: He, Bowei, et al.
Pubblicazione: (2023)
di: He, Bowei, et al.
Pubblicazione: (2023)
Deconfounding Scores and Representation Learning for Causal Effect Estimation with Weak Overlap
di: Clivio, Oscar, et al.
Pubblicazione: (2026)
di: Clivio, Oscar, et al.
Pubblicazione: (2026)
Myosotis: structured computation for attention like layer
di: Egorov, Evgenii, et al.
Pubblicazione: (2025)
di: Egorov, Evgenii, et al.
Pubblicazione: (2025)
Deconfounded Lifelong Learning for Autonomous Driving via Dynamic Knowledge Spaces
di: Du, Jiayuan, et al.
Pubblicazione: (2026)
di: Du, Jiayuan, et al.
Pubblicazione: (2026)
A Deep Dive into Scaling RL for Code Generation with Synthetic Data and Curricula
di: Sancaktar, Cansu, et al.
Pubblicazione: (2026)
di: Sancaktar, Cansu, et al.
Pubblicazione: (2026)
Discovery of Decision Synchronization Patterns from Event Logs
di: Kuijpers, Tijmen, et al.
Pubblicazione: (2026)
di: Kuijpers, Tijmen, et al.
Pubblicazione: (2026)
ClevrSkills: Compositional Language and Visual Reasoning in Robotics
di: Haresh, Sanjay, et al.
Pubblicazione: (2024)
di: Haresh, Sanjay, et al.
Pubblicazione: (2024)
Efficient RL Training for LLMs with Experience Replay
di: Arnal, Charles, et al.
Pubblicazione: (2026)
di: Arnal, Charles, et al.
Pubblicazione: (2026)
Deconfounded Reasoning for Multimodal Fake News Detection via Causal Intervention
di: Liu, Moyang, et al.
Pubblicazione: (2025)
di: Liu, Moyang, et al.
Pubblicazione: (2025)
Imitative Membership Inference Attack
di: Du, Yuntao, et al.
Pubblicazione: (2025)
di: Du, Yuntao, et al.
Pubblicazione: (2025)
Bayesian Robust Optimization for Imitation Learning
di: Brown, Daniel S., et al.
Pubblicazione: (2020)
di: Brown, Daniel S., et al.
Pubblicazione: (2020)
Cross-Domain Imitation Learning via Optimal Transport
di: Fickinger, Arnaud, et al.
Pubblicazione: (2021)
di: Fickinger, Arnaud, et al.
Pubblicazione: (2021)
RL-finetuning LLMs from on- and off-policy data with a single algorithm
di: Tang, Yunhao, et al.
Pubblicazione: (2025)
di: Tang, Yunhao, et al.
Pubblicazione: (2025)
Learning with Importance Weighted Variational Inference
di: Daudel, Kamélia, et al.
Pubblicazione: (2024)
di: Daudel, Kamélia, et al.
Pubblicazione: (2024)
Deconfounded Warm-Start Thompson Sampling with Applications to Precision Medicine
di: Jaiswal, Prateek, et al.
Pubblicazione: (2025)
di: Jaiswal, Prateek, et al.
Pubblicazione: (2025)
Differentiable and Learnable Wireless Simulation with Geometric Transformers
di: Hehn, Thomas, et al.
Pubblicazione: (2024)
di: Hehn, Thomas, et al.
Pubblicazione: (2024)
Excluding the Target Domain Improves Extrapolation: Deconfounded Hierarchical Physics Constraints
di: Okita, Tsuyoshi
Pubblicazione: (2026)
di: Okita, Tsuyoshi
Pubblicazione: (2026)
Synergizing Deconfounding and Temporal Generalization For Time-series Counterfactual Outcome Estimation
di: Liu, Yiling, et al.
Pubblicazione: (2025)
di: Liu, Yiling, et al.
Pubblicazione: (2025)
Consistent Zero-Shot Imitation with Contrastive Goal Inference
di: Wantlin, Kathryn, et al.
Pubblicazione: (2025)
di: Wantlin, Kathryn, et al.
Pubblicazione: (2025)
Deep Controlled Learning for Inventory Control
di: Temizöz, Tarkan, et al.
Pubblicazione: (2020)
di: Temizöz, Tarkan, et al.
Pubblicazione: (2020)
Interpretable Imitation Learning via Generative Adversarial STL Inference and Control
di: Liu, Wenliang, et al.
Pubblicazione: (2024)
di: Liu, Wenliang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Euclidean, Projective, Conformal: Choosing a Geometric Algebra for Equivariant Transformers
di: de Haan, Pim, et al.
Pubblicazione: (2023) -
Does equivariance matter at scale?
di: Brehmer, Johann, et al.
Pubblicazione: (2024) -
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
di: Grislain, Clémence, et al.
Pubblicazione: (2024) -
Information-driven Affordance Discovery for Efficient Robotic Manipulation
di: Mazzaglia, Pietro, et al.
Pubblicazione: (2024) -
A Bayesian Solution To The Imitation Gap
di: Vuorio, Risto, et al.
Pubblicazione: (2024)