Real-time Sampling-based Model Predictive Control based on Reverse Kullback-Leibler Divergence and Its Adaptive Acceleration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kobayashi, Taisuke, Fukumoto, Kota |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Real‐Time Sampling‐Based Model Predictive Control Based on Reverse Kullback–Leibler Divergence and Its Adaptive Acceleration
von: Taisuke Kobayashi, et al.
Veröffentlicht: (2026)
von: Taisuke Kobayashi, et al.
Veröffentlicht: (2026)
LiRA: Light-Robust Adversary for Model-based Reinforcement Learning in Real World
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
Robust Phase Retrieval via Reverse Kullback-Leibler Divergence
von: Choudhury, Nazia Afroz, et al.
Veröffentlicht: (2022)
von: Choudhury, Nazia Afroz, et al.
Veröffentlicht: (2022)
Towards Autonomous Driving of Personal Mobility with Small and Noisy Dataset using Tsallis-statistics-based Behavioral Cloning
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2021)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2021)
Variational Adaptive Noise and Dropout towards Stable Recurrent Neural Networks
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2025)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2025)
Diversity-Aware Reverse Kullback-Leibler Divergence for Large Language Model Distillation
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
Generalized Kullback-Leibler Divergence Loss
von: Cui, Jiequan, et al.
Veröffentlicht: (2025)
von: Cui, Jiequan, et al.
Veröffentlicht: (2025)
Decoupled Kullback-Leibler Divergence Loss
von: Cui, Jiequan, et al.
Veröffentlicht: (2023)
von: Cui, Jiequan, et al.
Veröffentlicht: (2023)
Intentionally-underestimated Value Function at Terminal State for Temporal-difference Learning with Mis-designed Reward
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
Better Estimation of the Kullback--Leibler Divergence Between Language Models
von: Amini, Afra, et al.
Veröffentlicht: (2025)
von: Amini, Afra, et al.
Veröffentlicht: (2025)
Evaluating Earth-Observing Satellite Sampling Effectiveness Using Kullback-Leibler Divergence
von: Esmaeili, Negin, et al.
Veröffentlicht: (2025)
von: Esmaeili, Negin, et al.
Veröffentlicht: (2025)
The Hellinger Bounds on the Kullback-Leibler Divergence and the Bernstein Norm
von: Kaji, Tetsuya
Veröffentlicht: (2026)
von: Kaji, Tetsuya
Veröffentlicht: (2026)
Fractional Programming for Kullback-Leibler Divergence in Hypothesis Testing
von: Park, Jeongwoo, et al.
Veröffentlicht: (2026)
von: Park, Jeongwoo, et al.
Veröffentlicht: (2026)
Statistically Adaptive Differential Protection for AC Microgrids Based on Kullback-Leibler Divergence
von: Torkashvand, Shahab Moradi, et al.
Veröffentlicht: (2025)
von: Torkashvand, Shahab Moradi, et al.
Veröffentlicht: (2025)
Parallelizing MCMC with Machine Learning Classifier and Its Criterion Based on Kullback-Leibler Divergence
von: Matsumoto, Tomoki
Veröffentlicht: (2024)
von: Matsumoto, Tomoki
Veröffentlicht: (2024)
Design of Restricted Normalizing Flow towards Arbitrary Stochastic Policy with Computational Efficiency
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
Rethinking Kullback-Leibler Divergence in Knowledge Distillation for Large Language Models
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
Sampling-based Model Predictive Control Leveraging Parallelizable Physics Simulations
von: Pezzato, Corrado, et al.
Veröffentlicht: (2023)
von: Pezzato, Corrado, et al.
Veröffentlicht: (2023)
Wasserstein Distance Rivals Kullback-Leibler Divergence for Knowledge Distillation
von: Lv, Jiaming, et al.
Veröffentlicht: (2024)
von: Lv, Jiaming, et al.
Veröffentlicht: (2024)
Central Limit Theorem on Symmetric Kullback-Leibler (KL) Divergence
von: Rojas, Helder, et al.
Veröffentlicht: (2024)
von: Rojas, Helder, et al.
Veröffentlicht: (2024)
Fast Kd-trees for the Kullback--Leibler Divergence and other Decomposable Bregman Divergences
von: Pham, Tuyen, et al.
Veröffentlicht: (2025)
von: Pham, Tuyen, et al.
Veröffentlicht: (2025)
Adaptive Learning-based Model Predictive Control Strategy for Drift Vehicles
von: Zhou, Bei, et al.
Veröffentlicht: (2025)
von: Zhou, Bei, et al.
Veröffentlicht: (2025)
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
Stein-based Optimization of Sampling Distributions in Model Predictive Path Integral Control
von: Aldrich, Jace, et al.
Veröffentlicht: (2025)
von: Aldrich, Jace, et al.
Veröffentlicht: (2025)
Quantum Causal Discovery via Amplitude Estimation of Kullback-Leibler Divergence
von: Sodagari, Shabnam
Veröffentlicht: (2026)
von: Sodagari, Shabnam
Veröffentlicht: (2026)
A New Estimator of Kullback--Leibler Divergence via Shannon Entropy
von: Cadirci, Mehmet Siddik, et al.
Veröffentlicht: (2026)
von: Cadirci, Mehmet Siddik, et al.
Veröffentlicht: (2026)
Domains as Objectives: Domain-Uncertainty-Aware Policy Optimization through Explicit Multi-Domain Convex Coverage Set Learning
von: Ilboudo, Wendyam Eric Lionel, et al.
Veröffentlicht: (2024)
von: Ilboudo, Wendyam Eric Lionel, et al.
Veröffentlicht: (2024)
Differentiable Annealed Importance Sampling Minimizes The Symmetrized Kullback-Leibler Divergence Between Initial and Target Distribution
von: Zenn, Johannes, et al.
Veröffentlicht: (2024)
von: Zenn, Johannes, et al.
Veröffentlicht: (2024)
Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings
von: Kishino, Ryo, et al.
Veröffentlicht: (2025)
von: Kishino, Ryo, et al.
Veröffentlicht: (2025)
Conformalized Non-uniform Sampling Strategies for Accelerated Sampling-based Motion Planning
von: Natraj, Shubham, et al.
Veröffentlicht: (2025)
von: Natraj, Shubham, et al.
Veröffentlicht: (2025)
Expected Kullback-Leibler-based characterizations of score-driven updates
von: de Punder, Ramon, et al.
Veröffentlicht: (2024)
von: de Punder, Ramon, et al.
Veröffentlicht: (2024)
Conic Reformulations for Kullback-Leibler Divergence Constrained Distributionally Robust Optimization and Applications
von: Kocuk, Burak
Veröffentlicht: (2020)
von: Kocuk, Burak
Veröffentlicht: (2020)
Relativistic Limits of Decoding: Critical Divergence of Kullback-Leibler Information and Free Energy
von: Tsuruyama, Tatsuaki
Veröffentlicht: (2025)
von: Tsuruyama, Tatsuaki
Veröffentlicht: (2025)
Nearly Minimax Discrete Distribution Estimation in Kullback-Leibler Divergence with High Probability
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2025)
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2025)
Learning with Delayed Payoffs in Population Games using Kullback-Leibler Divergence Regularization
von: Park, Shinkyu, et al.
Veröffentlicht: (2023)
von: Park, Shinkyu, et al.
Veröffentlicht: (2023)
Entropy and the Kullback-Leibler Divergence for Bayesian Networks: Computational Complexity and Efficient Implementation
von: Scutari, Marco
Veröffentlicht: (2023)
von: Scutari, Marco
Veröffentlicht: (2023)
Relaxed Triangle Inequality for Kullback-Leibler Divergence Between Multivariate Gaussian Distributions
von: Xiao, Shiji, et al.
Veröffentlicht: (2026)
von: Xiao, Shiji, et al.
Veröffentlicht: (2026)
Adaptive Complexity Model Predictive Control
von: Norby, Joseph, et al.
Veröffentlicht: (2022)
von: Norby, Joseph, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Real‐Time Sampling‐Based Model Predictive Control Based on Reverse Kullback–Leibler Divergence and Its Adaptive Acceleration
von: Taisuke Kobayashi, et al.
Veröffentlicht: (2026) -
LiRA: Light-Robust Adversary for Model-based Reinforcement Learning in Real World
von: Kobayashi, Taisuke
Veröffentlicht: (2024) -
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2022) -
Robust Phase Retrieval via Reverse Kullback-Leibler Divergence
von: Choudhury, Nazia Afroz, et al.
Veröffentlicht: (2022) -
Towards Autonomous Driving of Personal Mobility with Small and Noisy Dataset using Tsallis-statistics-based Behavioral Cloning
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2021)