Velocity-History-Based Soft Actor-Critic Tackling IROS'24 Competition "AI Olympics with RealAIGym"
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Faust, Tim Lukas, Maraqten, Habib, Aghadavoodi, Erfan, Belousov, Boris, Peters, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Learning for Robust Athletic Intelligence: Lessons from the 2nd 'AI Olympics with RealAIGym' Competition
von: Wiebe, Felix, et al.
Veröffentlicht: (2025)
von: Wiebe, Felix, et al.
Veröffentlicht: (2025)
AI Olympics challenge with Evolutionary Soft Actor Critic
von: Calì, Marco, et al.
Veröffentlicht: (2024)
von: Calì, Marco, et al.
Veröffentlicht: (2024)
TacEx: GelSight Tactile Simulation in Isaac Sim -- Combining Soft-Body and Visuotactile Simulators
von: Nguyen, Duc Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Duc Huy, et al.
Veröffentlicht: (2024)
MuJoCo MPC for Humanoid Control: Evaluation on HumanoidBench
von: Meser, Moritz, et al.
Veröffentlicht: (2024)
von: Meser, Moritz, et al.
Veröffentlicht: (2024)
In-Hand Object Pose Estimation via Visual-Tactile Fusion
von: Nonnengießer, Felix, et al.
Veröffentlicht: (2025)
von: Nonnengießer, Felix, et al.
Veröffentlicht: (2025)
Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards
von: Jolfaei, Erfan Aghadavoodi, et al.
Veröffentlicht: (2026)
von: Jolfaei, Erfan Aghadavoodi, et al.
Veröffentlicht: (2026)
IROS: A Dual-Process Architecture for Real-Time VLM-Based Indoor Navigation
von: Lee, Joonhee, et al.
Veröffentlicht: (2026)
von: Lee, Joonhee, et al.
Veröffentlicht: (2026)
Gradient Iterated Temporal-Difference Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2026)
von: Vincent, Théo, et al.
Veröffentlicht: (2026)
Revisiting Discrete Soft Actor-Critic
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
Safe Langevin Soft Actor Critic
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
Average-Reward Soft Actor-Critic
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Wasserstein Barycenter Soft Actor-Critic
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
AI-Olympics: Exploring the Generalization of Agents through Open Competitions
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
Learning Force Distribution Estimation for the GelSight Mini Optical Tactile Sensor Based on Finite Element Analysis
von: Helmut, Erik, et al.
Veröffentlicht: (2024)
von: Helmut, Erik, et al.
Veröffentlicht: (2024)
Distributional Soft Actor-Critic with Three Refinements
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
Generative Actor-Critic with Soft Bridge Policies
von: He, Ke, et al.
Veröffentlicht: (2026)
von: He, Ke, et al.
Veröffentlicht: (2026)
Distributional Soft Actor-Critic with Diffusion Policy
von: Liu, Tong, et al.
Veröffentlicht: (2025)
von: Liu, Tong, et al.
Veröffentlicht: (2025)
PAC-Bayesian Soft Actor-Critic Learning
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
The Role of Domain Randomization in Training Diffusion Policies for Whole-Body Humanoid Control
von: Kaidanov, Oleg, et al.
Veröffentlicht: (2024)
von: Kaidanov, Oleg, et al.
Veröffentlicht: (2024)
ISAACS: Iterative Soft Adversarial Actor-Critic for Safety
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2022)
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2022)
Unlocking the Potential of Soft Actor-Critic for Imitation Learning
von: Lessa, Nayari Marie, et al.
Veröffentlicht: (2025)
von: Lessa, Nayari Marie, et al.
Veröffentlicht: (2025)
SACn: Soft Actor-Critic with n-step Returns
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
MSSEAC: Multi‐State Soft Elastic Actor‐Critic
von: Yuwan Gu, et al.
Veröffentlicht: (2026)
von: Yuwan Gu, et al.
Veröffentlicht: (2026)
Real del Pezzo surfaces without points
von: Belousov, Grigory
Veröffentlicht: (2024)
von: Belousov, Grigory
Veröffentlicht: (2024)
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives
von: Asad, Reza, et al.
Veröffentlicht: (2025)
von: Asad, Reza, et al.
Veröffentlicht: (2025)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
Scalable Neighborhood-Based Multi-Agent Actor-Critic
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
One-dimensional quantum scattering from multiple Dirac delta potentials: A Python-based solution
von: Keshavarz, Erfan, et al.
Veröffentlicht: (2023)
von: Keshavarz, Erfan, et al.
Veröffentlicht: (2023)
XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies
von: Palenicek, Daniel, et al.
Veröffentlicht: (2026)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2026)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
Chunking the Critic: A Transformer-based Soft Actor-Critic with N-Step Returns
von: Tian, Dong, et al.
Veröffentlicht: (2025)
von: Tian, Dong, et al.
Veröffentlicht: (2025)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
Soft Actor-Critic with Backstepping-Pretrained DeepONet for control of PDEs
von: Wang, Chenchen, et al.
Veröffentlicht: (2025)
von: Wang, Chenchen, et al.
Veröffentlicht: (2025)
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
Effective Reinforcement Learning Control using Conservative Soft Actor-Critic
von: Shang, Zhiwei, et al.
Veröffentlicht: (2025)
von: Shang, Zhiwei, et al.
Veröffentlicht: (2025)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
Soft Actor-Critic with Beta Policy via Implicit Reparameterization Gradients
von: Della Libera, Luca
Veröffentlicht: (2024)
von: Della Libera, Luca
Veröffentlicht: (2024)
Actor-Critic without Actor
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reinforcement Learning for Robust Athletic Intelligence: Lessons from the 2nd 'AI Olympics with RealAIGym' Competition
von: Wiebe, Felix, et al.
Veröffentlicht: (2025) -
AI Olympics challenge with Evolutionary Soft Actor Critic
von: Calì, Marco, et al.
Veröffentlicht: (2024) -
TacEx: GelSight Tactile Simulation in Isaac Sim -- Combining Soft-Body and Visuotactile Simulators
von: Nguyen, Duc Huy, et al.
Veröffentlicht: (2024) -
MuJoCo MPC for Humanoid Control: Evaluation on HumanoidBench
von: Meser, Moritz, et al.
Veröffentlicht: (2024) -
In-Hand Object Pose Estimation via Visual-Tactile Fusion
von: Nonnengießer, Felix, et al.
Veröffentlicht: (2025)