Robust RL Control for Bipedal Locomotion with Closed Kinematic Chains
Fuente:
arXiv
Salvato in:
| Autori principali: | Maslennikov, Egor, Zaliaev, Eduard, Dudorov, Nikita, Shamanin, Oleg, Dmitry, Karanov, Afanasev, Gleb, Burkov, Alexey, Lygin, Egor, Nedelchev, Simeon, Ponomarev, Evgeny |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How to model Human Actions distribution with Event Sequence Data
di: Surkov, Egor, et al.
Pubblicazione: (2025)
di: Surkov, Egor, et al.
Pubblicazione: (2025)
Heuristic Step Planning for Learning Dynamic Bipedal Locomotion: A Comparative Study of Model-Based and Model-Free Approaches
di: Suliman, William, et al.
Pubblicazione: (2025)
di: Suliman, William, et al.
Pubblicazione: (2025)
Not Yet: Humans Outperform LLMs in a Colonel Blotto Tournament
di: Dagaev, Dmitry, et al.
Pubblicazione: (2026)
di: Dagaev, Dmitry, et al.
Pubblicazione: (2026)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
di: Cherepanov, Egor, et al.
Pubblicazione: (2024)
di: Cherepanov, Egor, et al.
Pubblicazione: (2024)
Full‐Body Optimization‐Based Bipedal Walking Control With Task‐Space Inverse Dynamics and Virtual Constraints
di: William Suliman, et al.
Pubblicazione: (2025)
di: William Suliman, et al.
Pubblicazione: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
di: Petrov, Egor, et al.
Pubblicazione: (2025)
di: Petrov, Egor, et al.
Pubblicazione: (2025)
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
di: Skopin, Egor, et al.
Pubblicazione: (2026)
di: Skopin, Egor, et al.
Pubblicazione: (2026)
Testing Higgs $CP$ properties at the CEPC with an additional ISR parameter
di: Drutskoy, Alexey, et al.
Pubblicazione: (2025)
di: Drutskoy, Alexey, et al.
Pubblicazione: (2025)
Optimal Control for Minimizing Inescapable Ellipsoids in Linear Periodically Time-Varying Systems Under Bounded Disturbances
di: Dogadin, Egor, et al.
Pubblicazione: (2025)
di: Dogadin, Egor, et al.
Pubblicazione: (2025)
Energy harvesting coil for circularly polarized fields in magnetic resonance imaging
di: Seregin, Pavel, et al.
Pubblicazione: (2021)
di: Seregin, Pavel, et al.
Pubblicazione: (2021)
Self-dual classical higher-spin multicopy
di: Misuna, Nikita, et al.
Pubblicazione: (2026)
di: Misuna, Nikita, et al.
Pubblicazione: (2026)
Magnetisation moment of a bounded 3D sample: asymptotic recovery from planar measurements on a large disk
di: Ponomarev, Dmitry
Pubblicazione: (2022)
di: Ponomarev, Dmitry
Pubblicazione: (2022)
A Method to Extrapolate the Data for the Inverse Magnetisation Problem with a Planar Sample
di: Ponomarev, Dmitry
Pubblicazione: (2024)
di: Ponomarev, Dmitry
Pubblicazione: (2024)
Chiral higher-spin double copy
di: Ponomarev, Dmitry
Pubblicazione: (2024)
di: Ponomarev, Dmitry
Pubblicazione: (2024)
RL-augmented MPC Framework for Agile and Robust Bipedal Footstep Locomotion Planning and Control
di: Bang, Seung Hyeon, et al.
Pubblicazione: (2024)
di: Bang, Seung Hyeon, et al.
Pubblicazione: (2024)
RL-augmented Adaptive Model Predictive Control for Bipedal Locomotion over Challenging Terrain
di: Kamohara, Junnosuke, et al.
Pubblicazione: (2025)
di: Kamohara, Junnosuke, et al.
Pubblicazione: (2025)
Affine models for Noetherian schemes
di: Gorinov, Alexey G., et al.
Pubblicazione: (2026)
di: Gorinov, Alexey G., et al.
Pubblicazione: (2026)
Generative Flow Networks as Entropy-Regularized RL
di: Tiapkin, Daniil, et al.
Pubblicazione: (2023)
di: Tiapkin, Daniil, et al.
Pubblicazione: (2023)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
Color Mismatches in Stereoscopic Video: Real-World Dataset and Deep Correction Method
di: Chistov, Egor, et al.
Pubblicazione: (2023)
di: Chistov, Egor, et al.
Pubblicazione: (2023)
QuantNAS for super resolution: searching for efficient quantization-friendly architectures against quantization noise
di: Shvetsov, Egor, et al.
Pubblicazione: (2022)
di: Shvetsov, Egor, et al.
Pubblicazione: (2022)
Achieving Precise and Reliable Locomotion with Differentiable Simulation-Based System Identification
di: Kovalev, Vyacheslav, et al.
Pubblicazione: (2025)
di: Kovalev, Vyacheslav, et al.
Pubblicazione: (2025)
EBES: Easy Benchmarking for Event Sequences
di: Osin, Dmitry, et al.
Pubblicazione: (2024)
di: Osin, Dmitry, et al.
Pubblicazione: (2024)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
di: Shchendrigin, Oleg, et al.
Pubblicazione: (2026)
di: Shchendrigin, Oleg, et al.
Pubblicazione: (2026)
VIMPPI: Enhancing Model Predictive Path Integral Control with Variational Integration for Underactuated Systems
di: Alentev, Igor, et al.
Pubblicazione: (2025)
di: Alentev, Igor, et al.
Pubblicazione: (2025)
Efficient Neural Controlled Differential Equations via Attentive Kernel Smoothing
di: Serov, Egor, et al.
Pubblicazione: (2026)
di: Serov, Egor, et al.
Pubblicazione: (2026)
Meson mass spectrum in Ising Field Theory
di: Litvinov, Alexey, et al.
Pubblicazione: (2025)
di: Litvinov, Alexey, et al.
Pubblicazione: (2025)
Optimal Control for Discrete-Time Systems under Bounded Disturbances
di: Dogadin, Egor, et al.
Pubblicazione: (2024)
di: Dogadin, Egor, et al.
Pubblicazione: (2024)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
di: Polyakov, Igor, et al.
Pubblicazione: (2025)
di: Polyakov, Igor, et al.
Pubblicazione: (2025)
Electromagnetic diffraction and bidirectional plasmon launching in partially gated 2d systems
di: Moiseenko, Ilia, et al.
Pubblicazione: (2025)
di: Moiseenko, Ilia, et al.
Pubblicazione: (2025)
Tensor-network approach to quantum optical state evolution beyond the Fock basis
di: Kapridov, Nikolay, et al.
Pubblicazione: (2025)
di: Kapridov, Nikolay, et al.
Pubblicazione: (2025)
Geometry of symplectic flux and Lagrangian torus fibrations
di: Egor Shelukhin, et al.
Pubblicazione: (2024)
di: Egor Shelukhin, et al.
Pubblicazione: (2024)
Oscillatory integrals with polynomial phase and regularity of distributions
di: Kosov, Egor
Pubblicazione: (2025)
di: Kosov, Egor
Pubblicazione: (2025)
Groups with $\mathsf A_\ell$-commutator relations
di: Voronetsky, Egor
Pubblicazione: (2022)
di: Voronetsky, Egor
Pubblicazione: (2022)
Root graded groups revisited
di: Voronetsky, Egor
Pubblicazione: (2024)
di: Voronetsky, Egor
Pubblicazione: (2024)
Locally isotropic Steinberg groups II. Schur multipliers
di: Voronetsky, Egor
Pubblicazione: (2025)
di: Voronetsky, Egor
Pubblicazione: (2025)
Weyl elements in isotropic reductive groups
di: Voronetsky, Egor
Pubblicazione: (2026)
di: Voronetsky, Egor
Pubblicazione: (2026)
Dataset for the article "Capping agent control over the physicochemical and antibacterial properties of ZnO nanoparticles".
di: Ukraintsev, Egor
Pubblicazione: (2024)
di: Ukraintsev, Egor
Pubblicazione: (2024)
Das »System Isler«
di: Lykov, Egor
Pubblicazione: (2025)
di: Lykov, Egor
Pubblicazione: (2025)
Documenti analoghi
-
How to model Human Actions distribution with Event Sequence Data
di: Surkov, Egor, et al.
Pubblicazione: (2025) -
Heuristic Step Planning for Learning Dynamic Bipedal Locomotion: A Comparative Study of Model-Based and Model-Free Approaches
di: Suliman, William, et al.
Pubblicazione: (2025) -
Not Yet: Humans Outperform LLMs in a Colonel Blotto Tournament
di: Dagaev, Dmitry, et al.
Pubblicazione: (2026) -
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
di: Cherepanov, Egor, et al.
Pubblicazione: (2024) -
Full‐Body Optimization‐Based Bipedal Walking Control With Task‐Space Inverse Dynamics and Virtual Constraints
di: William Suliman, et al.
Pubblicazione: (2025)