FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Donghu, Lee, Youngdo, Park, Minho, Kim, Kinam, Nahendra, I Made Aswin, Seno, Takuma, Min, Sehee, Palenicek, Daniel, Vogt, Florian, Kragic, Danica, Peters, Jan, Choo, Jaegul, Lee, Hojoon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hyperspherical Normalization for Scalable Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2025)
by: Lee, Hojoon, et al.
Published: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
by: Kim, Donghu, et al.
Published: (2024)
by: Kim, Donghu, et al.
Published: (2024)
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
by: Kim, Hyunseung, et al.
Published: (2024)
by: Kim, Hyunseung, et al.
Published: (2024)
ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models
by: Park, Minho, et al.
Published: (2025)
by: Park, Minho, et al.
Published: (2025)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise Networks
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
by: Hwang, Sungwon, et al.
Published: (2025)
by: Hwang, Sungwon, et al.
Published: (2025)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
by: Kim, Kinam, et al.
Published: (2025)
by: Kim, Kinam, et al.
Published: (2025)
PHUMA: Physically-Grounded Humanoid Locomotion Dataset
by: Lee, Kyungmin, et al.
Published: (2025)
by: Lee, Kyungmin, et al.
Published: (2025)
EgoX: Egocentric Video Generation from a Single Exocentric Video
by: Kang, Taewoong, et al.
Published: (2025)
by: Kang, Taewoong, et al.
Published: (2025)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
by: Hwang, Dongyoon, et al.
Published: (2024)
by: Hwang, Dongyoon, et al.
Published: (2024)
XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies
by: Palenicek, Daniel, et al.
Published: (2026)
by: Palenicek, Daniel, et al.
Published: (2026)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
by: Cho, Hojun, et al.
Published: (2025)
by: Cho, Hojun, et al.
Published: (2025)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
by: Hyung, Junha, et al.
Published: (2024)
by: Hyung, Junha, et al.
Published: (2024)
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
by: Hwang, Dongyoon, et al.
Published: (2025)
by: Hwang, Dongyoon, et al.
Published: (2025)
FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability-Plasticity Tradeoff
by: Han, Isaac, et al.
Published: (2026)
by: Han, Isaac, et al.
Published: (2026)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
by: Jeong, Hawon, et al.
Published: (2024)
by: Jeong, Hawon, et al.
Published: (2024)
Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target
by: Kim, Taesan, et al.
Published: (2026)
by: Kim, Taesan, et al.
Published: (2026)
Dynamic Mixture of Experts Against Severe Distribution Shifts
by: Kim, Donghu
Published: (2025)
by: Kim, Donghu
Published: (2025)
Scaling CrossQ with Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
by: Kim, Jeongho, et al.
Published: (2024)
by: Kim, Jeongho, et al.
Published: (2024)
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
by: Park, Minho, et al.
Published: (2024)
by: Park, Minho, et al.
Published: (2024)
BankMathBench: A Benchmark for Numerical Reasoning in Banking Scenarios
by: Lee, Yunseung, et al.
Published: (2026)
by: Lee, Yunseung, et al.
Published: (2026)
A new proof of non-Cohen-Macaulayness of Bertin's example
by: Seno, Takuma
Published: (2025)
by: Seno, Takuma
Published: (2025)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
by: Jin, Hoiyeong, et al.
Published: (2025)
by: Jin, Hoiyeong, et al.
Published: (2025)
EPIC: Effective Prompting for Imbalanced-Class Data Synthesis in Tabular Data Classification via Large Language Models
by: Kim, Jinhee, et al.
Published: (2024)
by: Kim, Jinhee, et al.
Published: (2024)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
by: Kim, Jinhee, et al.
Published: (2024)
by: Kim, Jinhee, et al.
Published: (2024)
TRG-planner: Traversal Risk Graph-Based Path Planning in Unstructured Environments for Safe and Efficient Navigation
by: Lee, Dongkyu, et al.
Published: (2025)
by: Lee, Dongkyu, et al.
Published: (2025)
DreamFLEX: Learning Fault-Aware Quadrupedal Locomotion Controller for Anomaly Situation in Rough Terrains
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
When Model Meets New Normals: Test-time Adaptation for Unsupervised Time-series Anomaly Detection
by: Kim, Dongmin, et al.
Published: (2023)
by: Kim, Dongmin, et al.
Published: (2023)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
by: Shim, Gyumin, et al.
Published: (2025)
by: Shim, Gyumin, et al.
Published: (2025)
Enabling Region-Specific Control via Lassos in Point-Based Colorization
by: Lee, Sanghyeon, et al.
Published: (2024)
by: Lee, Sanghyeon, et al.
Published: (2024)
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
by: Lee, Hojoon, et al.
Published: (2025)
by: Lee, Hojoon, et al.
Published: (2025)
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
by: Park, Minho, et al.
Published: (2025)
by: Park, Minho, et al.
Published: (2025)
Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method
by: Kim, Taehee, et al.
Published: (2026)
by: Kim, Taehee, et al.
Published: (2026)
OPRO: Orthogonal Panel-Relative Operators for Panel-Aware In-Context Image Generation
by: Lee, Sanghyeon, et al.
Published: (2026)
by: Lee, Sanghyeon, et al.
Published: (2026)
DreamWaQ++: Obstacle-Aware Quadrupedal Locomotion With Resilient Multi-Modal Reinforcement Learning
by: Nahrendra, I Made Aswin, et al.
Published: (2024)
by: Nahrendra, I Made Aswin, et al.
Published: (2024)
STRIDE‐Based Threat Modeling for Smart Hazard Analysis and Critical Control Points in the Korean Food Industry
by: Sehee Jung, et al.
Published: (2026)
by: Sehee Jung, et al.
Published: (2026)
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
by: Gwak, Daehoon, et al.
Published: (2024)
by: Gwak, Daehoon, et al.
Published: (2024)
Similar Items
-
Hyperspherical Normalization for Scalable Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2025) -
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
by: Kim, Donghu, et al.
Published: (2024) -
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
by: Kim, Hyunseung, et al.
Published: (2024) -
ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models
by: Park, Minho, et al.
Published: (2025) -
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)