Hyperspherical Normalization for Scalable Deep Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Hojoon, Lee, Youngdo, Seno, Takuma, Kim, Donghu, Stone, Peter, Choo, Jaegul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
by: Kim, Donghu, et al.
Published: (2024)
by: Kim, Donghu, et al.
Published: (2024)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
by: Kim, Donghu, et al.
Published: (2026)
by: Kim, Donghu, et al.
Published: (2026)
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
by: Kim, Hyunseung, et al.
Published: (2024)
by: Kim, Hyunseung, et al.
Published: (2024)
Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise Networks
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
by: Lee, Hojoon, et al.
Published: (2025)
by: Lee, Hojoon, et al.
Published: (2025)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
by: Hwang, Dongyoon, et al.
Published: (2024)
by: Hwang, Dongyoon, et al.
Published: (2024)
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
by: Hwang, Dongyoon, et al.
Published: (2025)
by: Hwang, Dongyoon, et al.
Published: (2025)
When Model Meets New Normals: Test-time Adaptation for Unsupervised Time-series Anomaly Detection
by: Kim, Dongmin, et al.
Published: (2023)
by: Kim, Dongmin, et al.
Published: (2023)
FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability-Plasticity Tradeoff
by: Han, Isaac, et al.
Published: (2026)
by: Han, Isaac, et al.
Published: (2026)
A Super-human Vision-based Reinforcement Learning Agent for Autonomous Racing in Gran Turismo
by: Vasco, Miguel, et al.
Published: (2024)
by: Vasco, Miguel, et al.
Published: (2024)
Dynamic Mixture of Experts Against Severe Distribution Shifts
by: Kim, Donghu
Published: (2025)
by: Kim, Donghu
Published: (2025)
EPIC: Effective Prompting for Imbalanced-Class Data Synthesis in Tabular Data Classification via Large Language Models
by: Kim, Jinhee, et al.
Published: (2024)
by: Kim, Jinhee, et al.
Published: (2024)
Revisiting LLMs as Zero-Shot Time-Series Forecasters: Small Noise Can Break Large Models
by: Park, Junwoo, et al.
Published: (2025)
by: Park, Junwoo, et al.
Published: (2025)
nGPT: Normalized Transformer with Representation Learning on the Hypersphere
by: Loshchilov, Ilya, et al.
Published: (2024)
by: Loshchilov, Ilya, et al.
Published: (2024)
Uncertainty Estimation via Hyperspherical Confidence Mapping
by: Choi, Eunseo, et al.
Published: (2026)
by: Choi, Eunseo, et al.
Published: (2026)
RDA: Reward Design Agent for Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2026)
by: Lee, Hojoon, et al.
Published: (2026)
Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target
by: Kim, Taesan, et al.
Published: (2026)
by: Kim, Taesan, et al.
Published: (2026)
Self-Supervised Contrastive Learning for Long-term Forecasting
by: Park, Junwoo, et al.
Published: (2024)
by: Park, Junwoo, et al.
Published: (2024)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
by: Cho, Hojun, et al.
Published: (2025)
by: Cho, Hojun, et al.
Published: (2025)
O$n$ Learning Deep O($n$)-Equivariant Hyperspheres
by: Melnyk, Pavlo, et al.
Published: (2023)
by: Melnyk, Pavlo, et al.
Published: (2023)
DecDEC: A Systems Approach to Advancing Low-Bit LLM Quantization
by: Park, Yeonhong, et al.
Published: (2024)
by: Park, Yeonhong, et al.
Published: (2024)
Deep Reinforcement Learning in Parameterized Action Space
by: Hausknecht, Matthew, et al.
Published: (2015)
by: Hausknecht, Matthew, et al.
Published: (2015)
MagiCapture: High-Resolution Multi-Concept Portrait Customization
by: Hyung, Junha, et al.
Published: (2023)
by: Hyung, Junha, et al.
Published: (2023)
Enhancing Diversity in Bayesian Deep Learning via Hyperspherical Energy Minimization of CKA
by: Smerkous, David, et al.
Published: (2024)
by: Smerkous, David, et al.
Published: (2024)
Single-shot reconstruction of three-dimensional morphology of biological cells in digital holographic microscopy using a physics-driven neural network
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Deep Orthogonal Hypersphere Compression for Anomaly Detection
by: Zhang, Yunhe, et al.
Published: (2023)
by: Zhang, Yunhe, et al.
Published: (2023)
Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow
by: Chao, Chen-Hao, et al.
Published: (2024)
by: Chao, Chen-Hao, et al.
Published: (2024)
Sparse autoencoders reveal selective remapping of visual concepts during adaptation
by: Lim, Hyesu, et al.
Published: (2024)
by: Lim, Hyesu, et al.
Published: (2024)
PHUMA: Physically-Grounded Humanoid Locomotion Dataset
by: Lee, Kyungmin, et al.
Published: (2025)
by: Lee, Kyungmin, et al.
Published: (2025)
Angular Regularization for Positive-Unlabeled Learning on the Hypersphere
by: Sevetlidis, Vasileios, et al.
Published: (2025)
by: Sevetlidis, Vasileios, et al.
Published: (2025)
Constrained Machine Learning Through Hyperspherical Representation
by: Signorelli, Gaetano, et al.
Published: (2025)
by: Signorelli, Gaetano, et al.
Published: (2025)
Deep Reinforcement Learning for Robotics: A Survey of Real-World Successes
by: Tang, Chen, et al.
Published: (2024)
by: Tang, Chen, et al.
Published: (2024)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
HYPO: Hyperspherical Out-of-Distribution Generalization
by: Bai, Haoyue, et al.
Published: (2024)
by: Bai, Haoyue, et al.
Published: (2024)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
by: Jeong, Hawon, et al.
Published: (2024)
by: Jeong, Hawon, et al.
Published: (2024)
AgenticCache: Cache-Driven Asynchronous Planning for Embodied AI Agents
by: Kim, Hojoon, et al.
Published: (2026)
by: Kim, Hojoon, et al.
Published: (2026)
Semi-gradient DICE for Offline Constrained Reinforcement Learning
by: Kim, Woosung, et al.
Published: (2025)
by: Kim, Woosung, et al.
Published: (2025)
Adaptive $Q$-Aid for Conditional Supervised Learning in Offline Reinforcement Learning
by: Kim, Jeonghye, et al.
Published: (2024)
by: Kim, Jeonghye, et al.
Published: (2024)
RL makes MLLMs see better than SFT
by: Song, Junha, et al.
Published: (2025)
by: Song, Junha, et al.
Published: (2025)
Similar Items
-
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024) -
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
by: Kim, Donghu, et al.
Published: (2024) -
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
by: Kim, Donghu, et al.
Published: (2026) -
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
by: Kim, Hyunseung, et al.
Published: (2024) -
Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise Networks
by: Lee, Hojoon, et al.
Published: (2024)