SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Hojoon, Hwang, Dongyoon, Kim, Donghu, Kim, Hyunseung, Tai, Jun Jet, Subramanian, Kaushik, Wurman, Peter R., Choo, Jaegul, Stone, Peter, Seno, Takuma |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hyperspherical Normalization for Scalable Deep Reinforcement Learning
di: Lee, Hojoon, et al.
Pubblicazione: (2025)
di: Lee, Hojoon, et al.
Pubblicazione: (2025)
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
di: Lee, Hojoon, et al.
Pubblicazione: (2025)
di: Lee, Hojoon, et al.
Pubblicazione: (2025)
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
di: Kim, Hyunseung, et al.
Pubblicazione: (2024)
di: Kim, Hyunseung, et al.
Pubblicazione: (2024)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
di: Kim, Donghu, et al.
Pubblicazione: (2024)
di: Kim, Donghu, et al.
Pubblicazione: (2024)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
A Super-human Vision-based Reinforcement Learning Agent for Autonomous Racing in Gran Turismo
di: Vasco, Miguel, et al.
Pubblicazione: (2024)
di: Vasco, Miguel, et al.
Pubblicazione: (2024)
PHUMA: Physically-Grounded Humanoid Locomotion Dataset
di: Lee, Kyungmin, et al.
Pubblicazione: (2025)
di: Lee, Kyungmin, et al.
Pubblicazione: (2025)
Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise Networks
di: Lee, Hojoon, et al.
Pubblicazione: (2024)
di: Lee, Hojoon, et al.
Pubblicazione: (2024)
Automated Reward Design for Gran Turismo
di: Ma, Michel, et al.
Pubblicazione: (2025)
di: Ma, Michel, et al.
Pubblicazione: (2025)
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
di: Hwang, Dongyoon, et al.
Pubblicazione: (2025)
di: Hwang, Dongyoon, et al.
Pubblicazione: (2025)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
di: Kim, Donghu, et al.
Pubblicazione: (2026)
di: Kim, Donghu, et al.
Pubblicazione: (2026)
Out-of-Distribution Generalization with a SPARC: Racing 100 Unseen Vehicles with a Single Policy
di: Grooten, Bram, et al.
Pubblicazione: (2025)
di: Grooten, Bram, et al.
Pubblicazione: (2025)
The Trajectory Alignment Coefficient in Two Acts: From Reward Tuning to Reward Learning
di: Muslimani, Calarina, et al.
Pubblicazione: (2026)
di: Muslimani, Calarina, et al.
Pubblicazione: (2026)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
di: Cho, Hojun, et al.
Pubblicazione: (2025)
di: Cho, Hojun, et al.
Pubblicazione: (2025)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
di: Yun, Jooyeol, et al.
Pubblicazione: (2024)
di: Yun, Jooyeol, et al.
Pubblicazione: (2024)
Dynamic Mixture of Experts Against Severe Distribution Shifts
di: Kim, Donghu
Pubblicazione: (2025)
di: Kim, Donghu
Pubblicazione: (2025)
A new proof of non-Cohen-Macaulayness of Bertin's example
di: Seno, Takuma
Pubblicazione: (2025)
di: Seno, Takuma
Pubblicazione: (2025)
EPIC: Effective Prompting for Imbalanced-Class Data Synthesis in Tabular Data Classification via Large Language Models
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
VisualScratchpad: Inference-time Visual Concepts Analysis in Vision Language Models
di: Lim, Hyesu, et al.
Pubblicazione: (2026)
di: Lim, Hyesu, et al.
Pubblicazione: (2026)
FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability-Plasticity Tradeoff
di: Han, Isaac, et al.
Pubblicazione: (2026)
di: Han, Isaac, et al.
Pubblicazione: (2026)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
di: Jeong, Hawon, et al.
Pubblicazione: (2024)
di: Jeong, Hawon, et al.
Pubblicazione: (2024)
RL makes MLLMs see better than SFT
di: Song, Junha, et al.
Pubblicazione: (2025)
di: Song, Junha, et al.
Pubblicazione: (2025)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
di: Kim, Kinam, et al.
Pubblicazione: (2025)
di: Kim, Kinam, et al.
Pubblicazione: (2025)
When Model Meets New Normals: Test-time Adaptation for Unsupervised Time-series Anomaly Detection
di: Kim, Dongmin, et al.
Pubblicazione: (2023)
di: Kim, Dongmin, et al.
Pubblicazione: (2023)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method
di: Kim, Taehee, et al.
Pubblicazione: (2026)
di: Kim, Taehee, et al.
Pubblicazione: (2026)
ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models
di: Park, Minho, et al.
Pubblicazione: (2025)
di: Park, Minho, et al.
Pubblicazione: (2025)
VEGS: View Extrapolation of Urban Scenes in 3D Gaussian Splatting using Learned Priors
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
di: Choi, Jinho, et al.
Pubblicazione: (2025)
di: Choi, Jinho, et al.
Pubblicazione: (2025)
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models
di: Song, Junha, et al.
Pubblicazione: (2026)
di: Song, Junha, et al.
Pubblicazione: (2026)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
di: Jo, Kyungmin, et al.
Pubblicazione: (2024)
di: Jo, Kyungmin, et al.
Pubblicazione: (2024)
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
di: Cho, Wonwoo, et al.
Pubblicazione: (2024)
di: Cho, Wonwoo, et al.
Pubblicazione: (2024)
Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target
di: Kim, Taesan, et al.
Pubblicazione: (2026)
di: Kim, Taesan, et al.
Pubblicazione: (2026)
BankMathBench: A Benchmark for Numerical Reasoning in Banking Scenarios
di: Lee, Yunseung, et al.
Pubblicazione: (2026)
di: Lee, Yunseung, et al.
Pubblicazione: (2026)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hyperspherical Normalization for Scalable Deep Reinforcement Learning
di: Lee, Hojoon, et al.
Pubblicazione: (2025) -
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
di: Lee, Hojoon, et al.
Pubblicazione: (2025) -
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
di: Kim, Hyunseung, et al.
Pubblicazione: (2024) -
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
di: Kim, Donghu, et al.
Pubblicazione: (2024) -
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)