Online Pre-Training for Offline-to-Online Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Shin, Yongjae, Kim, Jeonghye, Jung, Whiyoung, Hong, Sunghoon, Yoon, Deunsol, Jang, Youngsoo, Kim, Geonhyeong, Chae, Jongseong, Sung, Youngchul, Lee, Kanghoon, Lim, Woohyung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
di: Kim, Jeonghye, et al.
Pubblicazione: (2025)
di: Kim, Jeonghye, et al.
Pubblicazione: (2025)
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
Flow Actor-Critic for Offline Reinforcement Learning
di: Chae, Jongseong, et al.
Pubblicazione: (2026)
di: Chae, Jongseong, et al.
Pubblicazione: (2026)
Align While Search: Belief-Guided Exploratory Inference for World-Grounded Embodied Agents
di: Bae, Seohui, et al.
Pubblicazione: (2025)
di: Bae, Seohui, et al.
Pubblicazione: (2025)
Adaptive Action Chunking via Multi-Chunk Q Value Estimation
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
Adaptive $Q$-Aid for Conditional Supervised Learning in Offline Reinforcement Learning
di: Kim, Jeonghye, et al.
Pubblicazione: (2024)
di: Kim, Jeonghye, et al.
Pubblicazione: (2024)
LESSON: Learning to Integrate Exploration Strategies for Reinforcement Learning via an Option Framework
di: Kim, Woojun, et al.
Pubblicazione: (2023)
di: Kim, Woojun, et al.
Pubblicazione: (2023)
CADO: From Imitation to Cost Minimization for Heatmap-based Solvers in Combinatorial Optimization
di: Song, Hyungseok, et al.
Pubblicazione: (2026)
di: Song, Hyungseok, et al.
Pubblicazione: (2026)
Decision ConvFormer: Local Filtering in MetaFormer is Sufficient for Decision Making
di: Kim, Jeonghye, et al.
Pubblicazione: (2023)
di: Kim, Jeonghye, et al.
Pubblicazione: (2023)
Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach
di: Byeon, Woohyeon, et al.
Pubblicazione: (2025)
di: Byeon, Woohyeon, et al.
Pubblicazione: (2025)
ReflAct: World-Grounded Decision Making in LLM Agents via Goal-State Reflection
di: Kim, Jeonghye, et al.
Pubblicazione: (2025)
di: Kim, Jeonghye, et al.
Pubblicazione: (2025)
Unsupervised Training of Diffusion Models for Feasible Solution Generation in Neural Combinatorial Optimization
di: Hong, Seong-Hyun, et al.
Pubblicazione: (2024)
di: Hong, Seong-Hyun, et al.
Pubblicazione: (2024)
A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
di: Kim, JaeYoon, et al.
Pubblicazione: (2024)
di: Kim, JaeYoon, et al.
Pubblicazione: (2024)
STAIRS-Former: Spatio-Temporal Attention with Interleaved Recursive Structure Transformer for Offline Multi-task Multi-agent Reinforcement Learning
di: Jeon, Jiwon, et al.
Pubblicazione: (2026)
di: Jeon, Jiwon, et al.
Pubblicazione: (2026)
Locality-Aware Zero-Shot Human-Object Interaction Detection
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
Dynamic Multi-period Experts for Online Time Series Forecasting
di: Hong, Seungha, et al.
Pubblicazione: (2026)
di: Hong, Seungha, et al.
Pubblicazione: (2026)
Reward Dimension Reduction for Scalable Multi-Objective Reinforcement Learning
di: Park, Giseung, et al.
Pubblicazione: (2025)
di: Park, Giseung, et al.
Pubblicazione: (2025)
Pan-cancer gene set discovery via scRNA-seq for optimal deep learning based downstream tasks
di: Kim, Jong Hyun, et al.
Pubblicazione: (2024)
di: Kim, Jong Hyun, et al.
Pubblicazione: (2024)
Debiasing Online Preference Learning via Preference Feature Preservation
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
Self-Guided Robust Graph Structure Refinement
di: In, Yeonjun, et al.
Pubblicazione: (2024)
di: In, Yeonjun, et al.
Pubblicazione: (2024)
Training Robust Graph Neural Networks by Modeling Noise Dependencies
di: In, Yeonjun, et al.
Pubblicazione: (2025)
di: In, Yeonjun, et al.
Pubblicazione: (2025)
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images
di: Kim, Sangwook, et al.
Pubblicazione: (2025)
di: Kim, Sangwook, et al.
Pubblicazione: (2025)
Integrating Offline Pre-Training with Online Fine-Tuning: A Reinforcement Learning Approach for Robot Social Navigation
di: Su, Run, et al.
Pubblicazione: (2025)
di: Su, Run, et al.
Pubblicazione: (2025)
Lactate administration induces skeletal muscle synthesis by influencing Akt/mTOR and MuRF1 in non‐trained mice but not in trained mice
di: Sunghwan Kyun, et al.
Pubblicazione: (2024)
di: Sunghwan Kyun, et al.
Pubblicazione: (2024)
MINT: Molecularly Informed Training with Spatial Transcriptomics Supervision for Pathology Foundation Models
di: Lee, Minsoo, et al.
Pubblicazione: (2026)
di: Lee, Minsoo, et al.
Pubblicazione: (2026)
THEME: Enhancing Thematic Investing with Semantic Stock Representations and Temporal Dynamics
di: Lee, Hoyoung, et al.
Pubblicazione: (2025)
di: Lee, Hoyoung, et al.
Pubblicazione: (2025)
Personalized Autonomous Driving via Optimal Control with Clearance Constraints from Questionnaires
di: Lim, Yongjae, et al.
Pubblicazione: (2026)
di: Lim, Yongjae, et al.
Pubblicazione: (2026)
Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2023)
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2023)
Semantic Skill Grounding for Embodied Instruction-Following in Cross-Domain Environments
di: Shin, Sangwoo, et al.
Pubblicazione: (2024)
di: Shin, Sangwoo, et al.
Pubblicazione: (2024)
Oaken: Fast and Efficient LLM Serving with Online-Offline Hybrid KV Cache Quantization
di: Kim, Minsu, et al.
Pubblicazione: (2025)
di: Kim, Minsu, et al.
Pubblicazione: (2025)
Online Generic Event Boundary Detection
di: Jung, Hyungrok, et al.
Pubblicazione: (2025)
di: Jung, Hyungrok, et al.
Pubblicazione: (2025)
Online Optimization for Offline Safe Reinforcement Learning
di: Chemingui, Yassine, et al.
Pubblicazione: (2025)
di: Chemingui, Yassine, et al.
Pubblicazione: (2025)
The Three Regimes of Offline-to-Online Reinforcement Learning
di: Li, Lu, et al.
Pubblicazione: (2025)
di: Li, Lu, et al.
Pubblicazione: (2025)
Bridging Offline and Online Reinforcement Learning for LLMs
di: Lanchantin, Jack, et al.
Pubblicazione: (2025)
di: Lanchantin, Jack, et al.
Pubblicazione: (2025)
Information-Directed Offline-to-Online Reinforcement Learning
di: Chen, Keru
Pubblicazione: (2026)
di: Chen, Keru
Pubblicazione: (2026)
SU(4) Kondo Lattice in Semiconductor Moiré Materials
di: Kim, Sunghoon
Pubblicazione: (2025)
di: Kim, Sunghoon
Pubblicazione: (2025)
EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
The Max-Min Formulation of Multi-Objective Reinforcement Learning: From Theory to a Model-Free Algorithm
di: Park, Giseung, et al.
Pubblicazione: (2024)
di: Park, Giseung, et al.
Pubblicazione: (2024)
Semantic Diversity-aware Prototype-based Learning for Unbiased Scene Graph Generation
di: Jeon, Jaehyeong, et al.
Pubblicazione: (2024)
di: Jeon, Jaehyeong, et al.
Pubblicazione: (2024)
Functional Analysis of the Pepper RING ‐Type E3 Ligase CaANKR1 Involved in Drought Stress Tolerance via Modulation of Abscisic Acid Signaling
di: Mirim Kim, et al.
Pubblicazione: (2025)
di: Mirim Kim, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
di: Kim, Jeonghye, et al.
Pubblicazione: (2025) -
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
di: Shin, Yongjae, et al.
Pubblicazione: (2026) -
Flow Actor-Critic for Offline Reinforcement Learning
di: Chae, Jongseong, et al.
Pubblicazione: (2026) -
Align While Search: Belief-Guided Exploratory Inference for World-Grounded Embodied Agents
di: Bae, Seohui, et al.
Pubblicazione: (2025) -
Adaptive Action Chunking via Multi-Chunk Q Value Estimation
di: Shin, Yongjae, et al.
Pubblicazione: (2026)