Frictional Q-Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Hyunwoo, Lee, Hyo Kyung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PULSE-ICU: A Pretrained Unified Long-Sequence Encoder for Multi-task Prediction in Intensive Care Units
by: Jang, Sejeong, et al.
Published: (2025)
by: Jang, Sejeong, et al.
Published: (2025)
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
by: Lee, Hyunwoo, et al.
Published: (2025)
by: Lee, Hyunwoo, et al.
Published: (2025)
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
by: Lee, Hyunwoo, et al.
Published: (2024)
by: Lee, Hyunwoo, et al.
Published: (2024)
Adaptive Sparsified Graph Learning Framework for Vessel Behavior Anomalies
by: Kim, Jeehong, et al.
Published: (2025)
by: Kim, Jeehong, et al.
Published: (2025)
Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
by: Kim, Hyunwoo, et al.
Published: (2025)
by: Kim, Hyunwoo, et al.
Published: (2025)
Belief Aided Navigation using Bayesian Reinforcement Learning for Avoiding Humans in Blind Spots
by: Kim, Jinyeob, et al.
Published: (2024)
by: Kim, Jinyeob, et al.
Published: (2024)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
by: Yeom, Junghyuk, et al.
Published: (2024)
by: Yeom, Junghyuk, et al.
Published: (2024)
Inversion-based Latent Bayesian Optimization
by: Chu, Jaewon, et al.
Published: (2024)
by: Chu, Jaewon, et al.
Published: (2024)
Graph Elicitation for Guiding Multi-Step Reasoning in Large Language Models
by: Park, Jinyoung, et al.
Published: (2023)
by: Park, Jinyoung, et al.
Published: (2023)
Chunk-Guided Q-Learning
by: Song, Gwanwoo, et al.
Published: (2026)
by: Song, Gwanwoo, et al.
Published: (2026)
Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation
by: Lee, Seokju, et al.
Published: (2026)
by: Lee, Seokju, et al.
Published: (2026)
Suppressing Overestimation in Q-Learning through Adversarial Behaviors
by: Lee, HyeAnn, et al.
Published: (2023)
by: Lee, HyeAnn, et al.
Published: (2023)
Periodic Regularized Q-Learning
by: Yang, Hyukjun, et al.
Published: (2026)
by: Yang, Hyukjun, et al.
Published: (2026)
Wafer-Level Etch Spatial Profiling for Process Monitoring from Time-Series with Time-LLM
by: Kim, Hyunwoo, et al.
Published: (2026)
by: Kim, Hyunwoo, et al.
Published: (2026)
Incremental Learning of Retrievable Skills For Efficient Continual Task Adaptation
by: Lee, Daehee, et al.
Published: (2024)
by: Lee, Daehee, et al.
Published: (2024)
Latent Bayesian Optimization via Autoregressive Normalizing Flows
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
by: Park, Kwanyoung, et al.
Published: (2024)
by: Park, Kwanyoung, et al.
Published: (2024)
Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL
by: Kim, Sung-Hyun, et al.
Published: (2025)
by: Kim, Sung-Hyun, et al.
Published: (2025)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
by: Jain, Ayush, et al.
Published: (2024)
by: Jain, Ayush, et al.
Published: (2024)
Safe-Support Q-Learning: Learning without Unsafe Exploration
by: Lim, Yeeun, et al.
Published: (2026)
by: Lim, Yeeun, et al.
Published: (2026)
NegMerge: Sign-Consensual Weight Merging for Machine Unlearning
by: Kim, Hyo Seo, et al.
Published: (2024)
by: Kim, Hyo Seo, et al.
Published: (2024)
ScaleDiff: Higher-Resolution Image Synthesis via Efficient and Model-Agnostic Diffusion
by: Koh, Sungho, et al.
Published: (2025)
by: Koh, Sungho, et al.
Published: (2025)
Trust Region Q Adjoint Matching
by: Dong, Yonghoon, et al.
Published: (2026)
by: Dong, Yonghoon, et al.
Published: (2026)
FragFM: Hierarchical Framework for Efficient Molecule Generation via Fragment-Level Discrete Flow Matching
by: Lee, Joongwon, et al.
Published: (2025)
by: Lee, Joongwon, et al.
Published: (2025)
Spatio-Temporal Graphs Beyond Grids: Benchmark for Maritime Anomaly Detection
by: Kim, Jeehong, et al.
Published: (2025)
by: Kim, Jeehong, et al.
Published: (2025)
Learning the Model While Learning Q: Finite-Time Sample Complexity of Online SyncMBQ
by: Lim, Han-Dong, et al.
Published: (2024)
by: Lim, Han-Dong, et al.
Published: (2024)
ST-MTM: Masked Time Series Modeling with Seasonal-Trend Decomposition for Time Series Forecasting
by: Seo, Hyunwoo, et al.
Published: (2025)
by: Seo, Hyunwoo, et al.
Published: (2025)
SPQR: Controlling Q-ensemble Independence with Spiked Random Model for Reinforcement Learning
by: Lee, Dohyeok, et al.
Published: (2024)
by: Lee, Dohyeok, et al.
Published: (2024)
Adaptive Friction in Deep Learning: Enhancing Optimizers with Sigmoid and Tanh Function
by: Zheng, Hongye, et al.
Published: (2024)
by: Zheng, Hongye, et al.
Published: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
Semi-Supervised Graph Representation Learning with Human-centric Explanation for Predicting Fatty Liver Disease
by: Kim, So Yeon, et al.
Published: (2024)
by: Kim, So Yeon, et al.
Published: (2024)
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic
by: Lee, Jeong Woon, et al.
Published: (2026)
by: Lee, Jeong Woon, et al.
Published: (2026)
FlowPath: Learning Data-Driven Manifolds with Invertible Flows for Robust Irregularly-sampled Time Series Classification
by: Oh, YongKyung, et al.
Published: (2025)
by: Oh, YongKyung, et al.
Published: (2025)
Robust Policy Learning via Offline Skill Diffusion
by: Kim, Woo Kyung, et al.
Published: (2024)
by: Kim, Woo Kyung, et al.
Published: (2024)
DIAR: Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation
by: Park, Jaehyun, et al.
Published: (2024)
by: Park, Jaehyun, et al.
Published: (2024)
Prism: Spectral Parameter Sharing for Multi-Agent Reinforcement Learning
by: Kim, Kyungbeom, et al.
Published: (2026)
by: Kim, Kyungbeom, et al.
Published: (2026)
A Scalable and Transferable Time Series Prediction Framework for Demand Forecasting
by: Park, Young-Jin, et al.
Published: (2024)
by: Park, Young-Jin, et al.
Published: (2024)
PhysioME: A Robust Multimodal Self-Supervised Framework for Physiological Signals with Missing Modalities
by: Lee, Cheol-Hui, et al.
Published: (2025)
by: Lee, Cheol-Hui, et al.
Published: (2025)
Pretraining a Shared Q-Network for Data-Efficient Offline Reinforcement Learning
by: Park, Jongchan, et al.
Published: (2025)
by: Park, Jongchan, et al.
Published: (2025)
Uncertainty-Resilient Multimodal Learning via Consistency-Guided Cross-Modal Transfer
by: Jang, Hyo-Jeong
Published: (2025)
by: Jang, Hyo-Jeong
Published: (2025)
Similar Items
-
PULSE-ICU: A Pretrained Unified Long-Sequence Encoder for Multi-task Prediction in Intensive Care Units
by: Jang, Sejeong, et al.
Published: (2025) -
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
by: Lee, Hyunwoo, et al.
Published: (2025) -
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
by: Lee, Hyunwoo, et al.
Published: (2024) -
Adaptive Sparsified Graph Learning Framework for Vessel Behavior Anomalies
by: Kim, Jeehong, et al.
Published: (2025) -
Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
by: Kim, Hyunwoo, et al.
Published: (2025)