Saved in:
| Main Authors: | Chung, Hojun, Lee, Junseo, Kim, Minsoo, Kim, Dohyeong, Oh, Songhwai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.19715 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Offline Reinforcement Learning with Universal Horizon Models
by: Chung, Hojun, et al.
Published: (2026)
by: Chung, Hojun, et al.
Published: (2026)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
TRACED: Transition-aware Regret Approximation with Co-learnability for Environment Design
by: Cho, Geonwoo, et al.
Published: (2025)
by: Cho, Geonwoo, et al.
Published: (2025)
Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning
by: Kim, Junseok, et al.
Published: (2026)
by: Kim, Junseok, et al.
Published: (2026)
Stage-Wise Reward Shaping for Acrobatic Robots: A Constrained Multi-Objective Reinforcement Learning Approach
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Conflict-Averse Gradient Aggregation for Constrained Multi-Objective Reinforcement Learning
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Safe CoR: A Dual-Expert Approach to Integrating Imitation Learning and Safe Reinforcement Learning Using Constraint Rewards
by: Kwon, Hyeokjin, et al.
Published: (2024)
by: Kwon, Hyeokjin, et al.
Published: (2024)
Tidiness Score-Guided Monte Carlo Tree Search for Visual Tabletop Rearrangement
by: Kee, Hogun, et al.
Published: (2025)
by: Kee, Hogun, et al.
Published: (2025)
PreSto: An In-Storage Data Preprocessing System for Training Recommendation Models
by: Lee, Yunjae, et al.
Published: (2024)
by: Lee, Yunjae, et al.
Published: (2024)
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
by: Hwang, Junseo, et al.
Published: (2025)
by: Hwang, Junseo, et al.
Published: (2025)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
by: Cho, Taehyun, et al.
Published: (2024)
by: Cho, Taehyun, et al.
Published: (2024)
Policy-labeled Preference Learning: Is Preference Enough for RLHF?
by: Cho, Taehyun, et al.
Published: (2025)
by: Cho, Taehyun, et al.
Published: (2025)
ScaleDiff: Higher-Resolution Image Synthesis via Efficient and Model-Agnostic Diffusion
by: Koh, Sungho, et al.
Published: (2025)
by: Koh, Sungho, et al.
Published: (2025)
Semantic Environment Atlas for Object-Goal Navigation
by: Kim, Nuri, et al.
Published: (2024)
by: Kim, Nuri, et al.
Published: (2024)
Training-Free Restoration of Pruned Neural Networks
by: Lee, Keonho, et al.
Published: (2025)
by: Lee, Keonho, et al.
Published: (2025)
vTrain: A Simulation Framework for Evaluating Cost-effective and Compute-optimal Large Language Model Training
by: Bang, Jehyeon, et al.
Published: (2023)
by: Bang, Jehyeon, et al.
Published: (2023)
SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding
by: Bang, Jehyeon, et al.
Published: (2026)
by: Bang, Jehyeon, et al.
Published: (2026)
RILQ: Rank-Insensitive LoRA-based Quantization Error Compensation for Boosting 2-bit Large Language Model Accuracy
by: Lee, Geonho, et al.
Published: (2024)
by: Lee, Geonho, et al.
Published: (2024)
Regret-Based Defense in Adversarial Reinforcement Learning
by: Belaire, Roman, et al.
Published: (2023)
by: Belaire, Roman, et al.
Published: (2023)
A Ridge Too Far: Correcting Over-Shrinkage via Negative Regularization
by: Kim, Dongseok, et al.
Published: (2025)
by: Kim, Dongseok, et al.
Published: (2025)
Censored Sampling for Topology Design: Guiding Diffusion with Human Preferences
by: Kim, Euihyun, et al.
Published: (2025)
by: Kim, Euihyun, et al.
Published: (2025)
Uncertainty-Aware Multi-Objective Reinforcement Learning-Guided Diffusion Models for 3D De Novo Molecular Design
by: Chen, Lianghong, et al.
Published: (2025)
by: Chen, Lianghong, et al.
Published: (2025)
PREBA: A Hardware/Software Co-Design for Multi-Instance GPU based AI Inference Servers
by: Yeo, Gwangoo, et al.
Published: (2024)
by: Yeo, Gwangoo, et al.
Published: (2024)
Reinforcement Learning via Conservative Agent for Environments with Random Delays
by: Lee, Jongsoo, et al.
Published: (2025)
by: Lee, Jongsoo, et al.
Published: (2025)
Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics
by: Kim, Kwanyoung
Published: (2026)
by: Kim, Kwanyoung
Published: (2026)
Adversarial Reinforcement Learning Framework for ESP Cheater Simulation
by: Park, Inkyu, et al.
Published: (2025)
by: Park, Inkyu, et al.
Published: (2025)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
Diffusion Guided Adversarial State Perturbations in Reinforcement Learning
by: Sun, Xiaolin, et al.
Published: (2025)
by: Sun, Xiaolin, et al.
Published: (2025)
PnPXAI: A Universal XAI Framework Providing Automatic Explanations Across Diverse Modalities and Models
by: Kim, Seongun, et al.
Published: (2025)
by: Kim, Seongun, et al.
Published: (2025)
ODIM: Outlier Detection via Likelihood of Under-Fitted Generative Models
by: Kim, Dongha, et al.
Published: (2023)
by: Kim, Dongha, et al.
Published: (2023)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
by: Moon, Sang Bin, et al.
Published: (2024)
by: Moon, Sang Bin, et al.
Published: (2024)
DPAC: Distribution-Preserving Adversarial Control for Diffusion Sampling
by: Lee, Han-Jin, et al.
Published: (2025)
by: Lee, Han-Jin, et al.
Published: (2025)
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
by: Ganguly, Sourav, et al.
Published: (2026)
by: Ganguly, Sourav, et al.
Published: (2026)
Addressing Negative Transfer in Diffusion Models
by: Go, Hyojun, et al.
Published: (2023)
by: Go, Hyojun, et al.
Published: (2023)
Iterative Distillation for Reward-Guided Fine-Tuning of Diffusion Models in Biomolecular Design
by: Su, Xingyu, et al.
Published: (2025)
by: Su, Xingyu, et al.
Published: (2025)
EXAONE Path 2.0: Pathology Foundation Model with End-to-End Supervision
by: Pyeon, Myeongjang, et al.
Published: (2025)
by: Pyeon, Myeongjang, et al.
Published: (2025)
Generative Representation Learning on Hyper-relational Knowledge Graphs via Masked Discrete Diffusion
by: Lee, Jaejun, et al.
Published: (2026)
by: Lee, Jaejun, et al.
Published: (2026)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
by: Lee, Dohun, et al.
Published: (2024)
by: Lee, Dohun, et al.
Published: (2024)
Structure-Guided Adversarial Training of Diffusion Models
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Similar Items
-
Offline Reinforcement Learning with Universal Horizon Models
by: Chung, Hojun, et al.
Published: (2026) -
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024) -
TRACED: Transition-aware Regret Approximation with Co-learnability for Environment Design
by: Cho, Geonwoo, et al.
Published: (2025) -
Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning
by: Kim, Junseok, et al.
Published: (2026) -
Stage-Wise Reward Shaping for Acrobatic Robots: A Constrained Multi-Objective Reinforcement Learning Approach
by: Kim, Dohyeong, et al.
Published: (2024)