Enhancing Control Policy Smoothness by Aligning Actions with Predictions from Preceding States
Fuente:
arXiv
Saved in:
| Main Authors: | Kwak, Kyoleen, Hwang, Hyoseok |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic
by: Lee, Jeong Woon, et al.
Published: (2026)
by: Lee, Jeong Woon, et al.
Published: (2026)
Scale-Consistent State-Space Dynamics via Fractal of Stationary Transformations
by: Yu, Geunhyeok, et al.
Published: (2026)
by: Yu, Geunhyeok, et al.
Published: (2026)
Efficient Monte Carlo Tree Search via On-the-Fly State-Conditioned Action Abstraction
by: Kwak, Yunhyeok, et al.
Published: (2024)
by: Kwak, Yunhyeok, et al.
Published: (2024)
Policy Optimization with Smooth Guidance Learned from State-Only Demonstrations
by: Wang, Guojian, et al.
Published: (2023)
by: Wang, Guojian, et al.
Published: (2023)
PRISM: Breaking the O(n) Memory Wall in Long-Context LLM Inference via O(1) Photonic Block Selection
by: Park, Hyoseok, et al.
Published: (2026)
by: Park, Hyoseok, et al.
Published: (2026)
FAKER: Full-body Anonymization with Human Keypoint Extraction for Real-time Video Deidentification
by: Ban, Byunghyun, et al.
Published: (2024)
by: Ban, Byunghyun, et al.
Published: (2024)
Benchmarking Smoothness and Reducing High-Frequency Oscillations in Continuous Control Policies
by: Christmann, Guilherme, et al.
Published: (2024)
by: Christmann, Guilherme, et al.
Published: (2024)
ForceGrip: Reference-Free Curriculum Learning for Realistic Grip Force Control in VR Hand Manipulation
by: Han, DongHeun, et al.
Published: (2025)
by: Han, DongHeun, et al.
Published: (2025)
Enhancing Variational Autoencoders with Smooth Robust Latent Encoding
by: Lee, Hyomin, et al.
Published: (2025)
by: Lee, Hyomin, et al.
Published: (2025)
GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents
by: Wu, Xiongbin, et al.
Published: (2026)
by: Wu, Xiongbin, et al.
Published: (2026)
When Do Off-Policy and On-Policy Policy Gradient Methods Align?
by: Mambelli, Davide, et al.
Published: (2024)
by: Mambelli, Davide, et al.
Published: (2024)
Precedence-Constrained Decision Trees and Coverings
by: Szyfelbein, Michał, et al.
Published: (2026)
by: Szyfelbein, Michał, et al.
Published: (2026)
Precedence-Constrained Winter Value for Effective Graph Data Valuation
by: Chi, Hongliang, et al.
Published: (2024)
by: Chi, Hongliang, et al.
Published: (2024)
Sequential Off-Policy Learning with Logarithmic Smoothing
by: Haddouche, Maxime, et al.
Published: (2025)
by: Haddouche, Maxime, et al.
Published: (2025)
Enhancing Cost Efficiency in Active Learning with Candidate Set Query
by: Gwon, Yeho, et al.
Published: (2025)
by: Gwon, Yeho, et al.
Published: (2025)
On the Sample Complexity of Imitation Learning for Smoothed Model Predictive Control
by: Pfrommer, Daniel, et al.
Published: (2023)
by: Pfrommer, Daniel, et al.
Published: (2023)
Off-Policy Maximum Entropy RL with Future State and Action Visitation Measures
by: Bolland, Adrien, et al.
Published: (2024)
by: Bolland, Adrien, et al.
Published: (2024)
ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning
by: Shamass, Faiq
Published: (2026)
by: Shamass, Faiq
Published: (2026)
Why Alignment Must Precede Distillation: A Minimal Working Explanation
by: Cha, Sungmin, et al.
Published: (2025)
by: Cha, Sungmin, et al.
Published: (2025)
Observations Meet Actions: Learning Control-Sufficient Representations for Robust Policy Generalization
by: Gu, Yuliang, et al.
Published: (2025)
by: Gu, Yuliang, et al.
Published: (2025)
On the Identifiability of Latent Action Policies
by: Lachapelle, Sébastien
Published: (2025)
by: Lachapelle, Sébastien
Published: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
by: Neggatu, Natinael Solomon, et al.
Published: (2026)
by: Neggatu, Natinael Solomon, et al.
Published: (2026)
Deep Support Vectors
by: Lee, Junhoo, et al.
Published: (2024)
by: Lee, Junhoo, et al.
Published: (2024)
Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings
by: Chen, Wenxin, et al.
Published: (2026)
by: Chen, Wenxin, et al.
Published: (2026)
Landscape of Policy Optimization for Finite Horizon MDPs with General State and Action
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Logarithmic Smoothing for Pessimistic Off-Policy Evaluation, Selection and Learning
by: Sakhi, Otmane, et al.
Published: (2024)
by: Sakhi, Otmane, et al.
Published: (2024)
Aligning Flow Map Policies with Optimal Q-Guidance
by: Ziakas, Christos, et al.
Published: (2026)
by: Ziakas, Christos, et al.
Published: (2026)
Soft Deterministic Policy Gradient with Gaussian Smoothing
by: Na, Hyunjun, et al.
Published: (2026)
by: Na, Hyunjun, et al.
Published: (2026)
ActFusion: a Unified Diffusion Model for Action Segmentation and Anticipation
by: Gong, Dayoung, et al.
Published: (2024)
by: Gong, Dayoung, et al.
Published: (2024)
State-Action Inpainting Diffuser for Continuous Control with Delay
by: Han, Dongqi, et al.
Published: (2026)
by: Han, Dongqi, et al.
Published: (2026)
Predicting Social Media Engagement from Emotional and Temporal Features
by: Kim, Yunwoo, et al.
Published: (2025)
by: Kim, Yunwoo, et al.
Published: (2025)
SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control
by: Orfanoudakis, Stavros, et al.
Published: (2026)
by: Orfanoudakis, Stavros, et al.
Published: (2026)
Forward KL Regularized Preference Optimization for Aligning Diffusion Policies
by: Shan, Zhao, et al.
Published: (2024)
by: Shan, Zhao, et al.
Published: (2024)
CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval
by: Putta, Akshith Reddy, et al.
Published: (2026)
by: Putta, Akshith Reddy, et al.
Published: (2026)
Smoothed Online Learning for Prediction in Piecewise Affine Systems
by: Block, Adam, et al.
Published: (2023)
by: Block, Adam, et al.
Published: (2023)
Smoothing-Based Conformal Prediction for Balancing Efficiency and Interpretability
by: Zheng, Mingyi, et al.
Published: (2025)
by: Zheng, Mingyi, et al.
Published: (2025)
Imitate Optimal Policy: Prevail and Induce Action Collapse in Policy Gradient
by: Zhou, Zhongzhu, et al.
Published: (2025)
by: Zhou, Zhongzhu, et al.
Published: (2025)
Smooth Gate Functions for Soft Advantage Policy Optimization
by: Denisov, Egor, et al.
Published: (2026)
by: Denisov, Egor, et al.
Published: (2026)
Aligning the Evaluation of Probabilistic Predictions with Downstream Value
by: Shahroudi, Novin, et al.
Published: (2025)
by: Shahroudi, Novin, et al.
Published: (2025)
Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers
by: Hyoseok, Lee, et al.
Published: (2026)
by: Hyoseok, Lee, et al.
Published: (2026)
Similar Items
-
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic
by: Lee, Jeong Woon, et al.
Published: (2026) -
Scale-Consistent State-Space Dynamics via Fractal of Stationary Transformations
by: Yu, Geunhyeok, et al.
Published: (2026) -
Efficient Monte Carlo Tree Search via On-the-Fly State-Conditioned Action Abstraction
by: Kwak, Yunhyeok, et al.
Published: (2024) -
Policy Optimization with Smooth Guidance Learned from State-Only Demonstrations
by: Wang, Guojian, et al.
Published: (2023) -
PRISM: Breaking the O(n) Memory Wall in Long-Context LLM Inference via O(1) Photonic Block Selection
by: Park, Hyoseok, et al.
Published: (2026)