Score-Based One-step MeanFlow Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Kyungyoon, Ki, Donghyeon, Ahn, Hee-Jun, Lee, Byung-Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Direct Soft-Policy Sampling via Langevin Dynamics
by: Ki, Donghyeon, et al.
Published: (2026)
by: Ki, Donghyeon, et al.
Published: (2026)
Actor-Critic without Actor
by: Ki, Donghyeon, et al.
Published: (2025)
by: Ki, Donghyeon, et al.
Published: (2025)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
by: Wang, Zeyuan, et al.
Published: (2025)
by: Wang, Zeyuan, et al.
Published: (2025)
Stochastic MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent
by: Wang, Zeyuan, et al.
Published: (2026)
by: Wang, Zeyuan, et al.
Published: (2026)
Discrete MeanFlow: One-Step Generation via Conditional Transition Kernels
by: Khan, Fairoz Nower, et al.
Published: (2026)
by: Khan, Fairoz Nower, et al.
Published: (2026)
Understanding, Accelerating, and Improving MeanFlow Training
by: Kim, Jin-Young, et al.
Published: (2025)
by: Kim, Jin-Young, et al.
Published: (2025)
MeanFlow Transformers with Representation Autoencoders
by: Hu, Zheyuan, et al.
Published: (2025)
by: Hu, Zheyuan, et al.
Published: (2025)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
by: Zhan, Guojian, et al.
Published: (2026)
by: Zhan, Guojian, et al.
Published: (2026)
Riemannian MeanFlow
by: Woo, Dongyeop, et al.
Published: (2026)
by: Woo, Dongyeop, et al.
Published: (2026)
FlashMoE: Reducing SSD I/O Bottlenecks via ML-Based Cache Replacement for Mixture-of-Experts Inference on Edge Devices
by: Kim, Byeongju, et al.
Published: (2026)
by: Kim, Byeongju, et al.
Published: (2026)
MeanVoiceFlow: One-step Nonparallel Voice Conversion with Mean Flows
by: Kaneko, Takuhiro, et al.
Published: (2026)
by: Kaneko, Takuhiro, et al.
Published: (2026)
Offline Reinforcement Learning with Penalized Action Noise Injection
by: Oh, JunHyeok, et al.
Published: (2025)
by: Oh, JunHyeok, et al.
Published: (2025)
FairDICE: Fairness-Driven Offline Multi-Objective Reinforcement Learning
by: Kim, Woosung, et al.
Published: (2025)
by: Kim, Woosung, et al.
Published: (2025)
NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations
by: Kim, Myunsoo, et al.
Published: (2025)
by: Kim, Myunsoo, et al.
Published: (2025)
VPO: Leveraging the Number of Votes in Preference Optimization
by: Cho, Jae Hyeon, et al.
Published: (2024)
by: Cho, Jae Hyeon, et al.
Published: (2024)
Riemannian MeanFlow for One-Step Generation on Manifolds
by: Zhong, Zichen, et al.
Published: (2026)
by: Zhong, Zichen, et al.
Published: (2026)
TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning
by: Lee, Hayeong, et al.
Published: (2026)
by: Lee, Hayeong, et al.
Published: (2026)
MPruner: Optimizing Neural Network Size with CKA-Based Mutual Information Pruning
by: Hu, Seungbeom, et al.
Published: (2024)
by: Hu, Seungbeom, et al.
Published: (2024)
DIAR: Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation
by: Park, Jaehyun, et al.
Published: (2024)
by: Park, Jaehyun, et al.
Published: (2024)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
A Framework for Mining Collectively-Behaving Bots in MMORPGs
by: Kim, Hyunsoo, et al.
Published: (2025)
by: Kim, Hyunsoo, et al.
Published: (2025)
Prior-Guided Diffusion Planning for Offline Reinforcement Learning
by: Ki, Donghyeon, et al.
Published: (2025)
by: Ki, Donghyeon, et al.
Published: (2025)
Accelerating LMO-Based Optimization via Implicit Gradient Transport
by: Jang, Won-Jun, et al.
Published: (2026)
by: Jang, Won-Jun, et al.
Published: (2026)
Multiple Invertible and Partial-Equivariant Function for Latent Vector Transformation to Enhance Disentanglement in VAEs
by: Jung, Hee-Jun, et al.
Published: (2025)
by: Jung, Hee-Jun, et al.
Published: (2025)
CFASL: Composite Factor-Aligned Symmetry Learning for Disentanglement in Variational AutoEncoder
by: Jung, Hee-Jun, et al.
Published: (2024)
by: Jung, Hee-Jun, et al.
Published: (2024)
Syndrome-Flow Consistency Model Achieves One-step Denoising Error Correction Codes
by: Lei, Haoyu, et al.
Published: (2025)
by: Lei, Haoyu, et al.
Published: (2025)
Reparameterization Flow Policy Optimization
by: Zhong, Hai, et al.
Published: (2026)
by: Zhong, Hai, et al.
Published: (2026)
GDFlow: Anomaly Detection with NCDE-based Normalizing Flow for Advanced Driver Assistance System
by: Lee, Kangjun, et al.
Published: (2024)
by: Lee, Kangjun, et al.
Published: (2024)
Diffusion-Based Offline RL for Improved Decision-Making in Augmented ARC Task
by: Kim, Yunho, et al.
Published: (2024)
by: Kim, Yunho, et al.
Published: (2024)
Decision Flow Policy Optimization
by: Hu, Jifeng, et al.
Published: (2025)
by: Hu, Jifeng, et al.
Published: (2025)
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy
by: Doo, JaeHyeok, et al.
Published: (2026)
by: Doo, JaeHyeok, et al.
Published: (2026)
Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training
by: Kim, Myunsoo, et al.
Published: (2024)
by: Kim, Myunsoo, et al.
Published: (2024)
Unsupervised Training of Diffusion Models for Feasible Solution Generation in Neural Combinatorial Optimization
by: Hong, Seong-Hyun, et al.
Published: (2024)
by: Hong, Seong-Hyun, et al.
Published: (2024)
OM2P: Offline Multi-Agent Mean-Flow Policy
by: Li, Zhuoran, et al.
Published: (2025)
by: Li, Zhuoran, et al.
Published: (2025)
Target Circuit Matching in Large-Scale Netlists using GNN-Based Region Prediction
by: Seo, Sangwoo, et al.
Published: (2025)
by: Seo, Sangwoo, et al.
Published: (2025)
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods
by: Jeon, WooJae, et al.
Published: (2024)
by: Jeon, WooJae, et al.
Published: (2024)
One-Way Policy Optimization for Self-Evolving LLMs
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
Imagined Speech State Classification for Robust Brain-Computer Interface
by: Ko, Byung-Kwan, et al.
Published: (2024)
by: Ko, Byung-Kwan, et al.
Published: (2024)
Learning Design-Score Manifold to Guide Diffusion Models for Offline Optimization
by: Zhou, Tailin, et al.
Published: (2025)
by: Zhou, Tailin, et al.
Published: (2025)
Breaking the Curse of Repulsion: Optimistic Distributionally Robust Policy Optimization for Off-Policy Generative Recommendation
by: Jiang, Jie, et al.
Published: (2026)
by: Jiang, Jie, et al.
Published: (2026)
Similar Items
-
Direct Soft-Policy Sampling via Langevin Dynamics
by: Ki, Donghyeon, et al.
Published: (2026) -
Actor-Critic without Actor
by: Ki, Donghyeon, et al.
Published: (2025) -
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
by: Wang, Zeyuan, et al.
Published: (2025) -
Stochastic MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent
by: Wang, Zeyuan, et al.
Published: (2026) -
Discrete MeanFlow: One-Step Generation via Conditional Transition Kernels
by: Khan, Fairoz Nower, et al.
Published: (2026)