FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Sungha, Lee, Gawon, Lee, Jusuk, Park, Jonghae, Kim, H. Jin, Cho, Daesol |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Periodic Skill Discovery
di: Park, Jonghae, et al.
Pubblicazione: (2025)
di: Park, Jonghae, et al.
Pubblicazione: (2025)
Leveraging Temporally Extended Behavior Sharing for Multi-task Reinforcement Learning
di: Lee, Gawon, et al.
Pubblicazione: (2025)
di: Lee, Gawon, et al.
Pubblicazione: (2025)
DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation
di: Lee, Jusuk, et al.
Pubblicazione: (2026)
di: Lee, Jusuk, et al.
Pubblicazione: (2026)
Temporal Action Representation Learning for Tactical Resource Control and Subsequent Maneuver Generation
di: Jung, Hoseong, et al.
Pubblicazione: (2026)
di: Jung, Hoseong, et al.
Pubblicazione: (2026)
Latent Policy Steering through One-Step Flow Policies
di: Im, Hokyun, et al.
Pubblicazione: (2026)
di: Im, Hokyun, et al.
Pubblicazione: (2026)
Learning Generalizable Visuomotor Policy through Dynamics-Alignment
di: Lee, Dohyeok, et al.
Pubblicazione: (2025)
di: Lee, Dohyeok, et al.
Pubblicazione: (2025)
MaxEnt Loss: Constrained Maximum Entropy for Calibration under Out-of-Distribution Shift
di: Neo, Dexter, et al.
Pubblicazione: (2023)
di: Neo, Dexter, et al.
Pubblicazione: (2023)
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning
di: Lee, Sungyoung, et al.
Pubblicazione: (2026)
di: Lee, Sungyoung, et al.
Pubblicazione: (2026)
MEReQ: Max-Ent Residual-Q Inverse RL for Sample-Efficient Alignment from Intervention
di: Chen, Yuxin, et al.
Pubblicazione: (2024)
di: Chen, Yuxin, et al.
Pubblicazione: (2024)
ERPPO: Entropy Regularization-based Proximal Policy Optimization
di: Lee, Changha, et al.
Pubblicazione: (2026)
di: Lee, Changha, et al.
Pubblicazione: (2026)
Improving Generative Behavior Cloning via Self-Guidance and Adaptive Chunking
di: So, Junhyuk, et al.
Pubblicazione: (2025)
di: So, Junhyuk, et al.
Pubblicazione: (2025)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
di: Kang, Sinjae, et al.
Pubblicazione: (2026)
di: Kang, Sinjae, et al.
Pubblicazione: (2026)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
di: Kim, Changyeon, et al.
Pubblicazione: (2025)
di: Kim, Changyeon, et al.
Pubblicazione: (2025)
Can only LLMs do Reasoning?: Potential of Small Language Models in Task Planning
di: Choi, Gawon, et al.
Pubblicazione: (2024)
di: Choi, Gawon, et al.
Pubblicazione: (2024)
Behavior Generation with Latent Actions
di: Lee, Seungjae, et al.
Pubblicazione: (2024)
di: Lee, Seungjae, et al.
Pubblicazione: (2024)
AdaptManip: Learning Adaptive Whole-Body Object Lifting and Delivery with Online Recurrent State Estimation
di: Byrd, Morgan, et al.
Pubblicazione: (2026)
di: Byrd, Morgan, et al.
Pubblicazione: (2026)
Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation
di: Lee, Seokju, et al.
Pubblicazione: (2026)
di: Lee, Seokju, et al.
Pubblicazione: (2026)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
di: Park, Seohong, et al.
Pubblicazione: (2023)
di: Park, Seohong, et al.
Pubblicazione: (2023)
Flow Matching Policy Gradients
di: McAllister, David, et al.
Pubblicazione: (2025)
di: McAllister, David, et al.
Pubblicazione: (2025)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
di: Kim, Donghu, et al.
Pubblicazione: (2026)
di: Kim, Donghu, et al.
Pubblicazione: (2026)
Smoother Action Chunking Flow Policy via Prior-Corrected Orthogonal Trust-Region Guidance
di: Fang, Kai, et al.
Pubblicazione: (2026)
di: Fang, Kai, et al.
Pubblicazione: (2026)
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
di: Cho, Daesol, et al.
Pubblicazione: (2026)
di: Cho, Daesol, et al.
Pubblicazione: (2026)
Communication-Efficient Module-Wise Federated Learning for Grasp Pose Detection in Cluttered Environments
di: Kang, Woonsang, et al.
Pubblicazione: (2025)
di: Kang, Woonsang, et al.
Pubblicazione: (2025)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
di: Kim, Hyeonjun, et al.
Pubblicazione: (2025)
di: Kim, Hyeonjun, et al.
Pubblicazione: (2025)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
di: Kim, Dongwon, et al.
Pubblicazione: (2026)
di: Kim, Dongwon, et al.
Pubblicazione: (2026)
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
di: Song, Yunzhou, et al.
Pubblicazione: (2026)
di: Song, Yunzhou, et al.
Pubblicazione: (2026)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
di: Ankile, Lars, et al.
Pubblicazione: (2025)
di: Ankile, Lars, et al.
Pubblicazione: (2025)
MePoly: Max Entropy Polynomial Policy Optimization
di: Liu, Hang, et al.
Pubblicazione: (2026)
di: Liu, Hang, et al.
Pubblicazione: (2026)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
DIAR: Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
di: Liu, Xin, et al.
Pubblicazione: (2026)
di: Liu, Xin, et al.
Pubblicazione: (2026)
XPG-RL: Reinforcement Learning with Explainable Priority Guidance for Efficiency-Boosted Mechanical Search
di: Zhang, Yiting, et al.
Pubblicazione: (2025)
di: Zhang, Yiting, et al.
Pubblicazione: (2025)
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
di: Cho, Seongwoong, et al.
Pubblicazione: (2024)
di: Cho, Seongwoong, et al.
Pubblicazione: (2024)
Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models
di: Zhang, Hongyin, et al.
Pubblicazione: (2025)
di: Zhang, Hongyin, et al.
Pubblicazione: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
di: Choi, Wonhyeok, et al.
Pubblicazione: (2026)
di: Choi, Wonhyeok, et al.
Pubblicazione: (2026)
End-to-end RL Improves Dexterous Grasping Policies
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
Learning to Transfer Human Hand Skills for Robot Manipulations
di: Park, Sungjae, et al.
Pubblicazione: (2025)
di: Park, Sungjae, et al.
Pubblicazione: (2025)
Impedance Matching: Enabling an RL-Based Running Jump in a Quadruped Robot
di: Guan, Neil, et al.
Pubblicazione: (2024)
di: Guan, Neil, et al.
Pubblicazione: (2024)
ACFormer: Mitigating Non-linearity with Auto Convolutional Encoder for Time Series Forecasting
di: Lee, Gawon, et al.
Pubblicazione: (2026)
di: Lee, Gawon, et al.
Pubblicazione: (2026)
Refined Policy Distillation: From VLA Generalists to RL Experts
di: Jülg, Tobias, et al.
Pubblicazione: (2025)
di: Jülg, Tobias, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Periodic Skill Discovery
di: Park, Jonghae, et al.
Pubblicazione: (2025) -
Leveraging Temporally Extended Behavior Sharing for Multi-task Reinforcement Learning
di: Lee, Gawon, et al.
Pubblicazione: (2025) -
DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation
di: Lee, Jusuk, et al.
Pubblicazione: (2026) -
Temporal Action Representation Learning for Tactical Resource Control and Subsequent Maneuver Generation
di: Jung, Hoseong, et al.
Pubblicazione: (2026) -
Latent Policy Steering through One-Step Flow Policies
di: Im, Hokyun, et al.
Pubblicazione: (2026)