Flow Q-Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Park, Seohong, Li, Qiyang, Levine, Sergey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Q-learning with Adjoint Matching
por: Li, Qiyang, et al.
Publicado: (2026)
por: Li, Qiyang, et al.
Publicado: (2026)
Intention-Conditioned Flow Occupancy Models
por: Zheng, Chongyi, et al.
Publicado: (2025)
por: Zheng, Chongyi, et al.
Publicado: (2025)
Dual Goal Representations
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Foundation Policies with Hilbert Representations
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Transitive RL: Value Learning via Divide and Conquer
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
Is Value Learning Really the Main Bottleneck in Offline RL?
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
por: Frans, Kevin, et al.
Publicado: (2024)
por: Frans, Kevin, et al.
Publicado: (2024)
Scalable Offline Model-Based RL with Action Chunks
por: Park, Kwanyoung, et al.
Publicado: (2025)
por: Park, Kwanyoung, et al.
Publicado: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Horizon Reduction Makes RL Scalable
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
por: Xu, Charles, et al.
Publicado: (2024)
por: Xu, Charles, et al.
Publicado: (2024)
Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration
por: Wilcoxson, Max, et al.
Publicado: (2024)
por: Wilcoxson, Max, et al.
Publicado: (2024)
Unsupervised-to-Online Reinforcement Learning
por: Kim, Junsu, et al.
Publicado: (2024)
por: Kim, Junsu, et al.
Publicado: (2024)
Q-SFT: Q-Learning for Language Models via Supervised Fine-Tuning
por: Hong, Joey, et al.
Publicado: (2024)
por: Hong, Joey, et al.
Publicado: (2024)
Real-Time Execution of Action Chunking Flow Policies
por: Black, Kevin, et al.
Publicado: (2025)
por: Black, Kevin, et al.
Publicado: (2025)
Learning Visuotactile Skills with Two Multifingered Hands
por: Lin, Toru, et al.
Publicado: (2024)
por: Lin, Toru, et al.
Publicado: (2024)
Diffusion Guidance Is a Controllable Policy Improvement Operator
por: Frans, Kevin, et al.
Publicado: (2025)
por: Frans, Kevin, et al.
Publicado: (2025)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
por: Stachowicz, Kyle, et al.
Publicado: (2024)
por: Stachowicz, Kyle, et al.
Publicado: (2024)
Chunk-Guided Q-Learning
por: Song, Gwanwoo, et al.
Publicado: (2026)
por: Song, Gwanwoo, et al.
Publicado: (2026)
EXPO: Stable Reinforcement Learning with Expressive Policies
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
por: Li, Mingxuan, et al.
Publicado: (2026)
por: Li, Mingxuan, et al.
Publicado: (2026)
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy
por: Doo, JaeHyeok, et al.
Publicado: (2026)
por: Doo, JaeHyeok, et al.
Publicado: (2026)
ViVa: Video-Trained Value Functions for Guiding Online RL from Diverse Data
por: Dashora, Nitish, et al.
Publicado: (2025)
por: Dashora, Nitish, et al.
Publicado: (2025)
REFACTOR: Learning to Extract Theorems from Proofs
por: Zhou, Jin Peng, et al.
Publicado: (2024)
por: Zhou, Jin Peng, et al.
Publicado: (2024)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
por: Park, Kwanyoung, et al.
Publicado: (2024)
por: Park, Kwanyoung, et al.
Publicado: (2024)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
por: Tayal, Mumuksh, et al.
Publicado: (2026)
por: Tayal, Mumuksh, et al.
Publicado: (2026)
Pretraining a Shared Q-Network for Data-Efficient Offline Reinforcement Learning
por: Park, Jongchan, et al.
Publicado: (2025)
por: Park, Jongchan, et al.
Publicado: (2025)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
por: Wang, Zeyuan, et al.
Publicado: (2025)
por: Wang, Zeyuan, et al.
Publicado: (2025)
Cliqueformer: Model-Based Optimization with Structured Transformers
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
por: Myers, Vivek, et al.
Publicado: (2024)
por: Myers, Vivek, et al.
Publicado: (2024)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
por: Alles, Marvin, et al.
Publicado: (2025)
por: Alles, Marvin, et al.
Publicado: (2025)
Interactive Dialogue Agents via Reinforcement Learning on Hindsight Regenerations
por: Hong, Joey, et al.
Publicado: (2024)
por: Hong, Joey, et al.
Publicado: (2024)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
por: Hatch, Kyle B., et al.
Publicado: (2024)
por: Hatch, Kyle B., et al.
Publicado: (2024)
Learning Shortest Paths with Generative Flow Networks
por: Morozov, Nikita, et al.
Publicado: (2026)
por: Morozov, Nikita, et al.
Publicado: (2026)
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
por: Uehara, Masatoshi, et al.
Publicado: (2024)
por: Uehara, Masatoshi, et al.
Publicado: (2024)
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
Ejemplares similares
-
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025) -
Q-learning with Adjoint Matching
por: Li, Qiyang, et al.
Publicado: (2026) -
Intention-Conditioned Flow Occupancy Models
por: Zheng, Chongyi, et al.
Publicado: (2025) -
Dual Goal Representations
por: Park, Seohong, et al.
Publicado: (2025) -
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023)