Aligning Flow Map Policies with Optimal Q-Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Ziakas, Christos, Russo, Alessandra, Bose, Avishek Joey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grounding Generated Videos in Feasible Plans via World Models
by: Ziakas, Christos, et al.
Published: (2026)
by: Ziakas, Christos, et al.
Published: (2026)
Generalised Flow Maps for Few-Step Generative Modelling on Riemannian Manifolds
by: Davis, Oscar, et al.
Published: (2025)
by: Davis, Oscar, et al.
Published: (2025)
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
by: Ziakas, Christos, et al.
Published: (2025)
by: Ziakas, Christos, et al.
Published: (2025)
RETRO SYNFLOW: Discrete Flow Matching for Accurate and Diverse Single-Step Retrosynthesis
by: Yadav, Robin, et al.
Published: (2025)
by: Yadav, Robin, et al.
Published: (2025)
Fisher Flow Matching for Generative Modeling over Discrete Data
by: Davis, Oscar, et al.
Published: (2024)
by: Davis, Oscar, et al.
Published: (2024)
Coupling Models for One-Step Discrete Generation
by: Peng, Fred Zhangzhi, et al.
Published: (2026)
by: Peng, Fred Zhangzhi, et al.
Published: (2026)
On the Stability of Iterative Retraining of Generative Models on their own Data
by: Bertrand, Quentin, et al.
Published: (2023)
by: Bertrand, Quentin, et al.
Published: (2023)
The Superposition of Diffusion Models Using the Itô Density Estimator
by: Skreta, Marta, et al.
Published: (2024)
by: Skreta, Marta, et al.
Published: (2024)
Predicting Drug Effects from High-Dimensional, Asymmetric Drug Datasets by Using Graph Neural Networks: A Comprehensive Analysis of Multitarget Drug Effect Prediction
by: Bose, Avishek, et al.
Published: (2024)
by: Bose, Avishek, et al.
Published: (2024)
Metric Flow Matching for Smooth Interpolations on the Data Manifold
by: Kapuśniak, Kacper, et al.
Published: (2024)
by: Kapuśniak, Kacper, et al.
Published: (2024)
Curly Flow Matching for Learning Non-gradient Field Dynamics
by: Petrović, Katarina, et al.
Published: (2025)
by: Petrović, Katarina, et al.
Published: (2025)
Efficient Regression-Based Training of Normalizing Flows for Boltzmann Generators
by: Rehman, Danyal, et al.
Published: (2025)
by: Rehman, Danyal, et al.
Published: (2025)
Self-Consuming Generative Models with Curated Data Provably Optimize Human Preferences
by: Ferbach, Damien, et al.
Published: (2024)
by: Ferbach, Damien, et al.
Published: (2024)
Feature Likelihood Divergence: Evaluating the Generalization of Generative Models Using Samples
by: Jiralerspong, Marco, et al.
Published: (2023)
by: Jiralerspong, Marco, et al.
Published: (2023)
Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts
by: Ziakas, Christos, et al.
Published: (2025)
by: Ziakas, Christos, et al.
Published: (2025)
Scalable Equilibrium Sampling with Sequential Boltzmann Generators
by: Tan, Charlie B., et al.
Published: (2025)
by: Tan, Charlie B., et al.
Published: (2025)
SE(3)-Stochastic Flow Matching for Protein Backbone Generation
by: Bose, Avishek Joey, et al.
Published: (2023)
by: Bose, Avishek Joey, et al.
Published: (2023)
AlignIQL: Policy Alignment in Implicit Q-Learning through Constrained Optimization
by: He, Longxiang, et al.
Published: (2024)
by: He, Longxiang, et al.
Published: (2024)
From $r$ to $Q^*$: Your Language Model is Secretly a Q-Function
by: Rafailov, Rafael, et al.
Published: (2024)
by: Rafailov, Rafael, et al.
Published: (2024)
Optimal Regret for Single Index Bandits
by: Dey, Devdan, et al.
Published: (2026)
by: Dey, Devdan, et al.
Published: (2026)
Strategically Conservative Q-Learning
by: Shimizu, Yutaka, et al.
Published: (2024)
by: Shimizu, Yutaka, et al.
Published: (2024)
AlignFlow: Improving Flow-based Generative Models with Semi-Discrete Optimal Transport
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Sequence-Augmented SE(3)-Flow Matching For Conditional Protein Backbone Generation
by: Huguet, Guillaume, et al.
Published: (2024)
by: Huguet, Guillaume, et al.
Published: (2024)
Planner Aware Path Learning in Diffusion Language Models Training
by: Peng, Fred Zhangzhi, et al.
Published: (2025)
by: Peng, Fred Zhangzhi, et al.
Published: (2025)
FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance
by: Kim, Sungha, et al.
Published: (2026)
by: Kim, Sungha, et al.
Published: (2026)
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy
by: Doo, JaeHyeok, et al.
Published: (2026)
by: Doo, JaeHyeok, et al.
Published: (2026)
Near Optimal Best Arm Identification for Clustered Bandits
by: Yash, et al.
Published: (2025)
by: Yash, et al.
Published: (2025)
Q-SFT: Q-Learning for Language Models via Supervised Fine-Tuning
by: Hong, Joey, et al.
Published: (2024)
by: Hong, Joey, et al.
Published: (2024)
On the Guidance of Flow Matching
by: Feng, Ruiqi, et al.
Published: (2025)
by: Feng, Ruiqi, et al.
Published: (2025)
EP-GRPO: Entropy-Progress Aligned Group Relative Policy Optimization with Implicit Process Guidance
by: Yu, Song, et al.
Published: (2026)
by: Yu, Song, et al.
Published: (2026)
Path Planning for Masked Diffusion Model Sampling
by: Peng, Fred Zhangzhi, et al.
Published: (2025)
by: Peng, Fred Zhangzhi, et al.
Published: (2025)
Align Your Flow: Scaling Continuous-Time Flow Map Distillation
by: Sabour, Amirmojtaba, et al.
Published: (2025)
by: Sabour, Amirmojtaba, et al.
Published: (2025)
How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance
by: Huang, Jerry Y., et al.
Published: (2026)
by: Huang, Jerry Y., et al.
Published: (2026)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2025)
by: Alles, Marvin, et al.
Published: (2025)
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching
by: Wan, Zhengyan, et al.
Published: (2025)
by: Wan, Zhengyan, et al.
Published: (2025)
Neural DNF-MT: A Neuro-symbolic Approach for Learning Interpretable and Editable Policies
by: Baugh, Kexin Gu, et al.
Published: (2025)
by: Baugh, Kexin Gu, et al.
Published: (2025)
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
by: Chen, Gongpu, et al.
Published: (2024)
by: Chen, Gongpu, et al.
Published: (2024)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
by: Tayal, Mumuksh, et al.
Published: (2026)
by: Tayal, Mumuksh, et al.
Published: (2026)
Smoother Action Chunking Flow Policy via Prior-Corrected Orthogonal Trust-Region Guidance
by: Fang, Kai, et al.
Published: (2026)
by: Fang, Kai, et al.
Published: (2026)
Predictive Representations for Skill Transfer in Reinforcement Learning
by: Vereecken, Ruben, et al.
Published: (2026)
by: Vereecken, Ruben, et al.
Published: (2026)
Similar Items
-
Grounding Generated Videos in Feasible Plans via World Models
by: Ziakas, Christos, et al.
Published: (2026) -
Generalised Flow Maps for Few-Step Generative Modelling on Riemannian Manifolds
by: Davis, Oscar, et al.
Published: (2025) -
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
by: Ziakas, Christos, et al.
Published: (2025) -
RETRO SYNFLOW: Discrete Flow Matching for Accurate and Diverse Single-Step Retrosynthesis
by: Yadav, Robin, et al.
Published: (2025) -
Fisher Flow Matching for Generative Modeling over Discrete Data
by: Davis, Oscar, et al.
Published: (2024)