Random Policy Evaluation Uncovers Policies of Generative Flow Networks
Fuente:
arXiv
Saved in:
| Main Authors: | He, Haoran, Bengio, Emmanuel, Cai, Qingpeng, Pan, Ling |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
Bifurcated Generative Flow Networks
by: Li, Chunhui, et al.
Published: (2024)
by: Li, Chunhui, et al.
Published: (2024)
Investigating Generalization Behaviours of Generative Flow Networks
by: Atanackovic, Lazar, et al.
Published: (2024)
by: Atanackovic, Lazar, et al.
Published: (2024)
Pretraining Generative Flow Networks with Inexpensive Rewards for Molecular Graph Generation
by: Pandey, Mohit, et al.
Published: (2025)
by: Pandey, Mohit, et al.
Published: (2025)
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)
by: Krichel, Anas, et al.
Published: (2024)
State Regularized Policy Optimization on Data with Dynamics Shift
by: Xue, Zhenghai, et al.
Published: (2023)
by: Xue, Zhenghai, et al.
Published: (2023)
QGFN: Controllable Greediness with Action Values
by: Lau, Elaine, et al.
Published: (2024)
by: Lau, Elaine, et al.
Published: (2024)
Fast and Robust Visuomotor Riemannian Flow Matching Policy
by: Ding, Haoran, et al.
Published: (2024)
by: Ding, Haoran, et al.
Published: (2024)
Baking Symmetry into GFlowNets
by: Ma, George, et al.
Published: (2024)
by: Ma, George, et al.
Published: (2024)
Distributional GFlowNets with Quantile Flows
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Evolution Guided Generative Flow Networks
by: Ikram, Zarif, et al.
Published: (2024)
by: Ikram, Zarif, et al.
Published: (2024)
Learning to Scale Logits for Temperature-Conditional GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
FlowPG: Action-constrained Policy Gradient with Normalizing Flows
by: Brahmanage, Janaka Chathuranga, et al.
Published: (2024)
by: Brahmanage, Janaka Chathuranga, et al.
Published: (2024)
Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards
by: Takahashi, Tatsuki, et al.
Published: (2025)
by: Takahashi, Tatsuki, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Learning Intractable Multimodal Policies with Reparameterization and Diversity Regularization
by: Wang, Ziqi, et al.
Published: (2025)
by: Wang, Ziqi, et al.
Published: (2025)
Generalization on the Unseen, Logic Reasoning and Degree Curriculum
by: Abbe, Emmanuel, et al.
Published: (2023)
by: Abbe, Emmanuel, et al.
Published: (2023)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
Boosting Continuous Control with Consistency Policy
by: Chen, Yuhui, et al.
Published: (2023)
by: Chen, Yuhui, et al.
Published: (2023)
Off-Policy Evaluation for Ranking Policies under Deterministic Logging Policies
by: Tanaka, Koichi, et al.
Published: (2026)
by: Tanaka, Koichi, et al.
Published: (2026)
FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy
by: He, Qian, et al.
Published: (2026)
by: He, Qian, et al.
Published: (2026)
Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System
by: Zhang, Ruining, et al.
Published: (2023)
by: Zhang, Ruining, et al.
Published: (2023)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
by: Shen, Yi, et al.
Published: (2023)
by: Shen, Yi, et al.
Published: (2023)
Latent Policy Steering through One-Step Flow Policies
by: Im, Hokyun, et al.
Published: (2026)
by: Im, Hokyun, et al.
Published: (2026)
Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities
by: Xie, Ziwen, et al.
Published: (2026)
by: Xie, Ziwen, et al.
Published: (2026)
FlowCorrect: Efficient Interactive Correction of Generative Flow Policies for Robotic Manipulation
by: Welte, Edgar, et al.
Published: (2026)
by: Welte, Edgar, et al.
Published: (2026)
Flow Matching Policy Gradients
by: McAllister, David, et al.
Published: (2025)
by: McAllister, David, et al.
Published: (2025)
Bayesian Off-Policy Evaluation and Learning for Large Action Spaces
by: Aouali, Imad, et al.
Published: (2024)
by: Aouali, Imad, et al.
Published: (2024)
Looking Backward: Retrospective Backward Synthesis for Goal-Conditioned GFlowNets
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Reinforcement Learning for Flow-Matching Policies
by: Pfrommer, Samuel, et al.
Published: (2025)
by: Pfrommer, Samuel, et al.
Published: (2025)
Unleashing Flow Policies with Distributional Critics
by: Chen, Deshu, et al.
Published: (2025)
by: Chen, Deshu, et al.
Published: (2025)
Doubly-Robust Off-Policy Evaluation with Estimated Logging Policy
by: Lee, Kyungbok, et al.
Published: (2024)
by: Lee, Kyungbok, et al.
Published: (2024)
Local Search GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
by: Koo, Juil, et al.
Published: (2026)
by: Koo, Juil, et al.
Published: (2026)
RePO: Bridging On-Policy Learning and Off-Policy Knowledge through Rephrasing Policy Optimization
by: Xia, Linxuan, et al.
Published: (2026)
by: Xia, Linxuan, et al.
Published: (2026)
Decision Flow Policy Optimization
by: Hu, Jifeng, et al.
Published: (2025)
by: Hu, Jifeng, et al.
Published: (2025)
Reparameterization Flow Policy Optimization
by: Zhong, Hai, et al.
Published: (2026)
by: Zhong, Hai, et al.
Published: (2026)
GFlowNet Pretraining with Inexpensive Rewards
by: Pandey, Mohit, et al.
Published: (2024)
by: Pandey, Mohit, et al.
Published: (2024)
Efficient Biological Data Acquisition through Inference Set Design
by: Neporozhnii, Ihor, et al.
Published: (2024)
by: Neporozhnii, Ihor, et al.
Published: (2024)
Generative Actor-Critic with Soft Bridge Policies
by: He, Ke, et al.
Published: (2026)
by: He, Ke, et al.
Published: (2026)
Similar Items
-
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
by: He, Haoran, et al.
Published: (2025) -
Bifurcated Generative Flow Networks
by: Li, Chunhui, et al.
Published: (2024) -
Investigating Generalization Behaviours of Generative Flow Networks
by: Atanackovic, Lazar, et al.
Published: (2024) -
Pretraining Generative Flow Networks with Inexpensive Rewards for Molecular Graph Generation
by: Pandey, Mohit, et al.
Published: (2025) -
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)