Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Ruishuo, Wang, Xun, Hu, Rui, Li, Zhuoran, Huang, Longbo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GFlowNet Foundations
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
von: Wang, Xi, et al.
Veröffentlicht: (2026)
von: Wang, Xi, et al.
Veröffentlicht: (2026)
Order-Preserving GFlowNets
von: Chen, Yihang, et al.
Veröffentlicht: (2023)
von: Chen, Yihang, et al.
Veröffentlicht: (2023)
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
von: Chen, Ruishuo, et al.
Veröffentlicht: (2026)
von: Chen, Ruishuo, et al.
Veröffentlicht: (2026)
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
von: Wang, Xun, et al.
Veröffentlicht: (2025)
von: Wang, Xun, et al.
Veröffentlicht: (2025)
Beyond Squared Error: Exploring Loss Design for Enhanced Training of Generative Flow Networks
von: Hu, Rui, et al.
Veröffentlicht: (2024)
von: Hu, Rui, et al.
Veröffentlicht: (2024)
Distributional GFlowNets with Quantile Flows
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
GFlowNet Pretraining with Inexpensive Rewards
von: Pandey, Mohit, et al.
Veröffentlicht: (2024)
von: Pandey, Mohit, et al.
Veröffentlicht: (2024)
Global-Order GFlowNets
von: Pastor-Pérez, Lluís, et al.
Veröffentlicht: (2025)
von: Pastor-Pérez, Lluís, et al.
Veröffentlicht: (2025)
Offline Critic-Guided Diffusion Policy for Multi-User Delay-Constrained Scheduling
von: Li, Zhuoran, et al.
Veröffentlicht: (2025)
von: Li, Zhuoran, et al.
Veröffentlicht: (2025)
Improving GFlowNets with Monte Carlo Tree Search
von: Morozov, Nikita, et al.
Veröffentlicht: (2024)
von: Morozov, Nikita, et al.
Veröffentlicht: (2024)
OM2P: Offline Multi-Agent Mean-Flow Policy
von: Li, Zhuoran, et al.
Veröffentlicht: (2025)
von: Li, Zhuoran, et al.
Veröffentlicht: (2025)
Controlling Exploration-Exploitation in GFlowNets via Markov Chain Perspectives
von: Chen, Lin, et al.
Veröffentlicht: (2026)
von: Chen, Lin, et al.
Veröffentlicht: (2026)
On Divergence Measures for Training GFlowNets
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
Reparameterization Proximal Policy Optimization
von: Zhong, Hai, et al.
Veröffentlicht: (2025)
von: Zhong, Hai, et al.
Veröffentlicht: (2025)
Reparameterization Flow Policy Optimization
von: Zhong, Hai, et al.
Veröffentlicht: (2026)
von: Zhong, Hai, et al.
Veröffentlicht: (2026)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
$f$-Trajectory Balance: A Loss Family for Tuning GFlowNets, Generative Models, and LLMs with Off- and On-Policy Data
von: Fawkes, Jake, et al.
Veröffentlicht: (2026)
von: Fawkes, Jake, et al.
Veröffentlicht: (2026)
Interpreting GFlowNets for Drug Discovery: Extracting Actionable Insights for Medicinal Chemistry
von: S, Amirtha Varshini A, et al.
Veröffentlicht: (2025)
von: S, Amirtha Varshini A, et al.
Veröffentlicht: (2025)
AbFlowNet: Optimizing Antibody-Antigen Binding Energy via Diffusion-GFlowNet Fusion
von: Abir, Abrar Rahman, et al.
Veröffentlicht: (2025)
von: Abir, Abrar Rahman, et al.
Veröffentlicht: (2025)
EMERGENT: Efficient and Manipulation-resistant Matching using GFlowNets
von: Tasnim, Mayesha, et al.
Veröffentlicht: (2025)
von: Tasnim, Mayesha, et al.
Veröffentlicht: (2025)
Why Pool When You Can Flow? Active Learning with GFlowNets
von: Zhang, Renfei, et al.
Veröffentlicht: (2025)
von: Zhang, Renfei, et al.
Veröffentlicht: (2025)
GFlowNet Training by Policy Gradients
von: Niu, Puhua, et al.
Veröffentlicht: (2024)
von: Niu, Puhua, et al.
Veröffentlicht: (2024)
Offline-to-Online Multi-Agent Reinforcement Learning with Offline Value Function Memory and Sequential Exploration
von: Zhong, Hai, et al.
Veröffentlicht: (2024)
von: Zhong, Hai, et al.
Veröffentlicht: (2024)
Finite-time Convergence Analysis of Actor-Critic with Evolving Reward
von: Hu, Rui, et al.
Veröffentlicht: (2025)
von: Hu, Rui, et al.
Veröffentlicht: (2025)
Symmetry-Aware GFlowNets
von: Kim, Hohyun, et al.
Veröffentlicht: (2025)
von: Kim, Hohyun, et al.
Veröffentlicht: (2025)
Streaming Bayes GFlowNets
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
Embarrassingly Parallel GFlowNets
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
von: da Silva, Tiago, et al.
Veröffentlicht: (2024)
Baking Symmetry into GFlowNets
von: Ma, George, et al.
Veröffentlicht: (2024)
von: Ma, George, et al.
Veröffentlicht: (2024)
Local Search GFlowNets
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
Pessimistic Backward Policy for GFlowNets
von: Jang, Hyosoon, et al.
Veröffentlicht: (2024)
von: Jang, Hyosoon, et al.
Veröffentlicht: (2024)
Stable GFlowNets with Probabilistic Guarantees
von: Lei, Zengxiang, et al.
Veröffentlicht: (2026)
von: Lei, Zengxiang, et al.
Veröffentlicht: (2026)
Avoid What You Know: Divergent Trajectory Balance for GFlowNets
von: Dall'Antonia, Pedro, et al.
Veröffentlicht: (2026)
von: Dall'Antonia, Pedro, et al.
Veröffentlicht: (2026)
Optimizing Backward Policies in GFlowNets via Trajectory Likelihood Maximization
von: Gritsaev, Timofei, et al.
Veröffentlicht: (2024)
von: Gritsaev, Timofei, et al.
Veröffentlicht: (2024)
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
Multi-Fidelity Active Learning with GFlowNets
von: Hernandez-Garcia, Alex, et al.
Veröffentlicht: (2023)
von: Hernandez-Garcia, Alex, et al.
Veröffentlicht: (2023)
Revisiting Non-Acyclic GFlowNets in Discrete Environments
von: Morozov, Nikita, et al.
Veröffentlicht: (2025)
von: Morozov, Nikita, et al.
Veröffentlicht: (2025)
torchgfn: A PyTorch GFlowNet library
von: Viviano, Joseph D., et al.
Veröffentlicht: (2023)
von: Viviano, Joseph D., et al.
Veröffentlicht: (2023)
Maximum entropy GFlowNets with soft Q-learning
von: Mohammadpour, Sobhan, et al.
Veröffentlicht: (2023)
von: Mohammadpour, Sobhan, et al.
Veröffentlicht: (2023)
Learning to Scale Logits for Temperature-Conditional GFlowNets
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GFlowNet Foundations
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021) -
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
von: Wang, Xi, et al.
Veröffentlicht: (2026) -
Order-Preserving GFlowNets
von: Chen, Yihang, et al.
Veröffentlicht: (2023) -
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
von: Chen, Ruishuo, et al.
Veröffentlicht: (2026) -
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
von: Wang, Xun, et al.
Veröffentlicht: (2025)