Discrete Probabilistic Inference as Control in Multi-path Environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Deleu, Tristan, Nouri, Padideh, Malkin, Nikolay, Precup, Doina, Bengio, Yoshua |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Relative Trajectory Balance is equivalent to Trust-PCL
por: Deleu, Tristan, et al.
Publicado: (2025)
por: Deleu, Tristan, et al.
Publicado: (2025)
In-Context Parametric Inference: Point or Distribution Estimators?
por: Mittal, Sarthak, et al.
Publicado: (2025)
por: Mittal, Sarthak, et al.
Publicado: (2025)
On Generalization for Generative Flow Networks
por: Krichel, Anas, et al.
Publicado: (2024)
por: Krichel, Anas, et al.
Publicado: (2024)
Bayesian learning of Causal Structure and Mechanisms with GFlowNets and Variational Bayes
por: Nishikawa-Toomey, Mizu, et al.
Publicado: (2022)
por: Nishikawa-Toomey, Mizu, et al.
Publicado: (2022)
GFlowNet Foundations
por: Bengio, Yoshua, et al.
Publicado: (2021)
por: Bengio, Yoshua, et al.
Publicado: (2021)
Discrete, compositional, and symbolic representations through attractor dynamics
por: Nam, Andrew, et al.
Publicado: (2023)
por: Nam, Andrew, et al.
Publicado: (2023)
Learning Decision Trees as Amortized Structure Inference
por: Mahfoud, Mohammed, et al.
Publicado: (2025)
por: Mahfoud, Mohammed, et al.
Publicado: (2025)
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
por: Zhao, Mingde, et al.
Publicado: (2023)
por: Zhao, Mingde, et al.
Publicado: (2023)
Generative Flow Networks: Theory and Applications to Structure Learning
por: Deleu, Tristan
Publicado: (2025)
por: Deleu, Tristan
Publicado: (2025)
QGFN: Controllable Greediness with Action Values
por: Lau, Elaine, et al.
Publicado: (2024)
por: Lau, Elaine, et al.
Publicado: (2024)
Machine learning and information theory concepts towards an AI Mathematician
por: Bengio, Yoshua, et al.
Publicado: (2024)
por: Bengio, Yoshua, et al.
Publicado: (2024)
Rejecting Hallucinated State Targets during Planning
por: Zhao, Mingde, et al.
Publicado: (2024)
por: Zhao, Mingde, et al.
Publicado: (2024)
Functional Acceleration for Policy Mirror Descent
por: Chelu, Veronica, et al.
Publicado: (2024)
por: Chelu, Veronica, et al.
Publicado: (2024)
Diversity-Enriched Option-Critic
por: Kamat, Anand, et al.
Publicado: (2020)
por: Kamat, Anand, et al.
Publicado: (2020)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
por: Alver, Safa, et al.
Publicado: (2022)
por: Alver, Safa, et al.
Publicado: (2022)
Outsourced diffusion sampling: Efficient posterior inference in latent spaces of generative models
por: Venkatraman, Siddarth, et al.
Publicado: (2025)
por: Venkatraman, Siddarth, et al.
Publicado: (2025)
Simulation-free Schrödinger bridges via score and flow matching
por: Tong, Alexander, et al.
Publicado: (2023)
por: Tong, Alexander, et al.
Publicado: (2023)
Expected flow networks in stochastic environments and two-player zero-sum games
por: Jiralerspong, Marco, et al.
Publicado: (2023)
por: Jiralerspong, Marco, et al.
Publicado: (2023)
Amortizing intractable inference in large language models
por: Hu, Edward J., et al.
Publicado: (2023)
por: Hu, Edward J., et al.
Publicado: (2023)
Adaptive teachers for amortized samplers
por: Kim, Minsu, et al.
Publicado: (2024)
por: Kim, Minsu, et al.
Publicado: (2024)
Improving and generalizing flow-based generative models with minibatch optimal transport
por: Tong, Alexander, et al.
Publicado: (2023)
por: Tong, Alexander, et al.
Publicado: (2023)
Action abstractions for amortized sampling
por: Boussif, Oussama, et al.
Publicado: (2024)
por: Boussif, Oussama, et al.
Publicado: (2024)
PhyloGFN: Phylogenetic inference with generative flow networks
por: Zhou, Mingyang, et al.
Publicado: (2023)
por: Zhou, Mingyang, et al.
Publicado: (2023)
Balancing Plasticity and Stability with Fast and Slow Successor Features
por: Chua, Raymond, et al.
Publicado: (2026)
por: Chua, Raymond, et al.
Publicado: (2026)
Baking Symmetry into GFlowNets
por: Ma, George, et al.
Publicado: (2024)
por: Ma, George, et al.
Publicado: (2024)
Improved off-policy training of diffusion samplers
por: Sendera, Marcin, et al.
Publicado: (2024)
por: Sendera, Marcin, et al.
Publicado: (2024)
On the Privacy of Selection Mechanisms with Gaussian Noise
por: Lebensold, Jonathan, et al.
Publicado: (2024)
por: Lebensold, Jonathan, et al.
Publicado: (2024)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
por: Carr, Jonathan Colaço, et al.
Publicado: (2023)
por: Carr, Jonathan Colaço, et al.
Publicado: (2023)
Delta-AI: Local objectives for amortized inference in sparse graphical models
por: Falet, Jean-Pierre, et al.
Publicado: (2023)
por: Falet, Jean-Pierre, et al.
Publicado: (2023)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
por: Jain, Arushi, et al.
Publicado: (2024)
por: Jain, Arushi, et al.
Publicado: (2024)
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
por: Alver, Safa, et al.
Publicado: (2024)
por: Alver, Safa, et al.
Publicado: (2024)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
por: Arnob, Samin Yeasar, et al.
Publicado: (2025)
por: Arnob, Samin Yeasar, et al.
Publicado: (2025)
Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments
por: Luo, Ziyan, et al.
Publicado: (2025)
por: Luo, Ziyan, et al.
Publicado: (2025)
Data-to-Energy Stochastic Dynamics
por: Tamogashev, Kirill, et al.
Publicado: (2025)
por: Tamogashev, Kirill, et al.
Publicado: (2025)
Local Inconsistency Resolution: The Interplay between Attention and Control in Probabilistic Models
por: Richardson, Oliver E., et al.
Publicado: (2026)
por: Richardson, Oliver E., et al.
Publicado: (2026)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
Fluid-Agent Reinforcement Learning
por: Sharma, Shishir, et al.
Publicado: (2026)
por: Sharma, Shishir, et al.
Publicado: (2026)
Discrete Feynman-Kac Correctors
por: Hasan, Mohsin, et al.
Publicado: (2026)
por: Hasan, Mohsin, et al.
Publicado: (2026)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
por: Ishfaq, Haque, et al.
Publicado: (2025)
por: Ishfaq, Haque, et al.
Publicado: (2025)
Parseval Regularization for Continual Reinforcement Learning
por: Chung, Wesley, et al.
Publicado: (2024)
por: Chung, Wesley, et al.
Publicado: (2024)
Ejemplares similares
-
Relative Trajectory Balance is equivalent to Trust-PCL
por: Deleu, Tristan, et al.
Publicado: (2025) -
In-Context Parametric Inference: Point or Distribution Estimators?
por: Mittal, Sarthak, et al.
Publicado: (2025) -
On Generalization for Generative Flow Networks
por: Krichel, Anas, et al.
Publicado: (2024) -
Bayesian learning of Causal Structure and Mechanisms with GFlowNets and Variational Bayes
por: Nishikawa-Toomey, Mizu, et al.
Publicado: (2022) -
GFlowNet Foundations
por: Bengio, Yoshua, et al.
Publicado: (2021)