Amortizing intractable inference in large language models
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Edward J., Jain, Moksh, Elmoznino, Eric, Kaddar, Younesse, Lajoie, Guillaume, Bengio, Yoshua, Malkin, Nikolay |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024)
by: Venkatraman, Siddarth, et al.
Published: (2024)
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023)
by: Nam, Andrew, et al.
Published: (2023)
Likelihood hacking in probabilistic program synthesis
by: Karwowski, Jacek, et al.
Published: (2026)
by: Karwowski, Jacek, et al.
Published: (2026)
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Learning diverse attacks on large language models for robust red-teaming and safety tuning
by: Lee, Seanie, et al.
Published: (2024)
by: Lee, Seanie, et al.
Published: (2024)
PhyloGFN: Phylogenetic inference with generative flow networks
by: Zhou, Mingyang, et al.
Published: (2023)
by: Zhou, Mingyang, et al.
Published: (2023)
Can a Bayesian Oracle Prevent Harm from an Agent?
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Delta-AI: Local objectives for amortized inference in sparse graphical models
by: Falet, Jean-Pierre, et al.
Published: (2023)
by: Falet, Jean-Pierre, et al.
Published: (2023)
Learning Decision Trees as Amortized Structure Inference
by: Mahfoud, Mohammed, et al.
Published: (2025)
by: Mahfoud, Mohammed, et al.
Published: (2025)
Recursive Self-Aggregation Unlocks Deep Thinking in Large Language Models
by: Venkatraman, Siddarth, et al.
Published: (2025)
by: Venkatraman, Siddarth, et al.
Published: (2025)
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)
by: Krichel, Anas, et al.
Published: (2024)
Action abstractions for amortized sampling
by: Boussif, Oussama, et al.
Published: (2024)
by: Boussif, Oussama, et al.
Published: (2024)
A model of stochastic memoization and name generation in probabilistic programming: categorical semantics via monads on presheaf categories
by: Kaddar, Younesse, et al.
Published: (2023)
by: Kaddar, Younesse, et al.
Published: (2023)
Outsourced diffusion sampling: Efficient posterior inference in latent spaces of generative models
by: Venkatraman, Siddarth, et al.
Published: (2025)
by: Venkatraman, Siddarth, et al.
Published: (2025)
Discrete Probabilistic Inference as Control in Multi-path Environments
by: Deleu, Tristan, et al.
Published: (2024)
by: Deleu, Tristan, et al.
Published: (2024)
Multi-Fidelity Active Learning with GFlowNets
by: Hernandez-Garcia, Alex, et al.
Published: (2023)
by: Hernandez-Garcia, Alex, et al.
Published: (2023)
Improving and generalizing flow-based generative models with minibatch optimal transport
by: Tong, Alexander, et al.
Published: (2023)
by: Tong, Alexander, et al.
Published: (2023)
Machine learning and information theory concepts towards an AI Mathematician
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Simulation-free Schrödinger bridges via score and flow matching
by: Tong, Alexander, et al.
Published: (2023)
by: Tong, Alexander, et al.
Published: (2023)
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
by: Shah, Vedant, et al.
Published: (2025)
by: Shah, Vedant, et al.
Published: (2025)
A comparative study of zero-shot inference with large language models and supervised modeling in breast cancer pathology classification
by: Sushil, Madhumita, et al.
Published: (2024)
by: Sushil, Madhumita, et al.
Published: (2024)
Iterative Amortized Inference: Unifying In-Context Learning and Learned Optimizers
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Expected flow networks in stochastic environments and two-player zero-sum games
by: Jiralerspong, Marco, et al.
Published: (2023)
by: Jiralerspong, Marco, et al.
Published: (2023)
Amortized In-Context Bayesian Posterior Estimation
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Reasoning emerges from constrained inference manifolds in large language models
by: Ma, Yanbiao, et al.
Published: (2026)
by: Ma, Yanbiao, et al.
Published: (2026)
Does learning the right latent variables necessarily improve in-context learning?
by: Mittal, Sarthak, et al.
Published: (2024)
by: Mittal, Sarthak, et al.
Published: (2024)
Next-Token Prediction Should be Ambiguity-Sensitive: A Meta-Learning Perspective
by: Gagnon, Leo, et al.
Published: (2025)
by: Gagnon, Leo, et al.
Published: (2025)
Adaptive teachers for amortized samplers
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
GFlowNet Foundations
by: Bengio, Yoshua, et al.
Published: (2021)
by: Bengio, Yoshua, et al.
Published: (2021)
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
Proof Flow: Preliminary Study on Generative Flow Network Language Model Tuning for Formal Reasoning
by: Ho, Matthew, et al.
Published: (2024)
by: Ho, Matthew, et al.
Published: (2024)
Baking Symmetry into GFlowNets
by: Ma, George, et al.
Published: (2024)
by: Ma, George, et al.
Published: (2024)
Representation in large language models
by: Yetman, Cameron
Published: (2025)
by: Yetman, Cameron
Published: (2025)
Solving Bayesian inverse problems with diffusion priors and off-policy RL
by: Scimeca, Luca, et al.
Published: (2025)
by: Scimeca, Luca, et al.
Published: (2025)
Inducing anxiety in large language models can induce bias
by: Coda-Forno, Julian, et al.
Published: (2023)
by: Coda-Forno, Julian, et al.
Published: (2023)
Long-form factuality in large language models
by: Wei, Jerry, et al.
Published: (2024)
by: Wei, Jerry, et al.
Published: (2024)
In-context learning and Occam's razor
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Can large language models explore in-context?
by: Krishnamurthy, Akshay, et al.
Published: (2024)
by: Krishnamurthy, Akshay, et al.
Published: (2024)
Improved off-policy training of diffusion samplers
by: Sendera, Marcin, et al.
Published: (2024)
by: Sendera, Marcin, et al.
Published: (2024)
Similar Items
-
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024) -
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024) -
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023) -
Likelihood hacking in probabilistic program synthesis
by: Karwowski, Jacek, et al.
Published: (2026) -
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)