Learning Tractable Distributions Of Language Model Continuations
Fuente:
arXiv
Saved in:
| Main Authors: | Yidou-Weng, Gwen, Li, Ian, Liu, Anji, Broadrick, Oliver, Cui, Yuchen, Broeck, Guy Van den, Wang, Benjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language Generation
by: Weng, Gwen Yidou, et al.
Published: (2025)
by: Weng, Gwen Yidou, et al.
Published: (2025)
Polynomial Semantics of Tractable Probabilistic Circuits
by: Broadrick, Oliver, et al.
Published: (2024)
by: Broadrick, Oliver, et al.
Published: (2024)
The Limits of Tractable Marginalization
by: Broadrick, Oliver, et al.
Published: (2025)
by: Broadrick, Oliver, et al.
Published: (2025)
Probabilistic Circuits for Cumulative Distribution Functions
by: Broadrick, Oliver, et al.
Published: (2024)
by: Broadrick, Oliver, et al.
Published: (2024)
Restructuring Tractable Probabilistic Circuits
by: Zhang, Honghua, et al.
Published: (2024)
by: Zhang, Honghua, et al.
Published: (2024)
Breaking the Factorization Barrier in Diffusion Language Models
by: Li, Ian, et al.
Published: (2026)
by: Li, Ian, et al.
Published: (2026)
A Tractable Inference Perspective of Offline RL
by: Liu, Xuejie, et al.
Published: (2023)
by: Liu, Xuejie, et al.
Published: (2023)
On the Relationship Between Monotone and Squared Probabilistic Circuits
by: Wang, Benjie, et al.
Published: (2024)
by: Wang, Benjie, et al.
Published: (2024)
How to Marginalize in Causal Structure Learning?
by: Zhao, William, et al.
Published: (2025)
by: Zhao, William, et al.
Published: (2025)
Scaling Up Probabilistic Circuits by Latent Variable Distillation
by: Liu, Anji, et al.
Published: (2022)
by: Liu, Anji, et al.
Published: (2022)
Discrete Copula Diffusion
by: Liu, Anji, et al.
Published: (2024)
by: Liu, Anji, et al.
Published: (2024)
Tractable Transformers for Flexible Conditional Generation
by: Liu, Anji, et al.
Published: (2025)
by: Liu, Anji, et al.
Published: (2025)
Enabling Autoregressive Models to Fill In Masked Tokens
by: Israel, Daniel, et al.
Published: (2025)
by: Israel, Daniel, et al.
Published: (2025)
Prepacking: A Simple Method for Fast Prefilling and Increased Throughput in Large Language Models
by: Zhao, Siyan, et al.
Published: (2024)
by: Zhao, Siyan, et al.
Published: (2024)
Image Inpainting via Tractable Steering of Diffusion Models
by: Liu, Anji, et al.
Published: (2023)
by: Liu, Anji, et al.
Published: (2023)
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
by: Ahmed, Kareem, et al.
Published: (2023)
by: Ahmed, Kareem, et al.
Published: (2023)
Scaling Tractable Probabilistic Circuits: A Systems Perspective
by: Liu, Anji, et al.
Published: (2024)
by: Liu, Anji, et al.
Published: (2024)
Accelerating Diffusion LLMs via Adaptive Parallel Decoding
by: Israel, Daniel, et al.
Published: (2025)
by: Israel, Daniel, et al.
Published: (2025)
Adversarial Tokenization
by: Geh, Renato Lui, et al.
Published: (2025)
by: Geh, Renato Lui, et al.
Published: (2025)
Collapsed Inference for Bayesian Deep Learning
by: Zeng, Zhe, et al.
Published: (2023)
by: Zeng, Zhe, et al.
Published: (2023)
A Compositional Atlas for Algebraic Circuits
by: Wang, Benjie, et al.
Published: (2024)
by: Wang, Benjie, et al.
Published: (2024)
Probabilistic Programs of Thought
by: Garg, Poorva, et al.
Published: (2026)
by: Garg, Poorva, et al.
Published: (2026)
Combining Supervised Learning and Reinforcement Learning for Multi-Label Classification Tasks with Partial Labels
by: Jia, Zixia, et al.
Published: (2024)
by: Jia, Zixia, et al.
Published: (2024)
Mitigating Bias in Locally Constrained Decoding via Tractable Proposals
by: Dang, Meihua, et al.
Published: (2026)
by: Dang, Meihua, et al.
Published: (2026)
Where is the signal in tokenization space?
by: Geh, Renato Lui, et al.
Published: (2024)
by: Geh, Renato Lui, et al.
Published: (2024)
Rethinking Probabilistic Circuit Parameter Learning
by: Liu, Anji, et al.
Published: (2025)
by: Liu, Anji, et al.
Published: (2025)
Continual Learning Using Only Large Language Model Prompting
by: Qiu, Jiabao, et al.
Published: (2024)
by: Qiu, Jiabao, et al.
Published: (2024)
Natural Language Satisfiability: Exploring the Problem Distribution and Evaluating Transformer-based Language Models
by: Madusanka, Tharindu, et al.
Published: (2025)
by: Madusanka, Tharindu, et al.
Published: (2025)
Continual Learning in Large Language Models: Methods, Challenges, and Opportunities
by: Chen, Hongyang, et al.
Published: (2026)
by: Chen, Hongyang, et al.
Published: (2026)
KIF: Knowledge Identification and Fusion for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2024)
by: Feng, Yujie, et al.
Published: (2024)
SIMPLE: A Gradient Estimator for $k$-Subset Sampling
by: Ahmed, Kareem, et al.
Published: (2022)
by: Ahmed, Kareem, et al.
Published: (2022)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
by: Zhao, Heng, et al.
Published: (2026)
by: Zhao, Heng, et al.
Published: (2026)
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation
by: Wang, Zihao, et al.
Published: (2024)
by: Wang, Zihao, et al.
Published: (2024)
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
by: Cui, Ganqu, et al.
Published: (2025)
by: Cui, Ganqu, et al.
Published: (2025)
Tractable Offline Learning of Regular Decision Processes
by: Deb, Ahana, et al.
Published: (2024)
by: Deb, Ahana, et al.
Published: (2024)
SEM: Reinforcement Learning for Search-Efficient Large Language Models
by: Sha, Zeyang, et al.
Published: (2025)
by: Sha, Zeyang, et al.
Published: (2025)
CMT: A Memory Compression Method for Continual Knowledge Learning of Large Language Models
by: Li, Dongfang, et al.
Published: (2024)
by: Li, Dongfang, et al.
Published: (2024)
Zero-Variance Gradients for Variational Autoencoders
by: Shao, Zilei, et al.
Published: (2025)
by: Shao, Zilei, et al.
Published: (2025)
AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2025)
by: Feng, Yujie, et al.
Published: (2025)
A Distributional Perspective on Word Learning in Neural Language Models
by: Ficarra, Filippo, et al.
Published: (2025)
by: Ficarra, Filippo, et al.
Published: (2025)
Similar Items
-
TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language Generation
by: Weng, Gwen Yidou, et al.
Published: (2025) -
Polynomial Semantics of Tractable Probabilistic Circuits
by: Broadrick, Oliver, et al.
Published: (2024) -
The Limits of Tractable Marginalization
by: Broadrick, Oliver, et al.
Published: (2025) -
Probabilistic Circuits for Cumulative Distribution Functions
by: Broadrick, Oliver, et al.
Published: (2024) -
Restructuring Tractable Probabilistic Circuits
by: Zhang, Honghua, et al.
Published: (2024)