Massively Parallel Expectation Maximization For Approximate Posteriors
Fuente:
arXiv
Salvato in:
| Autori principali: | Heap, Thomas, Bowyer, Sam, Aitchison, Laurence |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Using Autodiff to Estimate Posterior Moments, Marginals and Samples
di: Bowyer, Sam, et al.
Pubblicazione: (2023)
di: Bowyer, Sam, et al.
Pubblicazione: (2023)
Position: Don't Use the CLT in LLM Evals With Fewer Than a Few Hundred Datapoints
di: Bowyer, Sam, et al.
Pubblicazione: (2025)
di: Bowyer, Sam, et al.
Pubblicazione: (2025)
Automated Interpretability Metrics Do Not Distinguish Trained and Random Transformers
di: Heap, Thomas, et al.
Pubblicazione: (2025)
di: Heap, Thomas, et al.
Pubblicazione: (2025)
Learning Generation Orders for Masked Discrete Diffusion Models via Variational Inference
di: Fox, David, et al.
Pubblicazione: (2026)
di: Fox, David, et al.
Pubblicazione: (2026)
Why you don't overfit, and don't need Bayes if you only train for one epoch
di: Aitchison, Laurence
Pubblicazione: (2024)
di: Aitchison, Laurence
Pubblicazione: (2024)
Controlling changes to attention logits
di: Anson, Ben, et al.
Pubblicazione: (2025)
di: Anson, Ben, et al.
Pubblicazione: (2025)
Batch size invariant Adam
di: Wang, Xi, et al.
Pubblicazione: (2024)
di: Wang, Xi, et al.
Pubblicazione: (2024)
Learning to Skip the Middle Layers of Transformers
di: Lawson, Tim, et al.
Pubblicazione: (2025)
di: Lawson, Tim, et al.
Pubblicazione: (2025)
How to set AdamW's weight decay as you scale model and dataset size
di: Wang, Xi, et al.
Pubblicazione: (2024)
di: Wang, Xi, et al.
Pubblicazione: (2024)
Stochastic Approximation with Biased MCMC for Expectation Maximization
di: Gruffaz, Samuel, et al.
Pubblicazione: (2024)
di: Gruffaz, Samuel, et al.
Pubblicazione: (2024)
Function-Space Learning Rates
di: Milsom, Edward, et al.
Pubblicazione: (2025)
di: Milsom, Edward, et al.
Pubblicazione: (2025)
Using Neural Networks for Data Cleaning in Weather Datasets
di: Hanslope, Jack R. P., et al.
Pubblicazione: (2024)
di: Hanslope, Jack R. P., et al.
Pubblicazione: (2024)
Flexible Infinite-Width Graph Convolutional Neural Networks
di: Anson, Ben, et al.
Pubblicazione: (2024)
di: Anson, Ben, et al.
Pubblicazione: (2024)
Convolutional Deep Kernel Machines
di: Milsom, Edward, et al.
Pubblicazione: (2023)
di: Milsom, Edward, et al.
Pubblicazione: (2023)
Stochastic Kernel Regularisation Improves Generalisation in Deep Kernel Machines
di: Milsom, Edward, et al.
Pubblicazione: (2024)
di: Milsom, Edward, et al.
Pubblicazione: (2024)
Allocating Variance to Maximize Expectation
di: Leme, Renato Purita Paes, et al.
Pubblicazione: (2025)
di: Leme, Renato Purita Paes, et al.
Pubblicazione: (2025)
Scale-invariant Attention
di: Anson, Ben, et al.
Pubblicazione: (2025)
di: Anson, Ben, et al.
Pubblicazione: (2025)
Diffusion Alignment as Variational Expectation-Maximization
di: Lee, Jaewoo, et al.
Pubblicazione: (2025)
di: Lee, Jaewoo, et al.
Pubblicazione: (2025)
MONGOOSE: Path-wise Smooth Bayesian Optimisation via Meta-learning
di: Yang, Adam X., et al.
Pubblicazione: (2023)
di: Yang, Adam X., et al.
Pubblicazione: (2023)
Bayesian Deep Learning Via Expectation Maximization and Turbo Deep Approximate Message Passing
di: Xu, Wei, et al.
Pubblicazione: (2024)
di: Xu, Wei, et al.
Pubblicazione: (2024)
Density Operator Expectation Maximization
di: Vishnu, Adit, et al.
Pubblicazione: (2025)
di: Vishnu, Adit, et al.
Pubblicazione: (2025)
Deep Generative Clustering with VAEs and Expectation-Maximization
di: Adipoetra, Michael, et al.
Pubblicazione: (2025)
di: Adipoetra, Michael, et al.
Pubblicazione: (2025)
Bayesian Low-rank Adaptation for Large Language Models
di: Yang, Adam X., et al.
Pubblicazione: (2023)
di: Yang, Adam X., et al.
Pubblicazione: (2023)
Learning to Reason in LLMs by Expectation Maximization
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
di: Lee, Junghyun, et al.
Pubblicazione: (2025)
Efficient Benchmarking Is Just Feature Selection and Multiple Regression
di: Bowyer, Sam, et al.
Pubblicazione: (2026)
di: Bowyer, Sam, et al.
Pubblicazione: (2026)
Residual Stream Analysis with Multi-Layer SAEs
di: Lawson, Tim, et al.
Pubblicazione: (2024)
di: Lawson, Tim, et al.
Pubblicazione: (2024)
SEMF: Supervised Expectation-Maximization Framework for Predicting Intervals
di: Azizi, Ilia, et al.
Pubblicazione: (2024)
di: Azizi, Ilia, et al.
Pubblicazione: (2024)
Learning Diffusion Priors from Observations by Expectation Maximization
di: Rozet, François, et al.
Pubblicazione: (2024)
di: Rozet, François, et al.
Pubblicazione: (2024)
Inverse-Free Sparse Variational Gaussian Processes
di: Cortinovis, Stefano, et al.
Pubblicazione: (2026)
di: Cortinovis, Stefano, et al.
Pubblicazione: (2026)
Importance Weighted Expectation-Maximization for Protein Sequence Design
di: Song, Zhenqiao, et al.
Pubblicazione: (2023)
di: Song, Zhenqiao, et al.
Pubblicazione: (2023)
Uncertainty-Aware Graph Self-Training with Expectation-Maximization Regularization
di: Wang, Emily, et al.
Pubblicazione: (2025)
di: Wang, Emily, et al.
Pubblicazione: (2025)
Learning Mixture Density via Natural Gradient Expectation Maximization
di: Chen, Yutao, et al.
Pubblicazione: (2026)
di: Chen, Yutao, et al.
Pubblicazione: (2026)
Expectation Maximization Pseudo Labels
di: Xu, Moucheng, et al.
Pubblicazione: (2023)
di: Xu, Moucheng, et al.
Pubblicazione: (2023)
Robust Classification with Noisy Labels Based on Posterior Maximization
di: Novello, Nicola, et al.
Pubblicazione: (2025)
di: Novello, Nicola, et al.
Pubblicazione: (2025)
Jacobian Sparse Autoencoders: Sparsify Computations, Not Just Activations
di: Farnik, Lucy, et al.
Pubblicazione: (2025)
di: Farnik, Lucy, et al.
Pubblicazione: (2025)
Rodent-Bench
di: Heap, Thomas, et al.
Pubblicazione: (2026)
di: Heap, Thomas, et al.
Pubblicazione: (2026)
Approximate Maximum Likelihood Inference for Acoustic Spatial Capture-Recapture with Unknown Identities, Using Monte Carlo Expectation Maximization
di: Wang, Yuheng, et al.
Pubblicazione: (2024)
di: Wang, Yuheng, et al.
Pubblicazione: (2024)
Convergence of Expectation-Maximization Algorithm with Mixed-Integer Optimization
di: Joseph, Geethu
Pubblicazione: (2024)
di: Joseph, Geethu
Pubblicazione: (2024)
Expectation Maximization (EM) Converges for General Agnostic Mixtures
di: Ghosh, Avishek
Pubblicazione: (2026)
di: Ghosh, Avishek
Pubblicazione: (2026)
Characterizing Evolution in Expectation-Maximization Estimates for Overspecified Mixed Linear Regression
di: Luo, Zhankun, et al.
Pubblicazione: (2025)
di: Luo, Zhankun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Using Autodiff to Estimate Posterior Moments, Marginals and Samples
di: Bowyer, Sam, et al.
Pubblicazione: (2023) -
Position: Don't Use the CLT in LLM Evals With Fewer Than a Few Hundred Datapoints
di: Bowyer, Sam, et al.
Pubblicazione: (2025) -
Automated Interpretability Metrics Do Not Distinguish Trained and Random Transformers
di: Heap, Thomas, et al.
Pubblicazione: (2025) -
Learning Generation Orders for Masked Discrete Diffusion Models via Variational Inference
di: Fox, David, et al.
Pubblicazione: (2026) -
Why you don't overfit, and don't need Bayes if you only train for one epoch
di: Aitchison, Laurence
Pubblicazione: (2024)