Transformers Can Do Bayesian Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Müller, Samuel, Hollmann, Noah, Arango, Sebastian Pineda, Grabocka, Josif, Hutter, Frank |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quick-Tune: Quickly Learning Which Pretrained Model to Finetune and How
by: Arango, Sebastian Pineda, et al.
Published: (2023)
by: Arango, Sebastian Pineda, et al.
Published: (2023)
Interpretable Mesomorphic Networks for Tabular Data
by: Kadra, Arlind, et al.
Published: (2023)
by: Kadra, Arlind, et al.
Published: (2023)
Regularized Neural Ensemblers
by: Arango, Sebastian Pineda, et al.
Published: (2024)
by: Arango, Sebastian Pineda, et al.
Published: (2024)
Ensembling Finetuned Language Models for Text Classification
by: Arango, Sebastian Pineda, et al.
Published: (2024)
by: Arango, Sebastian Pineda, et al.
Published: (2024)
Bayes' Power for Explaining In-Context Learning Generalizations
by: Müller, Samuel, et al.
Published: (2024)
by: Müller, Samuel, et al.
Published: (2024)
Position: The Future of Bayesian Prediction Is Prior-Fitted
by: Müller, Samuel, et al.
Published: (2025)
by: Müller, Samuel, et al.
Published: (2025)
FairPFN: Transformers Can do Counterfactual Fairness
by: Robertson, Jake, et al.
Published: (2024)
by: Robertson, Jake, et al.
Published: (2024)
Hierarchical Transformers are Efficient Meta-Reinforcement Learners
by: Shala, Gresa, et al.
Published: (2024)
by: Shala, Gresa, et al.
Published: (2024)
Multi-objective Differentiable Neural Architecture Search
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
FairPFN: A Tabular Foundation Model for Causal Fairness
by: Robertson, Jake, et al.
Published: (2025)
by: Robertson, Jake, et al.
Published: (2025)
When is Warmstarting Effective for Scaling Language Models?
by: Mallik, Neeratyoy, et al.
Published: (2026)
by: Mallik, Neeratyoy, et al.
Published: (2026)
Drift-Resilient TabPFN: In-Context Learning Temporal Distribution Shifts on Tabular Data
by: Helli, Kai, et al.
Published: (2024)
by: Helli, Kai, et al.
Published: (2024)
Do-PFN: In-Context Learning for Causal Effect Estimation
by: Robertson, Jake, et al.
Published: (2025)
by: Robertson, Jake, et al.
Published: (2025)
Warmstarting for Scaling Language Models
by: Mallik, Neeratyoy, et al.
Published: (2024)
by: Mallik, Neeratyoy, et al.
Published: (2024)
Real-TabPFN: Improving Tabular Foundation Models via Continued Pre-training With Real-World Data
by: Garg, Anurag, et al.
Published: (2025)
by: Garg, Anurag, et al.
Published: (2025)
Learning to Order: Task Sequencing as In-Context Optimization
by: Kobiolka, Jan, et al.
Published: (2026)
by: Kobiolka, Jan, et al.
Published: (2026)
Tabular Data: Is Deep Learning all you need?
by: Zabërgja, Guri, et al.
Published: (2024)
by: Zabërgja, Guri, et al.
Published: (2024)
POP: Prior-Fitted First-Order Optimization Policies
by: Kobiolka, Jan, et al.
Published: (2026)
by: Kobiolka, Jan, et al.
Published: (2026)
End-to-End Compression for Tabular Foundation Models
by: Zabërgja, Guri, et al.
Published: (2026)
by: Zabërgja, Guri, et al.
Published: (2026)
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
by: Abdalla, M. H. I., et al.
Published: (2025)
by: Abdalla, M. H. I., et al.
Published: (2025)
Lightweight Correlation-Aware Table Compression
by: Stoian, Mihail, et al.
Published: (2024)
by: Stoian, Mihail, et al.
Published: (2024)
Can Transformers Learn Full Bayesian Inference in Context?
by: Reuter, Arik, et al.
Published: (2025)
by: Reuter, Arik, et al.
Published: (2025)
Tune My Adam, Please!
by: Athanasiadis, Theodoros, et al.
Published: (2025)
by: Athanasiadis, Theodoros, et al.
Published: (2025)
Self-Correcting Bayesian Optimization through Bayesian Active Learning
by: Hvarfner, Carl, et al.
Published: (2023)
by: Hvarfner, Carl, et al.
Published: (2023)
A General Framework for User-Guided Bayesian Optimization
by: Hvarfner, Carl, et al.
Published: (2023)
by: Hvarfner, Carl, et al.
Published: (2023)
From Tables to Time: Extending TabPFN-v2 to Time Series Forecasting
by: Hoo, Shi Bin, et al.
Published: (2025)
by: Hoo, Shi Bin, et al.
Published: (2025)
Concise One-Layer Transformers Can Do Function Evaluation (Sometimes)
by: Strobl, Lena, et al.
Published: (2025)
by: Strobl, Lena, et al.
Published: (2025)
Can Transformers Do Enumerative Geometry?
by: Hashemi, Baran, et al.
Published: (2024)
by: Hashemi, Baran, et al.
Published: (2024)
Towards Improved Variational Inference for Deep Bayesian Models
by: Ober, Sebastian W.
Published: (2024)
by: Ober, Sebastian W.
Published: (2024)
Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch
by: Ferreira, Fabio, et al.
Published: (2026)
by: Ferreira, Fabio, et al.
Published: (2026)
In-Context Freeze-Thaw Bayesian Optimization for Hyperparameter Optimization
by: Rakotoarison, Herilalaina, et al.
Published: (2024)
by: Rakotoarison, Herilalaina, et al.
Published: (2024)
Transformers Can Do Arithmetic with the Right Embeddings
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
Unreflected Use of Tabular Data Repositories Can Undermine Research Quality
by: Tschalzev, Andrej, et al.
Published: (2025)
by: Tschalzev, Andrej, et al.
Published: (2025)
Robust Experimental Design via Generalised Bayesian Inference
by: Barlas, Yasir Zubayr, et al.
Published: (2025)
by: Barlas, Yasir Zubayr, et al.
Published: (2025)
Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation
by: Whittle, George, et al.
Published: (2025)
by: Whittle, George, et al.
Published: (2025)
Meta-Attention: Bayesian Per-Token Routing for Efficient Transformer Inference
by: Ferrari, Alan
Published: (2026)
by: Ferrari, Alan
Published: (2026)
ALINE: Joint Amortization for Bayesian Inference and Active Data Acquisition
by: Huang, Daolang, et al.
Published: (2025)
by: Huang, Daolang, et al.
Published: (2025)
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
On Sequential Bayesian Inference for Continual Learning
by: Kessler, Samuel, et al.
Published: (2023)
by: Kessler, Samuel, et al.
Published: (2023)
Generalization Can Emerge in Tabular Foundation Models From a Single Table
by: Ma, Junwei, et al.
Published: (2025)
by: Ma, Junwei, et al.
Published: (2025)
Similar Items
-
Quick-Tune: Quickly Learning Which Pretrained Model to Finetune and How
by: Arango, Sebastian Pineda, et al.
Published: (2023) -
Interpretable Mesomorphic Networks for Tabular Data
by: Kadra, Arlind, et al.
Published: (2023) -
Regularized Neural Ensemblers
by: Arango, Sebastian Pineda, et al.
Published: (2024) -
Ensembling Finetuned Language Models for Text Classification
by: Arango, Sebastian Pineda, et al.
Published: (2024) -
Bayes' Power for Explaining In-Context Learning Generalizations
by: Müller, Samuel, et al.
Published: (2024)