Transformers can do Bayesian Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | Bhaskaran, Prajit, Viering, Tom |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LCDB 1.1: A Database Illustrating Learning Curves Are More Ill-Behaved Than Previously Thought
by: Yan, Cheng, et al.
Published: (2025)
by: Yan, Cheng, et al.
Published: (2025)
X-Node: Self-Explanation is All We Need
by: Sengupta, Prajit, et al.
Published: (2025)
by: Sengupta, Prajit, et al.
Published: (2025)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026)
by: Rajendran, Prajit T, et al.
Published: (2026)
Fair Bayesian Model-Based Clustering
by: Lee, Jihu, et al.
Published: (2025)
by: Lee, Jihu, et al.
Published: (2025)
Kalman Bayesian Transformer
by: Jing, Haoming, et al.
Published: (2025)
by: Jing, Haoming, et al.
Published: (2025)
The Bayesian Geometry of Transformer Attention
by: Agarwal, Naman, et al.
Published: (2025)
by: Agarwal, Naman, et al.
Published: (2025)
Finding Clustering Algorithms in the Transformer Architecture
by: Clarkson, Kenneth L., et al.
Published: (2025)
by: Clarkson, Kenneth L., et al.
Published: (2025)
FireGNN: Neuro-Symbolic Graph Neural Networks with Trainable Fuzzy Rules for Interpretable Medical Image Classification
by: Sengupta, Prajit, et al.
Published: (2025)
by: Sengupta, Prajit, et al.
Published: (2025)
Incorporating Unlabelled Data into Bayesian Neural Networks
by: Sharma, Mrinank, et al.
Published: (2023)
by: Sharma, Mrinank, et al.
Published: (2023)
Clustered FedStack: Intermediate Global Models with Bayesian Information Criterion
by: Shaik, Thanveer, et al.
Published: (2023)
by: Shaik, Thanveer, et al.
Published: (2023)
Transformers trained on proteins can learn to attend to Euclidean distance
by: Ellmen, Isaac, et al.
Published: (2025)
by: Ellmen, Isaac, et al.
Published: (2025)
Ranking over Regression for Bayesian Optimization and Molecule Selection
by: Tom, Gary, et al.
Published: (2024)
by: Tom, Gary, et al.
Published: (2024)
Grokking as Structural Inference: Transformers Need Bayesian Lottery Tickets
by: Hidajat, Kai, et al.
Published: (2026)
by: Hidajat, Kai, et al.
Published: (2026)
Clustering by Attention: Leveraging Prior Fitted Transformers for Data Partitioning
by: Shokry, Ahmed, et al.
Published: (2025)
by: Shokry, Ahmed, et al.
Published: (2025)
Accident-Driven Congestion Prediction and Simulation: An Explainable Framework Using Advanced Clustering and Bayesian Networks
by: Talluri, Kranthi Kumar, et al.
Published: (2025)
by: Talluri, Kranthi Kumar, et al.
Published: (2025)
One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning
by: He, Bowen, et al.
Published: (2026)
by: He, Bowen, et al.
Published: (2026)
Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Transformers
by: Li, Albus Yizhuo, et al.
Published: (2026)
by: Li, Albus Yizhuo, et al.
Published: (2026)
Fraud Detection Through Large-Scale Graph Clustering with Heterogeneous Link Transformation
by: Liu, Chi
Published: (2025)
by: Liu, Chi
Published: (2025)
I can't see it but I can Fine-tune it: On Encrypted Fine-tuning of Transformers using Fully Homomorphic Encryption
by: Panzade, Prajwal, et al.
Published: (2024)
by: Panzade, Prajwal, et al.
Published: (2024)
Bayesian Inverse Problems Meet Flow Matching: Efficient and Flexible Inference via Transformers
by: Sherki, Daniil, et al.
Published: (2025)
by: Sherki, Daniil, et al.
Published: (2025)
Self-Clustering Graph Transformer Approach to Model Resting-State Functional Brain Activity
by: Thapaliya, Bishal, et al.
Published: (2025)
by: Thapaliya, Bishal, et al.
Published: (2025)
A Multi-target Bayesian Transformer Framework for Predicting Cardiovascular Disease Biomarkers during Pandemics
by: Inekwe, Trusting, et al.
Published: (2025)
by: Inekwe, Trusting, et al.
Published: (2025)
The Unreasonable Effectiveness Of Early Discarding After One Epoch In Neural Network Hyperparameter Optimization
by: Egele, Romain, et al.
Published: (2024)
by: Egele, Romain, et al.
Published: (2024)
Graph Diffusion that can Insert and Delete
by: Ninniri, Matteo, et al.
Published: (2025)
by: Ninniri, Matteo, et al.
Published: (2025)
Post-Norm can Resharpen Attention
by: Zsámboki, Pál, et al.
Published: (2025)
by: Zsámboki, Pál, et al.
Published: (2025)
Influence-based Attributions can be Manipulated
by: Yadav, Chhavi, et al.
Published: (2024)
by: Yadav, Chhavi, et al.
Published: (2024)
B-TGAT: A Bi-directional Temporal Graph Attention Transformer for Clustering Multivariate Spatiotemporal Data
by: Nji, Francis Ndikum, et al.
Published: (2025)
by: Nji, Francis Ndikum, et al.
Published: (2025)
Integrating Temporal and Structural Context in Graph Transformers for Relational Deep Learning
by: Lachi, Divyansha, et al.
Published: (2025)
by: Lachi, Divyansha, et al.
Published: (2025)
Learning Regularizers: Learning Optimizers that can Regularize
by: Sahoo, Suraj Kumar, et al.
Published: (2025)
by: Sahoo, Suraj Kumar, et al.
Published: (2025)
Dying Clusters Is All You Need -- Deep Clustering With an Unknown Number of Clusters
by: Leiber, Collin, et al.
Published: (2024)
by: Leiber, Collin, et al.
Published: (2024)
Unsupervised anomaly detection algorithms on real-world data: how many do we need?
by: Bouman, Roel, et al.
Published: (2023)
by: Bouman, Roel, et al.
Published: (2023)
Transformers Can Do Arithmetic with the Right Embeddings
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
In-Context Learning can Perform Continual Learning Like Humans
by: Kang, Liuwang, et al.
Published: (2025)
by: Kang, Liuwang, et al.
Published: (2025)
Causality can systematically address the monsters under the bench(marks)
by: Leeb, Felix, et al.
Published: (2025)
by: Leeb, Felix, et al.
Published: (2025)
Preference Learning with Lie Detectors can Induce Honesty or Evasion
by: Cundy, Chris, et al.
Published: (2025)
by: Cundy, Chris, et al.
Published: (2025)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
by: Helff, Lukas, et al.
Published: (2026)
by: Helff, Lukas, et al.
Published: (2026)
ADNF-Clustering: An Adaptive and Dynamic Neuro-Fuzzy Clustering for Leukemia Prediction
by: Aruta, Marco, et al.
Published: (2025)
by: Aruta, Marco, et al.
Published: (2025)
One-Shot Clustering for Federated Learning Under Clustering-Agnostic Assumption
by: Zuziak, Maciej Krzysztof, et al.
Published: (2025)
by: Zuziak, Maciej Krzysztof, et al.
Published: (2025)
Multivariate Beta Mixture Model: Probabilistic Clustering With Flexible Cluster Shapes
by: Hsu, Yung-Peng, et al.
Published: (2024)
by: Hsu, Yung-Peng, et al.
Published: (2024)
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved)
by: Qin, Chongli, et al.
Published: (2025)
by: Qin, Chongli, et al.
Published: (2025)
Similar Items
-
LCDB 1.1: A Database Illustrating Learning Curves Are More Ill-Behaved Than Previously Thought
by: Yan, Cheng, et al.
Published: (2025) -
X-Node: Self-Explanation is All We Need
by: Sengupta, Prajit, et al.
Published: (2025) -
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026) -
Fair Bayesian Model-Based Clustering
by: Lee, Jihu, et al.
Published: (2025) -
Kalman Bayesian Transformer
by: Jing, Haoming, et al.
Published: (2025)