Repurposing Language Models into Embedding Models: Finding the Compute-Optimal Recipe
Fuente:
arXiv
Saved in:
| Main Authors: | Ziarko, Alicja, Jiang, Albert Q., Piotrowski, Bartosz, Li, Wenda, Jamnik, Mateja, Miłoś, Piotr |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
End-to-End Ontology Learning with Large Language Models
by: Lo, Andy, et al.
Published: (2024)
by: Lo, Andy, et al.
Published: (2024)
Contrastive Representations for Temporal Reasoning
by: Ziarko, Alicja, et al.
Published: (2025)
by: Ziarko, Alicja, et al.
Published: (2025)
On the Role of Iterative Computation in Reinforcement Learning
by: Ghugare, Raj, et al.
Published: (2026)
by: Ghugare, Raj, et al.
Published: (2026)
How Well Does Your Tabular Generator Learn the Structure of Tabular Data?
by: Jiang, Xiangjian, et al.
Published: (2025)
by: Jiang, Xiangjian, et al.
Published: (2025)
TabStruct: Measuring Structural Fidelity of Tabular Data
by: Jiang, Xiangjian, et al.
Published: (2025)
by: Jiang, Xiangjian, et al.
Published: (2025)
TabEBM: A Tabular Data Augmentation Method with Distinct Class-Specific Energy-Based Models
by: Margeloiu, Andrei, et al.
Published: (2024)
by: Margeloiu, Andrei, et al.
Published: (2024)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
Tabular Foundation Model for Generative Modelling
by: Jiang, Xiangjian, et al.
Published: (2026)
by: Jiang, Xiangjian, et al.
Published: (2026)
Understanding Inter-Concept Relationships in Concept-Based Models
by: Raman, Naveen, et al.
Published: (2024)
by: Raman, Naveen, et al.
Published: (2024)
Multimodal Lego: Model Merging and Fine-Tuning Across Topologies and Modalities in Biomedicine
by: Hemker, Konstantin, et al.
Published: (2024)
by: Hemker, Konstantin, et al.
Published: (2024)
Hierarchical Concept-based Interpretable Models
by: Hill, Oscar, et al.
Published: (2026)
by: Hill, Oscar, et al.
Published: (2026)
Measuring Cross-Modal Interactions in Multimodal Models
by: Wenderoth, Laura, et al.
Published: (2024)
by: Wenderoth, Laura, et al.
Published: (2024)
RO-FIGS: Efficient and Expressive Tree-Based Ensembles for Tabular Data
by: Matjašec, Urška, et al.
Published: (2025)
by: Matjašec, Urška, et al.
Published: (2025)
ProtoGate: Prototype-based Neural Networks with Global-to-local Feature Selection for Tabular Biomedical Data
by: Jiang, Xiangjian, et al.
Published: (2023)
by: Jiang, Xiangjian, et al.
Published: (2023)
LLM Embeddings for Deep Learning on Tabular Data
by: Koloski, Boshko, et al.
Published: (2025)
by: Koloski, Boshko, et al.
Published: (2025)
Learning to Receive Help: Intervention-Aware Concept Embedding Models
by: Zarlenga, Mateo Espinosa, et al.
Published: (2023)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2023)
Do Concept Bottleneck Models Respect Localities?
by: Raman, Naveen, et al.
Published: (2024)
by: Raman, Naveen, et al.
Published: (2024)
HEALNet: Multimodal Fusion for Heterogeneous Biomedical Data
by: Hemker, Konstantin, et al.
Published: (2023)
by: Hemker, Konstantin, et al.
Published: (2023)
Digging Deeper: Learning Multi-Level Concept Hierarchies
by: Hill, Oscar, et al.
Published: (2026)
by: Hill, Oscar, et al.
Published: (2026)
GCondNet: A Novel Method for Improving Neural Networks on Small High-Dimensional Tabular Data
by: Margeloiu, Andrei, et al.
Published: (2022)
by: Margeloiu, Andrei, et al.
Published: (2022)
Foundations of Interpretable Models
by: Barbiero, Pietro, et al.
Published: (2025)
by: Barbiero, Pietro, et al.
Published: (2025)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
TabMDA: Tabular Manifold Data Augmentation for Any Classifier using Transformers with In-context Subsetting
by: Margeloiu, Andrei, et al.
Published: (2024)
by: Margeloiu, Andrei, et al.
Published: (2024)
Magnushammer: A Transformer-Based Approach to Premise Selection
by: Mikuła, Maciej, et al.
Published: (2023)
by: Mikuła, Maciej, et al.
Published: (2023)
ProofOptimizer: Training Language Models to Simplify Proofs without Human Demonstrations
by: Gu, Alex, et al.
Published: (2025)
by: Gu, Alex, et al.
Published: (2025)
Local Intrinsic Dimensionality for Dynamic Graph Embeddings
by: Knežević, Dušica, et al.
Published: (2024)
by: Knežević, Dušica, et al.
Published: (2024)
A Training Data Recipe to Accelerate A* Search with Language Models
by: Gupta, Devaansh, et al.
Published: (2024)
by: Gupta, Devaansh, et al.
Published: (2024)
Avoiding Leakage Poisoning: Concept Interventions Under Distribution Shifts
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
Efficient Bias Mitigation Without Privileged Information
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
Embedding Empirical Distributions for Computing Optimal Transport Maps
by: Jiang, Mingchen, et al.
Published: (2025)
by: Jiang, Mingchen, et al.
Published: (2025)
Memory Self-Regeneration: Uncovering Hidden Knowledge in Unlearned Models
by: Polowczyk, Agnieszka, et al.
Published: (2025)
by: Polowczyk, Agnieszka, et al.
Published: (2025)
Discrete Lagrangian Neural Networks with Automatic Symmetry Discovery
by: Lishkova, Yana, et al.
Published: (2022)
by: Lishkova, Yana, et al.
Published: (2022)
OpenThoughts: Data Recipes for Reasoning Models
by: Guha, Etash, et al.
Published: (2025)
by: Guha, Etash, et al.
Published: (2025)
Repurposing Protein Language Models for Latent Flow-Based Fitness Optimization
by: Arroyo, Amaru Caceres, et al.
Published: (2026)
by: Arroyo, Amaru Caceres, et al.
Published: (2026)
LongAlign: A Recipe for Long Context Alignment of Large Language Models
by: Bai, Yushi, et al.
Published: (2024)
by: Bai, Yushi, et al.
Published: (2024)
Reasoning-Finetuning Repurposes Latent Representations in Base Models
by: Ward, Jake, et al.
Published: (2025)
by: Ward, Jake, et al.
Published: (2025)
Training Compute-Optimal Protein Language Models
by: Cheng, Xingyi, et al.
Published: (2024)
by: Cheng, Xingyi, et al.
Published: (2024)
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
by: Xiao, Yuxin, et al.
Published: (2024)
by: Xiao, Yuxin, et al.
Published: (2024)
tsGT: Stochastic Time Series Modeling With Transformer
by: Kuciński, Łukasz, et al.
Published: (2024)
by: Kuciński, Łukasz, et al.
Published: (2024)
Similar Items
-
End-to-End Ontology Learning with Large Language Models
by: Lo, Andy, et al.
Published: (2024) -
Contrastive Representations for Temporal Reasoning
by: Ziarko, Alicja, et al.
Published: (2025) -
On the Role of Iterative Computation in Reinforcement Learning
by: Ghugare, Raj, et al.
Published: (2026) -
How Well Does Your Tabular Generator Learn the Structure of Tabular Data?
by: Jiang, Xiangjian, et al.
Published: (2025) -
TabStruct: Measuring Structural Fidelity of Tabular Data
by: Jiang, Xiangjian, et al.
Published: (2025)