LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | BehnamGhader, Parishad, Adlakha, Vaibhav, Mosbach, Marius, Bahdanau, Dzmitry, Chapados, Nicolas, Reddy, Siva |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM2Vec-Gen: Generative Embeddings from Large Language Models
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2026)
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2026)
Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering
von: Adlakha, Vaibhav, et al.
Veröffentlicht: (2023)
von: Adlakha, Vaibhav, et al.
Veröffentlicht: (2023)
Exploiting Instruction-Following Retrievers for Malicious Information Retrieval
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2025)
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2025)
Forecasting Downstream Performance of LLMs With Proxy Metrics
von: Patel, Arkil, et al.
Veröffentlicht: (2026)
von: Patel, Arkil, et al.
Veröffentlicht: (2026)
Understanding the Influence of Synthetic Data for Text Embedders
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2025)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2025)
How to Get Your LLM to Generate Challenging Problems for Evaluation
von: Patel, Arkil, et al.
Veröffentlicht: (2025)
von: Patel, Arkil, et al.
Veröffentlicht: (2025)
DeepSeek-R1 Thoughtology: Let's think about LLM Reasoning
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2025)
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2025)
BRIDGE: Predicting Human Task Completion Time From Model Performance
von: Liu, Fengyuan, et al.
Veröffentlicht: (2026)
von: Liu, Fengyuan, et al.
Veröffentlicht: (2026)
Evaluating In-Context Learning of Libraries for Code Generation
von: Patel, Arkil, et al.
Veröffentlicht: (2023)
von: Patel, Arkil, et al.
Veröffentlicht: (2023)
LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs
von: Krojer, Benno, et al.
Veröffentlicht: (2026)
von: Krojer, Benno, et al.
Veröffentlicht: (2026)
Not All Data Are Unlearned Equally
von: Krishnan, Aravind, et al.
Veröffentlicht: (2025)
von: Krishnan, Aravind, et al.
Veröffentlicht: (2025)
Build the web for agents, not agents for the web
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
Scope Ambiguities in Large Language Models
von: Kamath, Gaurav, et al.
Veröffentlicht: (2024)
von: Kamath, Gaurav, et al.
Veröffentlicht: (2024)
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
Language Models Largely Exhibit Human-like Constituent Ordering Preferences
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
Large Language Models Are Overparameterized Text Encoders
von: K, Thennal D, et al.
Veröffentlicht: (2024)
von: K, Thennal D, et al.
Veröffentlicht: (2024)
Value Drifts: Tracing Value Alignment During LLM Post-Training
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
New Encoders for German Trained from Scratch: Comparing ModernGBERT with Converted LLM2Vec Models
von: Wunderle, Julia, et al.
Veröffentlicht: (2025)
von: Wunderle, Julia, et al.
Veröffentlicht: (2025)
GT2Vec: Large Language Models as Multi-Modal Encoders for Text and Graph-Structured Data
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
Explaining Large Language Models Decisions Using Shapley Values
von: Mohammadi, Behnam
Veröffentlicht: (2024)
von: Mohammadi, Behnam
Veröffentlicht: (2024)
Large Language Models are Powerful Electronic Health Record Encoders
von: Hegselmann, Stefan, et al.
Veröffentlicht: (2025)
von: Hegselmann, Stefan, et al.
Veröffentlicht: (2025)
LLMSQL: Upgrading WikiSQL for the LLM Era of Text-to-SQL
von: Pihulski, Dzmitry, et al.
Veröffentlicht: (2025)
von: Pihulski, Dzmitry, et al.
Veröffentlicht: (2025)
Creativity Has Left the Chat: The Price of Debiasing Language Models
von: Mohammadi, Behnam
Veröffentlicht: (2024)
von: Mohammadi, Behnam
Veröffentlicht: (2024)
LLMs can learn self-restraint through iterative self-reflection
von: Piché, Alexandre, et al.
Veröffentlicht: (2024)
von: Piché, Alexandre, et al.
Veröffentlicht: (2024)
BELL: Benchmarking the Explainability of Large Language Models
von: Ahmed, Syed Quiser, et al.
Veröffentlicht: (2025)
von: Ahmed, Syed Quiser, et al.
Veröffentlicht: (2025)
NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild
von: Murty, Shikhar, et al.
Veröffentlicht: (2024)
von: Murty, Shikhar, et al.
Veröffentlicht: (2024)
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
von: Rizvi-Martel, Michael, et al.
Veröffentlicht: (2026)
von: Rizvi-Martel, Michael, et al.
Veröffentlicht: (2026)
Your Language Model Can Secretly Write Like Humans: Contrastive Paraphrase Attacks on LLM-Generated Text Detectors
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Societal Alignment Frameworks Can Improve LLM Alignment
von: Stańczak, Karolina, et al.
Veröffentlicht: (2025)
von: Stańczak, Karolina, et al.
Veröffentlicht: (2025)
Investigating Adversarial Trigger Transfer in Large Language Models
von: Meade, Nicholas, et al.
Veröffentlicht: (2024)
von: Meade, Nicholas, et al.
Veröffentlicht: (2024)
XC-Cache: Cross-Attending to Cached Context for Efficient LLM Inference
von: Monteiro, João, et al.
Veröffentlicht: (2024)
von: Monteiro, João, et al.
Veröffentlicht: (2024)
Sparse Auto-Encoders and Holism about Large Language Models
von: Grindrod, Jumbly
Veröffentlicht: (2026)
von: Grindrod, Jumbly
Veröffentlicht: (2026)
NAG: A Unified Native Architecture for Encoder-free Text-Graph Modeling in Language Models
von: Gong, Haisong, et al.
Veröffentlicht: (2026)
von: Gong, Haisong, et al.
Veröffentlicht: (2026)
Language, Culture, and Ideology: Personalizing Offensiveness Detection in Political Tweets with Reasoning LLMs
von: Pihulski, Dzmitry, et al.
Veröffentlicht: (2025)
von: Pihulski, Dzmitry, et al.
Veröffentlicht: (2025)
On Linearizing Structured Data in Encoder-Decoder Language Models: Insights from Text-to-SQL
von: Shao, Yutong, et al.
Veröffentlicht: (2024)
von: Shao, Yutong, et al.
Veröffentlicht: (2024)
TnT-LLM: Text Mining at Scale with Large Language Models
von: Wan, Mengting, et al.
Veröffentlicht: (2024)
von: Wan, Mengting, et al.
Veröffentlicht: (2024)
PLDR-LLM: Large Language Model from Power Law Decoder Representations
von: Gokden, Burc
Veröffentlicht: (2024)
von: Gokden, Burc
Veröffentlicht: (2024)
Are Large Language Models Truly Smarter Than Humans?
von: M, Eshwar Reddy, et al.
Veröffentlicht: (2026)
von: M, Eshwar Reddy, et al.
Veröffentlicht: (2026)
Pel, A Programming Language for Orchestrating AI Agents
von: Mohammadi, Behnam
Veröffentlicht: (2025)
von: Mohammadi, Behnam
Veröffentlicht: (2025)
CardiffNLP at CLEARS-2025: Prompting Large Language Models for Plain Language and Easy-to-Read Text Rewriting
von: Ayesh, Mutaz, et al.
Veröffentlicht: (2025)
von: Ayesh, Mutaz, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLM2Vec-Gen: Generative Embeddings from Large Language Models
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2026) -
Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering
von: Adlakha, Vaibhav, et al.
Veröffentlicht: (2023) -
Exploiting Instruction-Following Retrievers for Malicious Information Retrieval
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2025) -
Forecasting Downstream Performance of LLMs With Proxy Metrics
von: Patel, Arkil, et al.
Veröffentlicht: (2026) -
Understanding the Influence of Synthetic Data for Text Embedders
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2025)