Pooling And Attention: What Are Effective Designs For LLM-Based Embedding Models?
Fuente:
arXiv
Salvato in:
| Autori principali: | Tang, Yixuan, Yang, Yi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Do We Need Domain-Specific Embedding Models? An Empirical Investigation
di: Tang, Yixuan, et al.
Pubblicazione: (2024)
di: Tang, Yixuan, et al.
Pubblicazione: (2024)
FinMTEB: Finance Massive Text Embedding Benchmark
di: Tang, Yixuan, et al.
Pubblicazione: (2025)
di: Tang, Yixuan, et al.
Pubblicazione: (2025)
LMK > CLS: Landmark Pooling for Dense Embeddings
di: Doshi, Meet, et al.
Pubblicazione: (2026)
di: Doshi, Meet, et al.
Pubblicazione: (2026)
Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval
di: Gao, Hang, et al.
Pubblicazione: (2026)
di: Gao, Hang, et al.
Pubblicazione: (2026)
Lessons Learned on Information Retrieval in Electronic Health Records: A Comparison of Embedding Models and Pooling Strategies
di: Myers, Skatje, et al.
Pubblicazione: (2024)
di: Myers, Skatje, et al.
Pubblicazione: (2024)
LLM-based Embeddings: Attention Values Encode Sentence Semantics Better Than Hidden States
di: Zhang, Yeqin, et al.
Pubblicazione: (2026)
di: Zhang, Yeqin, et al.
Pubblicazione: (2026)
ReinPool: Reinforcement Learning Pooling Multi-Vector Embeddings for Retrieval System
di: Cha, Sungguk, et al.
Pubblicazione: (2026)
di: Cha, Sungguk, et al.
Pubblicazione: (2026)
Adapting General-Purpose Embedding Models to Private Datasets Using Keyword-based Retrieval
di: Wei, Yubai, et al.
Pubblicazione: (2025)
di: Wei, Yubai, et al.
Pubblicazione: (2025)
Evaluating the Effectiveness and Scalability of LLM-Based Data Augmentation for Retrieval
di: Chitale, Pranjal A., et al.
Pubblicazione: (2025)
di: Chitale, Pranjal A., et al.
Pubblicazione: (2025)
Improving Text Embeddings with Large Language Models
di: Wang, Liang, et al.
Pubblicazione: (2023)
di: Wang, Liang, et al.
Pubblicazione: (2023)
Diagnosing LLM Reranker Behavior Under Fixed Evidence Pools
di: Arat, Baris, et al.
Pubblicazione: (2026)
di: Arat, Baris, et al.
Pubblicazione: (2026)
Enhancing Lexicon-Based Text Embeddings with Large Language Models
di: Lei, Yibin, et al.
Pubblicazione: (2025)
di: Lei, Yibin, et al.
Pubblicazione: (2025)
Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction
di: Z., Erwin D. López, et al.
Pubblicazione: (2024)
di: Z., Erwin D. López, et al.
Pubblicazione: (2024)
Granite Embedding Models
di: Awasthy, Parul, et al.
Pubblicazione: (2025)
di: Awasthy, Parul, et al.
Pubblicazione: (2025)
TurkEmbed: Turkish Embedding Model on NLI & STS Tasks
di: Ezerceli, Özay, et al.
Pubblicazione: (2025)
di: Ezerceli, Özay, et al.
Pubblicazione: (2025)
Granite Embedding R2 Models
di: Awasthy, Parul, et al.
Pubblicazione: (2025)
di: Awasthy, Parul, et al.
Pubblicazione: (2025)
Rethinking LLM-Based Recommendations: A Personalized Query-Driven Parallel Integration
di: Han, Donghee, et al.
Pubblicazione: (2025)
di: Han, Donghee, et al.
Pubblicazione: (2025)
Leveraging Generative Models for Real-Time Query-Driven Text Summarization in Large-Scale Web Search
di: Xiong, Zeyu, et al.
Pubblicazione: (2025)
di: Xiong, Zeyu, et al.
Pubblicazione: (2025)
Uncovering the Bigger Picture: Comprehensive Event Understanding Via Diverse News Retrieval
di: Tang, Yixuan, et al.
Pubblicazione: (2025)
di: Tang, Yixuan, et al.
Pubblicazione: (2025)
EmbeddingRWKV: State-Centric Retrieval with Reusable States
di: Hou, Haowen, et al.
Pubblicazione: (2026)
di: Hou, Haowen, et al.
Pubblicazione: (2026)
Bagging-Based Model Merging for Robust General Text Embeddings
di: Zhang, Hengran, et al.
Pubblicazione: (2026)
di: Zhang, Hengran, et al.
Pubblicazione: (2026)
How Do LLM-Generated Texts Impact Term-Based Retrieval Models?
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
Llama-Embed-Nemotron-8B: A Universal Text Embedding Model for Multilingual and Cross-Lingual Tasks
di: Babakhin, Yauhen, et al.
Pubblicazione: (2025)
di: Babakhin, Yauhen, et al.
Pubblicazione: (2025)
Multilingual E5 Text Embeddings: A Technical Report
di: Wang, Liang, et al.
Pubblicazione: (2024)
di: Wang, Liang, et al.
Pubblicazione: (2024)
Text Embeddings by Weakly-Supervised Contrastive Pre-training
di: Wang, Liang, et al.
Pubblicazione: (2022)
di: Wang, Liang, et al.
Pubblicazione: (2022)
Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking
di: Roitero, Kevin, et al.
Pubblicazione: (2025)
di: Roitero, Kevin, et al.
Pubblicazione: (2025)
Zero-Shot Contextual Embeddings via Offline Synthetic Corpus Generation
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
SPAR: Personalized Content-Based Recommendation via Long Engagement Attention
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
di: Qiang, Minjie, et al.
Pubblicazione: (2026)
di: Qiang, Minjie, et al.
Pubblicazione: (2026)
AdaGATE: Adaptive Gap-Aware Token-Efficient Evidence Assembly for Multi-Hop Retrieval-Augmented Generation
di: Guo, Yilin, et al.
Pubblicazione: (2026)
di: Guo, Yilin, et al.
Pubblicazione: (2026)
A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models
di: Zhang, Gongbo, et al.
Pubblicazione: (2024)
di: Zhang, Gongbo, et al.
Pubblicazione: (2024)
INDUS: Effective and Efficient Language Models for Scientific Applications
di: Bhattacharjee, Bishwaranjan, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Bishwaranjan, et al.
Pubblicazione: (2024)
Leveraging Passage Embeddings for Efficient Listwise Reranking with Large Language Models
di: Liu, Qi, et al.
Pubblicazione: (2024)
di: Liu, Qi, et al.
Pubblicazione: (2024)
One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluation
di: Kostiuk, Yevhen, et al.
Pubblicazione: (2026)
di: Kostiuk, Yevhen, et al.
Pubblicazione: (2026)
Applying Text Embedding Models for Efficient Analysis in Labeled Property Graphs
di: Podstawski, Michal
Pubblicazione: (2025)
di: Podstawski, Michal
Pubblicazione: (2025)
LegalMALR:Multi-Agent Query Understanding and LLM-Based Reranking for Chinese Statute Retrieval
di: Li, Yunhan, et al.
Pubblicazione: (2026)
di: Li, Yunhan, et al.
Pubblicazione: (2026)
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking
di: Qiao, Yixuan, et al.
Pubblicazione: (2022)
di: Qiao, Yixuan, et al.
Pubblicazione: (2022)
Improving Topic Relevance Model by Mix-structured Summarization and LLM-based Data Augmentation
di: Liu, Yizhu, et al.
Pubblicazione: (2024)
di: Liu, Yizhu, et al.
Pubblicazione: (2024)
ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval
di: Chen, Jianlyu, et al.
Pubblicazione: (2025)
di: Chen, Jianlyu, et al.
Pubblicazione: (2025)
MMREC: LLM Based Multi-Modal Recommender System
di: Tian, Jiahao, et al.
Pubblicazione: (2024)
di: Tian, Jiahao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Do We Need Domain-Specific Embedding Models? An Empirical Investigation
di: Tang, Yixuan, et al.
Pubblicazione: (2024) -
FinMTEB: Finance Massive Text Embedding Benchmark
di: Tang, Yixuan, et al.
Pubblicazione: (2025) -
LMK > CLS: Landmark Pooling for Dense Embeddings
di: Doshi, Meet, et al.
Pubblicazione: (2026) -
Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval
di: Gao, Hang, et al.
Pubblicazione: (2026) -
Lessons Learned on Information Retrieval in Electronic Health Records: A Comparison of Embedding Models and Pooling Strategies
di: Myers, Skatje, et al.
Pubblicazione: (2024)