Can Cross Encoders Produce Useful Sentence Embeddings?
Fuente:
arXiv
Saved in:
| Main Authors: | Ananthakrishnan, Haritha, Dolby, Julian, Kokel, Harsha, Samulowitz, Horst, Srinivas, Kavitha |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SemStruct: Contextualizing Semantic Embeddings with Structural Information for Schema Matching
by: Kang, Inwon, et al.
Published: (2026)
by: Kang, Inwon, et al.
Published: (2026)
TOPJoin: A Context-Aware Multi-Criteria Approach for Joinable Column Search
by: Kokel, Harsha, et al.
Published: (2025)
by: Kokel, Harsha, et al.
Published: (2025)
Evaluating Joinable Column Discovery Approaches for Context-Aware Search
by: Kokel, Harsha, et al.
Published: (2025)
by: Kokel, Harsha, et al.
Published: (2025)
StructText: A Synthetic Table-to-Text Approach for Benchmark Generation with Multi-Dimensional Evaluation
by: Kashyap, Satyananda, et al.
Published: (2025)
by: Kashyap, Satyananda, et al.
Published: (2025)
Shallow Cross-Encoders for Low-Latency Retrieval
by: Petrov, Aleksandr V., et al.
Published: (2024)
by: Petrov, Aleksandr V., et al.
Published: (2024)
Modeling Sequential Sentence Relation to Improve Cross-lingual Dense Retrieval
by: Zhang, Shunyu, et al.
Published: (2023)
by: Zhang, Shunyu, et al.
Published: (2023)
Predicting Oscar-Nominated Screenplays with Sentence Embeddings
by: Gross, Francis
Published: (2025)
by: Gross, Francis
Published: (2025)
LLM-based Embeddings: Attention Values Encode Sentence Semantics Better Than Hidden States
by: Zhang, Yeqin, et al.
Published: (2026)
by: Zhang, Yeqin, et al.
Published: (2026)
UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024)
by: Ciancone, Mathieu, et al.
Published: (2024)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
by: Mozafari, Jamshid, et al.
Published: (2025)
by: Mozafari, Jamshid, et al.
Published: (2025)
Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents
by: Sohrabi, Shirin, et al.
Published: (2026)
by: Sohrabi, Shirin, et al.
Published: (2026)
SetCSE: Set Operations using Contrastive Learning of Sentence Embeddings
by: Liu, Kang
Published: (2024)
by: Liu, Kang
Published: (2024)
Estimating the Usefulness of Clarifying Questions and Answers for Conversational Search
by: Sekulić, Ivan, et al.
Published: (2024)
by: Sekulić, Ivan, et al.
Published: (2024)
Topic mining based on fine-tuning Sentence-BERT and LDA
by: Li, Jianheng, et al.
Published: (2025)
by: Li, Jianheng, et al.
Published: (2025)
SemCSE: Semantic Contrastive Sentence Embeddings Using LLM-Generated Summaries For Scientific Abstracts
by: Brinner, Marc, et al.
Published: (2025)
by: Brinner, Marc, et al.
Published: (2025)
Llama-Embed-Nemotron-8B: A Universal Text Embedding Model for Multilingual and Cross-Lingual Tasks
by: Babakhin, Yauhen, et al.
Published: (2025)
by: Babakhin, Yauhen, et al.
Published: (2025)
LLMEmb: Large Language Model Can Be a Good Embedding Generator for Sequential Recommendation
by: Liu, Qidong, et al.
Published: (2024)
by: Liu, Qidong, et al.
Published: (2024)
SAFE: Improving LLM Systems using Sentence-Level In-generation Attribution
by: Batista, João Eduardo, et al.
Published: (2025)
by: Batista, João Eduardo, et al.
Published: (2025)
Data-CUBE: Data Curriculum for Instruction-based Sentence Representation Learning
by: Min, Yingqian, et al.
Published: (2024)
by: Min, Yingqian, et al.
Published: (2024)
Planning in the LLM Era: Building for Reliability and Efficiency
by: Katz, Michael, et al.
Published: (2026)
by: Katz, Michael, et al.
Published: (2026)
Caraman at SemEval-2026 Task 8: Three-Stage Multi-Turn Retrieval with Query Rewriting, Hybrid Search, and Cross-Encoder Reranking
by: Caraman, David-Maximilian, et al.
Published: (2026)
by: Caraman, David-Maximilian, et al.
Published: (2026)
Semantic Novelty Trajectories in 80,000 Books: A Cross-Corpus Embedding Analysis
by: Zimmerman, Fred
Published: (2026)
by: Zimmerman, Fred
Published: (2026)
DS@GT eRisk 2024: Sentence Transformers for Social Media Risk Assessment
by: Guecha, David, et al.
Published: (2024)
by: Guecha, David, et al.
Published: (2024)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
by: Verma, Shashank, et al.
Published: (2025)
by: Verma, Shashank, et al.
Published: (2025)
Constructing Cross-lingual Consumer Health Vocabulary with Word-Embedding from Comparable User Generated Content
by: Chang, Chia-Hsuan, et al.
Published: (2022)
by: Chang, Chia-Hsuan, et al.
Published: (2022)
TurkEmbed: Turkish Embedding Model on NLI & STS Tasks
by: Ezerceli, Özay, et al.
Published: (2025)
by: Ezerceli, Özay, et al.
Published: (2025)
TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
by: Qiang, Minjie, et al.
Published: (2026)
by: Qiang, Minjie, et al.
Published: (2026)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
by: Abbes, Istabrak, et al.
Published: (2025)
by: Abbes, Istabrak, et al.
Published: (2025)
ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval
by: Chen, Jianlyu, et al.
Published: (2025)
by: Chen, Jianlyu, et al.
Published: (2025)
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders
by: Sarkar, Souvika, et al.
Published: (2025)
by: Sarkar, Souvika, et al.
Published: (2025)
Granite Embedding Models
by: Awasthy, Parul, et al.
Published: (2025)
by: Awasthy, Parul, et al.
Published: (2025)
DELTA: Pre-train a Discriminative Encoder for Legal Case Retrieval via Structural Word Alignment
by: Li, Haitao, et al.
Published: (2024)
by: Li, Haitao, et al.
Published: (2024)
Improving Bilingual Lexicon Induction with Cross-Encoder Reranking
by: Li, Yaoyiran, et al.
Published: (2022)
by: Li, Yaoyiran, et al.
Published: (2022)
ChEmbed: Enhancing Chemical Literature Search Through Domain-Specific Text Embeddings
by: Kasmaee, Ali Shiraee, et al.
Published: (2025)
by: Kasmaee, Ali Shiraee, et al.
Published: (2025)
Twitter Sentiment Analysis using Distributed Word and Sentence Representation
by: Reddy, Dwarampudi Mahidhar, et al.
Published: (2019)
by: Reddy, Dwarampudi Mahidhar, et al.
Published: (2019)
MuRAR: A Simple and Effective Multimodal Retrieval and Answer Refinement Framework for Multimodal Question Answering
by: Zhu, Zhengyuan, et al.
Published: (2024)
by: Zhu, Zhengyuan, et al.
Published: (2024)
Rethinking the Privacy of Text Embeddings: A Reproducibility Study of "Text Embeddings Reveal (Almost) As Much As Text"
by: Seputis, Dominykas, et al.
Published: (2025)
by: Seputis, Dominykas, et al.
Published: (2025)
Granite Embedding R2 Models
by: Awasthy, Parul, et al.
Published: (2025)
by: Awasthy, Parul, et al.
Published: (2025)
Context Embeddings for Efficient Answer Generation in RAG
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Similar Items
-
SemStruct: Contextualizing Semantic Embeddings with Structural Information for Schema Matching
by: Kang, Inwon, et al.
Published: (2026) -
TOPJoin: A Context-Aware Multi-Criteria Approach for Joinable Column Search
by: Kokel, Harsha, et al.
Published: (2025) -
Evaluating Joinable Column Discovery Approaches for Context-Aware Search
by: Kokel, Harsha, et al.
Published: (2025) -
StructText: A Synthetic Table-to-Text Approach for Benchmark Generation with Multi-Dimensional Evaluation
by: Kashyap, Satyananda, et al.
Published: (2025) -
Shallow Cross-Encoders for Low-Latency Retrieval
by: Petrov, Aleksandr V., et al.
Published: (2024)