Benchmarking pre-trained text embedding models in aligning built asset information
Fuente:
arXiv
Saved in:
| Main Authors: | Shahinmoghadam, Mehrzad, Motamedi, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Recent advances in text embedding: A Comprehensive Review of Top-Performing Methods on the MTEB Benchmark
by: Cao, Hongliu
Published: (2024)
by: Cao, Hongliu
Published: (2024)
Benchmarking large language models for biomedical natural language processing applications and recommendations
by: Chen, Qingyu, et al.
Published: (2023)
by: Chen, Qingyu, et al.
Published: (2023)
A Benchmark for Deep Information Synthesis
by: Paul, Debjit, et al.
Published: (2026)
by: Paul, Debjit, et al.
Published: (2026)
Coarse-Tuning for Ad-hoc Document Retrieval Using Pre-trained Language Models
by: Keyaki, Atsushi, et al.
Published: (2024)
by: Keyaki, Atsushi, et al.
Published: (2024)
LangProBe: a Language Programs Benchmark
by: Tan, Shangyin, et al.
Published: (2025)
by: Tan, Shangyin, et al.
Published: (2025)
CPRM: A LLM-based Continual Pre-training Framework for Relevance Modeling in Commercial Search
by: Wu, Kaixin, et al.
Published: (2024)
by: Wu, Kaixin, et al.
Published: (2024)
BIRCO: A Benchmark of Information Retrieval Tasks with Complex Objectives
by: Wang, Xiaoyue, et al.
Published: (2024)
by: Wang, Xiaoyue, et al.
Published: (2024)
REGEN: A Dataset and Benchmarks with Natural Language Critiques and Narratives
by: Su, Kun, et al.
Published: (2025)
by: Su, Kun, et al.
Published: (2025)
Improving Pre-trained Language Model Sensitivity via Mask Specific losses: A case study on Biomedical NER
by: Abaho, Micheal, et al.
Published: (2024)
by: Abaho, Micheal, et al.
Published: (2024)
IRPAPERS: A Visual Document Benchmark for Scientific Retrieval and Question Answering
by: Shorten, Connor, et al.
Published: (2026)
by: Shorten, Connor, et al.
Published: (2026)
Online and Offline Evaluations of Collaborative Filtering and Content Based Recommender Systems
by: Elahi, Ali, et al.
Published: (2024)
by: Elahi, Ali, et al.
Published: (2024)
Do LLMs Understand Collaborative Signals? Diagnosis and Repair
by: Pouryousef, Shahrooz, et al.
Published: (2025)
by: Pouryousef, Shahrooz, et al.
Published: (2025)
GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval
by: Fernandes, Peter, et al.
Published: (2026)
by: Fernandes, Peter, et al.
Published: (2026)
AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents
by: Hu, Lingxiang, et al.
Published: (2026)
by: Hu, Lingxiang, et al.
Published: (2026)
ArtistMus: A Globally Diverse, Artist-Centric Benchmark for Retrieval-Augmented Music Question Answering
by: Kwon, Daeyong, et al.
Published: (2025)
by: Kwon, Daeyong, et al.
Published: (2025)
OAEI-LLM-T: A TBox Benchmark Dataset for Understanding Large Language Model Hallucinations in Ontology Matching
by: Qiang, Zhangcheng, et al.
Published: (2025)
by: Qiang, Zhangcheng, et al.
Published: (2025)
MultiMind at SemEval-2025 Task 7: Crosslingual Fact-Checked Claim Retrieval via Multi-Source Alignment
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
RAIR: A Rule-Aware Benchmark Uniting Challenging Long-Tail and Visual Salience Subset for E-commerce Relevance Assessment
by: Lu, Chenji, et al.
Published: (2025)
by: Lu, Chenji, et al.
Published: (2025)
UQA: Corpus for Urdu Question Answering
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
Complex QA and language models hybrid architectures, Survey
by: Daull, Xavier, et al.
Published: (2023)
by: Daull, Xavier, et al.
Published: (2023)
Automated Literature Review Using NLP Techniques and LLM-Based Retrieval-Augmented Generation
by: Ali, Nurshat Fateh, et al.
Published: (2024)
by: Ali, Nurshat Fateh, et al.
Published: (2024)
Large language models can accurately predict searcher preferences
by: Thomas, Paul, et al.
Published: (2023)
by: Thomas, Paul, et al.
Published: (2023)
PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
by: Abhyankar, Nikhil, et al.
Published: (2025)
by: Abhyankar, Nikhil, et al.
Published: (2025)
Something's Fishy In The Data Lake: A Critical Re-evaluation of Table Union Search Benchmarks
by: Boutaleb, Allaa, et al.
Published: (2025)
by: Boutaleb, Allaa, et al.
Published: (2025)
Using text embedding models as text classifiers with medical data
by: Goel, Rishabh
Published: (2024)
by: Goel, Rishabh
Published: (2024)
KGLink: A column type annotation method that combines knowledge graph and pre-trained language model
by: Wang, Yubo, et al.
Published: (2024)
by: Wang, Yubo, et al.
Published: (2024)
LitSearch: A Retrieval Benchmark for Scientific Literature Search
by: Ajith, Anirudh, et al.
Published: (2024)
by: Ajith, Anirudh, et al.
Published: (2024)
SCOPE: A Lightweight-training LLM Framework for Air Traffic Control Readback Monitoring
by: Deng, Qihan, et al.
Published: (2026)
by: Deng, Qihan, et al.
Published: (2026)
Personalized Benchmarking: Evaluating LLMs by Individual Preferences
by: Garbacea, Cristina, et al.
Published: (2026)
by: Garbacea, Cristina, et al.
Published: (2026)
ChineseHarm-Bench: A Chinese Harmful Content Detection Benchmark
by: Liu, Kangwei, et al.
Published: (2025)
by: Liu, Kangwei, et al.
Published: (2025)
Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Task
by: Ghashami, Mina, et al.
Published: (2024)
by: Ghashami, Mina, et al.
Published: (2024)
Automated Justification Production for Claim Veracity in Fact Checking: A Survey on Architectures and Approaches
by: Eldifrawi, Islam, et al.
Published: (2024)
by: Eldifrawi, Islam, et al.
Published: (2024)
The Chronicles of RAG: The Retriever, the Chunk and the Generator
by: Finardi, Paulo, et al.
Published: (2024)
by: Finardi, Paulo, et al.
Published: (2024)
Case-Based Reasoning Approach for Solving Financial Question Answering
by: Kim, Yikyung, et al.
Published: (2024)
by: Kim, Yikyung, et al.
Published: (2024)
Generalized knowledge-enhanced framework for biomedical entity and relation extraction
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Towards Robust Evaluation: A Comprehensive Taxonomy of Datasets and Metrics for Open Domain Question Answering in the Era of Large Language Models
by: Srivastava, Akchay, et al.
Published: (2024)
by: Srivastava, Akchay, et al.
Published: (2024)
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
by: Fleshman, William, et al.
Published: (2024)
by: Fleshman, William, et al.
Published: (2024)
Retrieval Augmented Structured Generation: Business Document Information Extraction As Tool Use
by: Cesista, Franz Louis, et al.
Published: (2024)
by: Cesista, Franz Louis, et al.
Published: (2024)
ARL2: Aligning Retrievers for Black-box Large Language Models via Self-guided Adaptive Relevance Labeling
by: Zhang, Lingxi, et al.
Published: (2024)
by: Zhang, Lingxi, et al.
Published: (2024)
Similar Items
-
Recent advances in text embedding: A Comprehensive Review of Top-Performing Methods on the MTEB Benchmark
by: Cao, Hongliu
Published: (2024) -
Benchmarking large language models for biomedical natural language processing applications and recommendations
by: Chen, Qingyu, et al.
Published: (2023) -
A Benchmark for Deep Information Synthesis
by: Paul, Debjit, et al.
Published: (2026) -
Coarse-Tuning for Ad-hoc Document Retrieval Using Pre-trained Language Models
by: Keyaki, Atsushi, et al.
Published: (2024) -
LangProBe: a Language Programs Benchmark
by: Tan, Shangyin, et al.
Published: (2025)