Empirical Evaluation of Embedding Models in the Context of Text Classification in Document Review in Construction Delay Disputes
Fuente:
arXiv
Guardado en:
| Autores principales: | Wei, Fusheng, Neary, Robert, Qin, Han, Mao, Qiang, Zhang, Jianping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging Machine Learning and Large Language Models for Automated Image Clustering and Description in Legal Discovery
por: Mao, Qiang, et al.
Publicado: (2025)
por: Mao, Qiang, et al.
Publicado: (2025)
A Comparative Study of Retrieval Methods in Azure AI Search
por: Mao, Qiang, et al.
Publicado: (2025)
por: Mao, Qiang, et al.
Publicado: (2025)
Exploiting the Randomness of Large Language Models (LLM) in Text Classification Tasks: Locating Privileged Documents in Legal Matters
por: Huffman, Keith, et al.
Publicado: (2025)
por: Huffman, Keith, et al.
Publicado: (2025)
Detecting Privileged Documents by Ranking Connected Network Entities
por: Zhang, Jianping, et al.
Publicado: (2025)
por: Zhang, Jianping, et al.
Publicado: (2025)
Context is Gold to find the Gold Passage: Evaluating and Training Contextual Document Embeddings
por: Conti, Max, et al.
Publicado: (2025)
por: Conti, Max, et al.
Publicado: (2025)
Dewey Long Context Embedding Model: A Technical Report
por: Zhang, Dun, et al.
Publicado: (2025)
por: Zhang, Dun, et al.
Publicado: (2025)
Decoding Recommendation Behaviors of In-Context Learning LLMs Through Gradient Descent
por: Xu, Yi, et al.
Publicado: (2025)
por: Xu, Yi, et al.
Publicado: (2025)
RAKG:Document-level Retrieval Augmented Knowledge Graph Construction
por: Zhang, Hairong, et al.
Publicado: (2025)
por: Zhang, Hairong, et al.
Publicado: (2025)
ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval
por: Chen, Jianlyu, et al.
Publicado: (2025)
por: Chen, Jianlyu, et al.
Publicado: (2025)
Labeling Case Similarity based on Co-Citation of Legal Articles in Judgment Documents with Empirical Dispute-Based Evaluation
por: Liu, Chao-Lin, et al.
Publicado: (2025)
por: Liu, Chao-Lin, et al.
Publicado: (2025)
Queries Are Not Alone: Clustering Text Embeddings for Video Search
por: Liu, Peyang, et al.
Publicado: (2025)
por: Liu, Peyang, et al.
Publicado: (2025)
Exploring Selective Retrieval-Augmentation for Long-Tail Legal Text Classification
por: Mao, Boheng
Publicado: (2025)
por: Mao, Boheng
Publicado: (2025)
Make Any Collection Navigable: Methods for Constructing and Evaluating Hypergraph of Text
por: Alvarez, Dean E., et al.
Publicado: (2026)
por: Alvarez, Dean E., et al.
Publicado: (2026)
Improving Text Embeddings with Large Language Models
por: Wang, Liang, et al.
Publicado: (2023)
por: Wang, Liang, et al.
Publicado: (2023)
Are ID Embeddings Necessary? Whitening Pre-trained Text Embeddings for Effective Sequential Recommendation
por: Zhang, Lingzi, et al.
Publicado: (2024)
por: Zhang, Lingzi, et al.
Publicado: (2024)
Leveraging LLMs to Evaluate Usefulness of Document
por: Wang, Xingzhu, et al.
Publicado: (2025)
por: Wang, Xingzhu, et al.
Publicado: (2025)
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
por: Ma, Yubo, et al.
Publicado: (2025)
por: Ma, Yubo, et al.
Publicado: (2025)
From Text to Context: An Entailment Approach for News Stakeholder Classification
por: Kuila, Alapan, et al.
Publicado: (2024)
por: Kuila, Alapan, et al.
Publicado: (2024)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
por: Zhou, Yongjie, et al.
Publicado: (2026)
por: Zhou, Yongjie, et al.
Publicado: (2026)
Nemotron ColEmbed V2: Top-Performing Late Interaction Embedding Models for Visual Document Retrieval
por: Moreira, Gabriel de Souza P., et al.
Publicado: (2026)
por: Moreira, Gabriel de Souza P., et al.
Publicado: (2026)
An Empirical Study of Evaluating Long-form Question Answering
por: Xian, Ning, et al.
Publicado: (2025)
por: Xian, Ning, et al.
Publicado: (2025)
Research on Evaluation Methods for Patent Novelty Search Systems and Empirical Analysis
por: Zhang, Shu, et al.
Publicado: (2025)
por: Zhang, Shu, et al.
Publicado: (2025)
mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval
por: Zhang, Xin, et al.
Publicado: (2024)
por: Zhang, Xin, et al.
Publicado: (2024)
Do We Need Domain-Specific Embedding Models? An Empirical Investigation
por: Tang, Yixuan, et al.
Publicado: (2024)
por: Tang, Yixuan, et al.
Publicado: (2024)
RLTP: Reinforcement Learning to Pace for Delayed Impression Modeling in Preloaded Ads
por: Wei, Penghui, et al.
Publicado: (2023)
por: Wei, Penghui, et al.
Publicado: (2023)
Unifying Multimodal Retrieval via Document Screenshot Embedding
por: Ma, Xueguang, et al.
Publicado: (2024)
por: Ma, Xueguang, et al.
Publicado: (2024)
Retrieval Augmented Zero-Shot Text Classification
por: Abdullahi, Tassallah, et al.
Publicado: (2024)
por: Abdullahi, Tassallah, et al.
Publicado: (2024)
Token-Level Graphs for Short Text Classification
por: Donabauer, Gregor, et al.
Publicado: (2024)
por: Donabauer, Gregor, et al.
Publicado: (2024)
Rethinking the Privacy of Text Embeddings: A Reproducibility Study of "Text Embeddings Reveal (Almost) As Much As Text"
por: Seputis, Dominykas, et al.
Publicado: (2025)
por: Seputis, Dominykas, et al.
Publicado: (2025)
Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding
por: Cheng, Yiruo, et al.
Publicado: (2024)
por: Cheng, Yiruo, et al.
Publicado: (2024)
Beyond Benchmarks: Evaluating Embedding Model Similarity for Retrieval Augmented Generation Systems
por: Caspari, Laura, et al.
Publicado: (2024)
por: Caspari, Laura, et al.
Publicado: (2024)
Docs2KG: Unified Knowledge Graph Construction from Heterogeneous Documents Assisted by Large Language Models
por: Sun, Qiang, et al.
Publicado: (2024)
por: Sun, Qiang, et al.
Publicado: (2024)
Text Embeddings by Weakly-Supervised Contrastive Pre-training
por: Wang, Liang, et al.
Publicado: (2022)
por: Wang, Liang, et al.
Publicado: (2022)
HyQE: Ranking Contexts with Hypothetical Query Embeddings
por: Zhou, Weichao, et al.
Publicado: (2024)
por: Zhou, Weichao, et al.
Publicado: (2024)
Adapting General-Purpose Embedding Models to Private Datasets Using Keyword-based Retrieval
por: Wei, Yubai, et al.
Publicado: (2025)
por: Wei, Yubai, et al.
Publicado: (2025)
Enhancing Lexicon-Based Text Embeddings with Large Language Models
por: Lei, Yibin, et al.
Publicado: (2025)
por: Lei, Yibin, et al.
Publicado: (2025)
Multilingual E5 Text Embeddings: A Technical Report
por: Wang, Liang, et al.
Publicado: (2024)
por: Wang, Liang, et al.
Publicado: (2024)
Prompt Tuning on Graph-augmented Low-resource Text Classification
por: Wen, Zhihao, et al.
Publicado: (2023)
por: Wen, Zhihao, et al.
Publicado: (2023)
Bridging Search and Recommendation through Latent Cross Reasoning
por: Shi, Teng, et al.
Publicado: (2025)
por: Shi, Teng, et al.
Publicado: (2025)
When Text-as-Vision Meets Semantic IDs in Generative Recommendation: An Empirical Study
por: Qiao, Shutong, et al.
Publicado: (2026)
por: Qiao, Shutong, et al.
Publicado: (2026)
Ejemplares similares
-
Leveraging Machine Learning and Large Language Models for Automated Image Clustering and Description in Legal Discovery
por: Mao, Qiang, et al.
Publicado: (2025) -
A Comparative Study of Retrieval Methods in Azure AI Search
por: Mao, Qiang, et al.
Publicado: (2025) -
Exploiting the Randomness of Large Language Models (LLM) in Text Classification Tasks: Locating Privileged Documents in Legal Matters
por: Huffman, Keith, et al.
Publicado: (2025) -
Detecting Privileged Documents by Ranking Connected Network Entities
por: Zhang, Jianping, et al.
Publicado: (2025) -
Context is Gold to find the Gold Passage: Evaluating and Training Contextual Document Embeddings
por: Conti, Max, et al.
Publicado: (2025)