MTEB-NL and E5-NL: Embedding Benchmark and Models for Dutch
Fuente:
arXiv
Saved in:
| Main Authors: | Banar, Nikolay, Lotfi, Ehsan, Van Nooten, Jens, Arhiliuc, Cristina, Kliocaite, Marija, Daelemans, Walter |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BEIR-NL: Zero-shot Information Retrieval Benchmark for the Dutch Language
by: Banar, Nikolay, et al.
Published: (2024)
by: Banar, Nikolay, et al.
Published: (2024)
Bilingual BSARD: Extending Statutory Article Retrieval to Dutch
by: Lotfi, Ehsan, et al.
Published: (2024)
by: Lotfi, Ehsan, et al.
Published: (2024)
One Size Does Not Fit All: Exploring Variable Thresholds for Distance-Based Multi-Label Text Classification
by: Van Nooten, Jens, et al.
Published: (2025)
by: Van Nooten, Jens, et al.
Published: (2025)
An Agentic AI Framework for Training General Practitioner Student Skills
by: De Marez, Victor, et al.
Published: (2025)
by: De Marez, Victor, et al.
Published: (2025)
PersonalityChat: Conversation Distillation for Personalized Dialog Modeling with Facts and Traits
by: Lotfi, Ehsan, et al.
Published: (2024)
by: Lotfi, Ehsan, et al.
Published: (2024)
AfriMTEB and AfriE5: Benchmarking and Adapting Text Embedding Models for African Languages
by: Uemura, Kosei, et al.
Published: (2025)
by: Uemura, Kosei, et al.
Published: (2025)
PL-MTEB: Polish Massive Text Embedding Benchmark
by: Poświata, Rafał, et al.
Published: (2024)
by: Poświata, Rafał, et al.
Published: (2024)
Do You Get the Hint? Benchmarking LLMs on the Board Game Concept
by: Gevers, Ine, et al.
Published: (2025)
by: Gevers, Ine, et al.
Published: (2025)
Evaluating NL2SQL via SQL2NL
by: Safarzadeh, Mohammadtaher, et al.
Published: (2025)
by: Safarzadeh, Mohammadtaher, et al.
Published: (2025)
NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions
by: Hou, Shizheng, et al.
Published: (2026)
by: Hou, Shizheng, et al.
Published: (2026)
FinMTEB: Finance Massive Text Embedding Benchmark
by: Tang, Yixuan, et al.
Published: (2025)
by: Tang, Yixuan, et al.
Published: (2025)
VN-MTEB: Vietnamese Massive Text Embedding Benchmark
by: Pham, Loc, et al.
Published: (2025)
by: Pham, Loc, et al.
Published: (2025)
FD-NL2SQL: Feedback-Driven Clinical NL2SQL that Improves with Use
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration
by: Moore, Hayden, et al.
Published: (2026)
by: Moore, Hayden, et al.
Published: (2026)
GPT-NL Public Corpus: A Permissively Licensed, Dutch-First Dataset for LLM Pre-training
by: van Oort, Jesse, et al.
Published: (2026)
by: van Oort, Jesse, et al.
Published: (2026)
$R^3$-NL2GQL: A Model Coordination and Knowledge Graph Alignment Approach for NL2GQL
by: Zhou, Yuhang, et al.
Published: (2023)
by: Zhou, Yuhang, et al.
Published: (2023)
Hybrid-NL2SVA: Integrating RAG and Finetuning for LLM-based NL2SVA
by: Xiao, Weihua, et al.
Published: (2025)
by: Xiao, Weihua, et al.
Published: (2025)
Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification
by: Thatikonda, Ramya Keerthy, et al.
Published: (2024)
by: Thatikonda, Ramya Keerthy, et al.
Published: (2024)
FaMTEB: Massive Text Embedding Benchmark in Persian Language
by: Zinvandi, Erfan, et al.
Published: (2025)
by: Zinvandi, Erfan, et al.
Published: (2025)
NL-Eye: Abductive NLI for Images
by: Ventura, Mor, et al.
Published: (2024)
by: Ventura, Mor, et al.
Published: (2024)
Bag of Lies: Robustness in Continuous Pre-training BERT
by: Gevers, Ine, et al.
Published: (2024)
by: Gevers, Ine, et al.
Published: (2024)
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025)
by: Chung, Isaac, et al.
Published: (2025)
Swan and ArabicMTEB: Dialect-Aware, Arabic-Centric, Cross-Lingual, and Cross-Cultural Embedding Models and Benchmarks
by: Bhatia, Gagan, et al.
Published: (2024)
by: Bhatia, Gagan, et al.
Published: (2024)
SPENCE: A Syntactic Probe for Detecting Contamination in NL2SQL Benchmarks
by: Safarzadeh, Mohammadtaher, et al.
Published: (2026)
by: Safarzadeh, Mohammadtaher, et al.
Published: (2026)
Effectiveness of Prompt Optimization in NL2SQL Systems
by: Gurajada, Sairam, et al.
Published: (2025)
by: Gurajada, Sairam, et al.
Published: (2025)
CLARITY: A Framework and Benchmark for Conversational Language Ambiguity and Unanswerability in Interactive NL2SQL Systems
by: Sarwar, Tabinda, et al.
Published: (2026)
by: Sarwar, Tabinda, et al.
Published: (2026)
An Agentic System for Schema Aware NL2SQL Generation
by: Onyango, David, et al.
Published: (2026)
by: Onyango, David, et al.
Published: (2026)
NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging
by: Zhang, Weiming, et al.
Published: (2025)
by: Zhang, Weiming, et al.
Published: (2025)
Blar-SQL: Faster, Stronger, Smaller NL2SQL
by: Domínguez, José Manuel, et al.
Published: (2024)
by: Domínguez, José Manuel, et al.
Published: (2024)
ODIN: A NL2SQL Recommender to Handle Schema Ambiguity
by: Vaidya, Kapil, et al.
Published: (2025)
by: Vaidya, Kapil, et al.
Published: (2025)
NL2TL: Transforming Natural Languages to Temporal Logics using Large Language Models
by: Chen, Yongchao, et al.
Published: (2023)
by: Chen, Yongchao, et al.
Published: (2023)
VeriMinder: Mitigating Analytical Vulnerabilities in NL2SQL
by: Mohole, Shubham, et al.
Published: (2025)
by: Mohole, Shubham, et al.
Published: (2025)
NL2KQL: From Natural Language to Kusto Query
by: Tang, Xinye, et al.
Published: (2024)
by: Tang, Xinye, et al.
Published: (2024)
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024)
by: Ciancone, Mathieu, et al.
Published: (2024)
TailorSQL: An NL2SQL System Tailored to Your Query Workload
by: Vaidya, Kapil, et al.
Published: (2025)
by: Vaidya, Kapil, et al.
Published: (2025)
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models
by: Liao, Weibin, et al.
Published: (2025)
by: Liao, Weibin, et al.
Published: (2025)
Do LLMs Really Struggle at NL-FOL Translation? Revealing their Strengths via a Novel Benchmarking Strategy
by: Brunello, Andrea, et al.
Published: (2025)
by: Brunello, Andrea, et al.
Published: (2025)
Distill-C: Enhanced NL2SQL via Distilled Customization with LLMs
by: Hoang, Cong Duy Vu, et al.
Published: (2025)
by: Hoang, Cong Duy Vu, et al.
Published: (2025)
NL2Formula: Generating Spreadsheet Formulas from Natural Language Queries
by: Zhao, Wei, et al.
Published: (2024)
by: Zhao, Wei, et al.
Published: (2024)
NL2Repo-Bench: Towards Long-Horizon Repository Generation Evaluation of Coding Agents
by: Ding, Jingzhe, et al.
Published: (2025)
by: Ding, Jingzhe, et al.
Published: (2025)
Similar Items
-
BEIR-NL: Zero-shot Information Retrieval Benchmark for the Dutch Language
by: Banar, Nikolay, et al.
Published: (2024) -
Bilingual BSARD: Extending Statutory Article Retrieval to Dutch
by: Lotfi, Ehsan, et al.
Published: (2024) -
One Size Does Not Fit All: Exploring Variable Thresholds for Distance-Based Multi-Label Text Classification
by: Van Nooten, Jens, et al.
Published: (2025) -
An Agentic AI Framework for Training General Practitioner Student Skills
by: De Marez, Victor, et al.
Published: (2025) -
PersonalityChat: Conversation Distillation for Personalized Dialog Modeling with Facts and Traits
by: Lotfi, Ehsan, et al.
Published: (2024)