FiNERweb: Datasets and Artifacts for Scalable Multilingual Named Entity Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Golde, Jonas, Haller, Patrick, Akbik, Alan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What Matters When Building Universal Multilingual Named Entity Recognition Models?
by: Golde, Jonas, et al.
Published: (2026)
by: Golde, Jonas, et al.
Published: (2026)
Large-Scale Label Interpretation Learning for Few-Shot Named Entity Recognition
by: Golde, Jonas, et al.
Published: (2024)
by: Golde, Jonas, et al.
Published: (2024)
Familiarity: Better Evaluation of Zero-Shot Named Entity Recognition by Quantifying Label Shifts in Synthetic Training Data
by: Golde, Jonas, et al.
Published: (2024)
by: Golde, Jonas, et al.
Published: (2024)
BabyHGRN: Exploring RNNs for Sample-Efficient Training of Language Models
by: Haller, Patrick, et al.
Published: (2024)
by: Haller, Patrick, et al.
Published: (2024)
MastermindEval: A Simple But Scalable Reasoning Benchmark
by: Golde, Jonas, et al.
Published: (2025)
by: Golde, Jonas, et al.
Published: (2025)
Sample-Efficient Language Modeling with Linear Attention and Lightweight Enhancements
by: Haller, Patrick, et al.
Published: (2025)
by: Haller, Patrick, et al.
Published: (2025)
What Matters in Linearizing Language Models? A Comparative Study of Architecture, Scale, and Task Adaptation
by: Haller, Patrick, et al.
Published: (2025)
by: Haller, Patrick, et al.
Published: (2025)
Fabricator: An Open Source Toolkit for Generating Labeled Training Data with Teacher LLMs
by: Golde, Jonas, et al.
Published: (2023)
by: Golde, Jonas, et al.
Published: (2023)
PISA-Bench: The PISA Index as a Multilingual and Multimodal Metric for the Evaluation of Vision-Language Models
by: Haller, Patrick, et al.
Published: (2025)
by: Haller, Patrick, et al.
Published: (2025)
PECC: Problem Extraction and Coding Challenges
by: Haller, Patrick, et al.
Published: (2024)
by: Haller, Patrick, et al.
Published: (2024)
NoiseBench: Benchmarking the Impact of Real Label Noise on Named Entity Recognition
by: Merdjanovska, Elena, et al.
Published: (2024)
by: Merdjanovska, Elena, et al.
Published: (2024)
Question Decomposition for Retrieval-Augmented Generation
by: Ammann, Paul J. L., et al.
Published: (2025)
by: Ammann, Paul J. L., et al.
Published: (2025)
Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling
by: Aynetdinov, Ansar, et al.
Published: (2026)
by: Aynetdinov, Ansar, et al.
Published: (2026)
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
by: Meeus, Quentin, et al.
Published: (2024)
by: Meeus, Quentin, et al.
Published: (2024)
FiNER-ORD: Financial Named Entity Recognition Open Research Dataset
by: Shah, Agam, et al.
Published: (2023)
by: Shah, Agam, et al.
Published: (2023)
Evaluating Design Decisions for Dual Encoder-based Entity Disambiguation
by: Rücker, Susanna, et al.
Published: (2025)
by: Rücker, Susanna, et al.
Published: (2025)
NERdME: a Named Entity Recognition Dataset for Indexing Research Artifacts in Code Repositories
by: Gesese, Genet Asefa, et al.
Published: (2026)
by: Gesese, Genet Asefa, et al.
Published: (2026)
From Data to Knowledge: Evaluating How Efficiently Language Models Learn Facts
by: Christoph, Daniel, et al.
Published: (2025)
by: Christoph, Daniel, et al.
Published: (2025)
Universal NER: A Gold-Standard Multilingual Named Entity Recognition Benchmark
by: Mayhew, Stephen, et al.
Published: (2023)
by: Mayhew, Stephen, et al.
Published: (2023)
DynamicNER: A Dynamic, Multilingual, and Fine-Grained Dataset for LLM-based Named Entity Recognition
by: Luo, Hanjun, et al.
Published: (2024)
by: Luo, Hanjun, et al.
Published: (2024)
Named Entity Recognition in Context
by: Brisson, Colin, et al.
Published: (2025)
by: Brisson, Colin, et al.
Published: (2025)
Federated Incremental Named Entity Recognition
by: Zhang, Duzhen, et al.
Published: (2024)
by: Zhang, Duzhen, et al.
Published: (2024)
Universal NER v2: Towards a Massively Multilingual Named Entity Recognition Benchmark
by: Blevins, Terra, et al.
Published: (2026)
by: Blevins, Terra, et al.
Published: (2026)
Named Entity Recognition for the Kurdish Sorani Language: Dataset Creation and Comparative Analysis
by: Abdalla, Bakhtawar, et al.
Published: (2025)
by: Abdalla, Bakhtawar, et al.
Published: (2025)
MariNER: A Dataset for Historical Brazilian Portuguese Named Entity Recognition
by: Sarcinelli, João Lucas Luz Lima, et al.
Published: (2025)
by: Sarcinelli, João Lucas Luz Lima, et al.
Published: (2025)
A Benchmark Dataset and a Framework for Urdu Multimodal Named Entity Recognition
by: Ahmad, Hussain, et al.
Published: (2025)
by: Ahmad, Hussain, et al.
Published: (2025)
Learning to Rank Context for Named Entity Recognition Using a Synthetic Dataset
by: Amalvy, Arthur, et al.
Published: (2023)
by: Amalvy, Arthur, et al.
Published: (2023)
BioUNER: A Benchmark Dataset for Clinical Urdu Named Entity Recognition
by: Ali, Wazir, et al.
Published: (2026)
by: Ali, Wazir, et al.
Published: (2026)
Selecting and Merging: Towards Adaptable and Scalable Named Entity Recognition with Large Language Models
by: Ding, Zhuojun, et al.
Published: (2025)
by: Ding, Zhuojun, et al.
Published: (2025)
RetrieveAll: A Multilingual Named Entity Recognition Framework with Large Language Models
by: Zhang, Jin, et al.
Published: (2025)
by: Zhang, Jin, et al.
Published: (2025)
Nested Named-Entity Recognition on Vietnamese COVID-19: Dataset and Experiments
by: Lê, Ngoc C., et al.
Published: (2025)
by: Lê, Ngoc C., et al.
Published: (2025)
Beyond Boundaries: Learning a Universal Entity Taxonomy across Datasets and Languages for Open Named Entity Recognition
by: Yang, Yuming, et al.
Published: (2024)
by: Yang, Yuming, et al.
Published: (2024)
A Reasoning Paradigm for Named Entity Recognition
by: Huang, Hui, et al.
Published: (2025)
by: Huang, Hui, et al.
Published: (2025)
A Brief History of Named Entity Recognition
by: Munnangi, Monica
Published: (2024)
by: Munnangi, Monica
Published: (2024)
YoNER: A New Yorùbá Multi-domain Named Entity Recognition Dataset
by: Falola, Peace Busola, et al.
Published: (2026)
by: Falola, Peace Busola, et al.
Published: (2026)
PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature
by: Dao, An, et al.
Published: (2026)
by: Dao, An, et al.
Published: (2026)
ReProCon: Scalable and Resource-Efficient Few-Shot Biomedical Named Entity Recognition
by: Yoo, Jeongkyun, et al.
Published: (2025)
by: Yoo, Jeongkyun, et al.
Published: (2025)
Named Clinical Entity Recognition Benchmark
by: Abdul, Wadood M, et al.
Published: (2024)
by: Abdul, Wadood M, et al.
Published: (2024)
ANCHOLIK-NER: A Benchmark Dataset for Bangla Regional Named Entity Recognition
by: Paul, Bidyarthi, et al.
Published: (2025)
by: Paul, Bidyarthi, et al.
Published: (2025)
SemScore: Automated Evaluation of Instruction-Tuned LLMs based on Semantic Textual Similarity
by: Aynetdinov, Ansar, et al.
Published: (2024)
by: Aynetdinov, Ansar, et al.
Published: (2024)
Similar Items
-
What Matters When Building Universal Multilingual Named Entity Recognition Models?
by: Golde, Jonas, et al.
Published: (2026) -
Large-Scale Label Interpretation Learning for Few-Shot Named Entity Recognition
by: Golde, Jonas, et al.
Published: (2024) -
Familiarity: Better Evaluation of Zero-Shot Named Entity Recognition by Quantifying Label Shifts in Synthetic Training Data
by: Golde, Jonas, et al.
Published: (2024) -
BabyHGRN: Exploring RNNs for Sample-Efficient Training of Language Models
by: Haller, Patrick, et al.
Published: (2024) -
MastermindEval: A Simple But Scalable Reasoning Benchmark
by: Golde, Jonas, et al.
Published: (2025)