Saved in:
| Main Authors: | Mayhew, Stephen, Blevins, Terra, Liu, Shuheng, Šuppa, Marek, Gonen, Hila, Imperial, Joseph Marvin, Karlsson, Börje F., Lin, Peiqin, Ljubešić, Nikola, Miranda, LJ, Plank, Barbara, Riabi, Arij, Pinter, Yuval |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2311.09122 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Universal NER v2: Towards a Massively Multilingual Named Entity Recognition Benchmark
by: Blevins, Terra, et al.
Published: (2026)
by: Blevins, Terra, et al.
Published: (2026)
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
by: Limisiewicz, Tomasz, et al.
Published: (2024)
by: Limisiewicz, Tomasz, et al.
Published: (2024)
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
by: Gonen, Hila, et al.
Published: (2024)
by: Gonen, Hila, et al.
Published: (2024)
Demystifying Prompts in Language Models via Perplexity Estimation
by: Gonen, Hila, et al.
Published: (2022)
by: Gonen, Hila, et al.
Published: (2022)
HiligayNER: A Baseline Named Entity Recognition Model for Hiligaynon
by: Teves, James Ald, et al.
Published: (2025)
by: Teves, James Ald, et al.
Published: (2025)
Can Character-based Language Models Improve Downstream Task Performance in Low-Resource and Noisy Language Scenarios?
by: Riabi, Arij, et al.
Published: (2021)
by: Riabi, Arij, et al.
Published: (2021)
Enriching the NArabizi Treebank: A Multifaceted Approach to Supporting an Under-Resourced Language
by: Riabi, Arij, et al.
Published: (2023)
by: Riabi, Arij, et al.
Published: (2023)
Common Ground, Diverse Roots: The Difficulty of Classifying Common Examples in Spanish Varieties
by: Lopetegui, Javier A., et al.
Published: (2024)
by: Lopetegui, Javier A., et al.
Published: (2024)
Breaking the Curse of Multilinguality with Cross-lingual Expert Language Models
by: Blevins, Terra, et al.
Published: (2024)
by: Blevins, Terra, et al.
Published: (2024)
Cloaked Classifiers: Pseudonymization Strategies on Sensitive Classification Tasks
by: Riabi, Arij, et al.
Published: (2024)
by: Riabi, Arij, et al.
Published: (2024)
Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models
by: Blevins, Terra
Published: (2026)
by: Blevins, Terra
Published: (2026)
Beyond Dataset Creation: Critical View of Annotation Variation and Bias Probing of a Dataset for Online Radical Content Detection
by: Riabi, Arij, et al.
Published: (2024)
by: Riabi, Arij, et al.
Published: (2024)
CodeNER: Code Prompting for Named Entity Recognition
by: Han, Sungwoo, et al.
Published: (2025)
by: Han, Sungwoo, et al.
Published: (2025)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
by: Ahuja, Sanchit, et al.
Published: (2026)
by: Ahuja, Sanchit, et al.
Published: (2026)
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification
by: Kuzman, Taja, et al.
Published: (2024)
by: Kuzman, Taja, et al.
Published: (2024)
CLASSLA-web: Comparable Web Corpora of South Slavic Languages Enriched with Linguistic and Genre Annotation
by: Ljubešić, Nikola, et al.
Published: (2024)
by: Ljubešić, Nikola, et al.
Published: (2024)
OpeNER: Open Polarity Enhanced Named Entity Recognition
by: Rodrigo Agerri
Published: (2013)
by: Rodrigo Agerri
Published: (2013)
WhisperNER: Unified Open Named Entity and Speech Recognition
by: Ayache, Gil, et al.
Published: (2024)
by: Ayache, Gil, et al.
Published: (2024)
Wiki-TabNER: Integrating Named Entity Recognition into Wikipedia Tables
by: Koleva, Aneta, et al.
Published: (2024)
by: Koleva, Aneta, et al.
Published: (2024)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
by: Peng, Siyao, et al.
Published: (2024)
by: Peng, Siyao, et al.
Published: (2024)
NER Retriever: Zero-Shot Named Entity Retrieval with Type-Aware Embeddings
by: Shachar, Or, et al.
Published: (2025)
by: Shachar, Or, et al.
Published: (2025)
ANCHOLIK-NER: A Benchmark Dataset for Bangla Regional Named Entity Recognition
by: Paul, Bidyarthi, et al.
Published: (2025)
by: Paul, Bidyarthi, et al.
Published: (2025)
MFE-NER: Multi-feature Fusion Embedding for Chinese Named Entity Recognition
by: Li, Jiatong, et al.
Published: (2021)
by: Li, Jiatong, et al.
Published: (2021)
WojoodNER 2024: The Second Arabic Named Entity Recognition Shared Task
by: Jarrar, Mustafa, et al.
Published: (2024)
by: Jarrar, Mustafa, et al.
Published: (2024)
SAM-NER: Semantic Archetype Mediation for Zero-Shot Named Entity Recognition
by: Cai, Ruichu, et al.
Published: (2026)
by: Cai, Ruichu, et al.
Published: (2026)
TriG-NER: Triplet-Grid Framework for Discontinuous Named Entity Recognition
by: Cabral, Rina Carines, et al.
Published: (2024)
by: Cabral, Rina Carines, et al.
Published: (2024)
MariNER: A Dataset for Historical Brazilian Portuguese Named Entity Recognition
by: Sarcinelli, João Lucas Luz Lima, et al.
Published: (2025)
by: Sarcinelli, João Lucas Luz Lima, et al.
Published: (2025)
ToNER: Type-oriented Named Entity Recognition with Generative Language Model
by: Jiang, Guochao, et al.
Published: (2024)
by: Jiang, Guochao, et al.
Published: (2024)
CyberNER: A Harmonized STIX Corpus for Cybersecurity Named Entity Recognition
by: Ech-Chammakhy, Yasir, et al.
Published: (2025)
by: Ech-Chammakhy, Yasir, et al.
Published: (2025)
FiNER-ORD: Financial Named Entity Recognition Open Research Dataset
by: Shah, Agam, et al.
Published: (2023)
by: Shah, Agam, et al.
Published: (2023)
NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages
by: Abdullah, Abdulhady Abas, et al.
Published: (2024)
by: Abdullah, Abdulhady Abas, et al.
Published: (2024)
IYKYK: Using language models to decode extremist cryptolects
by: de Kock, Christine, et al.
Published: (2025)
by: de Kock, Christine, et al.
Published: (2025)
Do language models accommodate their users? A study of linguistic convergence
by: Blevins, Terra, et al.
Published: (2025)
by: Blevins, Terra, et al.
Published: (2025)
Comparing Hallucination Detection Metrics for Multilingual Generation
by: Kang, Haoqiang, et al.
Published: (2024)
by: Kang, Haoqiang, et al.
Published: (2024)
Identifying Primary Stress Across Related Languages and Dialects with Transformer-based Speech Encoder Models
by: Ljubešić, Nikola, et al.
Published: (2025)
by: Ljubešić, Nikola, et al.
Published: (2025)
Mići Princ -- A Little Boy Teaching Speech Technologies the Chakavian Dialect
by: Ljubešić, Nikola, et al.
Published: (2026)
by: Ljubešić, Nikola, et al.
Published: (2026)
The ParlaSent Multilingual Training Dataset for Sentiment Identification in Parliamentary Proceedings
by: Mochtak, Michal, et al.
Published: (2023)
by: Mochtak, Michal, et al.
Published: (2023)
The ParlaSpeech Collection of Automatically Generated Speech and Text Datasets from Parliamentary Proceedings
by: Ljubešić, Nikola, et al.
Published: (2024)
by: Ljubešić, Nikola, et al.
Published: (2024)
YoNER: A New Yorùbá Multi-domain Named Entity Recognition Dataset
by: Falola, Peace Busola, et al.
Published: (2026)
by: Falola, Peace Busola, et al.
Published: (2026)
PBa-LLM: Privacy- and Bias-aware NLP using Named-Entity Recognition (NER)
by: Mancera, Gonzalo, et al.
Published: (2025)
by: Mancera, Gonzalo, et al.
Published: (2025)
Similar Items
-
Universal NER v2: Towards a Massively Multilingual Named Entity Recognition Benchmark
by: Blevins, Terra, et al.
Published: (2026) -
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
by: Limisiewicz, Tomasz, et al.
Published: (2024) -
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
by: Gonen, Hila, et al.
Published: (2024) -
Demystifying Prompts in Language Models via Perplexity Estimation
by: Gonen, Hila, et al.
Published: (2022) -
HiligayNER: A Baseline Named Entity Recognition Model for Hiligaynon
by: Teves, James Ald, et al.
Published: (2025)