Annotation Errors and NER: A Study with OntoNotes 5.0
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bernier-Colborne, Gabriel, Vajjala, Sowmya |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Test Set Quality in Multilingual LLM Evaluation
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
IndicGEC: Powerful Models, or a Measurement Mirage?
par: Vajjala, Sowmya
Publié: (2025)
par: Vajjala, Sowmya
Publié: (2025)
The Problem with Safety Classification is not just the Models
par: Vajjala, Sowmya
Publié: (2025)
par: Vajjala, Sowmya
Publié: (2025)
MetricalARGS: A Taxonomy for Studying Metrical Poetry with LLMs
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
Dravidian language family through Universal Dependencies lens
par: Rama, Taraka, et autres
Publié: (2024)
par: Rama, Taraka, et autres
Publié: (2024)
Text Classification in the LLM Era -- Where do we stand?
par: Vajjala, Sowmya, et autres
Publié: (2025)
par: Vajjala, Sowmya, et autres
Publié: (2025)
Does Synthetic Data Help Named Entity Recognition for Low-Resource Languages?
par: Kamath, Gaurav, et autres
Publié: (2025)
par: Kamath, Gaurav, et autres
Publié: (2025)
MATA: Mindful Assessment of the Telugu Abilities of Large Language Models
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
par: Kranti, Chalamalasetti, et autres
Publié: (2025)
Scope Ambiguities in Large Language Models
par: Kamath, Gaurav, et autres
Publié: (2024)
par: Kamath, Gaurav, et autres
Publié: (2024)
LLMs in Education: Novel Perspectives, Challenges, and Opportunities
par: Alhafni, Bashar, et autres
Publié: (2024)
par: Alhafni, Bashar, et autres
Publié: (2024)
Opportunities and Challenges of LLMs in Education: An NLP Perspective
par: Vajjala, Sowmya, et autres
Publié: (2025)
par: Vajjala, Sowmya, et autres
Publié: (2025)
Human-Annotated NER Dataset for the Kyrgyz Language
par: Turatali, Timur, et autres
Publié: (2025)
par: Turatali, Timur, et autres
Publié: (2025)
OpenNER 1.0: Standardized Open-Access Named Entity Recognition Datasets in 50+ Languages
par: Palen-Michel, Chester, et autres
Publié: (2024)
par: Palen-Michel, Chester, et autres
Publié: (2024)
Augmenting NER Datasets with LLMs: Towards Automated and Refined Annotation
par: Naraki, Yuji, et autres
Publié: (2024)
par: Naraki, Yuji, et autres
Publié: (2024)
Towards DS-NER: Unveiling and Addressing Latent Noise in Distant Annotations
par: Ding, Yuyang, et autres
Publié: (2025)
par: Ding, Yuyang, et autres
Publié: (2025)
WikiNER-fr-gold: A Gold-Standard NER Corpus
par: Cao, Danrun, et autres
Publié: (2024)
par: Cao, Danrun, et autres
Publié: (2024)
The GELATO Dataset for Legislative NER
par: Flynn, Matthew, et autres
Publié: (2026)
par: Flynn, Matthew, et autres
Publié: (2026)
VerifiNER: Verification-augmented NER via Knowledge-grounded Reasoning with Large Language Models
par: Kim, Seoyeon, et autres
Publié: (2024)
par: Kim, Seoyeon, et autres
Publié: (2024)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
par: Bogdanov, Sergei, et autres
Publié: (2024)
par: Bogdanov, Sergei, et autres
Publié: (2024)
LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories
par: Vishnubhotla, Krishnapriya, et autres
Publié: (2026)
par: Vishnubhotla, Krishnapriya, et autres
Publié: (2026)
The Million-Label NER: Breaking Scale Barriers with GLiNER bi-encoder
par: Stepanov, Ihor, et autres
Publié: (2026)
par: Stepanov, Ihor, et autres
Publié: (2026)
Astro-NER -- Astronomy Named Entity Recognition: Is GPT a Good Domain Expert Annotator?
par: Evans, Julia, et autres
Publié: (2024)
par: Evans, Julia, et autres
Publié: (2024)
ERNIE 5.0 Technical Report
par: Wang, Haifeng, et autres
Publié: (2026)
par: Wang, Haifeng, et autres
Publié: (2026)
2M-NER: Contrastive Learning for Multilingual and Multimodal NER with Language and Modal Fusion
par: Wang, Dongsheng, et autres
Publié: (2024)
par: Wang, Dongsheng, et autres
Publié: (2024)
Comparative Analysis of Extrinsic Factors for NER in French
par: Yang, Grace, et autres
Publié: (2024)
par: Yang, Grace, et autres
Publié: (2024)
On-the-fly Definition Augmentation of LLMs for Biomedical NER
par: Munnangi, Monica, et autres
Publié: (2024)
par: Munnangi, Monica, et autres
Publié: (2024)
Do LLMs Surpass Encoders for Biomedical NER?
par: Obeidat, Motasem S, et autres
Publié: (2025)
par: Obeidat, Motasem S, et autres
Publié: (2025)
Novel Benchmark for NER in the Wastewater and Stormwater Domain
par: Cardillo, Franco Alberto, et autres
Publié: (2025)
par: Cardillo, Franco Alberto, et autres
Publié: (2025)
L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models
par: Chaudhari, Harsh, et autres
Publié: (2023)
par: Chaudhari, Harsh, et autres
Publié: (2023)
Comparative Study of Zero-Shot Cross-Lingual Transfer for Bodo POS and NER Tagging Using Gemini 2.0 Flash Thinking Experimental Model
par: Narzary, Sanjib, et autres
Publié: (2025)
par: Narzary, Sanjib, et autres
Publié: (2025)
PrOnto: Language Model Evaluations for 859 Languages
par: Gessler, Luke
Publié: (2023)
par: Gessler, Luke
Publié: (2023)
ErAConD : Error Annotated Conversational Dialog Dataset for Grammatical Error Correction
par: Yuan, Xun, et autres
Publié: (2021)
par: Yuan, Xun, et autres
Publié: (2021)
Donkii: Can Annotation Error Detection Methods Find Errors in Instruction-Tuning Datasets?
par: Weber-Genzel, Leon, et autres
Publié: (2023)
par: Weber-Genzel, Leon, et autres
Publié: (2023)
Label Unification for Cross-Dataset Generalization in Cybersecurity NER
par: Jalocha, Maciej, et autres
Publié: (2025)
par: Jalocha, Maciej, et autres
Publié: (2025)
Semantic Similarity in Radiology Reports via LLMs and NER
par: Pearson, Beth, et autres
Publié: (2025)
par: Pearson, Beth, et autres
Publié: (2025)
CMNER: A Chinese Multimodal NER Dataset based on Social Media
par: Ji, Yuanze, et autres
Publié: (2024)
par: Ji, Yuanze, et autres
Publié: (2024)
HiligayNER: A Baseline Named Entity Recognition Model for Hiligaynon
par: Teves, James Ald, et autres
Publié: (2025)
par: Teves, James Ald, et autres
Publié: (2025)
An Annotated Dataset of Errors in Premodern Greek and Baselines for Detecting Them
par: Brooks, Creston, et autres
Publié: (2024)
par: Brooks, Creston, et autres
Publié: (2024)
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits
par: Sonkar, Shashank, et autres
Publié: (2024)
par: Sonkar, Shashank, et autres
Publié: (2024)
NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages
par: Abdullah, Abdulhady Abas, et autres
Publié: (2024)
par: Abdullah, Abdulhady Abas, et autres
Publié: (2024)
Documents similaires
-
Test Set Quality in Multilingual LLM Evaluation
par: Kranti, Chalamalasetti, et autres
Publié: (2025) -
IndicGEC: Powerful Models, or a Measurement Mirage?
par: Vajjala, Sowmya
Publié: (2025) -
The Problem with Safety Classification is not just the Models
par: Vajjala, Sowmya
Publié: (2025) -
MetricalARGS: A Taxonomy for Studying Metrical Poetry with LLMs
par: Kranti, Chalamalasetti, et autres
Publié: (2025) -
Dravidian language family through Universal Dependencies lens
par: Rama, Taraka, et autres
Publié: (2024)