Towards DS-NER: Unveiling and Addressing Latent Noise in Distant Annotations
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Yuyang, Qiao, Dan, Li, Juntao, Xu, Jiajie, Chao, Pingfu, Zhou, Xiaofang, Zhang, Min |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Re-Examine Distantly Supervised NER: A New Benchmark and a Simple Approach
by: Li, Yuepei, et al.
Published: (2024)
by: Li, Yuepei, et al.
Published: (2024)
SCAN: Self-Denoising Monte Carlo Annotation for Robust Process Reward Learning
by: Ding, Yuyang, et al.
Published: (2025)
by: Ding, Yuyang, et al.
Published: (2025)
Augmenting NER Datasets with LLMs: Towards Automated and Refined Annotation
by: Naraki, Yuji, et al.
Published: (2024)
by: Naraki, Yuji, et al.
Published: (2024)
LongFlow: Efficient KV Cache Compression for Reasoning Models
by: Su, Yi, et al.
Published: (2026)
by: Su, Yi, et al.
Published: (2026)
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-Experts
by: Yang, Jiajie
Published: (2025)
by: Yang, Jiajie
Published: (2025)
Distantly-Supervised Joint Extraction with Noise-Robust Learning
by: Li, Yufei, et al.
Published: (2023)
by: Li, Yufei, et al.
Published: (2023)
The Million-Label NER: Breaking Scale Barriers with GLiNER bi-encoder
by: Stepanov, Ihor, et al.
Published: (2026)
by: Stepanov, Ihor, et al.
Published: (2026)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
by: Bogdanov, Sergei, et al.
Published: (2024)
by: Bogdanov, Sergei, et al.
Published: (2024)
L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models
by: Chaudhari, Harsh, et al.
Published: (2023)
by: Chaudhari, Harsh, et al.
Published: (2023)
Graph Neural Network and NER-Based Text Summarization
by: Khan, Imaad Zaffar, et al.
Published: (2024)
by: Khan, Imaad Zaffar, et al.
Published: (2024)
DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition
by: Kim, Siun, et al.
Published: (2026)
by: Kim, Siun, et al.
Published: (2026)
WhisperNER: Unified Open Named Entity and Speech Recognition
by: Ayache, Gil, et al.
Published: (2024)
by: Ayache, Gil, et al.
Published: (2024)
Unveiling and Addressing Pseudo Forgetting in Large Language Models
by: Sun, Huashan, et al.
Published: (2024)
by: Sun, Huashan, et al.
Published: (2024)
Adversarial Demonstration Learning for Low-resource NER Using Dual Similarity
by: Yuan, Guowen, et al.
Published: (2025)
by: Yuan, Guowen, et al.
Published: (2025)
Sentence Bag Graph Formulation for Biomedical Distant Supervision Relation Extraction
by: Zhang, Hao, et al.
Published: (2023)
by: Zhang, Hao, et al.
Published: (2023)
FAPO: Flawed-Aware Policy Optimization for Efficient and Reliable Reasoning
by: Ding, Yuyang, et al.
Published: (2025)
by: Ding, Yuyang, et al.
Published: (2025)
Regular-pattern-sensitive CRFs for Distant Label Interactions
by: Papay, Sean, et al.
Published: (2024)
by: Papay, Sean, et al.
Published: (2024)
Enhancing Latent Computation in Transformers with Latent Tokens
by: Sun, Yuchang, et al.
Published: (2025)
by: Sun, Yuchang, et al.
Published: (2025)
ANCHOLIK-NER: A Benchmark Dataset for Bangla Regional Named Entity Recognition
by: Paul, Bidyarthi, et al.
Published: (2025)
by: Paul, Bidyarthi, et al.
Published: (2025)
Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancement
by: Zhong, Qimin, et al.
Published: (2026)
by: Zhong, Qimin, et al.
Published: (2026)
3DS: Medical Domain Adaptation of LLMs via Decomposed Difficulty-based Data Selection
by: Ding, Hongxin, et al.
Published: (2024)
by: Ding, Hongxin, et al.
Published: (2024)
PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning
by: Chen, Xiaoyi, et al.
Published: (2026)
by: Chen, Xiaoyi, et al.
Published: (2026)
DIDS: Domain Impact-aware Data Sampling for Large Language Model Training
by: Shi, Weijie, et al.
Published: (2025)
by: Shi, Weijie, et al.
Published: (2025)
Mapping from Meaning: Addressing the Miscalibration of Prompt-Sensitive Language Models
by: Cox, Kyle, et al.
Published: (2025)
by: Cox, Kyle, et al.
Published: (2025)
mucAI at WojoodNER 2024: Arabic Named Entity Recognition with Nearest Neighbor Search
by: Abdou, Ahmed, et al.
Published: (2024)
by: Abdou, Ahmed, et al.
Published: (2024)
GLiNER-Relex: A Unified Framework for Joint Named Entity Recognition and Relation Extraction
by: Stepanov, Ihor, et al.
Published: (2026)
by: Stepanov, Ihor, et al.
Published: (2026)
Towards Principled Design of Mixture-of-Experts Language Models under Memory and Inference Constraints
by: Liew, Seng Pei, et al.
Published: (2026)
by: Liew, Seng Pei, et al.
Published: (2026)
Addressing Performance Saturation for LLM RL via Precise Entropy Curve Control
by: Li, Bolian, et al.
Published: (2026)
by: Li, Bolian, et al.
Published: (2026)
CrEst: Credibility Estimation for Contexts in LLMs via Weak Supervision
by: Adila, Dyah, et al.
Published: (2025)
by: Adila, Dyah, et al.
Published: (2025)
MorphNAS: Differentiable Architecture Search for Morphologically-Aware Multilingual NER
by: Devadiga, Prathamesh, et al.
Published: (2025)
by: Devadiga, Prathamesh, et al.
Published: (2025)
PIP: Perturbation-based Iterative Pruning for Large Language Models
by: Cao, Yi, et al.
Published: (2025)
by: Cao, Yi, et al.
Published: (2025)
Guided Distant Supervision for Multilingual Relation Extraction Data: Adapting to a New Language
by: Plum, Alistair, et al.
Published: (2024)
by: Plum, Alistair, et al.
Published: (2024)
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
by: Hu, Shengding, et al.
Published: (2024)
by: Hu, Shengding, et al.
Published: (2024)
Towards Understanding Multi-Round Large Language Model Reasoning: Approximability, Learnability and Generalizability
by: Xu, Chenhui, et al.
Published: (2025)
by: Xu, Chenhui, et al.
Published: (2025)
Improving Pre-trained Language Model Sensitivity via Mask Specific losses: A case study on Biomedical NER
by: Abaho, Micheal, et al.
Published: (2024)
by: Abaho, Micheal, et al.
Published: (2024)
Hierarchical Spatial-Temporal Graph-Enhanced Model for Map-Matching
by: Gao, Anjun, et al.
Published: (2026)
by: Gao, Anjun, et al.
Published: (2026)
ALMA: Alignment with Minimal Annotation
by: Yasunaga, Michihiro, et al.
Published: (2024)
by: Yasunaga, Michihiro, et al.
Published: (2024)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning
by: Deng, Jingcheng, et al.
Published: (2026)
by: Deng, Jingcheng, et al.
Published: (2026)
Latent Logic Tree Extraction for Event Sequence Explanation from LLMs
by: Song, Zitao, et al.
Published: (2024)
by: Song, Zitao, et al.
Published: (2024)
Similar Items
-
Re-Examine Distantly Supervised NER: A New Benchmark and a Simple Approach
by: Li, Yuepei, et al.
Published: (2024) -
SCAN: Self-Denoising Monte Carlo Annotation for Robust Process Reward Learning
by: Ding, Yuyang, et al.
Published: (2025) -
Augmenting NER Datasets with LLMs: Towards Automated and Refined Annotation
by: Naraki, Yuji, et al.
Published: (2024) -
LongFlow: Efficient KV Cache Compression for Reasoning Models
by: Su, Yi, et al.
Published: (2026) -
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-Experts
by: Yang, Jiajie
Published: (2025)