Reliable Part-of-Speech Tagging of Historical Corpora through Set-Valued Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Heid, Stefan, Wever, Marcel, Hüllermeier, Eyke |
|---|---|
| Format: | Preprint |
| Published: |
2020
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Introducing Three New Benchmark Datasets for Hierarchical Text Classification
by: Toit, Jaco du, et al.
Published: (2024)
by: Toit, Jaco du, et al.
Published: (2024)
ARAGOG: Advanced RAG Output Grading
by: Eibich, Matouš, et al.
Published: (2024)
by: Eibich, Matouš, et al.
Published: (2024)
Efficient Knowledge Feeding to Language Models: A Novel Integrated Encoder-Decoder Architecture
by: Kumar, S Santosh, et al.
Published: (2025)
by: Kumar, S Santosh, et al.
Published: (2025)
Retrieval-Enhanced Named Entity Recognition
by: Shiraishi, Enzo, et al.
Published: (2024)
by: Shiraishi, Enzo, et al.
Published: (2024)
Leveraging Retrieval Augmented Generative LLMs For Automated Metadata Description Generation to Enhance Data Catalogs
by: Singh, Mayank, et al.
Published: (2025)
by: Singh, Mayank, et al.
Published: (2025)
Building Entity Association Mining Framework for Knowledge Discovery
by: Rawal, Anshika, et al.
Published: (2025)
by: Rawal, Anshika, et al.
Published: (2025)
Document Understanding for Healthcare Referrals
by: Mistry, Jimit, et al.
Published: (2023)
by: Mistry, Jimit, et al.
Published: (2023)
Detection of ChatGPT Fake Science with the xFakeSci Learning Algorithm
by: Hamed, Ahmed Abdeen, et al.
Published: (2023)
by: Hamed, Ahmed Abdeen, et al.
Published: (2023)
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
by: Qiu, Jingxi, et al.
Published: (2026)
by: Qiu, Jingxi, et al.
Published: (2026)
Improving RAG Retrieval via Propositional Content Extraction: a Speech Act Theory Approach
by: Lima, João Alberto de Oliveira
Published: (2025)
by: Lima, João Alberto de Oliveira
Published: (2025)
Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
A Language Model based Framework for New Concept Placement in Ontologies
by: Dong, Hang, et al.
Published: (2024)
by: Dong, Hang, et al.
Published: (2024)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026)
by: Teixeira, Tiago, et al.
Published: (2026)
Exploring User Retrieval Integration towards Large Language Models for Cross-Domain Sequential Recommendation
by: Shen, Tingjia, et al.
Published: (2024)
by: Shen, Tingjia, et al.
Published: (2024)
RARE: Redundancy-Aware Retrieval Evaluation Framework for High-Similarity Corpora
by: Cho, Hanjun, et al.
Published: (2026)
by: Cho, Hanjun, et al.
Published: (2026)
SegNSP: Revisiting Next Sentence Prediction for Linear Text Segmentation
by: Isidro, José, et al.
Published: (2026)
by: Isidro, José, et al.
Published: (2026)
Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports
by: Jabal, Mohamed Sobhi, et al.
Published: (2024)
by: Jabal, Mohamed Sobhi, et al.
Published: (2024)
MUDY: Multi-Granular Dynamic Candidate Contextualization for Unsupervised Keyphrase Extraction
by: Kang, Hyeongu, et al.
Published: (2026)
by: Kang, Hyeongu, et al.
Published: (2026)
Using LLM-Based Approaches to Enhance and Automate Topic Labeling
by: Khandelwal, Trishia
Published: (2025)
by: Khandelwal, Trishia
Published: (2025)
Knowledge Distillation of Domain-adapted LLMs for Question-Answering in Telecom
by: Sen, Rishika, et al.
Published: (2025)
by: Sen, Rishika, et al.
Published: (2025)
A Method for Detecting Legal Article Competition for Korean Criminal Law Using a Case-augmented Mention Graph
by: An, Seonho, et al.
Published: (2024)
by: An, Seonho, et al.
Published: (2024)
Annif at the GermEval-2025 LLMs4Subjects Task: Traditional XMTC Augmented by Efficient LLMs
by: Suominen, Osma, et al.
Published: (2025)
by: Suominen, Osma, et al.
Published: (2025)
RecaLLM: Addressing the Lost-in-Thought Phenomenon with Explicit In-Context Retrieval
by: Whitecross, Kyle, et al.
Published: (2026)
by: Whitecross, Kyle, et al.
Published: (2026)
Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models
by: Couturier, Camille, et al.
Published: (2025)
by: Couturier, Camille, et al.
Published: (2025)
Learning variant product relationship and variation attributes from e-commerce website structures
by: Herrero-Vidal, Pedro, et al.
Published: (2024)
by: Herrero-Vidal, Pedro, et al.
Published: (2024)
Disaster Question Answering with LoRA Efficiency and Accurate End Position
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
LLMs in the Loop: Leveraging Large Language Model Annotations for Active Learning in Low-Resource Languages
by: Kholodna, Nataliia, et al.
Published: (2024)
by: Kholodna, Nataliia, et al.
Published: (2024)
From Millions of Tweets to Actionable Insights: Leveraging LLMs for User Profiling
by: Rahimzadeh, Vahid, et al.
Published: (2025)
by: Rahimzadeh, Vahid, et al.
Published: (2025)
Leveraging Translation For Optimal Recall: Tailoring LLM Personalization With User Profiles
by: Ravichandran, Karthik, et al.
Published: (2024)
by: Ravichandran, Karthik, et al.
Published: (2024)
A Study into Investigating Temporal Robustness of LLMs
by: Wallat, Jonas, et al.
Published: (2025)
by: Wallat, Jonas, et al.
Published: (2025)
Overview of the ClinIQLink 2025 Shared Task on Medical Question-Answering
by: Colelough, Brandon, et al.
Published: (2025)
by: Colelough, Brandon, et al.
Published: (2025)
EnterpriseEM: Fine-tuned Embeddings for Enterprise Semantic Search
by: Rathinasamy, Kamalkumar, et al.
Published: (2024)
by: Rathinasamy, Kamalkumar, et al.
Published: (2024)
Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models
by: Günther, Michael, et al.
Published: (2024)
by: Günther, Michael, et al.
Published: (2024)
DeepSlide: From Artifacts to Presentation Delivery
by: Yang, Ming, et al.
Published: (2026)
by: Yang, Ming, et al.
Published: (2026)
SomaliWeb v1: A Quality-Filtered Somali Web Corpus with a Matched Tokenizer and a Public Language-Identification Benchmark
by: Dahir, Khalid Yusuf
Published: (2026)
by: Dahir, Khalid Yusuf
Published: (2026)
Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection
by: Kopanov, Kalin
Published: (2024)
by: Kopanov, Kalin
Published: (2024)
A Comparative Analysis of Retrieval-Augmented Generation Techniques for Bengali Standard-to-Dialect Machine Translation Using LLMs
by: Sami, K. M. Jubair, et al.
Published: (2025)
by: Sami, K. M. Jubair, et al.
Published: (2025)
ConQRet: Benchmarking Fine-Grained Evaluation of Retrieval Augmented Argumentation with LLM Judges
by: Dhole, Kaustubh D., et al.
Published: (2024)
by: Dhole, Kaustubh D., et al.
Published: (2024)
Leveraging Large Language Models to Extract and Translate Medical Information in Doctors' Notes for Health Records and Diagnostic Billing Codes
by: Hartnett, Peter, et al.
Published: (2026)
by: Hartnett, Peter, et al.
Published: (2026)
Temporal Decay of Co-Citation Predictability: A 20-Year Statute Retrieval Benchmark from 396M Ukrainian Court Citations
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
Similar Items
-
Introducing Three New Benchmark Datasets for Hierarchical Text Classification
by: Toit, Jaco du, et al.
Published: (2024) -
ARAGOG: Advanced RAG Output Grading
by: Eibich, Matouš, et al.
Published: (2024) -
Efficient Knowledge Feeding to Language Models: A Novel Integrated Encoder-Decoder Architecture
by: Kumar, S Santosh, et al.
Published: (2025) -
Retrieval-Enhanced Named Entity Recognition
by: Shiraishi, Enzo, et al.
Published: (2024) -
Leveraging Retrieval Augmented Generative LLMs For Automated Metadata Description Generation to Enhance Data Catalogs
by: Singh, Mayank, et al.
Published: (2025)