Evaluating Shortest Edit Script Methods for Contextual Lemmatization
Fuente:
arXiv
Saved in:
| Main Authors: | Toporkov, Olia, Agerri, Rodrigo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lemma Dilemma: On Lemma Generation Without Domain- or Language-Specific Training Data
by: Toporkov, Olia, et al.
Published: (2025)
by: Toporkov, Olia, et al.
Published: (2025)
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
by: Dorkin, Aleksei, et al.
Published: (2024)
by: Dorkin, Aleksei, et al.
Published: (2024)
A Simple Joint Model for Improved Contextual Neural Lemmatization
by: Malaviya, Chaitanya, et al.
Published: (2019)
by: Malaviya, Chaitanya, et al.
Published: (2019)
A LLM-Based Ranking Method for the Evaluation of Automatic Counter-Narrative Generation
by: Zubiaga, Irune, et al.
Published: (2024)
by: Zubiaga, Irune, et al.
Published: (2024)
Cross-lingual Argument Mining in the Medical Domain
by: Yeginbergen, Anar, et al.
Published: (2023)
by: Yeginbergen, Anar, et al.
Published: (2023)
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
by: Bengoetxea, Jaione, et al.
Published: (2025)
by: Bengoetxea, Jaione, et al.
Published: (2025)
Joint Lemmatization and Morphological Tagging with LEMMING
by: Muller, Thomas, et al.
Published: (2024)
by: Muller, Thomas, et al.
Published: (2024)
Critical Questions Generation: Motivation and Challenges
by: Figueras, Blanca Calvo, et al.
Published: (2024)
by: Figueras, Blanca Calvo, et al.
Published: (2024)
Benchmarking Critical Questions Generation: A Challenging Reasoning Task for Large Language Models
by: Figueras, Banca Calvo, et al.
Published: (2025)
by: Figueras, Banca Calvo, et al.
Published: (2025)
RUMLEM: A Dictionary-Based Lemmatizer for Romansh
by: Fischer, Dominic P., et al.
Published: (2026)
by: Fischer, Dominic P., et al.
Published: (2026)
ScEdit: Script-based Assessment of Knowledge Editing
by: Li, Xinye, et al.
Published: (2025)
by: Li, Xinye, et al.
Published: (2025)
Basque and Spanish Counter Narrative Generation: Data Creation and Evaluation
by: Bengoetxea, Jaione, et al.
Published: (2024)
by: Bengoetxea, Jaione, et al.
Published: (2024)
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering
by: Alonso, Iñigo, et al.
Published: (2024)
by: Alonso, Iñigo, et al.
Published: (2024)
Effects of Cross-lingual Evidence in Multilingual Medical Question Answering
by: Yeginbergen, Anar, et al.
Published: (2026)
by: Yeginbergen, Anar, et al.
Published: (2026)
Metaphor and Large Language Models: When Surface Features Matter More than Deep Understanding
by: Sanchez-Bayona, Elisa, et al.
Published: (2025)
by: Sanchez-Bayona, Elisa, et al.
Published: (2025)
Context Aware Lemmatization and Morphological Tagging Method in Turkish
by: Sayallar, Cagri
Published: (2025)
by: Sayallar, Cagri
Published: (2025)
Comparison of Current Approaches to Lemmatization: A Case Study in Estonian
by: Dorkin, Aleksei, et al.
Published: (2024)
by: Dorkin, Aleksei, et al.
Published: (2024)
A State-of-the-Art Morphosyntactic Parser and Lemmatizer for Ancient Greek
by: Celano, Giuseppe G. A.
Published: (2024)
by: Celano, Giuseppe G. A.
Published: (2024)
A Catalog of Basque Dialectal Resources: Online Collections and Standard-to-Dialectal Adaptations
by: Bengoetxea, Jaione, et al.
Published: (2026)
by: Bengoetxea, Jaione, et al.
Published: (2026)
Physical Commonsense Reasoning for Lower-Resourced Languages and Dialects: a Study on Basque
by: Bengoetxea, Jaione, et al.
Published: (2026)
by: Bengoetxea, Jaione, et al.
Published: (2026)
Meta4XNLI: A Crosslingual Parallel Corpus for Metaphor Detection and Interpretation
by: Sanchez-Bayona, Elisa, et al.
Published: (2024)
by: Sanchez-Bayona, Elisa, et al.
Published: (2024)
Argument Mining in Data Scarce Settings: Cross-lingual Transfer and Few-shot Techniques
by: Yeginbergen, Anar, et al.
Published: (2024)
by: Yeginbergen, Anar, et al.
Published: (2024)
Language Independent Stance Detection: Social Interaction-based Embeddings and Large Language Models
by: de Landa, Joseba Fernandez, et al.
Published: (2022)
by: de Landa, Joseba Fernandez, et al.
Published: (2022)
Dynamic Knowledge Integration for Evidence-Driven Counter-Argument Generation with Large Language Models
by: Yeginbergen, Anar, et al.
Published: (2025)
by: Yeginbergen, Anar, et al.
Published: (2025)
Multilingual Medical Reasoning for Question Answering with Large Language Models
by: Ferrazzi, Pietro, et al.
Published: (2025)
by: Ferrazzi, Pietro, et al.
Published: (2025)
Political Leaning Inference through Plurinational Scenarios
by: de Landa, Joseba Fernandez, et al.
Published: (2024)
by: de Landa, Joseba Fernandez, et al.
Published: (2024)
Lemmatization as a Classification Task: Results from Arabic across Multiple Genres
by: Saeed, Mostafa, et al.
Published: (2025)
by: Saeed, Mostafa, et al.
Published: (2025)
How Well Can Knowledge Edit Methods Edit Perplexing Knowledge?
by: Ge, Huaizhi, et al.
Published: (2024)
by: Ge, Huaizhi, et al.
Published: (2024)
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
by: Pattnayak, Priyaranjan
Published: (2026)
by: Pattnayak, Priyaranjan
Published: (2026)
eFontes. Part of Speech Tagging and Lemmatization of Medieval Latin Texts.A Cross-Genre Survey
by: Nowak, Krzysztof, et al.
Published: (2024)
by: Nowak, Krzysztof, et al.
Published: (2024)
Truth Knows No Language: Evaluating Truthfulness Beyond English
by: Figueras, Blanca Calvo, et al.
Published: (2025)
by: Figueras, Blanca Calvo, et al.
Published: (2025)
Unknown Script: Impact of Script on Cross-Lingual Transfer
by: Tufa, Wondimagegnhue Tsegaye, et al.
Published: (2024)
by: Tufa, Wondimagegnhue Tsegaye, et al.
Published: (2024)
Automatic Fact-checking in English and Telugu
by: Chikkala, Ravi Kiran, et al.
Published: (2025)
by: Chikkala, Ravi Kiran, et al.
Published: (2025)
Sentence Smith: Controllable Edits for Evaluating Text Embeddings
by: Li, Hongji, et al.
Published: (2025)
by: Li, Hongji, et al.
Published: (2025)
CasiMedicos-Arg: A Medical Question Answering Dataset Annotated with Explanatory Argumentative Structures
by: Sviridova, Ekaterina, et al.
Published: (2024)
by: Sviridova, Ekaterina, et al.
Published: (2024)
SkyScript-100M: 1,000,000,000 Pairs of Scripts and Shooting Scripts for Short Drama
by: Tang, Jing, et al.
Published: (2024)
by: Tang, Jing, et al.
Published: (2024)
GoLLIE: Annotation Guidelines improve Zero-Shot Information-Extraction
by: Sainz, Oscar, et al.
Published: (2023)
by: Sainz, Oscar, et al.
Published: (2023)
Grammatical Error Correction Evaluation by Optimally Transporting Edit Representation
by: Goto, Takumi, et al.
Published: (2026)
by: Goto, Takumi, et al.
Published: (2026)
One Language, Two Scripts: Probing Script-Invariance in LLM Concept Representations
by: Karne, Sripad
Published: (2026)
by: Karne, Sripad
Published: (2026)
Script Gap: Evaluating LLM Triage on Indian Languages in Native vs Romanized Scripts in a Real World Setting
by: Khullar, Manurag, et al.
Published: (2025)
by: Khullar, Manurag, et al.
Published: (2025)
Similar Items
-
Lemma Dilemma: On Lemma Generation Without Domain- or Language-Specific Training Data
by: Toporkov, Olia, et al.
Published: (2025) -
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
by: Dorkin, Aleksei, et al.
Published: (2024) -
A Simple Joint Model for Improved Contextual Neural Lemmatization
by: Malaviya, Chaitanya, et al.
Published: (2019) -
A LLM-Based Ranking Method for the Evaluation of Automatic Counter-Narrative Generation
by: Zubiaga, Irune, et al.
Published: (2024) -
Cross-lingual Argument Mining in the Medical Domain
by: Yeginbergen, Anar, et al.
Published: (2023)