SlovKE: A Large-Scale Dataset and LLM Evaluation for Slovak Keyphrase Extraction
Fuente:
arXiv
Salvato in:
| Autori principali: | Števaňák, David, Šuppa, Marek |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SloPal: A 60-Million-Word Slovak Parliamentary Corpus with Aligned Speech and Fine-Tuned ASR Models
di: Božík, Erik, et al.
Pubblicazione: (2025)
di: Božík, Erik, et al.
Pubblicazione: (2025)
Examining the Metrics for Document-Level Claim Extraction in Czech and Slovak
di: Makaiova, Lucia, et al.
Pubblicazione: (2025)
di: Makaiova, Lucia, et al.
Pubblicazione: (2025)
skLEP: A Slovak General Language Understanding Benchmark
di: Šuppa, Marek, et al.
Pubblicazione: (2025)
di: Šuppa, Marek, et al.
Pubblicazione: (2025)
EUROPA: A Legal Multilingual Keyphrase Generation Dataset
di: Salaün, Olivier, et al.
Pubblicazione: (2024)
di: Salaün, Olivier, et al.
Pubblicazione: (2024)
LongKey: Keyphrase Extraction for Long Documents
di: Alves, Jeovane Honorio, et al.
Pubblicazione: (2024)
di: Alves, Jeovane Honorio, et al.
Pubblicazione: (2024)
$μ$KE: Matryoshka Unstructured Knowledge Editing of Large Language Models
di: Su, Zian, et al.
Pubblicazione: (2025)
di: Su, Zian, et al.
Pubblicazione: (2025)
One2set + Large Language Model: Best Partners for Keyphrase Generation
di: Shao, Liangying, et al.
Pubblicazione: (2024)
di: Shao, Liangying, et al.
Pubblicazione: (2024)
OneKE: A Dockerized Schema-Guided LLM Agent-based Knowledge Extraction System
di: Luo, Yujie, et al.
Pubblicazione: (2024)
di: Luo, Yujie, et al.
Pubblicazione: (2024)
Slovak Conceptual Dictionary
di: Blšták, Miroslav
Pubblicazione: (2025)
di: Blšták, Miroslav
Pubblicazione: (2025)
Pre-Trained Language Models for Keyphrase Prediction: A Review
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task
di: Fujisaki, Yuya, et al.
Pubblicazione: (2024)
di: Fujisaki, Yuya, et al.
Pubblicazione: (2024)
Bryndza at ClimateActivism 2024: Stance, Target and Hate Event Detection via Retrieval-Augmented GPT-4 and LLaMA
di: Šuppa, Marek, et al.
Pubblicazione: (2024)
di: Šuppa, Marek, et al.
Pubblicazione: (2024)
OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets
di: Shen, Jiyuan, et al.
Pubblicazione: (2026)
di: Shen, Jiyuan, et al.
Pubblicazione: (2026)
Zero-Shot Keyphrase Generation: Investigating Specialized Instructions and Multi-Sample Aggregation on Large Language Models
di: Mohan, Jayanth, et al.
Pubblicazione: (2025)
di: Mohan, Jayanth, et al.
Pubblicazione: (2025)
Bench4KE: Benchmarking Automated Competency Question Generation
di: Lippolis, Anna Sofia, et al.
Pubblicazione: (2025)
di: Lippolis, Anna Sofia, et al.
Pubblicazione: (2025)
RideKE: Leveraging Low-Resource, User-Generated Twitter Content for Sentiment and Emotion Detection in Kenyan Code-Switched Dataset
di: Etori, Naome A., et al.
Pubblicazione: (2025)
di: Etori, Naome A., et al.
Pubblicazione: (2025)
Detecting Relevant Information in High-Volume Chat Logs: Keyphrase Extraction for Grooming and Drug Dealing Forensic Analysis
di: Alves, Jeovane Honório, et al.
Pubblicazione: (2023)
di: Alves, Jeovane Honório, et al.
Pubblicazione: (2023)
WilKE: Wise-Layer Knowledge Editor for Lifelong Knowledge Editing
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
di: Zheng, Lianmin, et al.
Pubblicazione: (2023)
di: Zheng, Lianmin, et al.
Pubblicazione: (2023)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
di: Khatun, Aisha, et al.
Pubblicazione: (2024)
di: Khatun, Aisha, et al.
Pubblicazione: (2024)
CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports
di: Zhang, Xiao Yu Cindy, et al.
Pubblicazione: (2025)
di: Zhang, Xiao Yu Cindy, et al.
Pubblicazione: (2025)
Overestimation in LLM Evaluation: A Controlled Large-Scale Study on Data Contamination's Impact on Machine Translation
di: Kocyigit, Muhammed Yusuf, et al.
Pubblicazione: (2025)
di: Kocyigit, Muhammed Yusuf, et al.
Pubblicazione: (2025)
Knots: A Large-Scale Multi-Agent Enhanced Expert-Annotated Dataset and LLM Prompt Optimization for NOTAM Semantic Parsing
di: Liu, Maoqi, et al.
Pubblicazione: (2025)
di: Liu, Maoqi, et al.
Pubblicazione: (2025)
MetaKE: Meta-Learning for Knowledge Editing Toward a Better Accuracy-Editability Trade-off
di: Liu, Shuxin, et al.
Pubblicazione: (2026)
di: Liu, Shuxin, et al.
Pubblicazione: (2026)
Ace-CEFR -- A Dataset for Automated Evaluation of the Linguistic Difficulty of Conversational Texts for LLM Applications
di: Kogan, David, et al.
Pubblicazione: (2025)
di: Kogan, David, et al.
Pubblicazione: (2025)
MessIRve: A Large-Scale Spanish Information Retrieval Dataset
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
A Large-Scale Dataset and Citation Intent Classification in Turkish with LLMs
di: Karaca, Kemal Sami, et al.
Pubblicazione: (2025)
di: Karaca, Kemal Sami, et al.
Pubblicazione: (2025)
CFEVER: A Chinese Fact Extraction and VERification Dataset
di: Lin, Ying-Jia, et al.
Pubblicazione: (2024)
di: Lin, Ying-Jia, et al.
Pubblicazione: (2024)
ALHD: A Large-Scale and Multigenre Benchmark Dataset for Arabic LLM-Generated Text Detection
di: Khairallah, Ali, et al.
Pubblicazione: (2025)
di: Khairallah, Ali, et al.
Pubblicazione: (2025)
BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset
di: Xi, Zhiheng, et al.
Pubblicazione: (2025)
di: Xi, Zhiheng, et al.
Pubblicazione: (2025)
Scaling Open-Weight Large Language Models for Hydropower Regulatory Information Extraction: A Systematic Analysis
di: Yoon, Hong-Jun, et al.
Pubblicazione: (2025)
di: Yoon, Hong-Jun, et al.
Pubblicazione: (2025)
MixRED: A Mix-lingual Relation Extraction Dataset
di: Kong, Lingxing, et al.
Pubblicazione: (2024)
di: Kong, Lingxing, et al.
Pubblicazione: (2024)
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
di: Wang, Zilong, et al.
Pubblicazione: (2025)
di: Wang, Zilong, et al.
Pubblicazione: (2025)
VietLyrics: A Large-Scale Dataset and Models for Vietnamese Automatic Lyrics Transcription
di: Nguyen, Quoc Anh, et al.
Pubblicazione: (2025)
di: Nguyen, Quoc Anh, et al.
Pubblicazione: (2025)
GenRES: Rethinking Evaluation for Generative Relation Extraction in the Era of Large Language Models
di: Jiang, Pengcheng, et al.
Pubblicazione: (2024)
di: Jiang, Pengcheng, et al.
Pubblicazione: (2024)
Meta-Reasoning Improves Tool Use in Large Language Models
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
$\texttt{BluePrint}$: A Social Media User Dataset for LLM Persona Evaluation and Training
di: Bück-Kaeffer, Aurélien, et al.
Pubblicazione: (2025)
di: Bück-Kaeffer, Aurélien, et al.
Pubblicazione: (2025)
PersianPunc: A Large-Scale Dataset and BERT-Based Approach for Persian Punctuation Restoration
di: Kalahroodi, Mohammad Javad Ranjbar, et al.
Pubblicazione: (2026)
di: Kalahroodi, Mohammad Javad Ranjbar, et al.
Pubblicazione: (2026)
JMultiWOZ: A Large-Scale Japanese Multi-Domain Task-Oriented Dialogue Dataset
di: Ohashi, Atsumoto, et al.
Pubblicazione: (2024)
di: Ohashi, Atsumoto, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SloPal: A 60-Million-Word Slovak Parliamentary Corpus with Aligned Speech and Fine-Tuned ASR Models
di: Božík, Erik, et al.
Pubblicazione: (2025) -
Examining the Metrics for Document-Level Claim Extraction in Czech and Slovak
di: Makaiova, Lucia, et al.
Pubblicazione: (2025) -
skLEP: A Slovak General Language Understanding Benchmark
di: Šuppa, Marek, et al.
Pubblicazione: (2025) -
EUROPA: A Legal Multilingual Keyphrase Generation Dataset
di: Salaün, Olivier, et al.
Pubblicazione: (2024) -
LongKey: Keyphrase Extraction for Long Documents
di: Alves, Jeovane Honorio, et al.
Pubblicazione: (2024)