NLPre: a revised approach towards language-centric benchmarking of Natural Language Preprocessing systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wiącek, Martyna, Rybak, Piotr, Pszenny, Łukasz, Wróblewska, Alina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating large language models for their competence in extracting grammatically sound sentences from transcribed noisy utterances
von: Wróblewska, Alina
Veröffentlicht: (2024)
von: Wróblewska, Alina
Veröffentlicht: (2024)
Integrating gender inclusivity into large language models via instruction tuning
von: Wróblewska, Alina, et al.
Veröffentlicht: (2025)
von: Wróblewska, Alina, et al.
Veröffentlicht: (2025)
Transferring BERT Capabilities from High-Resource to Low-Resource Languages Using Vocabulary Matching
von: Rybak, Piotr
Veröffentlicht: (2024)
von: Rybak, Piotr
Veröffentlicht: (2024)
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
Silver Retriever: Advancing Neural Passage Retrieval for Polish Question Answering
von: Rybak, Piotr, et al.
Veröffentlicht: (2023)
von: Rybak, Piotr, et al.
Veröffentlicht: (2023)
Enriching language models with graph-based context information to better understand textual data
von: Roethel, Albert, et al.
Veröffentlicht: (2023)
von: Roethel, Albert, et al.
Veröffentlicht: (2023)
On the error of the Euler scheme for approximation of solutions of nonlinear DDEs under inexact information
von: Przybyłowicz, Paweł, et al.
Veröffentlicht: (2026)
von: Przybyłowicz, Paweł, et al.
Veröffentlicht: (2026)
Linguini: A benchmark for language-agnostic linguistic reasoning
von: Sánchez, Eduardo, et al.
Veröffentlicht: (2024)
von: Sánchez, Eduardo, et al.
Veröffentlicht: (2024)
A thorough benchmark of automatic text classification: From traditional approaches to large language models
von: Cunha, Washington, et al.
Veröffentlicht: (2025)
von: Cunha, Washington, et al.
Veröffentlicht: (2025)
Deepfake tweets automatic detection
von: Frej, Adam, et al.
Veröffentlicht: (2024)
von: Frej, Adam, et al.
Veröffentlicht: (2024)
GRASP: A novel benchmark for evaluating language GRounding And Situated Physics understanding in multimodal language models
von: Jassim, Serwan, et al.
Veröffentlicht: (2023)
von: Jassim, Serwan, et al.
Veröffentlicht: (2023)
\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding
von: Min, Junghyun, et al.
Veröffentlicht: (2025)
von: Min, Junghyun, et al.
Veröffentlicht: (2025)
Language Models Model Language
von: Borchmann, Łukasz
Veröffentlicht: (2025)
von: Borchmann, Łukasz
Veröffentlicht: (2025)
Improving Indigenous Language Machine Translation with Synthetic Data and Language-Specific Preprocessing
von: Dhawan, Aashish, et al.
Veröffentlicht: (2026)
von: Dhawan, Aashish, et al.
Veröffentlicht: (2026)
Anchor function: a type of benchmark functions for studying language models
von: Zhang, Zhongwang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongwang, et al.
Veröffentlicht: (2024)
AIDBench: A benchmark for evaluating the authorship identification capability of large language models
von: Wen, Zichen, et al.
Veröffentlicht: (2024)
von: Wen, Zichen, et al.
Veröffentlicht: (2024)
A benchmark dataset for evaluating Syndrome Differentiation and Treatment in large language models
von: Li, Kunning, et al.
Veröffentlicht: (2025)
von: Li, Kunning, et al.
Veröffentlicht: (2025)
DevBench: A multimodal developmental benchmark for language learning
von: Tan, Alvin Wei Ming, et al.
Veröffentlicht: (2024)
von: Tan, Alvin Wei Ming, et al.
Veröffentlicht: (2024)
MultiLoKo: a multilingual local knowledge benchmark for LLMs spanning 31 languages
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2025)
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2025)
Clinical named entity recognition in the Portuguese language: a benchmark of modern BERT models and LLMs
von: de Almeida, Vinicius Anjos, et al.
Veröffentlicht: (2026)
von: de Almeida, Vinicius Anjos, et al.
Veröffentlicht: (2026)
Polish-English medical knowledge transfer: A new benchmark and results
von: Grzybowski, Łukasz, et al.
Veröffentlicht: (2024)
von: Grzybowski, Łukasz, et al.
Veröffentlicht: (2024)
LongTail-Swap: benchmarking language models' abilities on rare words
von: Algayres, Robin, et al.
Veröffentlicht: (2025)
von: Algayres, Robin, et al.
Veröffentlicht: (2025)
Investigating Large Language Models' Linguistic Abilities for Text Preprocessing
von: Braga, Marco, et al.
Veröffentlicht: (2025)
von: Braga, Marco, et al.
Veröffentlicht: (2025)
Tgea: An error-annotated dataset and benchmark tasks for text generation from pretrained language models
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Creativity Benchmark: A benchmark for marketing creativity for large language models
von: Bhat, Ninad, et al.
Veröffentlicht: (2025)
von: Bhat, Ninad, et al.
Veröffentlicht: (2025)
Natural language processing for African languages
von: Adelani, David Ifeoluwa
Veröffentlicht: (2025)
von: Adelani, David Ifeoluwa
Veröffentlicht: (2025)
Evaluating LLM-Generated Q&A Test: a Student-Centered Study
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
von: Barboule, Camille, et al.
Veröffentlicht: (2024)
von: Barboule, Camille, et al.
Veröffentlicht: (2024)
CoPrUS: Consistency Preserving Utterance Synthesis towards more realistic benchmark dialogues
von: Steindl, Sebastian, et al.
Veröffentlicht: (2024)
von: Steindl, Sebastian, et al.
Veröffentlicht: (2024)
The SMeL Test: A simple benchmark for media literacy in language models
von: Ahdritz, Gustaf, et al.
Veröffentlicht: (2025)
von: Ahdritz, Gustaf, et al.
Veröffentlicht: (2025)
Click it or Leave it: Detecting and Spoiling Clickbait with Informativeness Measures and Large Language Models
von: Michaluk, Wojciech, et al.
Veröffentlicht: (2026)
von: Michaluk, Wojciech, et al.
Veröffentlicht: (2026)
Fake News Detection: It's All in the Data!
von: Kuntur, Soveatin, et al.
Veröffentlicht: (2024)
von: Kuntur, Soveatin, et al.
Veröffentlicht: (2024)
BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
von: Lavechin, Marvin, et al.
Veröffentlicht: (2023)
von: Lavechin, Marvin, et al.
Veröffentlicht: (2023)
LLMzSzŁ: a comprehensive LLM benchmark for Polish
von: Jassem, Krzysztof, et al.
Veröffentlicht: (2025)
von: Jassem, Krzysztof, et al.
Veröffentlicht: (2025)
BgGPT 1.0: Extending English-centric LLMs to other languages
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
Vision-centric Token Compression in Large Language Model
von: Xing, Ling, et al.
Veröffentlicht: (2025)
von: Xing, Ling, et al.
Veröffentlicht: (2025)
Jellyfish: A Large Language Model for Data Preprocessing
von: Zhang, Haochen, et al.
Veröffentlicht: (2023)
von: Zhang, Haochen, et al.
Veröffentlicht: (2023)
LLM Essay Scoring Under Holistic and Analytic Rubrics: Prompt Effects and Bias
von: Kucia, Filip J., et al.
Veröffentlicht: (2026)
von: Kucia, Filip J., et al.
Veröffentlicht: (2026)
On the steerability of large language models toward data-driven personas
von: Li, Junyi, et al.
Veröffentlicht: (2023)
von: Li, Junyi, et al.
Veröffentlicht: (2023)
Towards Better Chinese-centric Neural Machine Translation for Low-resource Languages
von: Li, Bin, et al.
Veröffentlicht: (2022)
von: Li, Bin, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Investigating large language models for their competence in extracting grammatically sound sentences from transcribed noisy utterances
von: Wróblewska, Alina
Veröffentlicht: (2024) -
Integrating gender inclusivity into large language models via instruction tuning
von: Wróblewska, Alina, et al.
Veröffentlicht: (2025) -
Transferring BERT Capabilities from High-Resource to Low-Resource Languages Using Vocabulary Matching
von: Rybak, Piotr
Veröffentlicht: (2024) -
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022) -
Silver Retriever: Advancing Neural Passage Retrieval for Polish Question Answering
von: Rybak, Piotr, et al.
Veröffentlicht: (2023)