LengClaro2023: A Dataset of Administrative Texts in Spanish with Plain Language adaptations
Fuente:
arXiv
Salvato in:
| Autori principali: | Agüera-Marco, Belén, Gonzalez-Dios, Itziar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Text Adaptation to Plain Language and Easy Read via Automatic Post-Editing Cycles
di: Calleja, Jesús, et al.
Pubblicazione: (2025)
di: Calleja, Jesús, et al.
Pubblicazione: (2025)
BenchOverflow: Measuring Overflow in Large Language Models via Plain-Text Prompts
di: Feiglin, Erin, et al.
Pubblicazione: (2026)
di: Feiglin, Erin, et al.
Pubblicazione: (2026)
CardiffNLP at CLEARS-2025: Prompting Large Language Models for Plain Language and Easy-to-Read Text Rewriting
di: Ayesh, Mutaz, et al.
Pubblicazione: (2025)
di: Ayesh, Mutaz, et al.
Pubblicazione: (2025)
SimplifyMyText: An LLM-Based System for Inclusive Plain Language Text Simplification
di: Färber, Michael, et al.
Pubblicazione: (2025)
di: Färber, Michael, et al.
Pubblicazione: (2025)
NoticIA: A Clickbait Article Summarization Dataset in Spanish
di: García-Ferrero, Iker, et al.
Pubblicazione: (2024)
di: García-Ferrero, Iker, et al.
Pubblicazione: (2024)
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
MessIRve: A Large-Scale Spanish Information Retrieval Dataset
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
RigoChat 2: an adapted language model to Spanish using a bounded dataset and reduced hardware
di: Gómez, Gonzalo Santamaría, et al.
Pubblicazione: (2025)
di: Gómez, Gonzalo Santamaría, et al.
Pubblicazione: (2025)
Online Social Support Detection in Spanish Social Media Texts
di: Tash, Moein Shahiki, et al.
Pubblicazione: (2025)
di: Tash, Moein Shahiki, et al.
Pubblicazione: (2025)
The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models
di: Graichen, Nora, et al.
Pubblicazione: (2026)
di: Graichen, Nora, et al.
Pubblicazione: (2026)
Classification of Radiological Text in Small and Imbalanced Datasets in a Non-English Language
di: Beliveau, Vincent, et al.
Pubblicazione: (2024)
di: Beliveau, Vincent, et al.
Pubblicazione: (2024)
Classification of Human- and AI-Generated Texts for English, French, German, and Spanish
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
di: Schaaff, Kristina, et al.
Pubblicazione: (2023)
MuSeD: A Multimodal Spanish Dataset for Sexism Detection in Social Media Videos
di: De Grazia, Laura, et al.
Pubblicazione: (2025)
di: De Grazia, Laura, et al.
Pubblicazione: (2025)
Investigating Large Language Models' Linguistic Abilities for Text Preprocessing
di: Braga, Marco, et al.
Pubblicazione: (2025)
di: Braga, Marco, et al.
Pubblicazione: (2025)
Seventeenth-Century Spanish American Notary Records for Fine-Tuning Spanish Large Language Models
di: Sarker, Shraboni, et al.
Pubblicazione: (2024)
di: Sarker, Shraboni, et al.
Pubblicazione: (2024)
Automatic Histograms: Leveraging Language Models for Text Dataset Exploration
di: Reif, Emily, et al.
Pubblicazione: (2024)
di: Reif, Emily, et al.
Pubblicazione: (2024)
LSTM-Based Text Generation: A Study on Historical Datasets
di: Hussein, Mustafa Abbas Hussein, et al.
Pubblicazione: (2024)
di: Hussein, Mustafa Abbas Hussein, et al.
Pubblicazione: (2024)
REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning
di: Lee, Seungmin, et al.
Pubblicazione: (2026)
di: Lee, Seungmin, et al.
Pubblicazione: (2026)
KGGen: Extracting Knowledge Graphs from Plain Text with Language Models
di: Mo, Belinda, et al.
Pubblicazione: (2025)
di: Mo, Belinda, et al.
Pubblicazione: (2025)
A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension
di: Hai, Toan Nguyen, et al.
Pubblicazione: (2025)
di: Hai, Toan Nguyen, et al.
Pubblicazione: (2025)
Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation
di: Abdullah, Sharif Mohammad, et al.
Pubblicazione: (2025)
di: Abdullah, Sharif Mohammad, et al.
Pubblicazione: (2025)
PulseLM: A Foundation Dataset and Benchmark for PPG-Text Learning
di: Pham, Hung Manh, et al.
Pubblicazione: (2026)
di: Pham, Hung Manh, et al.
Pubblicazione: (2026)
Spanish TrOCR: Leveraging Transfer Learning for Language Adaptation
di: Lauar, Filipe, et al.
Pubblicazione: (2024)
di: Lauar, Filipe, et al.
Pubblicazione: (2024)
Ontology-Free General-Domain Knowledge Graph-to-Text Generation Dataset Synthesis using Large Language Model
di: Kim, Daehee, et al.
Pubblicazione: (2024)
di: Kim, Daehee, et al.
Pubblicazione: (2024)
FMI@SU ToxHabits: Evaluating LLMs Performance on Toxic Habit Extraction in Spanish Clinical Texts
di: Vassileva, Sylvia, et al.
Pubblicazione: (2026)
di: Vassileva, Sylvia, et al.
Pubblicazione: (2026)
Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding
di: De, Anik, et al.
Pubblicazione: (2025)
di: De, Anik, et al.
Pubblicazione: (2025)
FinTextQA: A Dataset for Long-form Financial Question Answering
di: Chen, Jian, et al.
Pubblicazione: (2024)
di: Chen, Jian, et al.
Pubblicazione: (2024)
Text2VLM: Adapting Text-Only Datasets to Evaluate Alignment Training in Visual Language Models
di: Downer, Gabriel, et al.
Pubblicazione: (2025)
di: Downer, Gabriel, et al.
Pubblicazione: (2025)
ClinText-SP and RigoBERTa Clinical: a new set of open resources for Spanish Clinical NLP
di: Subies, Guillem García, et al.
Pubblicazione: (2025)
di: Subies, Guillem García, et al.
Pubblicazione: (2025)
Plain language adaptations of biomedical text using LLMs: Comparision of evaluation metrics
di: Kocbek, Primoz, et al.
Pubblicazione: (2025)
di: Kocbek, Primoz, et al.
Pubblicazione: (2025)
AceParse: A Comprehensive Dataset with Diverse Structured Texts for Academic Literature Parsing
di: Ji, Huawei, et al.
Pubblicazione: (2024)
di: Ji, Huawei, et al.
Pubblicazione: (2024)
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
di: Gwak, Daehoon, et al.
Pubblicazione: (2024)
di: Gwak, Daehoon, et al.
Pubblicazione: (2024)
HeSum: a Novel Dataset for Abstractive Text Summarization in Hebrew
di: Paz-Argaman, Tzuf, et al.
Pubblicazione: (2024)
di: Paz-Argaman, Tzuf, et al.
Pubblicazione: (2024)
Ace-CEFR -- A Dataset for Automated Evaluation of the Linguistic Difficulty of Conversational Texts for LLM Applications
di: Kogan, David, et al.
Pubblicazione: (2025)
di: Kogan, David, et al.
Pubblicazione: (2025)
MultiBanAbs: A Comprehensive Multi-Domain Bangla Abstractive Text Summarization Dataset
di: Ferdous, Md. Tanzim, et al.
Pubblicazione: (2025)
di: Ferdous, Md. Tanzim, et al.
Pubblicazione: (2025)
Datasets for Large Language Models: A Comprehensive Survey
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Lost-in-the-Middle in Long-Text Generation: Synthetic Dataset, Evaluation Framework, and Mitigation
di: Zhang, Junhao, et al.
Pubblicazione: (2025)
di: Zhang, Junhao, et al.
Pubblicazione: (2025)
BIRDTurk: Adaptation of the BIRD Text-to-SQL Dataset to Turkish
di: Aktaş, Burak, et al.
Pubblicazione: (2026)
di: Aktaş, Burak, et al.
Pubblicazione: (2026)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
Text2Zinc: A Cross-Domain Dataset for Modeling Optimization and Satisfaction Problems in MiniZinc
di: Singirikonda, Akash, et al.
Pubblicazione: (2025)
di: Singirikonda, Akash, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Text Adaptation to Plain Language and Easy Read via Automatic Post-Editing Cycles
di: Calleja, Jesús, et al.
Pubblicazione: (2025) -
BenchOverflow: Measuring Overflow in Large Language Models via Plain-Text Prompts
di: Feiglin, Erin, et al.
Pubblicazione: (2026) -
CardiffNLP at CLEARS-2025: Prompting Large Language Models for Plain Language and Easy-to-Read Text Rewriting
di: Ayesh, Mutaz, et al.
Pubblicazione: (2025) -
SimplifyMyText: An LLM-Based System for Inclusive Plain Language Text Simplification
di: Färber, Michael, et al.
Pubblicazione: (2025) -
NoticIA: A Clickbait Article Summarization Dataset in Spanish
di: García-Ferrero, Iker, et al.
Pubblicazione: (2024)