MuTSE: A Human-in-the-Loop Multi-use Text Simplification Evaluator
Fuente:
arXiv
Salvato in:
| Autori principali: | Roscan, Rares-Alexandru, Petre1, Gabriel, Dumitran, Adrian-Marius, Dumitran, Angela-Liliana |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Agentic Evaluation Architecture for Historical Bias Detection in Educational Textbooks
di: Stefan, Gabriel, et al.
Pubblicazione: (2026)
di: Stefan, Gabriel, et al.
Pubblicazione: (2026)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2024)
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2024)
A Cross-Lingual Analysis of Bias in Large Language Models Using Romanian History
di: Cocu, Matei-Iulian, et al.
Pubblicazione: (2025)
di: Cocu, Matei-Iulian, et al.
Pubblicazione: (2025)
From Struggle (06-2024) to Mastery (02-2025) LLMs Conquer Advanced Algorithm Exams and Pave the Way for Editorial Generation
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
RoMathExam: A Longitudinal Dataset of Romanian Math Exams (1895-2025) with a Seven-Decade Core (1957-2025)
di: Cuclea, Luca-Ncolae, et al.
Pubblicazione: (2026)
di: Cuclea, Luca-Ncolae, et al.
Pubblicazione: (2026)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
di: Marius, Dumitran Adrian, et al.
Pubblicazione: (2025)
di: Marius, Dumitran Adrian, et al.
Pubblicazione: (2025)
Leveraging Generative AI for Enhancing Automated Assessment in Programming Education Contests
di: Dascalescu, Stefan, et al.
Pubblicazione: (2025)
di: Dascalescu, Stefan, et al.
Pubblicazione: (2025)
BacPrep: Lessons from Deploying an LLM-Based Bacalaureat Assessment Platform
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
A Culturally-Rich Romanian NLP Dataset from "Who Wants to Be a Millionaire?" Videos
di: Ganea, Alexandru-Gabriel, et al.
Pubblicazione: (2025)
di: Ganea, Alexandru-Gabriel, et al.
Pubblicazione: (2025)
Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
Algorithmical Aspects of Some Bio Inspired Operations
di: Dumitran, Marius
Pubblicazione: (2025)
di: Dumitran, Marius
Pubblicazione: (2025)
Exploring Large Language Models for Translating Romanian Computational Problems into English
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)
RoBiologyDataChoiceQA: A Romanian Dataset for improving Biology understanding of Large Language Models
di: Ghinea, Dragos-Dumitru, et al.
Pubblicazione: (2025)
di: Ghinea, Dragos-Dumitru, et al.
Pubblicazione: (2025)
SiTSE: Sinhala Text Simplification Dataset and Evaluation
di: Ranathunga, Surangika, et al.
Pubblicazione: (2024)
di: Ranathunga, Surangika, et al.
Pubblicazione: (2024)
Automated Feedback Loops to Protect Text Simplification with Generative AI from Information Loss
di: Nandiraju, Abhay Kumara Sri Krishna, et al.
Pubblicazione: (2025)
di: Nandiraju, Abhay Kumara Sri Krishna, et al.
Pubblicazione: (2025)
Text and Audio Simplification: Human vs. ChatGPT
di: Leroy, Gondy, et al.
Pubblicazione: (2024)
di: Leroy, Gondy, et al.
Pubblicazione: (2024)
Difficulty Estimation and Simplification of French Text Using LLMs
di: Jamet, Henri, et al.
Pubblicazione: (2024)
di: Jamet, Henri, et al.
Pubblicazione: (2024)
An In-depth Evaluation of Large Language Models in Sentence Simplification with Error-based Human Assessment
di: Wu, Xuanxin, et al.
Pubblicazione: (2024)
di: Wu, Xuanxin, et al.
Pubblicazione: (2024)
Two-Pronged Human Evaluation of ChatGPT Self-Correction in Radiology Report Simplification
di: Yang, Ziyu, et al.
Pubblicazione: (2024)
di: Yang, Ziyu, et al.
Pubblicazione: (2024)
Large Language Models for Biomedical Text Simplification: Promising But Not There Yet
di: Li, Zihao, et al.
Pubblicazione: (2024)
di: Li, Zihao, et al.
Pubblicazione: (2024)
Evaluating the Effectiveness of Direct Preference Optimization for Personalizing German Automatic Text Simplifications for Persons with Intellectual Disabilities
di: Gao, Yingqiang, et al.
Pubblicazione: (2025)
di: Gao, Yingqiang, et al.
Pubblicazione: (2025)
Beyond Repetition: Text Simplification and Curriculum Learning for Data-Constrained Pretraining
di: Roque, Matthew Theodore, et al.
Pubblicazione: (2025)
di: Roque, Matthew Theodore, et al.
Pubblicazione: (2025)
Scaling, Simplification, and Adaptation: Lessons from Pretraining on Machine-Translated Text
di: Velasco, Dan John, et al.
Pubblicazione: (2025)
di: Velasco, Dan John, et al.
Pubblicazione: (2025)
MultiLS: A Multi-task Lexical Simplification Framework
di: North, Kai, et al.
Pubblicazione: (2024)
di: North, Kai, et al.
Pubblicazione: (2024)
Towards Optimizing and Evaluating a Retrieval Augmented QA Chatbot using LLMs with Human in the Loop
di: Afzal, Anum, et al.
Pubblicazione: (2024)
di: Afzal, Anum, et al.
Pubblicazione: (2024)
SimplifyMyText: An LLM-Based System for Inclusive Plain Language Text Simplification
di: Färber, Michael, et al.
Pubblicazione: (2025)
di: Färber, Michael, et al.
Pubblicazione: (2025)
Edit-Constrained Decoding for Sentence Simplification
di: Zetsu, Tatsuya, et al.
Pubblicazione: (2024)
di: Zetsu, Tatsuya, et al.
Pubblicazione: (2024)
Knowledge Graph-Driven Retrieval-Augmented Generation: Integrating Deepseek-R1 with Weaviate for Advanced Chatbot Applications
di: Lecu, Alexandru, et al.
Pubblicazione: (2025)
di: Lecu, Alexandru, et al.
Pubblicazione: (2025)
Towards Emotionally Consistent Text-Based Speech Editing: Introducing EmoCorrector and The ECD-TSE Dataset
di: Liu, Rui, et al.
Pubblicazione: (2025)
di: Liu, Rui, et al.
Pubblicazione: (2025)
GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human
di: Lin, Yihang, et al.
Pubblicazione: (2026)
di: Lin, Yihang, et al.
Pubblicazione: (2026)
Navigating Large-Scale Document Collections: MuDABench for Multi-Document Analytical QA
di: Li, Zhanli, et al.
Pubblicazione: (2026)
di: Li, Zhanli, et al.
Pubblicazione: (2026)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
di: Huang, Hui, et al.
Pubblicazione: (2025)
di: Huang, Hui, et al.
Pubblicazione: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
di: Lior, Gili, et al.
Pubblicazione: (2025)
di: Lior, Gili, et al.
Pubblicazione: (2025)
HumBEL: A Human-in-the-Loop Approach for Evaluating Demographic Factors of Language Models in Human-Machine Conversations
di: Sicilia, Anthony, et al.
Pubblicazione: (2023)
di: Sicilia, Anthony, et al.
Pubblicazione: (2023)
Block-Based Pathfinding: A Minecraft System for Visualizing Graph Algorithms
di: Pirvu, Luca-Stefan, et al.
Pubblicazione: (2026)
di: Pirvu, Luca-Stefan, et al.
Pubblicazione: (2026)
Parallel Algorithms for the One Sided Crossing Minimization Problem
di: Popa, Bogdan-Ioan, et al.
Pubblicazione: (2025)
di: Popa, Bogdan-Ioan, et al.
Pubblicazione: (2025)
MuLD: The Multitask Long Document Benchmark
di: Hudson, G Thomas, et al.
Pubblicazione: (2022)
di: Hudson, G Thomas, et al.
Pubblicazione: (2022)
APIO: Automatic Prompt Induction and Optimization for Grammatical Error Correction and Text Simplification
di: Chernodub, Artem, et al.
Pubblicazione: (2025)
di: Chernodub, Artem, et al.
Pubblicazione: (2025)
Health Text Simplification: An Annotated Corpus for Digestive Cancer Education and Novel Strategies for Reinforcement Learning
di: Rahman, Md Mushfiqur, et al.
Pubblicazione: (2024)
di: Rahman, Md Mushfiqur, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An Agentic Evaluation Architecture for Historical Bias Detection in Educational Textbooks
di: Stefan, Gabriel, et al.
Pubblicazione: (2026) -
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025) -
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2024) -
A Cross-Lingual Analysis of Bias in Large Language Models Using Romanian History
di: Cocu, Matei-Iulian, et al.
Pubblicazione: (2025) -
From Struggle (06-2024) to Mastery (02-2025) LLMs Conquer Advanced Algorithm Exams and Pave the Way for Editorial Generation
di: Dumitran, Adrian Marius, et al.
Pubblicazione: (2025)