A Comparative Study of Controlled Text Generation Systems Using Level-Playing-Field Evaluation Principles
Fuente:
arXiv
Guardado en:
| Autores principales: | Lorandi, Michela, Belz, Anya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation
por: Lorandi, Michela, et al.
Publicado: (2026)
por: Lorandi, Michela, et al.
Publicado: (2026)
Reproducing the Metric-Based Evaluation of a Set of Controllable Text Generation Techniques
por: Lorandi, Michela, et al.
Publicado: (2024)
por: Lorandi, Michela, et al.
Publicado: (2024)
High-quality Data-to-Text Generation for Severely Under-Resourced Languages with Out-of-the-box Large Language Models
por: Lorandi, Michela, et al.
Publicado: (2024)
por: Lorandi, Michela, et al.
Publicado: (2024)
Enhancing Study-Level Inference from Clinical Trial Papers via Reinforcement Learning-Based Numeric Reasoning
por: Pronesti, Massimiliano, et al.
Publicado: (2025)
por: Pronesti, Massimiliano, et al.
Publicado: (2025)
QRA++: Quantified Reproducibility Assessment for Common Types of Results in Natural Language Processing
por: Belz, Anya
Publicado: (2025)
por: Belz, Anya
Publicado: (2025)
HEDS 3.0: The Human Evaluation Data Sheet Version 3.0
por: Belz, Anya, et al.
Publicado: (2024)
por: Belz, Anya, et al.
Publicado: (2024)
Induction Signatures Are Not Enough: A Matched-Compute Study of Load-Bearing Structure in In-Context Learning
por: Sabry, Mohammed, et al.
Publicado: (2025)
por: Sabry, Mohammed, et al.
Publicado: (2025)
The QCET Taxonomy of Standard Quality Criterion Names and Definitions for the Evaluation of NLP Systems
por: Belz, Anya, et al.
Publicado: (2025)
por: Belz, Anya, et al.
Publicado: (2025)
Budgeted LoRA: Distillation as Structured Compute Allocation for Efficient Inference
por: Sabry, Mohammed, et al.
Publicado: (2026)
por: Sabry, Mohammed, et al.
Publicado: (2026)
Assessing the Portability of Parameter Matrices Trained by Parameter-Efficient Finetuning Methods
por: Sabry, Mohammed, et al.
Publicado: (2024)
por: Sabry, Mohammed, et al.
Publicado: (2024)
News is More than a Collection of Facts: Moral Frame Preserving News Summarization
por: Liscio, Enrico, et al.
Publicado: (2025)
por: Liscio, Enrico, et al.
Publicado: (2025)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
por: Pronesti, Massimiliano, et al.
Publicado: (2026)
por: Pronesti, Massimiliano, et al.
Publicado: (2026)
Query-driven Document-level Scientific Evidence Extraction from Biomedical Studies
por: Pronesti, Massimiliano, et al.
Publicado: (2025)
por: Pronesti, Massimiliano, et al.
Publicado: (2025)
Comparative Evaluation of Machine Translation Systems on Images with Text
por: Puchol, Blai, et al.
Publicado: (2026)
por: Puchol, Blai, et al.
Publicado: (2026)
Plug and Play with Prompts: A Prompt Tuning Approach for Controlling Text Generation
por: Ajwani, Rohan Deepak, et al.
Publicado: (2024)
por: Ajwani, Rohan Deepak, et al.
Publicado: (2024)
A Comparative Study of Quality Evaluation Methods for Text Summarization
por: Nguyen, Huyen, et al.
Publicado: (2024)
por: Nguyen, Huyen, et al.
Publicado: (2024)
Creation of a Numerical Scoring System to Objectively Measure and Compare the Level of Rhetoric in Arabic Texts: A Feasibility Study, and A Working Prototype
por: Marathe, Mandar
Publicado: (2025)
por: Marathe, Mandar
Publicado: (2025)
Evaluating Large Language Models for Diacritic Restoration in Romanian Texts: A Comparative Study
por: Nadas, Mihai, et al.
Publicado: (2025)
por: Nadas, Mihai, et al.
Publicado: (2025)
A Comparative Study of Decoding Strategies in Medical Text Generation
por: Presacan, Oriana, et al.
Publicado: (2025)
por: Presacan, Oriana, et al.
Publicado: (2025)
Evaluating the Smooth Control of Attribute Intensity in Text Generation with LLMs
por: Zhou, Shang, et al.
Publicado: (2024)
por: Zhou, Shang, et al.
Publicado: (2024)
LLMs vs. Chinese Anime Enthusiasts: A Comparative Study on Emotionally Supportive Role-Playing
por: Qiu, Lanlan, et al.
Publicado: (2025)
por: Qiu, Lanlan, et al.
Publicado: (2025)
Facet-Level Persona Control by Trait-Activated Routing with Contrastive SAE for Role-Playing LLMs
por: Tang, Wenqiu, et al.
Publicado: (2026)
por: Tang, Wenqiu, et al.
Publicado: (2026)
Evaluating Prompt Engineering Strategies for Sentiment Control in AI-Generated Texts
por: Sahler, Kerstin, et al.
Publicado: (2026)
por: Sahler, Kerstin, et al.
Publicado: (2026)
Noise Steering for Controlled Text Generation: Improving Diversity and Reading-Level Fidelity in Arabic Educational Story Generation
por: Khalid, Haziq Mohammad, et al.
Publicado: (2026)
por: Khalid, Haziq Mohammad, et al.
Publicado: (2026)
Irony Detection in Urdu Text: A Comparative Study Using Machine Learning Models and Large Language Models
por: Ahmad, Fiaz, et al.
Publicado: (2025)
por: Ahmad, Fiaz, et al.
Publicado: (2025)
Robust Detection of LLM-Generated Text: A Comparative Analysis
por: Su, Yongye, et al.
Publicado: (2024)
por: Su, Yongye, et al.
Publicado: (2024)
Fine-Grained Detection of AI-Generated Text Using Sentence-Level Segmentation
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
Can Prompt Modifiers Control Bias? A Comparative Analysis of Text-to-Image Generative Models
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
por: Shin, Philip Wootaek, et al.
Publicado: (2024)
DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Agents
por: Zhang, Rongsheng, et al.
Publicado: (2026)
por: Zhang, Rongsheng, et al.
Publicado: (2026)
Towards Fine-Grained Citation Evaluation in Generated Text: A Comparative Analysis of Faithfulness Metrics
por: Zhang, Weijia, et al.
Publicado: (2024)
por: Zhang, Weijia, et al.
Publicado: (2024)
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds
por: Wang, Lei, et al.
Publicado: (2024)
por: Wang, Lei, et al.
Publicado: (2024)
Harnessing the Plug-and-Play Controller by Prompting
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Input Matters: Evaluating Input Structure's Impact on LLM Summaries of Sports Play-by-Play
por: Sundararajan, Barkavi, et al.
Publicado: (2025)
por: Sundararajan, Barkavi, et al.
Publicado: (2025)
Modeling Comparative Logical Relation with Contrastive Learning for Text Generation
por: Dan, Yuhao, et al.
Publicado: (2024)
por: Dan, Yuhao, et al.
Publicado: (2024)
A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems
por: Borgholt, Lasse, et al.
Publicado: (2025)
por: Borgholt, Lasse, et al.
Publicado: (2025)
FAID: Fine-Grained AI-Generated Text Detection Using Multi-Task Auxiliary and Multi-Level Contrastive Learning
por: Ta, Minh Ngoc, et al.
Publicado: (2025)
por: Ta, Minh Ngoc, et al.
Publicado: (2025)
Beyond Turing: A Comparative Analysis of Approaches for Detecting Machine-Generated Text
por: Adilazuarda, Muhammad Farid
Publicado: (2023)
por: Adilazuarda, Muhammad Farid
Publicado: (2023)
CIE: Controlling Language Model Text Generations Using Continuous Signals
por: Samuel, Vinay, et al.
Publicado: (2025)
por: Samuel, Vinay, et al.
Publicado: (2025)
Principled Gradient-based Markov Chain Monte Carlo for Text Generation
por: Du, Li, et al.
Publicado: (2023)
por: Du, Li, et al.
Publicado: (2023)
Sentence Smith: Controllable Edits for Evaluating Text Embeddings
por: Li, Hongji, et al.
Publicado: (2025)
por: Li, Hongji, et al.
Publicado: (2025)
Ejemplares similares
-
Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation
por: Lorandi, Michela, et al.
Publicado: (2026) -
Reproducing the Metric-Based Evaluation of a Set of Controllable Text Generation Techniques
por: Lorandi, Michela, et al.
Publicado: (2024) -
High-quality Data-to-Text Generation for Severely Under-Resourced Languages with Out-of-the-box Large Language Models
por: Lorandi, Michela, et al.
Publicado: (2024) -
Enhancing Study-Level Inference from Clinical Trial Papers via Reinforcement Learning-Based Numeric Reasoning
por: Pronesti, Massimiliano, et al.
Publicado: (2025) -
QRA++: Quantified Reproducibility Assessment for Common Types of Results in Natural Language Processing
por: Belz, Anya
Publicado: (2025)