PREF: Reference-Free Evaluation of Personalised Text Generation in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Fu, Xiao, Rahmani, Hossein A., Wu, Bin, Ramos, Jerome, Yilmaz, Emine, Lipani, Aldo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ink and Individuality: Crafting a Personalised Narrative in the Age of LLMs
por: Wasi, Azmine Toushik, et al.
Publicado: (2024)
por: Wasi, Azmine Toushik, et al.
Publicado: (2024)
Using Generative Text Models to Create Qualitative Codebooks for Student Evaluations of Teaching
por: Katz, Andrew, et al.
Publicado: (2024)
por: Katz, Andrew, et al.
Publicado: (2024)
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
por: Ellinger, Lukas, et al.
Publicado: (2026)
por: Ellinger, Lukas, et al.
Publicado: (2026)
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
por: Zhao, Runcong, et al.
Publicado: (2025)
por: Zhao, Runcong, et al.
Publicado: (2025)
RAG-based EEG-to-Text Translation Using Deep Learning and LLMs
por: Collautti, Enrico, et al.
Publicado: (2026)
por: Collautti, Enrico, et al.
Publicado: (2026)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
por: Tian, Yuan, et al.
Publicado: (2025)
por: Tian, Yuan, et al.
Publicado: (2025)
Evaluating LLMs as Human Surrogates in Controlled Experiments
por: Hoq, Adnan, et al.
Publicado: (2026)
por: Hoq, Adnan, et al.
Publicado: (2026)
Collaborative Evaluation of Deepfake Text with Deliberation-Enhancing Dialogue Systems
por: Lee, Jooyoung, et al.
Publicado: (2025)
por: Lee, Jooyoung, et al.
Publicado: (2025)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
por: Kim, Tae Soo, et al.
Publicado: (2025)
por: Kim, Tae Soo, et al.
Publicado: (2025)
Can LLMs Generate Visualizations with Dataless Prompts?
por: Coelho, Darius, et al.
Publicado: (2024)
por: Coelho, Darius, et al.
Publicado: (2024)
Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain
por: Chiang, Shang-Hsuan, et al.
Publicado: (2026)
por: Chiang, Shang-Hsuan, et al.
Publicado: (2026)
From Text to Self: Users' Perceptions of Potential of AI on Interpersonal Communication and Self
por: Fu, Yue, et al.
Publicado: (2023)
por: Fu, Yue, et al.
Publicado: (2023)
PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users
por: Kirk, Hannah Rose, et al.
Publicado: (2026)
por: Kirk, Hannah Rose, et al.
Publicado: (2026)
Meta-Evaluating Local LLMs: Rethinking Performance Metrics for Serious Games
por: Isaza-Giraldo, Andrés, et al.
Publicado: (2025)
por: Isaza-Giraldo, Andrés, et al.
Publicado: (2025)
Unpacking Human Preference for LLMs: Demographically Aware Evaluation with the HUMAINE Framework
por: Petrova, Nora, et al.
Publicado: (2026)
por: Petrova, Nora, et al.
Publicado: (2026)
An AI-Powered Research Assistant in the Lab: A Practical Guide for Text Analysis Through Iterative Collaboration with LLMs
por: Carmona-Díaz, Gino, et al.
Publicado: (2025)
por: Carmona-Díaz, Gino, et al.
Publicado: (2025)
VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents
por: Lee, Sam Yu-Te, et al.
Publicado: (2025)
por: Lee, Sam Yu-Te, et al.
Publicado: (2025)
LalaEval: A Holistic Human Evaluation Framework for Domain-Specific Large Language Models
por: Sun, Chongyan, et al.
Publicado: (2024)
por: Sun, Chongyan, et al.
Publicado: (2024)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
por: Liu, Hongtao, et al.
Publicado: (2025)
por: Liu, Hongtao, et al.
Publicado: (2025)
Dehumanizing Machines: Mitigating Anthropomorphic Behaviors in Text Generation Systems
por: Cheng, Myra, et al.
Publicado: (2025)
por: Cheng, Myra, et al.
Publicado: (2025)
Towards Stable and Personalised Profiles for Lexical Alignment in Spoken Human-Agent Dialogue
por: Schaaij, Keara, et al.
Publicado: (2025)
por: Schaaij, Keara, et al.
Publicado: (2025)
HARGPT: Are LLMs Zero-Shot Human Activity Recognizers?
por: Ji, Sijie, et al.
Publicado: (2024)
por: Ji, Sijie, et al.
Publicado: (2024)
Human Evaluation of Procedural Knowledge Graph Extraction from Text with Large Language Models
por: Carriero, Valentina Anita, et al.
Publicado: (2024)
por: Carriero, Valentina Anita, et al.
Publicado: (2024)
Navigating the Path of Writing: Outline-guided Text Generation with Large Language Models
por: Lee, Yukyung, et al.
Publicado: (2024)
por: Lee, Yukyung, et al.
Publicado: (2024)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
por: Oh, Juhyun, et al.
Publicado: (2024)
por: Oh, Juhyun, et al.
Publicado: (2024)
Logic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMs
por: Rahdari, Behnam, et al.
Publicado: (2023)
por: Rahdari, Behnam, et al.
Publicado: (2023)
Aligning LLMs with Individual Preferences via Interaction
por: Wu, Shujin, et al.
Publicado: (2024)
por: Wu, Shujin, et al.
Publicado: (2024)
Generation Z's Ability to Discriminate Between AI-generated and Human-Authored Text on Discord
por: Ramu, Dhruv, et al.
Publicado: (2023)
por: Ramu, Dhruv, et al.
Publicado: (2023)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
por: Zengaffinen, Yanick, et al.
Publicado: (2026)
por: Zengaffinen, Yanick, et al.
Publicado: (2026)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
por: Joshi, Ishika, et al.
Publicado: (2024)
por: Joshi, Ishika, et al.
Publicado: (2024)
BADGE: BADminton report Generation and Evaluation with LLM
por: Chiang, Shang-Hsuan, et al.
Publicado: (2024)
por: Chiang, Shang-Hsuan, et al.
Publicado: (2024)
Alignment Drift in Multimodal LLMs: A Two-Phase, Longitudinal Evaluation of Harm Across Eight Model Releases
por: Ford, Casey, et al.
Publicado: (2026)
por: Ford, Casey, et al.
Publicado: (2026)
Human Bias in the Face of AI: Examining Human Judgment Against Text Labeled as AI Generated
por: Zhu, Tiffany, et al.
Publicado: (2024)
por: Zhu, Tiffany, et al.
Publicado: (2024)
AgentCTG: Harnessing Multi-Agent Collaboration for Fine-Grained Precise Control in Text Generation
por: Zhou, Xinxu, et al.
Publicado: (2025)
por: Zhou, Xinxu, et al.
Publicado: (2025)
Consistency of Responses and Continuations Generated by Large Language Models on Social Media
por: Xu, Wentao, et al.
Publicado: (2025)
por: Xu, Wentao, et al.
Publicado: (2025)
Learning to Generate and Evaluate Fact-checking Explanations with Transformers
por: Feher, Darius, et al.
Publicado: (2024)
por: Feher, Darius, et al.
Publicado: (2024)
Evaluating Behavioral Alignment in Conflict Dialogue: A Multi-Dimensional Comparison of LLM Agents and Humans
por: Kwon, Deuksin, et al.
Publicado: (2025)
por: Kwon, Deuksin, et al.
Publicado: (2025)
Performance Gains of LLMs With Humans in a World of LLMs Versus Humans
por: McCullum, Lucas, et al.
Publicado: (2025)
por: McCullum, Lucas, et al.
Publicado: (2025)
Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
por: Wang, Junling, et al.
Publicado: (2025)
por: Wang, Junling, et al.
Publicado: (2025)
Can we Debias Social Stereotypes in AI-Generated Images? Examining Text-to-Image Outputs and User Perceptions
por: Barve, Saharsh, et al.
Publicado: (2025)
por: Barve, Saharsh, et al.
Publicado: (2025)
Ejemplares similares
-
Ink and Individuality: Crafting a Personalised Narrative in the Age of LLMs
por: Wasi, Azmine Toushik, et al.
Publicado: (2024) -
Using Generative Text Models to Create Qualitative Codebooks for Student Evaluations of Teaching
por: Katz, Andrew, et al.
Publicado: (2024) -
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
por: Ellinger, Lukas, et al.
Publicado: (2026) -
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
por: Zhao, Runcong, et al.
Publicado: (2025) -
RAG-based EEG-to-Text Translation Using Deep Learning and LLMs
por: Collautti, Enrico, et al.
Publicado: (2026)