Large Language Model Agent Personality and Response Appropriateness: Evaluation by Human Linguistic Experts, LLM-as-Judge, and Natural Language Processing Model
Fuente:
arXiv
Guardado en:
| Autores principales: | Jayakumar, Eswari, Dash, Niladri Sekhar, Mukherjee, Debasmita |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models
por: Bo, Jessica Y., et al.
Publicado: (2024)
por: Bo, Jessica Y., et al.
Publicado: (2024)
Limitations of the LLM-as-a-Judge Approach for Evaluating LLM Outputs in Expert Knowledge Tasks
por: Szymanski, Annalisa, et al.
Publicado: (2024)
por: Szymanski, Annalisa, et al.
Publicado: (2024)
PhysioLLM: Supporting Personalized Health Insights with Wearables and Large Language Models
por: Fang, Cathy Mengying, et al.
Publicado: (2024)
por: Fang, Cathy Mengying, et al.
Publicado: (2024)
Subjective Code Preferences in Experts and Large Language Models
por: Mokhova, Anna, et al.
Publicado: (2026)
por: Mokhova, Anna, et al.
Publicado: (2026)
LLM-for-X: Application-agnostic Integration of Large Language Models to Support Personal Writing Workflows
por: Teufelberger, Lukas, et al.
Publicado: (2024)
por: Teufelberger, Lukas, et al.
Publicado: (2024)
When Large Language Models are Reliable for Judging Empathic Communication
por: Kumar, Aakriti, et al.
Publicado: (2025)
por: Kumar, Aakriti, et al.
Publicado: (2025)
Appropriateness of LLM-equipped Robotic Well-being Coach Language in the Workplace: A Qualitative Evaluation
por: Spitale, Micol, et al.
Publicado: (2024)
por: Spitale, Micol, et al.
Publicado: (2024)
Fostering Appropriate Reliance on Large Language Models: The Role of Explanations, Sources, and Inconsistencies
por: Kim, Sunnie S. Y., et al.
Publicado: (2025)
por: Kim, Sunnie S. Y., et al.
Publicado: (2025)
Understanding Large-Language Model (LLM)-powered Human-Robot Interaction
por: Kim, Callie Y., et al.
Publicado: (2024)
por: Kim, Callie Y., et al.
Publicado: (2024)
An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering Systems
por: Martin-Boyle, Anna, et al.
Publicado: (2026)
por: Martin-Boyle, Anna, et al.
Publicado: (2026)
Natural Language Dataset Generation Framework for Visualizations Powered by Large Language Models
por: Ko, Hyung-Kwon, et al.
Publicado: (2023)
por: Ko, Hyung-Kwon, et al.
Publicado: (2023)
Visualization Generation with Large Language Models: An Evaluation
por: Wang, Xinyu, et al.
Publicado: (2024)
por: Wang, Xinyu, et al.
Publicado: (2024)
Harmony: A Human-Aware, Responsive, Modular Assistant with a Locally Deployed Large Language Model
por: Yin, Ziqi, et al.
Publicado: (2024)
por: Yin, Ziqi, et al.
Publicado: (2024)
Large Language Model-based Human-Agent Collaboration for Complex Task Solving
por: Feng, Xueyang, et al.
Publicado: (2024)
por: Feng, Xueyang, et al.
Publicado: (2024)
Are Humans as Brittle as Large Language Models?
por: Li, Jiahui, et al.
Publicado: (2025)
por: Li, Jiahui, et al.
Publicado: (2025)
Evaluation of Large Language Model-Driven AutoML in Data and Model Management from Human-Centered Perspective
por: Yao, Jiapeng, et al.
Publicado: (2025)
por: Yao, Jiapeng, et al.
Publicado: (2025)
Leveraging Large Language Models to Enhance Domain Expert Inclusion in Data Science Workflows
por: Shih, Jasmine Y., et al.
Publicado: (2024)
por: Shih, Jasmine Y., et al.
Publicado: (2024)
Task Supportive and Personalized Human-Large Language Model Interaction: A User Study
por: Wang, Ben, et al.
Publicado: (2024)
por: Wang, Ben, et al.
Publicado: (2024)
Human-Centered Design Recommendations for LLM-as-a-Judge
por: Pan, Qian, et al.
Publicado: (2024)
por: Pan, Qian, et al.
Publicado: (2024)
Generating Analytic Specifications for Data Visualization from Natural Language Queries using Large Language Models
por: Sah, Subham, et al.
Publicado: (2024)
por: Sah, Subham, et al.
Publicado: (2024)
Natural Language but Omitted? On the Ineffectiveness of Large Language Models' privacy policy from End-users' Perspective
por: Zhang, Shuning, et al.
Publicado: (2024)
por: Zhang, Shuning, et al.
Publicado: (2024)
Human-Computer Interaction and Visualization in Natural Language Generation Models: Applications, Challenges, and Opportunities
por: Wang, Yunchao, et al.
Publicado: (2024)
por: Wang, Yunchao, et al.
Publicado: (2024)
Abstract Operations Research Modeling Using Natural Language Inputs
por: Li, Junxuan, et al.
Publicado: (2024)
por: Li, Junxuan, et al.
Publicado: (2024)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
por: Jiang, Hang, et al.
Publicado: (2023)
por: Jiang, Hang, et al.
Publicado: (2023)
Leveraging Large Language Models for Generating Mobile Sensing Strategies in Human Behavior Modeling
por: Gao, Nan, et al.
Publicado: (2023)
por: Gao, Nan, et al.
Publicado: (2023)
Do Language Model Agents Align with Humans in Rating Visualizations? An Empirical Study
por: Shao, Zekai, et al.
Publicado: (2025)
por: Shao, Zekai, et al.
Publicado: (2025)
GPTutor: Great Personalized Tutor with Large Language Models for Personalized Learning Content Generation
por: Chen, Eason, et al.
Publicado: (2024)
por: Chen, Eason, et al.
Publicado: (2024)
TalkToAgent: A Human-centric Explanation of Reinforcement Learning Agents with Large Language Models
por: Kim, Haechang, et al.
Publicado: (2025)
por: Kim, Haechang, et al.
Publicado: (2025)
Learning in Context: Personalizing Educational Content with Large Language Models to Enhance Student Learning
por: Lim, Joy Jia Yin, et al.
Publicado: (2025)
por: Lim, Joy Jia Yin, et al.
Publicado: (2025)
Large Language Models Predict Human Well-being -- But Not Equally Everywhere
por: Pataranutaporn, Pat, et al.
Publicado: (2025)
por: Pataranutaporn, Pat, et al.
Publicado: (2025)
ChatVis: Large Language Model Agent for Generating Scientific Visualizations
por: Peterka, Tom, et al.
Publicado: (2025)
por: Peterka, Tom, et al.
Publicado: (2025)
Comparing Human Expertise and Large Language Models Embeddings in Content Validity Assessment of Personality Tests
por: Milano, Nicola, et al.
Publicado: (2025)
por: Milano, Nicola, et al.
Publicado: (2025)
Insights from Social Shaping Theory: The Appropriation of Large Language Models in an Undergraduate Programming Course
por: Padiyath, Aadarsh, et al.
Publicado: (2024)
por: Padiyath, Aadarsh, et al.
Publicado: (2024)
If You Had to Pitch Your Ideal Software -- Evaluating Large Language Models to Support User Scenario Writing for User Experience Experts and Laypersons
por: Stadler, Patrick, et al.
Publicado: (2025)
por: Stadler, Patrick, et al.
Publicado: (2025)
Human-AI Alignment of Multimodal Large Language Models with Speech-Language Pathologists in Parent-Child Interactions
por: Shi, Weiyan, et al.
Publicado: (2025)
por: Shi, Weiyan, et al.
Publicado: (2025)
Large Language Models for Virtual Human Gesture Selection
por: Torshizi, Parisa Ghanad, et al.
Publicado: (2025)
por: Torshizi, Parisa Ghanad, et al.
Publicado: (2025)
An Optimized Framework for Processing Large-scale Polysomnographic Data Incorporating Expert Human Oversight
por: Holm, Benedikt, et al.
Publicado: (2024)
por: Holm, Benedikt, et al.
Publicado: (2024)
Can Large Language Model Agents Simulate Human Trust Behavior?
por: Xie, Chengxing, et al.
Publicado: (2024)
por: Xie, Chengxing, et al.
Publicado: (2024)
Large Language Models with Human-In-The-Loop Validation for Systematic Review Data Extraction
por: Schroeder, Noah L., et al.
Publicado: (2025)
por: Schroeder, Noah L., et al.
Publicado: (2025)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
por: Lian, Zheng, et al.
Publicado: (2025)
por: Lian, Zheng, et al.
Publicado: (2025)
Ejemplares similares
-
To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models
por: Bo, Jessica Y., et al.
Publicado: (2024) -
Limitations of the LLM-as-a-Judge Approach for Evaluating LLM Outputs in Expert Knowledge Tasks
por: Szymanski, Annalisa, et al.
Publicado: (2024) -
PhysioLLM: Supporting Personalized Health Insights with Wearables and Large Language Models
por: Fang, Cathy Mengying, et al.
Publicado: (2024) -
Subjective Code Preferences in Experts and Large Language Models
por: Mokhova, Anna, et al.
Publicado: (2026) -
LLM-for-X: Application-agnostic Integration of Large Language Models to Support Personal Writing Workflows
por: Teufelberger, Lukas, et al.
Publicado: (2024)