Exploring the features used for summary evaluation by Human and GPT
Fuente:
arXiv
Guardado en:
| Autores principales: | Sadeghi, Zahra, Milios, Evangelos, Rudzicz, Frank |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Stance Reasoner: Zero-Shot Stance Detection on Social Media with Explicit Reasoning
por: Taranukhin, Maksym, et al.
Publicado: (2024)
por: Taranukhin, Maksym, et al.
Publicado: (2024)
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights
por: Taranukhin, Maksym, et al.
Publicado: (2024)
por: Taranukhin, Maksym, et al.
Publicado: (2024)
Scenarios and Approaches for Situated Natural Language Explanations
por: Qiu, Pengshuo, et al.
Publicado: (2024)
por: Qiu, Pengshuo, et al.
Publicado: (2024)
Show, Don't Tell: Uncovering Implicit Character Portrayal using LLMs
por: Jaipersaud, Brandon, et al.
Publicado: (2024)
por: Jaipersaud, Brandon, et al.
Publicado: (2024)
VISLA Benchmark: Evaluating Embedding Sensitivity to Semantic and Lexical Alterations
por: Dumpala, Sri Harsha, et al.
Publicado: (2024)
por: Dumpala, Sri Harsha, et al.
Publicado: (2024)
MedSynth: Realistic, Synthetic Medical Dialogue-Note Pairs
por: Mianroodi, Ahmad Rezaie, et al.
Publicado: (2025)
por: Mianroodi, Ahmad Rezaie, et al.
Publicado: (2025)
Plug and Play with Prompts: A Prompt Tuning Approach for Controlling Text Generation
por: Ajwani, Rohan Deepak, et al.
Publicado: (2024)
por: Ajwani, Rohan Deepak, et al.
Publicado: (2024)
Exploring the Capability of ChatGPT to Reproduce Human Labels for Social Computing Tasks (Extended Version)
por: Zhu, Yiming, et al.
Publicado: (2024)
por: Zhu, Yiming, et al.
Publicado: (2024)
Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4
por: Uhler, Lydia, et al.
Publicado: (2024)
por: Uhler, Lydia, et al.
Publicado: (2024)
An empirical evaluation of using ChatGPT to summarize disputes for recommending similar labor and employment cases in Chinese
por: Wu, Po-Hsien, et al.
Publicado: (2024)
por: Wu, Po-Hsien, et al.
Publicado: (2024)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
por: Mao, Rui, et al.
Publicado: (2023)
por: Mao, Rui, et al.
Publicado: (2023)
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning
por: Hameed, Marawan Gamal Abdel, et al.
Publicado: (2024)
por: Hameed, Marawan Gamal Abdel, et al.
Publicado: (2024)
The Human and the Mechanical: logos, truthfulness, and ChatGPT
por: Giannakidou, Anastasia, et al.
Publicado: (2024)
por: Giannakidou, Anastasia, et al.
Publicado: (2024)
Utilizing deep learning models for the identification of enhancers and super-enhancers based on genomic and epigenomic features
por: Ahani, Zahra, et al.
Publicado: (2024)
por: Ahani, Zahra, et al.
Publicado: (2024)
Can we trust the evaluation on ChatGPT?
por: Aiyappa, Rachith, et al.
Publicado: (2023)
por: Aiyappa, Rachith, et al.
Publicado: (2023)
Differentiate ChatGPT-generated and Human-written Medical Texts
por: Liao, Wenxiong, et al.
Publicado: (2023)
por: Liao, Wenxiong, et al.
Publicado: (2023)
Exploring the Capabilities of ChatGPT in Ancient Chinese Translation and Person Name Recognition
por: Si, Shijing, et al.
Publicado: (2023)
por: Si, Shijing, et al.
Publicado: (2023)
A study on general visual categorization of objects into animal and plant groups using global shape descriptors with a focus on category-specific deficits
por: Sadeghi, Zahra
Publicado: (2019)
por: Sadeghi, Zahra
Publicado: (2019)
Generative Floor Plan Design with LLMs via Reinforcement Learning with Verifiable Rewards
por: Lara, Luis, et al.
Publicado: (2026)
por: Lara, Luis, et al.
Publicado: (2026)
Is ChatGPT More Empathetic than Humans?
por: Welivita, Anuradha, et al.
Publicado: (2024)
por: Welivita, Anuradha, et al.
Publicado: (2024)
Can Agent Conquer Web? Exploring the Frontiers of ChatGPT Atlas Agent in Web Games
por: Zhang, Jingran, et al.
Publicado: (2025)
por: Zhang, Jingran, et al.
Publicado: (2025)
Plancraft: an evaluation dataset for planning with LLM agents
por: Dagan, Gautier, et al.
Publicado: (2024)
por: Dagan, Gautier, et al.
Publicado: (2024)
CAMS: A CityGPT-Powered Agentic Framework for Urban Human Mobility Simulation
por: Du, Yuwei, et al.
Publicado: (2025)
por: Du, Yuwei, et al.
Publicado: (2025)
ChatGPT Rates Natural Language Explanation Quality Like Humans: But on Which Scales?
por: Huang, Fan, et al.
Publicado: (2024)
por: Huang, Fan, et al.
Publicado: (2024)
HC3 Plus: A Semantic-Invariant Human ChatGPT Comparison Corpus
por: Su, Zhenpeng, et al.
Publicado: (2023)
por: Su, Zhenpeng, et al.
Publicado: (2023)
Two-Pronged Human Evaluation of ChatGPT Self-Correction in Radiology Report Simplification
por: Yang, Ziyu, et al.
Publicado: (2024)
por: Yang, Ziyu, et al.
Publicado: (2024)
An evaluation of LLMs for generating movie reviews: GPT-4o, Gemini-2.0 and DeepSeek-V3
por: Sands, Brendan, et al.
Publicado: (2025)
por: Sands, Brendan, et al.
Publicado: (2025)
HumanEval on Latest GPT Models -- 2024
por: Li, Daniel, et al.
Publicado: (2024)
por: Li, Daniel, et al.
Publicado: (2024)
A Linguistic Comparison between Human and ChatGPT-Generated Conversations
por: Sandler, Morgan, et al.
Publicado: (2024)
por: Sandler, Morgan, et al.
Publicado: (2024)
Alpha-GPT: Human-AI Interactive Alpha Mining for Quantitative Investment
por: Wang, Saizhuo, et al.
Publicado: (2023)
por: Wang, Saizhuo, et al.
Publicado: (2023)
Exploring ChatGPT and its Impact on Society
por: Haque, Md. Asraful, et al.
Publicado: (2024)
por: Haque, Md. Asraful, et al.
Publicado: (2024)
Silver-Tongued and Sundry: Exploring Intersectional Pronouns with ChatGPT
por: Fujii, Takao, et al.
Publicado: (2024)
por: Fujii, Takao, et al.
Publicado: (2024)
Can GPT models Follow Human Summarization Guidelines? A Study for Targeted Communication Goals
por: Zhou, Yongxin, et al.
Publicado: (2023)
por: Zhou, Yongxin, et al.
Publicado: (2023)
Monte Carlo Tree Search for Recipe Generation using GPT-2
por: Taneja, Karan, et al.
Publicado: (2024)
por: Taneja, Karan, et al.
Publicado: (2024)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
por: Urchs, Stefanie, et al.
Publicado: (2023)
por: Urchs, Stefanie, et al.
Publicado: (2023)
Exploring Diversity, Novelty, and Popularity Bias in ChatGPT's Recommendations
por: Di Palma, Dario, et al.
Publicado: (2026)
por: Di Palma, Dario, et al.
Publicado: (2026)
IIMedGPT: Promoting Large Language Model Capabilities of Medical Tasks by Efficient Human Preference Alignment
por: Zhang, Yiming, et al.
Publicado: (2025)
por: Zhang, Yiming, et al.
Publicado: (2025)
Text and Audio Simplification: Human vs. ChatGPT
por: Leroy, Gondy, et al.
Publicado: (2024)
por: Leroy, Gondy, et al.
Publicado: (2024)
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels
por: Yan, Jianhao, et al.
Publicado: (2024)
por: Yan, Jianhao, et al.
Publicado: (2024)
ACCORD: Closing the Commonsense Measurability Gap
por: Roewer-Després, François, et al.
Publicado: (2024)
por: Roewer-Després, François, et al.
Publicado: (2024)
Ejemplares similares
-
Stance Reasoner: Zero-Shot Stance Detection on Social Media with Explicit Reasoning
por: Taranukhin, Maksym, et al.
Publicado: (2024) -
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights
por: Taranukhin, Maksym, et al.
Publicado: (2024) -
Scenarios and Approaches for Situated Natural Language Explanations
por: Qiu, Pengshuo, et al.
Publicado: (2024) -
Show, Don't Tell: Uncovering Implicit Character Portrayal using LLMs
por: Jaipersaud, Brandon, et al.
Publicado: (2024) -
VISLA Benchmark: Evaluating Embedding Sensitivity to Semantic and Lexical Alterations
por: Dumpala, Sri Harsha, et al.
Publicado: (2024)