GPTEval: A Survey on Assessments of ChatGPT and GPT-4
Fuente:
arXiv
Saved in:
| Main Authors: | Mao, Rui, Chen, Guanyi, Zhang, Xulang, Guerin, Frank, Cambria, Erik |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Systematic Analysis of Biases in Large Language Models
by: Zhang, Xulang, et al.
Published: (2025)
by: Zhang, Xulang, et al.
Published: (2025)
A Survey on the Real Power of ChatGPT
by: Liu, Ming, et al.
Published: (2024)
by: Liu, Ming, et al.
Published: (2024)
Primacy Effect of ChatGPT
by: Wang, Yiwei, et al.
Published: (2023)
by: Wang, Yiwei, et al.
Published: (2023)
Fairness of ChatGPT
by: Li, Yunqi, et al.
Published: (2023)
by: Li, Yunqi, et al.
Published: (2023)
Performance Assessment of ChatGPT vs Bard in Detecting Alzheimer's Dementia
by: T, Balamurali B, et al.
Published: (2024)
by: T, Balamurali B, et al.
Published: (2024)
ChatGPT as a Math Questioner? Evaluating ChatGPT on Generating Pre-university Math Questions
by: Van Long, Phuoc Pham, et al.
Published: (2023)
by: Van Long, Phuoc Pham, et al.
Published: (2023)
ChatGPT Alternative Solutions: Large Language Models Survey
by: Alipour, Hanieh, et al.
Published: (2024)
by: Alipour, Hanieh, et al.
Published: (2024)
Comparative Analysis of ChatGPT, GPT-4, and Microsoft Bing Chatbots for GRE Test
by: Abu-Haifa, Mohammad, et al.
Published: (2023)
by: Abu-Haifa, Mohammad, et al.
Published: (2023)
AuditGPT: Auditing Smart Contracts with ChatGPT
by: Xia, Shihao, et al.
Published: (2024)
by: Xia, Shihao, et al.
Published: (2024)
ChatGPT as speechwriter for the French presidents
by: Labbé, Dominique, et al.
Published: (2024)
by: Labbé, Dominique, et al.
Published: (2024)
The Human and the Mechanical: logos, truthfulness, and ChatGPT
by: Giannakidou, Anastasia, et al.
Published: (2024)
by: Giannakidou, Anastasia, et al.
Published: (2024)
Does ChatGPT Have a Mind?
by: Goldstein, Simon, et al.
Published: (2024)
by: Goldstein, Simon, et al.
Published: (2024)
Can ChatGPT Learn to Count Letters?
by: Conde, Javier, et al.
Published: (2025)
by: Conde, Javier, et al.
Published: (2025)
On Prompt Sensitivity of ChatGPT in Affective Computing
by: Amin, Mostafa M., et al.
Published: (2024)
by: Amin, Mostafa M., et al.
Published: (2024)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
by: Urchs, Stefanie, et al.
Published: (2023)
by: Urchs, Stefanie, et al.
Published: (2023)
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study
by: Wang, Zengzhi, et al.
Published: (2023)
by: Wang, Zengzhi, et al.
Published: (2023)
ChatGPT or A Silent Everywhere Helper: A Survey of Large Language Models
by: Akhtarshenas, Azim, et al.
Published: (2025)
by: Akhtarshenas, Azim, et al.
Published: (2025)
"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
by: Li, Lingyao, et al.
Published: (2023)
by: Li, Lingyao, et al.
Published: (2023)
Benchmarking ChatGPT on Algorithmic Reasoning
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
ChatLog: Carefully Evaluating the Evolution of ChatGPT Across Time
by: Tu, Shangqing, et al.
Published: (2023)
by: Tu, Shangqing, et al.
Published: (2023)
Demystifying ChatGPT: How It Masters Genre Recognition
by: Raj, Subham, et al.
Published: (2025)
by: Raj, Subham, et al.
Published: (2025)
What is the Best Way for ChatGPT to Translate Poetry?
by: Wang, Shanshan, et al.
Published: (2024)
by: Wang, Shanshan, et al.
Published: (2024)
Evaluating ChatGPT on Nuclear Domain-Specific Data
by: Anwar, Muhammad, et al.
Published: (2024)
by: Anwar, Muhammad, et al.
Published: (2024)
Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
by: Chen, Shan, et al.
Published: (2023)
by: Chen, Shan, et al.
Published: (2023)
Differentiate ChatGPT-generated and Human-written Medical Texts
by: Liao, Wenxiong, et al.
Published: (2023)
by: Liao, Wenxiong, et al.
Published: (2023)
RecGPT: Generative Personalized Prompts for Sequential Recommendation via ChatGPT Training Paradigm
by: Zhang, Yabin, et al.
Published: (2024)
by: Zhang, Yabin, et al.
Published: (2024)
Is ChatGPT More Empathetic than Humans?
by: Welivita, Anuradha, et al.
Published: (2024)
by: Welivita, Anuradha, et al.
Published: (2024)
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context
by: Li, Victoria R., et al.
Published: (2024)
by: Li, Victoria R., et al.
Published: (2024)
Can we trust the evaluation on ChatGPT?
by: Aiyappa, Rachith, et al.
Published: (2023)
by: Aiyappa, Rachith, et al.
Published: (2023)
Exploring ChatGPT's Capabilities on Vulnerability Management
by: Liu, Peiyu, et al.
Published: (2023)
by: Liu, Peiyu, et al.
Published: (2023)
Can ChatGPT Really Understand Modern Chinese Poetry?
by: Wang, Shanshan, et al.
Published: (2026)
by: Wang, Shanshan, et al.
Published: (2026)
AI and the Law: Evaluating ChatGPT's Performance in Legal Classification
by: Weichbroth, Pawel
Published: (2025)
by: Weichbroth, Pawel
Published: (2025)
Experimental evidence of progressive ChatGPT models self-convergence
by: Xylogiannopoulos, Konstantinos F., et al.
Published: (2026)
by: Xylogiannopoulos, Konstantinos F., et al.
Published: (2026)
RogueGPT: dis-ethical tuning transforms ChatGPT4 into a Rogue AI in 158 Words
by: Buscemi, Alessio, et al.
Published: (2024)
by: Buscemi, Alessio, et al.
Published: (2024)
Exploring ChatGPT and its Impact on Society
by: Haque, Md. Asraful, et al.
Published: (2024)
by: Haque, Md. Asraful, et al.
Published: (2024)
Evaluating the Performance of ChatGPT for Spam Email Detection
by: Si, Shijing, et al.
Published: (2024)
by: Si, Shijing, et al.
Published: (2024)
Evaluating ChatGPT-4 Vision on Brazil's National Undergraduate Computer Science Exam
by: Mendonça, Nabor C.
Published: (2024)
by: Mendonça, Nabor C.
Published: (2024)
Exploring the Capabilities of ChatGPT in Ancient Chinese Translation and Person Name Recognition
by: Si, Shijing, et al.
Published: (2023)
by: Si, Shijing, et al.
Published: (2023)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
by: Liu, Yi, et al.
Published: (2023)
by: Liu, Yi, et al.
Published: (2023)
Working Memory Capacity of ChatGPT: An Empirical Study
by: Gong, Dongyu, et al.
Published: (2023)
by: Gong, Dongyu, et al.
Published: (2023)
Similar Items
-
A Systematic Analysis of Biases in Large Language Models
by: Zhang, Xulang, et al.
Published: (2025) -
A Survey on the Real Power of ChatGPT
by: Liu, Ming, et al.
Published: (2024) -
Primacy Effect of ChatGPT
by: Wang, Yiwei, et al.
Published: (2023) -
Fairness of ChatGPT
by: Li, Yunqi, et al.
Published: (2023) -
Performance Assessment of ChatGPT vs Bard in Detecting Alzheimer's Dementia
by: T, Balamurali B, et al.
Published: (2024)