From NLG Evaluation to Modern Student Assessment in the Era of ChatGPT: The Great Misalignment Problem and Pedagogical Multi-Factor Assessment (P-MFA)
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hämäläinen, Mika, Leiviskä, Kimmo |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
On Psychology of AI -- Does Primacy Effect Affect ChatGPT and Other LLMs?
par: Hämäläinen, Mika
Publié: (2025)
par: Hämäläinen, Mika
Publié: (2025)
DAG: Dictionary-Augmented Generation for Disambiguation of Sentences in Endangered Uralic Languages using ChatGPT
par: Hämäläinen, Mika
Publié: (2024)
par: Hämäläinen, Mika
Publié: (2024)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
par: Mao, Rui, et autres
Publié: (2023)
par: Mao, Rui, et autres
Publié: (2023)
Cross-Language Assessment of Mathematical Capability of ChatGPT
par: Sathe, Gargi, et autres
Publié: (2024)
par: Sathe, Gargi, et autres
Publié: (2024)
Evaluating OpenAI GPT Models for Translation of Endangered Uralic Languages: A Comparison of Reasoning and Non-Reasoning Architectures
par: Tereshchenko, Yehor, et autres
Publié: (2025)
par: Tereshchenko, Yehor, et autres
Publié: (2025)
Gradable ChatGPT Translation Evaluation
par: Jiao, Hui, et autres
Publié: (2024)
par: Jiao, Hui, et autres
Publié: (2024)
Analyzing Pokémon and Mario Streamers' Twitch Chat with LLM-based User Embeddings
par: Hämäläinen, Mika, et autres
Publié: (2024)
par: Hämäläinen, Mika, et autres
Publié: (2024)
Performance Assessment of ChatGPT vs Bard in Detecting Alzheimer's Dementia
par: T, Balamurali B, et autres
Publié: (2024)
par: T, Balamurali B, et autres
Publié: (2024)
Efficient Toxicity Detection in Gaming Chats: A Comparative Study of Embeddings, Fine-Tuned Transformers and LLMs
par: Tereshchenko, Yehor, et autres
Publié: (2025)
par: Tereshchenko, Yehor, et autres
Publié: (2025)
A Comprehensive Survey of Sentence Representations: From the BERT Epoch to the ChatGPT Era and Beyond
par: Kashyap, Abhinav Ramesh, et autres
Publié: (2023)
par: Kashyap, Abhinav Ramesh, et autres
Publié: (2023)
Can ChatGPT Really Understand Modern Chinese Poetry?
par: Wang, Shanshan, et autres
Publié: (2026)
par: Wang, Shanshan, et autres
Publié: (2026)
ChatGPT as a Math Questioner? Evaluating ChatGPT on Generating Pre-university Math Questions
par: Van Long, Phuoc Pham, et autres
Publié: (2023)
par: Van Long, Phuoc Pham, et autres
Publié: (2023)
Is ChatGPT Involved in Texts? Measure the Polish Ratio to Detect ChatGPT-Generated Text
par: Yang, Lingyi, et autres
Publié: (2023)
par: Yang, Lingyi, et autres
Publié: (2023)
ChEDDAR: Student-ChatGPT Dialogue in EFL Writing Education
par: Han, Jieun, et autres
Publié: (2023)
par: Han, Jieun, et autres
Publié: (2023)
Fairness of ChatGPT
par: Li, Yunqi, et autres
Publié: (2023)
par: Li, Yunqi, et autres
Publié: (2023)
From Instruction to Output: The Role of Prompting in Modern NLG
par: Zaib, Munazza, et autres
Publié: (2026)
par: Zaib, Munazza, et autres
Publié: (2026)
ChatLog: Carefully Evaluating the Evolution of ChatGPT Across Time
par: Tu, Shangqing, et autres
Publié: (2023)
par: Tu, Shangqing, et autres
Publié: (2023)
A Comparative Analysis of Ethical and Safety Gaps in LLMs using Relative Danger Coefficient
par: Tereshchenko, Yehor, et autres
Publié: (2025)
par: Tereshchenko, Yehor, et autres
Publié: (2025)
Evaluating ChatGPT on Nuclear Domain-Specific Data
par: Anwar, Muhammad, et autres
Publié: (2024)
par: Anwar, Muhammad, et autres
Publié: (2024)
Primacy Effect of ChatGPT
par: Wang, Yiwei, et autres
Publié: (2023)
par: Wang, Yiwei, et autres
Publié: (2023)
Mapping the Challenges of HCI: An Application and Evaluation of ChatGPT for Mining Insights at Scale
par: Oppenlaender, Jonas, et autres
Publié: (2023)
par: Oppenlaender, Jonas, et autres
Publié: (2023)
Evaluating the Performance of ChatGPT for Spam Email Detection
par: Si, Shijing, et autres
Publié: (2024)
par: Si, Shijing, et autres
Publié: (2024)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
par: Urchs, Stefanie, et autres
Publié: (2023)
par: Urchs, Stefanie, et autres
Publié: (2023)
Automated Coding of Communications in Collaborative Problem-solving Tasks Using ChatGPT
par: Hao, Jiangang, et autres
Publié: (2024)
par: Hao, Jiangang, et autres
Publié: (2024)
AI and the Law: Evaluating ChatGPT's Performance in Legal Classification
par: Weichbroth, Pawel
Publié: (2025)
par: Weichbroth, Pawel
Publié: (2025)
ChatGPT as speechwriter for the French presidents
par: Labbé, Dominique, et autres
Publié: (2024)
par: Labbé, Dominique, et autres
Publié: (2024)
Emergence of a phonological bias in ChatGPT
par: Toro, Juan Manuel
Publié: (2023)
par: Toro, Juan Manuel
Publié: (2023)
Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
par: Chen, Shan, et autres
Publié: (2023)
par: Chen, Shan, et autres
Publié: (2023)
Evaluating Privacy Questions From Stack Overflow: Can ChatGPT Compete?
par: Delile, Zack, et autres
Publié: (2023)
par: Delile, Zack, et autres
Publié: (2023)
ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT
par: Wei, Xiang, et autres
Publié: (2023)
par: Wei, Xiang, et autres
Publié: (2023)
RECIPE4U: Student-ChatGPT Interaction Dataset in EFL Writing Education
par: Han, Jieun, et autres
Publié: (2024)
par: Han, Jieun, et autres
Publié: (2024)
Convergences and Divergences between Automatic Assessment and Human Evaluation: Insights from Comparing ChatGPT-Generated Translation and Neural Machine Translation
par: Jiang, Zhaokun, et autres
Publié: (2024)
par: Jiang, Zhaokun, et autres
Publié: (2024)
WildChat: 1M ChatGPT Interaction Logs in the Wild
par: Zhao, Wenting, et autres
Publié: (2024)
par: Zhao, Wenting, et autres
Publié: (2024)
"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
par: Li, Lingyao, et autres
Publié: (2023)
par: Li, Lingyao, et autres
Publié: (2023)
Is ChatGPT the Future of Causal Text Mining? A Comprehensive Evaluation and Analysis
par: Takayanagi, Takehiro, et autres
Publié: (2024)
par: Takayanagi, Takehiro, et autres
Publié: (2024)
Evaluating ChatGPT on Medical Information Extraction Tasks: Performance, Explainability and Beyond
par: Li, Liz, et autres
Publié: (2026)
par: Li, Liz, et autres
Publié: (2026)
Using ChatGPT for Data Science Analyses
par: Evkaya, Ozan, et autres
Publié: (2024)
par: Evkaya, Ozan, et autres
Publié: (2024)
AuditGPT: Auditing Smart Contracts with ChatGPT
par: Xia, Shihao, et autres
Publié: (2024)
par: Xia, Shihao, et autres
Publié: (2024)
Can ChatGPT Read Who You Are?
par: Derner, Erik, et autres
Publié: (2023)
par: Derner, Erik, et autres
Publié: (2023)
Enhancing SLM via ChatGPT and Dataset Augmentation
par: Pieper, Tom, et autres
Publié: (2024)
par: Pieper, Tom, et autres
Publié: (2024)
Documents similaires
-
On Psychology of AI -- Does Primacy Effect Affect ChatGPT and Other LLMs?
par: Hämäläinen, Mika
Publié: (2025) -
DAG: Dictionary-Augmented Generation for Disambiguation of Sentences in Endangered Uralic Languages using ChatGPT
par: Hämäläinen, Mika
Publié: (2024) -
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
par: Mao, Rui, et autres
Publié: (2023) -
Cross-Language Assessment of Mathematical Capability of ChatGPT
par: Sathe, Gargi, et autres
Publié: (2024) -
Evaluating OpenAI GPT Models for Translation of Endangered Uralic Languages: A Comparison of Reasoning and Non-Reasoning Architectures
par: Tereshchenko, Yehor, et autres
Publié: (2025)