Evaluating Large Language Models on the GMAT: Implications for the Future of Business Education
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ashrafimoghari, Vahid, Gürkan, Necdet, Suchow, Jordan W. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
von: Gurkan, Necdet, et al.
Veröffentlicht: (2025)
von: Gurkan, Necdet, et al.
Veröffentlicht: (2025)
Exploring Public Opinion on Responsible AI Through The Lens of Cultural Consensus Theory
von: Gurkan, Necdet, et al.
Veröffentlicht: (2024)
von: Gurkan, Necdet, et al.
Veröffentlicht: (2024)
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
von: Sadanandan, Binesh, et al.
Veröffentlicht: (2026)
von: Sadanandan, Binesh, et al.
Veröffentlicht: (2026)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
An Exploration of Higher Education Course Evaluation by Large Language Models
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
von: Yuan, Bo, et al.
Veröffentlicht: (2024)
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
von: Hu, Yiran, et al.
Veröffentlicht: (2026)
von: Hu, Yiran, et al.
Veröffentlicht: (2026)
Evaluating Neural Language Models as Cognitive Models of Language Acquisition
von: Martínez, Héctor Javier Vázquez, et al.
Veröffentlicht: (2023)
von: Martínez, Héctor Javier Vázquez, et al.
Veröffentlicht: (2023)
Benchmark for Assessing Olfactory Perception of Large Language Models
von: Makri, Eftychia, et al.
Veröffentlicht: (2026)
von: Makri, Eftychia, et al.
Veröffentlicht: (2026)
IndicEval: A Bilingual Indian Educational Evaluation Framework for Large Language Models
von: Bharti, Saurabh, et al.
Veröffentlicht: (2026)
von: Bharti, Saurabh, et al.
Veröffentlicht: (2026)
Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications
von: Li, Anran, et al.
Veröffentlicht: (2025)
von: Li, Anran, et al.
Veröffentlicht: (2025)
Towards a Benchmark for Large Language Models for Business Process Management Tasks
von: Busch, Kiran, et al.
Veröffentlicht: (2024)
von: Busch, Kiran, et al.
Veröffentlicht: (2024)
On Adversarial Robustness and Out-of-Distribution Robustness of Large Language Models
von: Yang, April, et al.
Veröffentlicht: (2024)
von: Yang, April, et al.
Veröffentlicht: (2024)
EduEval: A Hierarchical Cognitive Benchmark for Evaluating Large Language Models in Chinese Education
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
Harnessing Business and Media Insights with Large Language Models
von: Bao, Yujia, et al.
Veröffentlicht: (2024)
von: Bao, Yujia, et al.
Veröffentlicht: (2024)
Evaluating Quantized Large Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
TextMineX: Data, Evaluation Framework and Ontology-guided LLM Pipeline for Humanitarian Mine Action
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Evaluating Small Language Models for News Summarization: Implications and Factors Influencing Performance
von: Xu, Borui, et al.
Veröffentlicht: (2025)
von: Xu, Borui, et al.
Veröffentlicht: (2025)
Large Language Models for Causal Discovery: Current Landscape and Future Directions
von: Wan, Guangya, et al.
Veröffentlicht: (2024)
von: Wan, Guangya, et al.
Veröffentlicht: (2024)
Large Language Models for Education: A Survey and Outlook
von: Wang, Shen, et al.
Veröffentlicht: (2024)
von: Wang, Shen, et al.
Veröffentlicht: (2024)
Multilingual Performance Biases of Large Language Models in Education
von: Gupta, Vansh, et al.
Veröffentlicht: (2025)
von: Gupta, Vansh, et al.
Veröffentlicht: (2025)
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs
von: Colakoglu, Gaye, et al.
Veröffentlicht: (2025)
von: Colakoglu, Gaye, et al.
Veröffentlicht: (2025)
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
Large Language Models Vote: Prompting for Rare Disease Identification
von: Oniani, David, et al.
Veröffentlicht: (2023)
von: Oniani, David, et al.
Veröffentlicht: (2023)
Automated Educational Question Generation at Different Bloom's Skill Levels using Large Language Models: Strategies and Evaluation
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
Realistic Evaluation of Toxicity in Large Language Models
von: Luong, Tinh Son, et al.
Veröffentlicht: (2024)
von: Luong, Tinh Son, et al.
Veröffentlicht: (2024)
Evaluating the Retrieval Robustness of Large Language Models
von: Cao, Shuyang, et al.
Veröffentlicht: (2025)
von: Cao, Shuyang, et al.
Veröffentlicht: (2025)
Evaluating Spatial Understanding of Large Language Models
von: Yamada, Yutaro, et al.
Veröffentlicht: (2023)
von: Yamada, Yutaro, et al.
Veröffentlicht: (2023)
Role-Playing Evaluation for Large Language Models
von: Boudouri, Yassine El, et al.
Veröffentlicht: (2025)
von: Boudouri, Yassine El, et al.
Veröffentlicht: (2025)
A Survey on Evaluation of Large Language Models
von: Chang, Yupeng, et al.
Veröffentlicht: (2023)
von: Chang, Yupeng, et al.
Veröffentlicht: (2023)
Large Language Models for Education: A Survey
von: Xu, Hanyi, et al.
Veröffentlicht: (2024)
von: Xu, Hanyi, et al.
Veröffentlicht: (2024)
Right to be Forgotten in the Era of Large Language Models: Implications, Challenges, and Solutions
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
Leveraging Large Language Model as Simulated Patients for Clinical Education
von: Li, Yanzeng, et al.
Veröffentlicht: (2024)
von: Li, Yanzeng, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for Abstract Evaluation Tasks: An Empirical Study
von: Liu, Yinuo, et al.
Veröffentlicht: (2026)
von: Liu, Yinuo, et al.
Veröffentlicht: (2026)
Evaluating Consistency and Reasoning Capabilities of Large Language Models
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
Steering Large Language Models to Evaluate and Amplify Creativity
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2024)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2024)
CriticEval: Evaluating Large Language Model as Critic
von: Lan, Tian, et al.
Veröffentlicht: (2024)
von: Lan, Tian, et al.
Veröffentlicht: (2024)
Evaluating Morphological Compositional Generalization in Large Language Models
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2024)
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2024)
An Empirical Analysis on Large Language Models in Debate Evaluation
von: Liu, Xinyi, et al.
Veröffentlicht: (2024)
von: Liu, Xinyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
von: Gurkan, Necdet, et al.
Veröffentlicht: (2025) -
Exploring Public Opinion on Responsible AI Through The Lens of Cultural Consensus Theory
von: Gurkan, Necdet, et al.
Veröffentlicht: (2024) -
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
von: Sadanandan, Binesh, et al.
Veröffentlicht: (2026) -
Evaluating and Optimizing Educational Content with Large Language Model Judgments
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024) -
Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions
von: Zhang, Kun, et al.
Veröffentlicht: (2025)