Automated Educational Question Generation at Different Bloom's Skill Levels using Large Language Models: Strategies and Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Scaria, Nicy, Chenna, Suma Dharani, Subramani, Deepak |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sensitivity of Small Language Models to Fine-tuning Data Contamination
por: Scaria, Nicy, et al.
Publicado: (2025)
por: Scaria, Nicy, et al.
Publicado: (2025)
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
por: Scaria, Nicy, et al.
Publicado: (2024)
por: Scaria, Nicy, et al.
Publicado: (2024)
Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning
por: Scaria, Nicy, et al.
Publicado: (2026)
por: Scaria, Nicy, et al.
Publicado: (2026)
Dissecting Physics Reasoning in Small Language Models: A Multi-Dimensional Analysis from an Educational Perspective
por: Scaria, Nicy, et al.
Publicado: (2025)
por: Scaria, Nicy, et al.
Publicado: (2025)
Harnessing Structured Knowledge: A Concept Map-Based Approach for High-Quality Multiple Choice Question Generation with Effective Distractors
por: Scaria, Nicy, et al.
Publicado: (2025)
por: Scaria, Nicy, et al.
Publicado: (2025)
Can Small Language Models Learn, Unlearn, and Retain Noise Patterns?
por: Scaria, Nicy, et al.
Publicado: (2024)
por: Scaria, Nicy, et al.
Publicado: (2024)
Applications of Large Language Model Reasoning in Feature Generation
por: Chandra, Dharani
Publicado: (2025)
por: Chandra, Dharani
Publicado: (2025)
Automated Analysis of Learning Outcomes and Exam Questions Based on Bloom's Taxonomy
por: Kumar, Ramya, et al.
Publicado: (2025)
por: Kumar, Ramya, et al.
Publicado: (2025)
Evaluating the Meta- and Object-Level Reasoning of Large Language Models for Question Answering
por: Ferguson, Nick, et al.
Publicado: (2025)
por: Ferguson, Nick, et al.
Publicado: (2025)
How Teachers Can Use Large Language Models and Bloom's Taxonomy to Create Educational Quizzes
por: Elkins, Sabina, et al.
Publicado: (2024)
por: Elkins, Sabina, et al.
Publicado: (2024)
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology
por: Tonga, Junior Cedric, et al.
Publicado: (2024)
por: Tonga, Junior Cedric, et al.
Publicado: (2024)
How Effective is GPT-4 Turbo in Generating School-Level Questions from Textbooks Based on Bloom's Revised Taxonomy?
por: Maity, Subhankar, et al.
Publicado: (2024)
por: Maity, Subhankar, et al.
Publicado: (2024)
The Future of Learning in the Age of Generative AI: Automated Question Generation and Assessment with Large Language Models
por: Maity, Subhankar, et al.
Publicado: (2024)
por: Maity, Subhankar, et al.
Publicado: (2024)
Automated Benchmark Generation from Domain Guidelines Informed by Bloom's Taxonomy
por: Chen, Si, et al.
Publicado: (2026)
por: Chen, Si, et al.
Publicado: (2026)
Multiple-Choice Question Generation Using Large Language Models: Methodology and Educator Insights
por: Biancini, Giorgio, et al.
Publicado: (2025)
por: Biancini, Giorgio, et al.
Publicado: (2025)
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
por: Alawwad, Hessa A., et al.
Publicado: (2025)
por: Alawwad, Hessa A., et al.
Publicado: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
por: Kubesch, Jonas, et al.
Publicado: (2026)
por: Kubesch, Jonas, et al.
Publicado: (2026)
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models
por: Chen, Yuyan, et al.
Publicado: (2024)
por: Chen, Yuyan, et al.
Publicado: (2024)
From Answers to Questions: EQGBench for Evaluating LLMs' Educational Question Generation
por: Zhou, Chengliang, et al.
Publicado: (2025)
por: Zhou, Chengliang, et al.
Publicado: (2025)
An Empirical Evaluation of Large Language Models on Consumer Health Questions
por: Abrar, Moaiz, et al.
Publicado: (2024)
por: Abrar, Moaiz, et al.
Publicado: (2024)
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
por: Song, Seok Hwan, et al.
Publicado: (2025)
por: Song, Seok Hwan, et al.
Publicado: (2025)
Evaluation Methodology for Large Language Models for Multilingual Document Question and Answer
por: Kahana, Adar, et al.
Publicado: (2024)
por: Kahana, Adar, et al.
Publicado: (2024)
Beyond Questions: Evaluating What Large Language Models (Actually) Know
por: Giordano, Luca, et al.
Publicado: (2026)
por: Giordano, Luca, et al.
Publicado: (2026)
Automatic Question & Answer Generation Using Generative Large Language Model (LLM)
por: Ehsan, Md. Alvee, et al.
Publicado: (2025)
por: Ehsan, Md. Alvee, et al.
Publicado: (2025)
BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting
por: Wang, Zhensheng, et al.
Publicado: (2026)
por: Wang, Zhensheng, et al.
Publicado: (2026)
ChartInsights: Evaluating Multimodal Large Language Models for Low-Level Chart Question Answering
por: Wu, Yifan, et al.
Publicado: (2024)
por: Wu, Yifan, et al.
Publicado: (2024)
TS-Skill: A Benchmark for Evaluating Analytical Skills in Time-Series Question Answering
por: Han, Liying, et al.
Publicado: (2026)
por: Han, Liying, et al.
Publicado: (2026)
Follow-Up Questions Improve Documents Generated by Large Language Models
por: Tix, Bernadette J
Publicado: (2024)
por: Tix, Bernadette J
Publicado: (2024)
Contrastive Learning for Knowledge-Based Question Generation in Large Language Models
por: Zhang, Zhenhong, et al.
Publicado: (2024)
por: Zhang, Zhenhong, et al.
Publicado: (2024)
Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets
por: Xenofontos, Andreas, et al.
Publicado: (2026)
por: Xenofontos, Andreas, et al.
Publicado: (2026)
MIRROR: A Novel Approach for the Automated Evaluation of Open-Ended Question Generation
por: Deroy, Aniket, et al.
Publicado: (2024)
por: Deroy, Aniket, et al.
Publicado: (2024)
Large Language Models for Automated Literature Review: An Evaluation of Reference Generation, Abstract Writing, and Review Composition
por: Tang, Xuemei, et al.
Publicado: (2024)
por: Tang, Xuemei, et al.
Publicado: (2024)
A Question on the Explainability of Large Language Models and the Word-Level Univariate First-Order Plausibility Assumption
por: Bogaert, Jeremie, et al.
Publicado: (2024)
por: Bogaert, Jeremie, et al.
Publicado: (2024)
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction
por: Bailis, Suma, et al.
Publicado: (2024)
por: Bailis, Suma, et al.
Publicado: (2024)
A Comprehensive Evaluation of Quantization Strategies for Large Language Models
por: Jin, Renren, et al.
Publicado: (2024)
por: Jin, Renren, et al.
Publicado: (2024)
Dynamic Strategy Planning for Efficient Question Answering with Large Language Models
por: Parekh, Tanmay, et al.
Publicado: (2024)
por: Parekh, Tanmay, et al.
Publicado: (2024)
Prompting Strategies for Language Model-Based Item Generation in K-12 Education: Bridging the Gap Between Small and Large Language Models
por: Amini, Mohammad, et al.
Publicado: (2025)
por: Amini, Mohammad, et al.
Publicado: (2025)
Ask Good Questions for Large Language Models
por: Wu, Qi, et al.
Publicado: (2025)
por: Wu, Qi, et al.
Publicado: (2025)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
por: He-Yueya, Joy, et al.
Publicado: (2024)
por: He-Yueya, Joy, et al.
Publicado: (2024)
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education
por: Henkel, Owen, et al.
Publicado: (2024)
por: Henkel, Owen, et al.
Publicado: (2024)
Ejemplares similares
-
Sensitivity of Small Language Models to Fine-tuning Data Contamination
por: Scaria, Nicy, et al.
Publicado: (2025) -
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
por: Scaria, Nicy, et al.
Publicado: (2024) -
Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning
por: Scaria, Nicy, et al.
Publicado: (2026) -
Dissecting Physics Reasoning in Small Language Models: A Multi-Dimensional Analysis from an Educational Perspective
por: Scaria, Nicy, et al.
Publicado: (2025) -
Harnessing Structured Knowledge: A Concept Map-Based Approach for High-Quality Multiple Choice Question Generation with Effective Distractors
por: Scaria, Nicy, et al.
Publicado: (2025)