Question Type, Cognitive Load, and CEFR Alignment: Evaluating LLM-Generated EFL Grammar Drill Exercises
Fuente:
arXiv
Guardado en:
| Autores principales: | Woollaston, Steve, Flanagan, Brendan, Toyokawa, Yuko, Ogata, Hiroaki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Training-Free Private Synthesis with Validation: A New Frontier for Practical Educational Data Sharing
por: Ito, Hibiki, et al.
Publicado: (2026)
por: Ito, Hibiki, et al.
Publicado: (2026)
Cyclic Adaptive Private Synthesis for Sharing Real-World Data in Education
por: Ito, Hibiki, et al.
Publicado: (2026)
por: Ito, Hibiki, et al.
Publicado: (2026)
The Third-Party Access Effect: An Overlooked Challenge in Secondary Use of Educational Real-World Data
por: Ito, Hibiki, et al.
Publicado: (2026)
por: Ito, Hibiki, et al.
Publicado: (2026)
Defining the Scope of Learning Analytics: An Axiomatic Approach for Analytic Practice and Measurable Learning Phenomena
por: Takii, Kensuke, et al.
Publicado: (2025)
por: Takii, Kensuke, et al.
Publicado: (2025)
An Evaluation of Cultural Value Alignment in LLM
por: Sukiennik, Nicholas, et al.
Publicado: (2025)
por: Sukiennik, Nicholas, et al.
Publicado: (2025)
Evaluating Intra-firm LLM Alignment Strategies in Business Contexts
por: Broestl, Noah, et al.
Publicado: (2025)
por: Broestl, Noah, et al.
Publicado: (2025)
The Ghost in the Grammar: Methodological Anthropomorphism in AI Safety Evaluations
por: Costa, Mariana Lins
Publicado: (2026)
por: Costa, Mariana Lins
Publicado: (2026)
Open Datasets in Learning Analytics: Trends, Challenges, and Best PRACTICE
por: Švábenský, Valdemar, et al.
Publicado: (2026)
por: Švábenský, Valdemar, et al.
Publicado: (2026)
Evaluating Contextually Personalized Programming Exercises Created with Generative AI
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
Sacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
por: Atif, Farah, et al.
Publicado: (2025)
por: Atif, Farah, et al.
Publicado: (2025)
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments?
por: Ocheja, Patrick, et al.
Publicado: (2024)
por: Ocheja, Patrick, et al.
Publicado: (2024)
Simulation as Reality? The Effectiveness of LLM-Generated Data in Open-ended Question Assessment
por: Zhang, Long, et al.
Publicado: (2025)
por: Zhang, Long, et al.
Publicado: (2025)
Hedging and Non-Affirmation: Quantifying LLM Alignment on Questions of Human Rights
por: Javed, Rafiya, et al.
Publicado: (2025)
por: Javed, Rafiya, et al.
Publicado: (2025)
Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ) Generation and Evaluation
por: An, Yuan
Publicado: (2026)
por: An, Yuan
Publicado: (2026)
Exploring Safety Alignment Evaluation of LLMs in Chinese Mental Health Dialogues via LLM-as-Judge
por: Cai, Yunna, et al.
Publicado: (2025)
por: Cai, Yunna, et al.
Publicado: (2025)
Probabilistic Analysis of Copyright Disputes and Generative AI Safety
por: Chiba-Okabe, Hiroaki
Publicado: (2024)
por: Chiba-Okabe, Hiroaki
Publicado: (2024)
Monitoring Human Dependence On AI Systems With Reliance Drills
por: Hunter, Rosco, et al.
Publicado: (2024)
por: Hunter, Rosco, et al.
Publicado: (2024)
Societal Alignment Frameworks Can Improve LLM Alignment
por: Stańczak, Karolina, et al.
Publicado: (2025)
por: Stańczak, Karolina, et al.
Publicado: (2025)
Wearable Device-Based Real-Time Monitoring of Physiological Signals: Evaluating Cognitive Load Across Different Tasks
por: He, Ling, et al.
Publicado: (2024)
por: He, Ling, et al.
Publicado: (2024)
Alignment Drift in CEFR-prompted LLMs for Interactive Spanish Tutoring
por: Almasi, Mina, et al.
Publicado: (2025)
por: Almasi, Mina, et al.
Publicado: (2025)
Evaluation of Systems Programming Exercises through Tailored Static Analysis
por: Natella, Roberto
Publicado: (2024)
por: Natella, Roberto
Publicado: (2024)
The Value of Disagreement in AI Design, Evaluation, and Alignment
por: Fazelpour, Sina, et al.
Publicado: (2025)
por: Fazelpour, Sina, et al.
Publicado: (2025)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
por: Hu, Yueqing, et al.
Publicado: (2026)
por: Hu, Yueqing, et al.
Publicado: (2026)
LLM-Driven Personalized Answer Generation and Evaluation
por: Molavi, Mohammadreza, et al.
Publicado: (2025)
por: Molavi, Mohammadreza, et al.
Publicado: (2025)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
por: Shankar, Hari, et al.
Publicado: (2026)
por: Shankar, Hari, et al.
Publicado: (2026)
How Jungian Cognitive Functions Explain MBTI Type Prevalence in Computer Industry Careers
por: VarastehNezhad, Arya, et al.
Publicado: (2025)
por: VarastehNezhad, Arya, et al.
Publicado: (2025)
Randomness, Not Representation: The Unreliability of Evaluating Cultural Alignment in LLMs
por: Khan, Ariba, et al.
Publicado: (2025)
por: Khan, Ariba, et al.
Publicado: (2025)
Exploring Prosocial Irrationality for LLM Agents: A Social Cognition View
por: Liu, Xuan, et al.
Publicado: (2024)
por: Liu, Xuan, et al.
Publicado: (2024)
The Urban Toolkit: A Grammar-based Framework for Urban Visual Analytics
por: Moreira, Gustavo, et al.
Publicado: (2023)
por: Moreira, Gustavo, et al.
Publicado: (2023)
Unintended Impacts of LLM Alignment on Global Representation
por: Ryan, Michael J., et al.
Publicado: (2024)
por: Ryan, Michael J., et al.
Publicado: (2024)
SPA: Achieving Consensus in LLM Alignment via Self-Priority Optimization
por: Huang, Yue, et al.
Publicado: (2025)
por: Huang, Yue, et al.
Publicado: (2025)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
por: Lee, Jaehyeok, et al.
Publicado: (2026)
por: Lee, Jaehyeok, et al.
Publicado: (2026)
Dean of LLM Tutors: Exploring Comprehensive and Automated Evaluation of LLM-generated Educational Feedback via LLM Feedback Evaluators
por: Qian, Keyang, et al.
Publicado: (2025)
por: Qian, Keyang, et al.
Publicado: (2025)
On the Credibility of Evaluating LLMs using Survey Questions
por: Libovický, Jindřich
Publicado: (2026)
por: Libovický, Jindřich
Publicado: (2026)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
por: Dumitran, Adrian-Marius, et al.
Publicado: (2025)
por: Dumitran, Adrian-Marius, et al.
Publicado: (2025)
Beyond Static Question Banks: Dynamic Knowledge Expansion via LLM-Automated Graph Construction and Adaptive Generation
por: Wang, Yingquan, et al.
Publicado: (2026)
por: Wang, Yingquan, et al.
Publicado: (2026)
How to Drill Into Silos: Creating a Free-to-Use Dataset of Data Subject Access Packages
por: Leschke, Nicola, et al.
Publicado: (2024)
por: Leschke, Nicola, et al.
Publicado: (2024)
PICA: A Data-driven Synthesis of Peer Instruction and Continuous Assessment
por: Geinitz, Steve
Publicado: (2024)
por: Geinitz, Steve
Publicado: (2024)
Moral Alignment for LLM Agents
por: Tennant, Elizaveta, et al.
Publicado: (2024)
por: Tennant, Elizaveta, et al.
Publicado: (2024)
StreetWeave: A Declarative Grammar for Street-Overlaid Visualization of Multivariate Data
por: Srabanti, Sanjana, et al.
Publicado: (2025)
por: Srabanti, Sanjana, et al.
Publicado: (2025)
Ejemplares similares
-
Training-Free Private Synthesis with Validation: A New Frontier for Practical Educational Data Sharing
por: Ito, Hibiki, et al.
Publicado: (2026) -
Cyclic Adaptive Private Synthesis for Sharing Real-World Data in Education
por: Ito, Hibiki, et al.
Publicado: (2026) -
The Third-Party Access Effect: An Overlooked Challenge in Secondary Use of Educational Real-World Data
por: Ito, Hibiki, et al.
Publicado: (2026) -
Defining the Scope of Learning Analytics: An Axiomatic Approach for Analytic Practice and Measurable Learning Phenomena
por: Takii, Kensuke, et al.
Publicado: (2025) -
An Evaluation of Cultural Value Alignment in LLM
por: Sukiennik, Nicholas, et al.
Publicado: (2025)