Instructional Goal-Aligned Question Generation for Student Evaluation in Virtual Lab Settings: How Closely Do LLMs Actually Align?
Fuente:
arXiv
Saved in:
| Main Authors: | Knipper, R. Alexander, Dey, Indrani, Sarkar, Souvika, Narayanan, Hari, Puntambekar, Sadhana, Karmaker, Santu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SimPal: Towards a Meta-Conversational Framework to Understand Teacher's Instructional Goals for K-12 Physics
by: Farhana, Effat, et al.
Published: (2024)
by: Farhana, Effat, et al.
Published: (2024)
Middle School Students' Application of Science Learning From Physical Versus Virtual Labs to New Contexts
by: Dana Gnesdilow, et al.
Published: (2025)
by: Dana Gnesdilow, et al.
Published: (2025)
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders
by: Sarkar, Souvika, et al.
Published: (2025)
by: Sarkar, Souvika, et al.
Published: (2025)
The Bias is in the Details: An Assessment of Cognitive Bias in LLMs
by: Knipper, R. Alexander, et al.
Published: (2025)
by: Knipper, R. Alexander, et al.
Published: (2025)
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform?
by: Sarkar, Souvika, et al.
Published: (2023)
by: Sarkar, Souvika, et al.
Published: (2023)
SNNLP: Energy-Efficient Natural Language Processing Using Spiking Neural Networks
by: Knipper, R. Alexander, et al.
Published: (2024)
by: Knipper, R. Alexander, et al.
Published: (2024)
LLMs as On-demand Customizable Service
by: Sarkar, Souvika, et al.
Published: (2024)
by: Sarkar, Souvika, et al.
Published: (2024)
Pitfalls of Evaluating Language Models with Open Benchmarks
by: Hasan, Md. Najib, et al.
Published: (2025)
by: Hasan, Md. Najib, et al.
Published: (2025)
Benchmarking LLMs on the Semantic Overlap Summarization Task
by: Salvador, John, et al.
Published: (2024)
by: Salvador, John, et al.
Published: (2024)
Redundancy Aware Multi-Reference Based Gainwise Evaluation of Extractive Summarization
by: Akter, Mousumi, et al.
Published: (2023)
by: Akter, Mousumi, et al.
Published: (2023)
LLMs as Meta-Reviewers' Assistants: A Case Study
by: Hossain, Eftekhar, et al.
Published: (2024)
by: Hossain, Eftekhar, et al.
Published: (2024)
Set-Theoretic Compositionality of Sentence Embeddings
by: Bansal, Naman, et al.
Published: (2025)
by: Bansal, Naman, et al.
Published: (2025)
How Well Can You Articulate that Idea? Insights from Automated Formative Assessment
by: Karizaki, Mahsa Sheikhi, et al.
Published: (2024)
by: Karizaki, Mahsa Sheikhi, et al.
Published: (2024)
A Survey on Evaluation Metrics for Music Generation
by: Kader, Faria Binte, et al.
Published: (2025)
by: Kader, Faria Binte, et al.
Published: (2025)
FaNS: a Facet-based Narrative Similarity Metric
by: Akter, Mousumi, et al.
Published: (2023)
by: Akter, Mousumi, et al.
Published: (2023)
Are LLMs Ready to Replace Bangla Annotators?
by: Hasan, Md. Najib, et al.
Published: (2026)
by: Hasan, Md. Najib, et al.
Published: (2026)
FinTradeBench: A Financial Reasoning Benchmark for LLMs
by: Agrawal, Yogesh, et al.
Published: (2026)
by: Agrawal, Yogesh, et al.
Published: (2026)
Instruction-tuning Aligns LLMs to the Human Brain
by: Aw, Khai Loong, et al.
Published: (2023)
by: Aw, Khai Loong, et al.
Published: (2023)
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
by: Hussonnois, Maxence, et al.
Published: (2025)
by: Hussonnois, Maxence, et al.
Published: (2025)
Revisiting Word Embeddings in the LLM Era
by: Mahajan, Yash, et al.
Published: (2025)
by: Mahajan, Yash, et al.
Published: (2025)
The Critical Role of Aspects in Measuring Document Similarity
by: Hossain, Eftekhar, et al.
Published: (2026)
by: Hossain, Eftekhar, et al.
Published: (2026)
SMART: Simulated Students Aligned with Item Response Theory for Question Difficulty Prediction
by: Scarlatos, Alexander, et al.
Published: (2025)
by: Scarlatos, Alexander, et al.
Published: (2025)
Improving Multilingual Language Models by Aligning Representations through Steering
by: Mahmoud, Omar, et al.
Published: (2025)
by: Mahmoud, Omar, et al.
Published: (2025)
Multi-Party Conversational Agents: A Survey
by: Sapkota, Sagar, et al.
Published: (2025)
by: Sapkota, Sagar, et al.
Published: (2025)
Large Language Models for IT Automation Tasks: Are We There Yet?
by: Hassan, Md Mahadi, et al.
Published: (2025)
by: Hassan, Md Mahadi, et al.
Published: (2025)
Teaching an Old LLM Secure Coding: Localized Preference Optimization on Distilled Preferences
by: Hasan, Mohammad Saqib, et al.
Published: (2025)
by: Hasan, Mohammad Saqib, et al.
Published: (2025)
Retrieval Challenges in Low-Resource Public Service Information: A Case Study on Food Pantry Access
by: Hasan, Touseef, et al.
Published: (2026)
by: Hasan, Touseef, et al.
Published: (2026)
Open-Set Object Detection By Aligning Known Class Representations
by: Sarkar, Hiran, et al.
Published: (2024)
by: Sarkar, Hiran, et al.
Published: (2024)
One Instruction Does Not Fit All: How Well Do Embeddings Align Personas and Instructions in Low-Resource Indian Languages?
by: Shah, Arya, et al.
Published: (2026)
by: Shah, Arya, et al.
Published: (2026)
Aligning the Energy Transition with the Sustainable Development Goals
Published: (2024)
Published: (2024)
Spectrally Distilled Representations Aligned with Instruction-Augmented LLMs for Satellite Imagery
by: Do, Minh Kha, et al.
Published: (2026)
by: Do, Minh Kha, et al.
Published: (2026)
Divide-Verify-Refine: Can LLMs Self-Align with Complex Instructions?
by: Zhang, Xianren, et al.
Published: (2024)
by: Zhang, Xianren, et al.
Published: (2024)
The Path Not Taken: Duality in Reasoning about Program Execution
by: Hasanov, Eshgin, et al.
Published: (2026)
by: Hasanov, Eshgin, et al.
Published: (2026)
Personality-Driven Student Agent-Based Modeling in Mathematics Education: How Well Do Student Agents Align with Human Learners?
by: Xiao, Bushi, et al.
Published: (2026)
by: Xiao, Bushi, et al.
Published: (2026)
Aligning What LLMs Do and Say: Towards Self-Consistent Explanations
by: Admoni, Sahar, et al.
Published: (2025)
by: Admoni, Sahar, et al.
Published: (2025)
Generative Agents and Expectations: Do LLMs Align with Heterogeneous Agent Models?
by: Gusella, Filippo, et al.
Published: (2025)
by: Gusella, Filippo, et al.
Published: (2025)
Sparks of Rationality: Do Reasoning LLMs Align with Human Judgment and Choice?
by: Tak, Ala N., et al.
Published: (2026)
by: Tak, Ala N., et al.
Published: (2026)
How Language Directions Align with Token Geometry in Multilingual LLMs
by: Kim, JaeSeong, et al.
Published: (2025)
by: Kim, JaeSeong, et al.
Published: (2025)
Align2Act: Instruction-Tuned Models for Human-Aligned Autonomous Driving
by: Jaisankar, Kanishkha, et al.
Published: (2025)
by: Jaisankar, Kanishkha, et al.
Published: (2025)
Aligning LLMs with Human Instructions and Stock Market Feedback in Financial Sentiment Analysis
by: Zhao, Zijie, et al.
Published: (2024)
by: Zhao, Zijie, et al.
Published: (2024)
Similar Items
-
SimPal: Towards a Meta-Conversational Framework to Understand Teacher's Instructional Goals for K-12 Physics
by: Farhana, Effat, et al.
Published: (2024) -
Middle School Students' Application of Science Learning From Physical Versus Virtual Labs to New Contexts
by: Dana Gnesdilow, et al.
Published: (2025) -
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders
by: Sarkar, Souvika, et al.
Published: (2025) -
The Bias is in the Details: An Assessment of Cognitive Bias in LLMs
by: Knipper, R. Alexander, et al.
Published: (2025) -
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform?
by: Sarkar, Souvika, et al.
Published: (2023)