Automated Evaluation of Classroom Instructional Support with LLMs and BoWs: Connecting Global Predictions to Specific Feedback
Fuente:
arXiv
Guardado en:
| Autores principales: | Whitehill, Jacob, LoCasale-Crouch, Jennifer |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
por: Trinh, Viet Anh, et al.
Publicado: (2025)
por: Trinh, Viet Anh, et al.
Publicado: (2025)
Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation
por: Guan, Yiwen, et al.
Publicado: (2025)
por: Guan, Yiwen, et al.
Publicado: (2025)
Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio
por: He, Xinlu, et al.
Publicado: (2025)
por: He, Xinlu, et al.
Publicado: (2025)
LLMs to Support a Domain Specific Knowledge Assistant
por: Lovin, Maria-Flavia
Publicado: (2025)
por: Lovin, Maria-Flavia
Publicado: (2025)
Agent-based Automated Claim Matching with Instruction-following LLMs
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
Evaluating the Evaluator: Measuring LLMs' Adherence to Task Evaluation Instructions
por: Murugadoss, Bhuvanashree, et al.
Publicado: (2024)
por: Murugadoss, Bhuvanashree, et al.
Publicado: (2024)
Qworld: Question-Specific Evaluation Criteria for LLMs
por: Gao, Shanghua, et al.
Publicado: (2026)
por: Gao, Shanghua, et al.
Publicado: (2026)
Argumentation for Explainable and Globally Contestable Decision Support with LLMs
por: Dejl, Adam, et al.
Publicado: (2026)
por: Dejl, Adam, et al.
Publicado: (2026)
DIALEVAL: Automated Type-Theoretic Evaluation of LLM Instruction Following
por: Basta, Nardine, et al.
Publicado: (2026)
por: Basta, Nardine, et al.
Publicado: (2026)
Offscript: Automated Auditing of Instruction Adherence in LLMs
por: Clark, Nicholas, et al.
Publicado: (2025)
por: Clark, Nicholas, et al.
Publicado: (2025)
How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs?
por: Doostmohammadi, Ehsan, et al.
Publicado: (2024)
por: Doostmohammadi, Ehsan, et al.
Publicado: (2024)
RAGalyst: Automated Human-Aligned Agentic Evaluation for Domain-Specific RAG
por: Gao, Joshua, et al.
Publicado: (2025)
por: Gao, Joshua, et al.
Publicado: (2025)
Automating Legal Interpretation with LLMs: Retrieval, Generation, and Evaluation
por: Luo, Kangcheng, et al.
Publicado: (2025)
por: Luo, Kangcheng, et al.
Publicado: (2025)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
por: Lissak, Shir, et al.
Publicado: (2024)
por: Lissak, Shir, et al.
Publicado: (2024)
DARE-bench: Evaluating Modeling and Instruction Fidelity of LLMs in Data Science
por: Shu, Fan, et al.
Publicado: (2026)
por: Shu, Fan, et al.
Publicado: (2026)
BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation
por: Sun, Peng, et al.
Publicado: (2026)
por: Sun, Peng, et al.
Publicado: (2026)
Connecting the Dots: Evaluating Abstract Reasoning Capabilities of LLMs Using the New York Times Connections Word Game
por: Samadarshi, Prisha, et al.
Publicado: (2024)
por: Samadarshi, Prisha, et al.
Publicado: (2024)
Visual Reasoning Benchmark: Evaluating Multimodal LLMs on Classroom-Authentic Visual Problems from Primary Education
por: Huti, Mohamed, et al.
Publicado: (2026)
por: Huti, Mohamed, et al.
Publicado: (2026)
AXCEL: Automated eXplainable Consistency Evaluation using LLMs
por: Sreekar, P Aditya, et al.
Publicado: (2024)
por: Sreekar, P Aditya, et al.
Publicado: (2024)
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
From Guidelines to Guarantees: A Graph-Based Evaluation Harness for Domain-Specific Evaluation of LLMs
por: Lundin, Jessica M., et al.
Publicado: (2025)
por: Lundin, Jessica M., et al.
Publicado: (2025)
The Instruction Gap: LLMs get lost in Following Instruction
por: Tripathi, Vishesh, et al.
Publicado: (2025)
por: Tripathi, Vishesh, et al.
Publicado: (2025)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
por: Xu, Wenda, et al.
Publicado: (2025)
por: Xu, Wenda, et al.
Publicado: (2025)
Automated Essay Scoring Incorporating Annotations from Automated Feedback Systems
por: Ormerod, Christopher
Publicado: (2025)
por: Ormerod, Christopher
Publicado: (2025)
Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs
por: Bouchard, Dylan
Publicado: (2024)
por: Bouchard, Dylan
Publicado: (2024)
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
por: Wu, Jiaxing, et al.
Publicado: (2024)
por: Wu, Jiaxing, et al.
Publicado: (2024)
Transfer Learning for Automated Feedback Generation on Small Datasets
por: Morris, Oscar
Publicado: (2025)
por: Morris, Oscar
Publicado: (2025)
Time-Reversal Provides Unsupervised Feedback to LLMs
por: Varun, Yerram, et al.
Publicado: (2024)
por: Varun, Yerram, et al.
Publicado: (2024)
Classroom AI: Large Language Models as Grade-Specific Teachers
por: Oh, Jio, et al.
Publicado: (2026)
por: Oh, Jio, et al.
Publicado: (2026)
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
por: Allaham, Mowafak, et al.
Publicado: (2024)
por: Allaham, Mowafak, et al.
Publicado: (2024)
Taming LLMs with Negative Samples: A Reference-Free Framework to Evaluate Presentation Content with Actionable Feedback
por: Muppidi, Ananth, et al.
Publicado: (2025)
por: Muppidi, Ananth, et al.
Publicado: (2025)
FB-Bench: A Fine-Grained Multi-Task Benchmark for Evaluating LLMs' Responsiveness to Human Feedback
por: Li, Youquan, et al.
Publicado: (2024)
por: Li, Youquan, et al.
Publicado: (2024)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
por: Banerjee, Tanushree, et al.
Publicado: (2024)
por: Banerjee, Tanushree, et al.
Publicado: (2024)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
por: Wang, Xingyao, et al.
Publicado: (2023)
por: Wang, Xingyao, et al.
Publicado: (2023)
LLMs can be easily Confused by Instructional Distractions
por: Hwang, Yerin, et al.
Publicado: (2025)
por: Hwang, Yerin, et al.
Publicado: (2025)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
por: Buscemi, Alessio, et al.
Publicado: (2025)
por: Buscemi, Alessio, et al.
Publicado: (2025)
Provable Interactive Learning with Hindsight Instruction Feedback
por: Misra, Dipendra, et al.
Publicado: (2024)
por: Misra, Dipendra, et al.
Publicado: (2024)
Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback
por: Ji, Jiaming, et al.
Publicado: (2024)
por: Ji, Jiaming, et al.
Publicado: (2024)
From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms
por: Jiang, Zhaokun, et al.
Publicado: (2025)
por: Jiang, Zhaokun, et al.
Publicado: (2025)
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
por: Merino, Tim, et al.
Publicado: (2024)
por: Merino, Tim, et al.
Publicado: (2024)
Ejemplares similares
-
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
por: Trinh, Viet Anh, et al.
Publicado: (2025) -
Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation
por: Guan, Yiwen, et al.
Publicado: (2025) -
Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio
por: He, Xinlu, et al.
Publicado: (2025) -
LLMs to Support a Domain Specific Knowledge Assistant
por: Lovin, Maria-Flavia
Publicado: (2025) -
Agent-based Automated Claim Matching with Instruction-following LLMs
por: Pisarevskaya, Dina, et al.
Publicado: (2025)