The Promises and Pitfalls of Using Language Models to Measure Instruction Quality in Education
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Paiheng, Liu, Jing, Jones, Nathan, Cohen, Julie, Ai, Wei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Explore Spurious Correlations at the Concept Level in Language Models for Text Classification
di: Zhou, Yuhang, et al.
Pubblicazione: (2023)
di: Zhou, Yuhang, et al.
Pubblicazione: (2023)
Multi-Stage Balanced Distillation: Addressing Long-Tail Challenges in Sequence-Level Knowledge Distillation
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
Emojis Decoded: Leveraging ChatGPT for Enhanced Understanding in Social Media Communications
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
Large Language Models and Causal Inference in Collaboration: A Survey
di: Liu, Xiaoyu, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyu, et al.
Pubblicazione: (2024)
Towards Understanding In-Context Learning with Contrastive Demonstrations and Saliency Maps
di: Liu, Fuxiao, et al.
Pubblicazione: (2023)
di: Liu, Fuxiao, et al.
Pubblicazione: (2023)
DISCO Balances the Scales: Adaptive Domain- and Difficulty-Aware Reinforcement Learning on Imbalanced Data
di: Zhou, Yuhang, et al.
Pubblicazione: (2025)
di: Zhou, Yuhang, et al.
Pubblicazione: (2025)
Measuring Pragmatic Influence in Large Language Model Instructions
di: Geng, Yilin, et al.
Pubblicazione: (2026)
di: Geng, Yilin, et al.
Pubblicazione: (2026)
Instruction Multi-Constraint Molecular Generation Using a Teacher-Student Large Language Model
di: Zhou, Peng, et al.
Pubblicazione: (2024)
di: Zhou, Peng, et al.
Pubblicazione: (2024)
Enhancing Instructional Quality: Leveraging Computer-Assisted Textual Analysis to Generate In-Depth Insights from Educational Artifacts
di: Tian, Zewei, et al.
Pubblicazione: (2024)
di: Tian, Zewei, et al.
Pubblicazione: (2024)
It is Too Many Options: Pitfalls of Multiple-Choice Questions in Generative AI and Medical Education
di: Singh, Shrutika, et al.
Pubblicazione: (2025)
di: Singh, Shrutika, et al.
Pubblicazione: (2025)
Measuring and Controlling Instruction (In)Stability in Language Model Dialogs
di: Li, Kenneth, et al.
Pubblicazione: (2024)
di: Li, Kenneth, et al.
Pubblicazione: (2024)
It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers
di: Clavié, Benjamin, et al.
Pubblicazione: (2025)
di: Clavié, Benjamin, et al.
Pubblicazione: (2025)
Promises, Outlooks and Challenges of Diffusion Language Modeling
di: Deschenaux, Justin, et al.
Pubblicazione: (2024)
di: Deschenaux, Justin, et al.
Pubblicazione: (2024)
Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models
di: McDaniel, Emma, et al.
Pubblicazione: (2024)
di: McDaniel, Emma, et al.
Pubblicazione: (2024)
Finetuning Generative Large Language Models with Discrimination Instructions for Knowledge Graph Completion
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Benefits and Pitfalls of Reinforcement Learning for Language Model Planning: A Theoretical Perspective
di: Wang, Siwei, et al.
Pubblicazione: (2025)
di: Wang, Siwei, et al.
Pubblicazione: (2025)
Automatic Instruction Evolving for Large Language Models
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
A Common Pitfall of Margin-based Language Model Alignment: Gradient Entanglement
di: Yuan, Hui, et al.
Pubblicazione: (2024)
di: Yuan, Hui, et al.
Pubblicazione: (2024)
Phased Instruction Fine-Tuning for Large Language Models
di: Pang, Wei, et al.
Pubblicazione: (2024)
di: Pang, Wei, et al.
Pubblicazione: (2024)
Towards Robust Instruction Tuning on Multimodal Large Language Models
di: Han, Wei, et al.
Pubblicazione: (2024)
di: Han, Wei, et al.
Pubblicazione: (2024)
Large Language Models for Biomedical Text Simplification: Promising But Not There Yet
di: Li, Zihao, et al.
Pubblicazione: (2024)
di: Li, Zihao, et al.
Pubblicazione: (2024)
Advancing Reasoning in Large Language Models: Promising Methods and Approaches
di: Patil, Avinash, et al.
Pubblicazione: (2025)
di: Patil, Avinash, et al.
Pubblicazione: (2025)
GraphGPT: Graph Instruction Tuning for Large Language Models
di: Tang, Jiabin, et al.
Pubblicazione: (2023)
di: Tang, Jiabin, et al.
Pubblicazione: (2023)
CA*: Addressing Evaluation Pitfalls in Computation-Aware Latency for Simultaneous Speech Translation
di: Xu, Xi, et al.
Pubblicazione: (2024)
di: Xu, Xi, et al.
Pubblicazione: (2024)
AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios
di: Qi, Yunjia, et al.
Pubblicazione: (2025)
di: Qi, Yunjia, et al.
Pubblicazione: (2025)
LIFEBench: Evaluating Length Instruction Following in Large Language Models
di: Zhang, Wei, et al.
Pubblicazione: (2025)
di: Zhang, Wei, et al.
Pubblicazione: (2025)
Investigating Instruction Tuning Large Language Models on Graphs
di: Zhu, Kerui, et al.
Pubblicazione: (2024)
di: Zhu, Kerui, et al.
Pubblicazione: (2024)
Small but Significant: On the Promise of Small Language Models for Accessible AIED
di: Wei, Yumou, et al.
Pubblicazione: (2025)
di: Wei, Yumou, et al.
Pubblicazione: (2025)
Pitfalls of Conversational LLMs on News Debiasing
di: Schlicht, Ipek Baris, et al.
Pubblicazione: (2024)
di: Schlicht, Ipek Baris, et al.
Pubblicazione: (2024)
Revisiting the Reliability of Language Models in Instruction-Following
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
Improve Large Language Model Systems with User Logs
di: Wang, Changyue, et al.
Pubblicazione: (2026)
di: Wang, Changyue, et al.
Pubblicazione: (2026)
Beyond Instruction Following: Evaluating Inferential Rule Following of Large Language Models
di: Sun, Wangtao, et al.
Pubblicazione: (2024)
di: Sun, Wangtao, et al.
Pubblicazione: (2024)
X-Instruction: Aligning Language Model in Low-resource Languages with Self-curated Cross-lingual Instructions
di: Li, Chong, et al.
Pubblicazione: (2024)
di: Li, Chong, et al.
Pubblicazione: (2024)
QCRD: Quality-guided Contrastive Rationale Distillation for Large Language Models
di: Wang, Wei, et al.
Pubblicazione: (2024)
di: Wang, Wei, et al.
Pubblicazione: (2024)
Constraint Back-translation Improves Complex Instruction Following of Large Language Models
di: Qi, Yunjia, et al.
Pubblicazione: (2024)
di: Qi, Yunjia, et al.
Pubblicazione: (2024)
One-Shot Learning as Instruction Data Prospector for Large Language Models
di: Li, Yunshui, et al.
Pubblicazione: (2023)
di: Li, Yunshui, et al.
Pubblicazione: (2023)
OctoPack: Instruction Tuning Code Large Language Models
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models
di: Ren, Huimin, et al.
Pubblicazione: (2025)
di: Ren, Huimin, et al.
Pubblicazione: (2025)
Teaching-Assistant-in-the-Loop: Improving Knowledge Distillation from Imperfect Teacher Models in Low-Budget Scenarios
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
AI Generated Text Detection Using Instruction Fine-tuned Large Language and Transformer-Based Models
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Explore Spurious Correlations at the Concept Level in Language Models for Text Classification
di: Zhou, Yuhang, et al.
Pubblicazione: (2023) -
Multi-Stage Balanced Distillation: Addressing Long-Tail Challenges in Sequence-Level Knowledge Distillation
di: Zhou, Yuhang, et al.
Pubblicazione: (2024) -
Emojis Decoded: Leveraging ChatGPT for Enhanced Understanding in Social Media Communications
di: Zhou, Yuhang, et al.
Pubblicazione: (2024) -
Large Language Models and Causal Inference in Collaboration: A Survey
di: Liu, Xiaoyu, et al.
Pubblicazione: (2024) -
Towards Understanding In-Context Learning with Contrastive Demonstrations and Saliency Maps
di: Liu, Fuxiao, et al.
Pubblicazione: (2023)