Automated Alignment of Math Items to Content Standards in Large-Scale Assessments Using Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Qingshu, Jiao, Hong, Zhou, Tianyi, Li, Ming, Zhang, Nan, Peters, Sydney, Fu, Yanbin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Text-Based Approaches to Item Alignment to Content Standards in Large-Scale Reading & Writing Tests
por: Fu, Yanbin, et al.
Publicado: (2025)
por: Fu, Yanbin, et al.
Publicado: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025)
por: Peters, Sydney, et al.
Publicado: (2025)
Understanding the Thinking Process of Reasoning Models: A Perspective from Schoenfeld's Episode Theory
por: Li, Ming, et al.
Publicado: (2025)
por: Li, Ming, et al.
Publicado: (2025)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
por: Li, Ming, et al.
Publicado: (2025)
por: Li, Ming, et al.
Publicado: (2025)
Pronunciation Assessment with Multi-modal Large Language Models
por: Fu, Kaiqi, et al.
Publicado: (2024)
por: Fu, Kaiqi, et al.
Publicado: (2024)
Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory
por: Song, Dan, et al.
Publicado: (2025)
por: Song, Dan, et al.
Publicado: (2025)
TabularMath: Understanding Math Reasoning over Tables with Large Language Models
por: Tian, Shi-Yu, et al.
Publicado: (2025)
por: Tian, Shi-Yu, et al.
Publicado: (2025)
InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
por: Ying, Huaiyuan, et al.
Publicado: (2024)
por: Ying, Huaiyuan, et al.
Publicado: (2024)
MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models
por: Peng, Shuai, et al.
Publicado: (2024)
por: Peng, Shuai, et al.
Publicado: (2024)
Implicit Grading Bias in Large Language Models: How Writing Style Affects Automated Assessment Across Math, Programming, and Essay Tasks
por: Jadhav, Rudra, et al.
Publicado: (2026)
por: Jadhav, Rudra, et al.
Publicado: (2026)
Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
por: Albalak, Alon, et al.
Publicado: (2025)
por: Albalak, Alon, et al.
Publicado: (2025)
Exploring Automated Distractor Generation for Math Multiple-choice Questions via Large Language Models
por: Feng, Wanyong, et al.
Publicado: (2024)
por: Feng, Wanyong, et al.
Publicado: (2024)
Neuron-based Personality Trait Induction in Large Language Models
por: Deng, Jia, et al.
Publicado: (2024)
por: Deng, Jia, et al.
Publicado: (2024)
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
por: Fang, Meng, et al.
Publicado: (2024)
por: Fang, Meng, et al.
Publicado: (2024)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
por: Wu, Sirui, et al.
Publicado: (2025)
por: Wu, Sirui, et al.
Publicado: (2025)
Scaling Item-to-Standard Alignment with Large Language Models: Accuracy, Limits, and Solutions
por: Karimi-Malekabadi, Farzan, et al.
Publicado: (2025)
por: Karimi-Malekabadi, Farzan, et al.
Publicado: (2025)
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
por: Xu, Yifan, et al.
Publicado: (2024)
por: Xu, Yifan, et al.
Publicado: (2024)
Pre-trained Large Language Models Use Fourier Features to Compute Addition
por: Zhou, Tianyi, et al.
Publicado: (2024)
por: Zhou, Tianyi, et al.
Publicado: (2024)
Large Language Models Struggle with Unreasonability in Math Problems
por: Ma, Jingyuan, et al.
Publicado: (2024)
por: Ma, Jingyuan, et al.
Publicado: (2024)
Standardize: Aligning Language Models with Expert-Defined Standards for Content Generation
por: Imperial, Joseph Marvin, et al.
Publicado: (2024)
por: Imperial, Joseph Marvin, et al.
Publicado: (2024)
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
por: Zeng, Liang, et al.
Publicado: (2024)
por: Zeng, Liang, et al.
Publicado: (2024)
A Survey on Knowledge Distillation of Large Language Models
por: Xu, Xiaohan, et al.
Publicado: (2024)
por: Xu, Xiaohan, et al.
Publicado: (2024)
MathGLM-Vision: Solving Mathematical Problems with Multi-Modal Large Language Model
por: Yang, Zhen, et al.
Publicado: (2024)
por: Yang, Zhen, et al.
Publicado: (2024)
Lost in Benchmarks? Rethinking Large Language Model Benchmarking with Item Response Theory
por: Zhou, Hongli, et al.
Publicado: (2025)
por: Zhou, Hongli, et al.
Publicado: (2025)
A Survey on Training-free Alignment of Large Language Models
por: Pan, Birong, et al.
Publicado: (2025)
por: Pan, Birong, et al.
Publicado: (2025)
Toward Automated Cognitive Assessment in Parkinson's Disease Using Pretrained Language Models
por: Khanna, Varada, et al.
Publicado: (2025)
por: Khanna, Varada, et al.
Publicado: (2025)
Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays
por: Hua, Haowei, et al.
Publicado: (2025)
por: Hua, Haowei, et al.
Publicado: (2025)
Multi-Objective Linguistic Control of Large Language Models
por: Nguyen, Dang, et al.
Publicado: (2024)
por: Nguyen, Dang, et al.
Publicado: (2024)
Aligning Large Language Models with Implicit Preferences from User-Generated Content
por: Tan, Zhaoxuan, et al.
Publicado: (2025)
por: Tan, Zhaoxuan, et al.
Publicado: (2025)
Automated Knowledge Graph Construction using Large Language Models and Sentence Complexity Modelling
por: Anuyah, Sydney, et al.
Publicado: (2025)
por: Anuyah, Sydney, et al.
Publicado: (2025)
Proof Automation with Large Language Models
por: Lu, Minghai, et al.
Publicado: (2024)
por: Lu, Minghai, et al.
Publicado: (2024)
Problematic Tokens: Tokenizer Bias in Large Language Models
por: Yang, Jin, et al.
Publicado: (2024)
por: Yang, Jin, et al.
Publicado: (2024)
Benchmarking Large Language Models for Math Reasoning Tasks
por: Seßler, Kathrin, et al.
Publicado: (2024)
por: Seßler, Kathrin, et al.
Publicado: (2024)
Arrows of Math Reasoning Data Synthesis for Large Language Models: Diversity, Complexity and Correctness
por: Chen, Sirui, et al.
Publicado: (2025)
por: Chen, Sirui, et al.
Publicado: (2025)
Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection
por: Gao, Yifan, et al.
Publicado: (2025)
por: Gao, Yifan, et al.
Publicado: (2025)
Using Large Language Models to Understand Telecom Standards
por: Karapantelakis, Athanasios, et al.
Publicado: (2024)
por: Karapantelakis, Athanasios, et al.
Publicado: (2024)
Automate Knowledge Concept Tagging on Math Questions with LLMs
por: Li, Hang, et al.
Publicado: (2024)
por: Li, Hang, et al.
Publicado: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
por: Christ, Bryan R., et al.
Publicado: (2024)
por: Christ, Bryan R., et al.
Publicado: (2024)
Can LLMs Master Math? Investigating Large Language Models on Math Stack Exchange
por: Satpute, Ankit, et al.
Publicado: (2024)
por: Satpute, Ankit, et al.
Publicado: (2024)
Adaptable and Reliable Text Classification using Large Language Models
por: Wang, Zhiqiang, et al.
Publicado: (2024)
por: Wang, Zhiqiang, et al.
Publicado: (2024)
Ejemplares similares
-
Text-Based Approaches to Item Alignment to Content Standards in Large-Scale Reading & Writing Tests
por: Fu, Yanbin, et al.
Publicado: (2025) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025) -
Understanding the Thinking Process of Reasoning Models: A Perspective from Schoenfeld's Episode Theory
por: Li, Ming, et al.
Publicado: (2025) -
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
por: Li, Ming, et al.
Publicado: (2025) -
Pronunciation Assessment with Multi-modal Large Language Models
por: Fu, Kaiqi, et al.
Publicado: (2024)