Using Vision + Language Models to Predict Item Difficulty
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Khan, Samin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automating Autograding: Large Language Models as Test Suite Generators for Introductory Programming
von: Alkafaween, Umar, et al.
Veröffentlicht: (2024)
von: Alkafaween, Umar, et al.
Veröffentlicht: (2024)
Evaluating LLMs for Career Guidance: Comparative Analysis of Computing Competency Recommendations Across Ten African Countries
von: Eze, Precious, et al.
Veröffentlicht: (2025)
von: Eze, Precious, et al.
Veröffentlicht: (2025)
Ensemble ToT of LLMs and Its Application to Automatic Grading System for Supporting Self-Learning
von: Ito, Yuki, et al.
Veröffentlicht: (2025)
von: Ito, Yuki, et al.
Veröffentlicht: (2025)
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
von: Song, Dingjie, et al.
Veröffentlicht: (2026)
von: Song, Dingjie, et al.
Veröffentlicht: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Self-hosted Lecture-to-Quiz: Local LLM MCQ Generation with Deterministic Quality Control
von: Shintani, Seine A.
Veröffentlicht: (2026)
von: Shintani, Seine A.
Veröffentlicht: (2026)
Enhancing Computer Programming Education with LLMs: A Study on Effective Prompt Engineering for Python Code Generation
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
VLEU: a Method for Automatic Evaluation for Generalizability of Text-to-Image Models
von: Cao, Jingtao, et al.
Veröffentlicht: (2024)
von: Cao, Jingtao, et al.
Veröffentlicht: (2024)
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective
von: Weng, Zhaotian, et al.
Veröffentlicht: (2024)
von: Weng, Zhaotian, et al.
Veröffentlicht: (2024)
On the Limitations of Vision-Language Models in Understanding Image Transforms
von: Anis, Ahmad Mustafa, et al.
Veröffentlicht: (2025)
von: Anis, Ahmad Mustafa, et al.
Veröffentlicht: (2025)
An Eye for an AI: Evaluating GPT-4o's Visual Perception Skills and Geometric Reasoning Skills Using Computer Graphics Questions
von: Feng, Tony Haoran, et al.
Veröffentlicht: (2024)
von: Feng, Tony Haoran, et al.
Veröffentlicht: (2024)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
von: Ramakrishnan, Aashish Anantha, et al.
Veröffentlicht: (2025)
von: Ramakrishnan, Aashish Anantha, et al.
Veröffentlicht: (2025)
Integrating Generative AI in Cybersecurity Education: Case Study Insights on Pedagogical Strategies, Critical Thinking, and Responsible AI Use
von: Elkhodr, Mahmoud, et al.
Veröffentlicht: (2025)
von: Elkhodr, Mahmoud, et al.
Veröffentlicht: (2025)
Robust Uncertainty Quantification for Factual Generation of Large Language Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
jina-vlm: Small Multilingual Vision Language Model
von: Koukounas, Andreas, et al.
Veröffentlicht: (2025)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2025)
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
Using Deep Learning to Generate Semantically Correct Hindi Captions
von: Khan, Wasim Akram, et al.
Veröffentlicht: (2026)
von: Khan, Wasim Akram, et al.
Veröffentlicht: (2026)
Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking
von: Khurdula, Harsha Vardhan, et al.
Veröffentlicht: (2024)
von: Khurdula, Harsha Vardhan, et al.
Veröffentlicht: (2024)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
von: Yan, Lingyong, et al.
Veröffentlicht: (2026)
von: Yan, Lingyong, et al.
Veröffentlicht: (2026)
Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
von: Rudman, William, et al.
Veröffentlicht: (2026)
von: Rudman, William, et al.
Veröffentlicht: (2026)
Safeguarding Vision-Language Models Against Patched Visual Prompt Injectors
von: Sun, Jiachen, et al.
Veröffentlicht: (2024)
von: Sun, Jiachen, et al.
Veröffentlicht: (2024)
DISCO: Document Intelligence Suite for COmparative Evaluation
von: Benkirane, Kenza, et al.
Veröffentlicht: (2026)
von: Benkirane, Kenza, et al.
Veröffentlicht: (2026)
Relative Drawing Identification Complexity is Invariant to Modality in Vision-Language Models
von: Freitas, Diogo, et al.
Veröffentlicht: (2025)
von: Freitas, Diogo, et al.
Veröffentlicht: (2025)
Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model
von: Georgiou, Athos
Veröffentlicht: (2026)
von: Georgiou, Athos
Veröffentlicht: (2026)
Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment
von: Burleigh, Tyler
Veröffentlicht: (2026)
von: Burleigh, Tyler
Veröffentlicht: (2026)
Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models
von: Wang, Junxin, et al.
Veröffentlicht: (2026)
von: Wang, Junxin, et al.
Veröffentlicht: (2026)
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
von: Sanders, Kate, et al.
Veröffentlicht: (2024)
von: Sanders, Kate, et al.
Veröffentlicht: (2024)
Synthetic Student Responses: LLM-Extracted Features for IRT Difficulty Parameter Estimation
von: Hoyl, Matias
Veröffentlicht: (2026)
von: Hoyl, Matias
Veröffentlicht: (2026)
More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage
von: He, Wei
Veröffentlicht: (2026)
von: He, Wei
Veröffentlicht: (2026)
The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark
von: Liu, Hao, et al.
Veröffentlicht: (2026)
von: Liu, Hao, et al.
Veröffentlicht: (2026)
Step-CoT: Stepwise Visual Chain-of-Thought for Medical Visual Question Answering
von: Fan, Lin, et al.
Veröffentlicht: (2026)
von: Fan, Lin, et al.
Veröffentlicht: (2026)
Transformers Get Stable: An End-to-End Signal Propagation Theory for Language Models
von: Kedia, Akhil, et al.
Veröffentlicht: (2024)
von: Kedia, Akhil, et al.
Veröffentlicht: (2024)
Seeing the Forest and the Trees: Solving Visual Graph and Tree Based Data Structure Problems using Large Multimodal Models
von: Gutierrez, Sebastian, et al.
Veröffentlicht: (2024)
von: Gutierrez, Sebastian, et al.
Veröffentlicht: (2024)
Scaling Large Vision-Language Models for Enhanced Multimodal Comprehension In Biomedical Image Analysis
von: Umeike, Robinson, et al.
Veröffentlicht: (2025)
von: Umeike, Robinson, et al.
Veröffentlicht: (2025)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
von: Putra, Rizky Ramadhana, et al.
Veröffentlicht: (2026)
von: Putra, Rizky Ramadhana, et al.
Veröffentlicht: (2026)
Fine-Tuning Vision-Language Models for Markdown Conversion of Financial Tables in Malaysian Audited Financial Reports
von: Tan, Jin Khye, et al.
Veröffentlicht: (2025)
von: Tan, Jin Khye, et al.
Veröffentlicht: (2025)
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
von: Han, Jongwook, et al.
Veröffentlicht: (2025)
DriveMRP: Enhancing Vision-Language Models with Synthetic Motion Data for Motion Risk Prediction
von: Hou, Zhiyi, et al.
Veröffentlicht: (2025)
von: Hou, Zhiyi, et al.
Veröffentlicht: (2025)
myMNIST: Benchmark of PETNN, KAN, and Classical Deep Learning Models for Burmese Handwritten Digit Recognition
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2026)
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Automating Autograding: Large Language Models as Test Suite Generators for Introductory Programming
von: Alkafaween, Umar, et al.
Veröffentlicht: (2024) -
Evaluating LLMs for Career Guidance: Comparative Analysis of Computing Competency Recommendations Across Ten African Countries
von: Eze, Precious, et al.
Veröffentlicht: (2025) -
Ensemble ToT of LLMs and Its Application to Automatic Grading System for Supporting Self-Learning
von: Ito, Yuki, et al.
Veröffentlicht: (2025) -
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
von: Song, Dingjie, et al.
Veröffentlicht: (2026) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)