Using Vision + Language Models to Predict Item Difficulty
Fuente:
arXiv
Salvato in:
| Autore principale: | Khan, Samin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Automating Autograding: Large Language Models as Test Suite Generators for Introductory Programming
di: Alkafaween, Umar, et al.
Pubblicazione: (2024)
di: Alkafaween, Umar, et al.
Pubblicazione: (2024)
Evaluating LLMs for Career Guidance: Comparative Analysis of Computing Competency Recommendations Across Ten African Countries
di: Eze, Precious, et al.
Pubblicazione: (2025)
di: Eze, Precious, et al.
Pubblicazione: (2025)
Ensemble ToT of LLMs and Its Application to Automatic Grading System for Supporting Self-Learning
di: Ito, Yuki, et al.
Pubblicazione: (2025)
di: Ito, Yuki, et al.
Pubblicazione: (2025)
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
di: Song, Dingjie, et al.
Pubblicazione: (2026)
di: Song, Dingjie, et al.
Pubblicazione: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Self-hosted Lecture-to-Quiz: Local LLM MCQ Generation with Deterministic Quality Control
di: Shintani, Seine A.
Pubblicazione: (2026)
di: Shintani, Seine A.
Pubblicazione: (2026)
Enhancing Computer Programming Education with LLMs: A Study on Effective Prompt Engineering for Python Code Generation
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
VLEU: a Method for Automatic Evaluation for Generalizability of Text-to-Image Models
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective
di: Weng, Zhaotian, et al.
Pubblicazione: (2024)
di: Weng, Zhaotian, et al.
Pubblicazione: (2024)
On the Limitations of Vision-Language Models in Understanding Image Transforms
di: Anis, Ahmad Mustafa, et al.
Pubblicazione: (2025)
di: Anis, Ahmad Mustafa, et al.
Pubblicazione: (2025)
An Eye for an AI: Evaluating GPT-4o's Visual Perception Skills and Geometric Reasoning Skills Using Computer Graphics Questions
di: Feng, Tony Haoran, et al.
Pubblicazione: (2024)
di: Feng, Tony Haoran, et al.
Pubblicazione: (2024)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
di: Nwatu, Joan, et al.
Pubblicazione: (2024)
di: Nwatu, Joan, et al.
Pubblicazione: (2024)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
Integrating Generative AI in Cybersecurity Education: Case Study Insights on Pedagogical Strategies, Critical Thinking, and Responsible AI Use
di: Elkhodr, Mahmoud, et al.
Pubblicazione: (2025)
di: Elkhodr, Mahmoud, et al.
Pubblicazione: (2025)
Robust Uncertainty Quantification for Factual Generation of Large Language Models
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
jina-vlm: Small Multilingual Vision Language Model
di: Koukounas, Andreas, et al.
Pubblicazione: (2025)
di: Koukounas, Andreas, et al.
Pubblicazione: (2025)
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
di: Nwatu, Joan, et al.
Pubblicazione: (2025)
di: Nwatu, Joan, et al.
Pubblicazione: (2025)
Using Deep Learning to Generate Semantically Correct Hindi Captions
di: Khan, Wasim Akram, et al.
Pubblicazione: (2026)
di: Khan, Wasim Akram, et al.
Pubblicazione: (2026)
Beyond Visual Understanding: Introducing PARROT-360V for Vision Language Model Benchmarking
di: Khurdula, Harsha Vardhan, et al.
Pubblicazione: (2024)
di: Khurdula, Harsha Vardhan, et al.
Pubblicazione: (2024)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
di: Rudman, William, et al.
Pubblicazione: (2026)
di: Rudman, William, et al.
Pubblicazione: (2026)
Safeguarding Vision-Language Models Against Patched Visual Prompt Injectors
di: Sun, Jiachen, et al.
Pubblicazione: (2024)
di: Sun, Jiachen, et al.
Pubblicazione: (2024)
DISCO: Document Intelligence Suite for COmparative Evaluation
di: Benkirane, Kenza, et al.
Pubblicazione: (2026)
di: Benkirane, Kenza, et al.
Pubblicazione: (2026)
Relative Drawing Identification Complexity is Invariant to Modality in Vision-Language Models
di: Freitas, Diogo, et al.
Pubblicazione: (2025)
di: Freitas, Diogo, et al.
Pubblicazione: (2025)
Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model
di: Georgiou, Athos
Pubblicazione: (2026)
di: Georgiou, Athos
Pubblicazione: (2026)
Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment
di: Burleigh, Tyler
Pubblicazione: (2026)
di: Burleigh, Tyler
Pubblicazione: (2026)
Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models
di: Wang, Junxin, et al.
Pubblicazione: (2026)
di: Wang, Junxin, et al.
Pubblicazione: (2026)
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
di: Sanders, Kate, et al.
Pubblicazione: (2024)
di: Sanders, Kate, et al.
Pubblicazione: (2024)
Synthetic Student Responses: LLM-Extracted Features for IRT Difficulty Parameter Estimation
di: Hoyl, Matias
Pubblicazione: (2026)
di: Hoyl, Matias
Pubblicazione: (2026)
More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage
di: He, Wei
Pubblicazione: (2026)
di: He, Wei
Pubblicazione: (2026)
The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark
di: Liu, Hao, et al.
Pubblicazione: (2026)
di: Liu, Hao, et al.
Pubblicazione: (2026)
Step-CoT: Stepwise Visual Chain-of-Thought for Medical Visual Question Answering
di: Fan, Lin, et al.
Pubblicazione: (2026)
di: Fan, Lin, et al.
Pubblicazione: (2026)
Transformers Get Stable: An End-to-End Signal Propagation Theory for Language Models
di: Kedia, Akhil, et al.
Pubblicazione: (2024)
di: Kedia, Akhil, et al.
Pubblicazione: (2024)
Seeing the Forest and the Trees: Solving Visual Graph and Tree Based Data Structure Problems using Large Multimodal Models
di: Gutierrez, Sebastian, et al.
Pubblicazione: (2024)
di: Gutierrez, Sebastian, et al.
Pubblicazione: (2024)
Scaling Large Vision-Language Models for Enhanced Multimodal Comprehension In Biomedical Image Analysis
di: Umeike, Robinson, et al.
Pubblicazione: (2025)
di: Umeike, Robinson, et al.
Pubblicazione: (2025)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
di: Putra, Rizky Ramadhana, et al.
Pubblicazione: (2026)
di: Putra, Rizky Ramadhana, et al.
Pubblicazione: (2026)
Fine-Tuning Vision-Language Models for Markdown Conversion of Financial Tables in Malaysian Audited Financial Reports
di: Tan, Jin Khye, et al.
Pubblicazione: (2025)
di: Tan, Jin Khye, et al.
Pubblicazione: (2025)
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
di: Han, Jongwook, et al.
Pubblicazione: (2025)
di: Han, Jongwook, et al.
Pubblicazione: (2025)
DriveMRP: Enhancing Vision-Language Models with Synthetic Motion Data for Motion Risk Prediction
di: Hou, Zhiyi, et al.
Pubblicazione: (2025)
di: Hou, Zhiyi, et al.
Pubblicazione: (2025)
myMNIST: Benchmark of PETNN, KAN, and Classical Deep Learning Models for Burmese Handwritten Digit Recognition
di: Thu, Ye Kyaw, et al.
Pubblicazione: (2026)
di: Thu, Ye Kyaw, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Automating Autograding: Large Language Models as Test Suite Generators for Introductory Programming
di: Alkafaween, Umar, et al.
Pubblicazione: (2024) -
Evaluating LLMs for Career Guidance: Comparative Analysis of Computing Competency Recommendations Across Ten African Countries
di: Eze, Precious, et al.
Pubblicazione: (2025) -
Ensemble ToT of LLMs and Its Application to Automatic Grading System for Supporting Self-Learning
di: Ito, Yuki, et al.
Pubblicazione: (2025) -
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
di: Song, Dingjie, et al.
Pubblicazione: (2026) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)