Decision-Level Ordinal Modeling for Multimodal Essay Scoring with Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Han, Su, Jiamin, liu, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
CAFES: A Collaborative Multi-Agent Framework for Multi-Granular Multimodal Essay Scoring
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025)
von: Cai, Yida, et al.
Veröffentlicht: (2025)
Can Large Language Models Automatically Score Proficiency of Written Essays?
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
von: Gao, Fan, et al.
Veröffentlicht: (2025)
von: Gao, Fan, et al.
Veröffentlicht: (2025)
LCES: Zero-shot Automated Essay Scoring via Pairwise Comparisons Using Large Language Models
von: Shibata, Takumi, et al.
Veröffentlicht: (2025)
von: Shibata, Takumi, et al.
Veröffentlicht: (2025)
GLoRE: Evaluating Logical Reasoning of Large Language Models
von: liu, Hanmeng, et al.
Veröffentlicht: (2023)
von: liu, Hanmeng, et al.
Veröffentlicht: (2023)
Are Large Language Models Good Essay Graders?
von: Kundu, Anindita, et al.
Veröffentlicht: (2024)
von: Kundu, Anindita, et al.
Veröffentlicht: (2024)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
Autoregressive Score Generation for Multi-trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
Towards Robust Instruction Tuning on Multimodal Large Language Models
von: Han, Wei, et al.
Veröffentlicht: (2024)
von: Han, Wei, et al.
Veröffentlicht: (2024)
Enhancing Marker Scoring Accuracy through Ordinal Confidence Modelling in Educational Assessments
von: Chakravarty, Abhirup, et al.
Veröffentlicht: (2025)
von: Chakravarty, Abhirup, et al.
Veröffentlicht: (2025)
Levels of Analysis for Large Language Models
von: Ku, Alexander Y., et al.
Veröffentlicht: (2025)
von: Ku, Alexander Y., et al.
Veröffentlicht: (2025)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
LAMPO: Large Language Models as Preference Machines for Few-shot Ordinal Classification
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
SGM: Safety Glasses for Multimodal Large Language Models via Neuron-Level Detoxification
von: Wang, Hongbo, et al.
Veröffentlicht: (2025)
von: Wang, Hongbo, et al.
Veröffentlicht: (2025)
Cross-Modal Consistency in Multimodal Large Language Models
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
Exploring and Evaluating Multimodal Knowledge Reasoning Consistency of Multimodal Large Language Models
von: Jia, Boyu, et al.
Veröffentlicht: (2025)
von: Jia, Boyu, et al.
Veröffentlicht: (2025)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
Teach-to-Reason with Scoring: Self-Explainable Rationale-Driven Multi-Trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
Unveiling Trust in Multimodal Large Language Models: Evaluation, Analysis, and Mitigation
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
Enhancing Decision-Making of Large Language Models via Actor-Critic
von: Dong, Heng, et al.
Veröffentlicht: (2025)
von: Dong, Heng, et al.
Veröffentlicht: (2025)
Applying Large Language Models and Chain-of-Thought for Automatic Scoring
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
On the Decision-Making Abilities in Role-Playing using Large Language Models
von: Shen, Chenglei, et al.
Veröffentlicht: (2024)
von: Shen, Chenglei, et al.
Veröffentlicht: (2024)
ChatSR: Multimodal Large Language Models for Scientific Formula Discovery
von: Li, Yanjie, et al.
Veröffentlicht: (2024)
von: Li, Yanjie, et al.
Veröffentlicht: (2024)
Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
Autoregressive Multi-trait Essay Scoring via Reinforcement Learning with Scoring-aware Multiple Rewards
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
InfiR : Crafting Effective Small Language Models and Multimodal Small Language Models in Reasoning
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
Human-AI Collaborative Essay Scoring: A Dual-Process Framework with LLMs
von: Xiao, Changrong, et al.
Veröffentlicht: (2024)
von: Xiao, Changrong, et al.
Veröffentlicht: (2024)
Measuring Meaning Composition in the Human Brain with Composition Scores from Large Language Models
von: Gao, Changjiang, et al.
Veröffentlicht: (2024)
von: Gao, Changjiang, et al.
Veröffentlicht: (2024)
Phrase-Level Adversarial Training for Mitigating Bias in Neural Network-based Automatic Essay Scoring
von: Philip, Haddad, et al.
Veröffentlicht: (2024)
von: Philip, Haddad, et al.
Veröffentlicht: (2024)
Cross-modal Information Flow in Multimodal Large Language Models
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
R-Capsule: Compressing High-Level Plans for Efficient Large Language Model Reasoning
von: Shan, Hongyu, et al.
Veröffentlicht: (2025)
von: Shan, Hongyu, et al.
Veröffentlicht: (2025)
On Fairness of Unified Multimodal Large Language Model for Image Generation
von: Liu, Ming, et al.
Veröffentlicht: (2025)
von: Liu, Ming, et al.
Veröffentlicht: (2025)
DIM: Dynamic Integration of Multimodal Entity Linking with Large Language Model
von: Song, Shezheng, et al.
Veröffentlicht: (2024)
von: Song, Shezheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
von: Su, Jiamin, et al.
Veröffentlicht: (2025) -
CAFES: A Collaborative Multi-Agent Framework for Multi-Granular Multimodal Essay Scoring
von: Su, Jiamin, et al.
Veröffentlicht: (2025) -
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026) -
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025) -
Can Large Language Models Automatically Score Proficiency of Written Essays?
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)