SteLLA: A Structured Grading System Using LLMs with RAG
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Qiu, Hefei, White, Brian, Ding, Ashley, Costa, Reinaldo, Hachem, Ali, Ding, Wei, Chen, Ping |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Assessing GPT Performance in a Proof-Based University-Level Course Under Blind Grading
par: Ding, Ming, et autres
Publié: (2025)
par: Ding, Ming, et autres
Publié: (2025)
Urban Mobility Assessment Using LLMs
par: Bhandari, Prabin, et autres
Publié: (2024)
par: Bhandari, Prabin, et autres
Publié: (2024)
Answering Students' Questions on Course Forums Using Multiple Chain-of-Thought Reasoning and Finetuning RAG-Enabled LLM
par: Wang, Neo, et autres
Publié: (2025)
par: Wang, Neo, et autres
Publié: (2025)
"Amazing, They All Lean Left" -- Analyzing the Political Temperaments of Current LLMs
par: Neuman, W. Russell, et autres
Publié: (2025)
par: Neuman, W. Russell, et autres
Publié: (2025)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
par: Russo, Daniel, et autres
Publié: (2024)
par: Russo, Daniel, et autres
Publié: (2024)
Persuasion Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PD
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2025)
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2025)
"They parted illusions -- they parted disclaim marinade": Misalignment as structural fidelity in LLMs
par: Costa, Mariana Lins
Publié: (2025)
par: Costa, Mariana Lins
Publié: (2025)
Towards Automated Situation Awareness: A RAG-Based Framework for Peacebuilding Reports
par: Nemkova, Poli A., et autres
Publié: (2025)
par: Nemkova, Poli A., et autres
Publié: (2025)
Using LLMs for Knowledge Component-level Correctness Labeling in Open-ended Coding Problems
par: Duan, Zhangqi, et autres
Publié: (2026)
par: Duan, Zhangqi, et autres
Publié: (2026)
Explore the Potential of LLMs in Misinformation Detection: An Empirical Study
par: Chen, Mengyang, et autres
Publié: (2023)
par: Chen, Mengyang, et autres
Publié: (2023)
Using LLMs to create analytical datasets: A case study of reconstructing the historical memory of Colombia
par: Anderson, David, et autres
Publié: (2025)
par: Anderson, David, et autres
Publié: (2025)
The Statistical Signature of LLMs
par: Hadad, Ortal, et autres
Publié: (2026)
par: Hadad, Ortal, et autres
Publié: (2026)
EulerESG: Automating ESG Disclosure Analysis with LLMs
par: Ding, Yi, et autres
Publié: (2025)
par: Ding, Yi, et autres
Publié: (2025)
LLMs for Low-Resource Dialect Translation Using Context-Aware Prompting: A Case Study on Sylheti
par: Prama, Tabia Tanzin, et autres
Publié: (2025)
par: Prama, Tabia Tanzin, et autres
Publié: (2025)
The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
par: Xu, Rongwu, et autres
Publié: (2023)
par: Xu, Rongwu, et autres
Publié: (2023)
A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extraction
par: Lincoln, Nicole, et autres
Publié: (2026)
par: Lincoln, Nicole, et autres
Publié: (2026)
Counterfactual LLM-based Framework for Measuring Rhetorical Style
par: Qiu, Jingyi, et autres
Publié: (2025)
par: Qiu, Jingyi, et autres
Publié: (2025)
Survival at Any Cost? LLMs and the Choice Between Self-Preservation and Human Harm
par: Mohamadi, Alireza, et autres
Publié: (2025)
par: Mohamadi, Alireza, et autres
Publié: (2025)
Which Type of Students can LLMs Act? Investigating Authentic Simulation with Graph-based Human-AI Collaborative System
par: Li, Haoxuan, et autres
Publié: (2025)
par: Li, Haoxuan, et autres
Publié: (2025)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
par: Ding, Junchen, et autres
Publié: (2025)
par: Ding, Junchen, et autres
Publié: (2025)
Benchmarking LLMs for Political Science: A United Nations Perspective
par: Liang, Yueqing, et autres
Publié: (2025)
par: Liang, Yueqing, et autres
Publié: (2025)
Adaptive Learning Systems: Personalized Curriculum Design Using LLM-Powered Analytics
par: Li, Yongjie, et autres
Publié: (2025)
par: Li, Yongjie, et autres
Publié: (2025)
Expressing Social Emotions: Misalignment Between LLMs and Human Cultural Emotion Norms
par: Bhattacharyya, Sree, et autres
Publié: (2026)
par: Bhattacharyya, Sree, et autres
Publié: (2026)
From Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents
par: Mou, Xinyi, et autres
Publié: (2024)
par: Mou, Xinyi, et autres
Publié: (2024)
Evaluating GPT-4 at Grading Handwritten Solutions in Math Exams
par: Caraeni, Adriana, et autres
Publié: (2024)
par: Caraeni, Adriana, et autres
Publié: (2024)
Classroom AI: Large Language Models as Grade-Specific Teachers
par: Oh, Jio, et autres
Publié: (2026)
par: Oh, Jio, et autres
Publié: (2026)
ArzEn-LLM: Code-Switched Egyptian Arabic-English Translation and Speech Recognition Using LLMs
par: Heakl, Ahmed, et autres
Publié: (2024)
par: Heakl, Ahmed, et autres
Publié: (2024)
Generative AI Advertising as a Problem of Trustworthy Commercial Intervention
par: Qiu, Jingyi, et autres
Publié: (2026)
par: Qiu, Jingyi, et autres
Publié: (2026)
None of the Above, Less of the Right: Parallel Patterns between Humans and LLMs on Multi-Choice Questions Answering
par: Tam, Zhi Rui, et autres
Publié: (2025)
par: Tam, Zhi Rui, et autres
Publié: (2025)
Academically intelligent LLMs are not necessarily socially intelligent
par: Xu, Ruoxi, et autres
Publié: (2024)
par: Xu, Ruoxi, et autres
Publié: (2024)
The Thin Line Between Comprehension and Persuasion in LLMs
par: de Wynter, Adrian, et autres
Publié: (2025)
par: de Wynter, Adrian, et autres
Publié: (2025)
LLMs Provide Unstable Answers to Legal Questions
par: Blair-Stanek, Andrew, et autres
Publié: (2025)
par: Blair-Stanek, Andrew, et autres
Publié: (2025)
On the Diminishing Returns of Complex Robust RAG Training in the Era of Powerful LLMs
par: Ding, Hanxing, et autres
Publié: (2025)
par: Ding, Hanxing, et autres
Publié: (2025)
Meaning Is Not A Metric: Using LLMs to make cultural context legible at scale
par: Kommers, Cody, et autres
Publié: (2025)
par: Kommers, Cody, et autres
Publié: (2025)
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
par: Zhao, Chengshuai, et autres
Publié: (2026)
par: Zhao, Chengshuai, et autres
Publié: (2026)
Unsupervised Concept Vector Extraction for Bias Control in LLMs
par: Cyberey, Hannah, et autres
Publié: (2025)
par: Cyberey, Hannah, et autres
Publié: (2025)
Surface Reading LLMs: Synthetic Text and its Styles
par: Bajohr, Hannes
Publié: (2025)
par: Bajohr, Hannes
Publié: (2025)
Hidden Persuaders: LLMs' Political Leaning and Their Influence on Voters
par: Potter, Yujin, et autres
Publié: (2024)
par: Potter, Yujin, et autres
Publié: (2024)
Hate Personified: Investigating the role of LLMs in content moderation
par: Masud, Sarah, et autres
Publié: (2024)
par: Masud, Sarah, et autres
Publié: (2024)
Promotional Language and the Adoption of Innovative Ideas in Science
par: Peng, Hao, et autres
Publié: (2024)
par: Peng, Hao, et autres
Publié: (2024)
Documents similaires
-
Assessing GPT Performance in a Proof-Based University-Level Course Under Blind Grading
par: Ding, Ming, et autres
Publié: (2025) -
Urban Mobility Assessment Using LLMs
par: Bhandari, Prabin, et autres
Publié: (2024) -
Answering Students' Questions on Course Forums Using Multiple Chain-of-Thought Reasoning and Finetuning RAG-Enabled LLM
par: Wang, Neo, et autres
Publié: (2025) -
"Amazing, They All Lean Left" -- Analyzing the Political Temperaments of Current LLMs
par: Neuman, W. Russell, et autres
Publié: (2025) -
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
par: Russo, Daniel, et autres
Publié: (2024)