Alignment Drift in CEFR-prompted LLMs for Interactive Spanish Tutoring
Fuente:
arXiv
Saved in:
| Main Authors: | Almasi, Mina, Kristensen-McLachlan, Ross Deans |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
I only read it for the plot! Maturity Ratings Affect Fanfiction Style and Community Engagement
by: Jacobsen, Mia, et al.
Published: (2025)
by: Jacobsen, Mia, et al.
Published: (2025)
Multilingual Embedding Probes Fail to Generalize Across Learner Corpora
by: Lyngbaek, Laurits, et al.
Published: (2026)
by: Lyngbaek, Laurits, et al.
Published: (2026)
Science is Exploration: Computational Frontiers for Conceptual Metaphor Theory
by: Hicke, Rebecca M. M., et al.
Published: (2024)
by: Hicke, Rebecca M. M., et al.
Published: (2024)
Context is Key(NMF): Modelling Topical Information Dynamics in Chinese Diaspora Media
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2024)
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2024)
Are Chatbots Reliable Text Annotators? Sometimes
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)
Says Who? Effective Zero-Shot Annotation of Focalization
by: Hicke, Rebecca M. M., et al.
Published: (2024)
by: Hicke, Rebecca M. M., et al.
Published: (2024)
Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries
by: Hicke, Rebecca M. M., et al.
Published: (2026)
by: Hicke, Rebecca M. M., et al.
Published: (2026)
UniversalCEFR: Enabling Open Multilingual Research on Language Proficiency Assessment
by: Imperial, Joseph Marvin, et al.
Published: (2025)
by: Imperial, Joseph Marvin, et al.
Published: (2025)
CEFR-Annotated WordNet: LLM-Based Proficiency-Guided Semantic Database for Language Learning
by: Kikuchi, Masato, et al.
Published: (2025)
by: Kikuchi, Masato, et al.
Published: (2025)
LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring
by: Lee, Unggi, et al.
Published: (2026)
by: Lee, Unggi, et al.
Published: (2026)
Prompting ChatGPT for Chinese Learning as L2: A CEFR and EBCL Level Study
by: Lin-Zucker, Miao, et al.
Published: (2025)
by: Lin-Zucker, Miao, et al.
Published: (2025)
Ace-CEFR -- A Dataset for Automated Evaluation of the Linguistic Difficulty of Conversational Texts for LLM Applications
by: Kogan, David, et al.
Published: (2025)
by: Kogan, David, et al.
Published: (2025)
Efficient multi-prompt evaluation of LLMs
by: Polo, Felipe Maia, et al.
Published: (2024)
by: Polo, Felipe Maia, et al.
Published: (2024)
Exploring LLMs for Predicting Tutor Strategy and Student Outcomes in Dialogues
by: Ikram, Fareya, et al.
Published: (2025)
by: Ikram, Fareya, et al.
Published: (2025)
Demystifying optimized prompts in language models
by: Melamed, Rimon, et al.
Published: (2025)
by: Melamed, Rimon, et al.
Published: (2025)
Beyond prompt brittleness: Evaluating the reliability and consistency of political worldviews in LLMs
by: Ceron, Tanise, et al.
Published: (2024)
by: Ceron, Tanise, et al.
Published: (2024)
The meaning of prompts and the prompts of meaning: Semiotic reflections and modelling
by: Thellefsen, Martin, et al.
Published: (2025)
by: Thellefsen, Martin, et al.
Published: (2025)
Towards interpretable models for language proficiency assessment: Predicting the CEFR level of Estonian learner texts
by: Allkivi, Kais
Published: (2026)
by: Allkivi, Kais
Published: (2026)
Training Turn-by-Turn Verifiers for Dialogue Tutoring Agents: The Curious Case of LLMs as Your Coding Tutors
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Position: LLMs Can be Good Tutors in English Education
by: Ye, Jingheng, et al.
Published: (2025)
by: Ye, Jingheng, et al.
Published: (2025)
Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework
by: Yao, Xintong
Published: (2026)
by: Yao, Xintong
Published: (2026)
Few-shot clinical entity recognition in English, French and Spanish: masked language models outperform generative model prompting
by: Naguib, Marco, et al.
Published: (2024)
by: Naguib, Marco, et al.
Published: (2024)
Exploring Knowledge Tracing in Tutor-Student Dialogues using LLMs
by: Scarlatos, Alexander, et al.
Published: (2024)
by: Scarlatos, Alexander, et al.
Published: (2024)
Drift: Decoding-time Personalized Alignments with Implicit User Preferences
by: Kim, Minbeom, et al.
Published: (2025)
by: Kim, Minbeom, et al.
Published: (2025)
SafeTutors: Benchmarking Pedagogical Safety in AI Tutoring Systems
by: Hazra, Rima, et al.
Published: (2026)
by: Hazra, Rima, et al.
Published: (2026)
What's in a prompt? Language models encode literary style in prompt embeddings
by: Sarfati, Raphaël, et al.
Published: (2025)
by: Sarfati, Raphaël, et al.
Published: (2025)
SSLfmm: An R Package for Semi-Supervised Learning with a Mixed-Missingness Mechanism in Finite Mixture Models
by: McLachlan, Geoffrey J., et al.
Published: (2025)
by: McLachlan, Geoffrey J., et al.
Published: (2025)
Using LLMs as prompt modifier to avoid biases in AI image generators
by: Peinl, René
Published: (2025)
by: Peinl, René
Published: (2025)
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
by: Thomas, Danielle R., et al.
Published: (2025)
by: Thomas, Danielle R., et al.
Published: (2025)
DeepTutor: Towards Agentic Personalized Tutoring
by: Zhao, Bingxi, et al.
Published: (2026)
by: Zhao, Bingxi, et al.
Published: (2026)
Alignment Drift in Multimodal LLMs: A Two-Phase, Longitudinal Evaluation of Harm Across Eight Model Releases
by: Ford, Casey, et al.
Published: (2026)
by: Ford, Casey, et al.
Published: (2026)
ClinTutor-R1: Advancing Scalable and Robust One-to-Many Alignment in Clinical Socratic Education
by: He, Zhitao, et al.
Published: (2025)
by: He, Zhitao, et al.
Published: (2025)
An asiatic Psychopsis (Ps. birmana, n. sp.)
by: McLachlan, Robert
Published: (1891)
by: McLachlan, Robert
Published: (1891)
A remarkable new mimetic species of Mantispa from Borneo
by: McLachlan, Robert
Published: (1900)
by: McLachlan, Robert
Published: (1900)
On the genus Meleoma A. Fitch
by: McLachlan, Robert
Published: (1896)
by: McLachlan, Robert
Published: (1896)
Neuroptera observed in the Channel Islands in September 1891
by: McLachlan, Robert
Published: (1892)
by: McLachlan, Robert
Published: (1892)
From Solver to Tutor: Evaluating the Pedagogical Intelligence of LLMs with KMP-Bench
by: Shi, Weikang, et al.
Published: (2026)
by: Shi, Weikang, et al.
Published: (2026)
Developing a Tutoring Dialog Dataset to Optimize LLMs for Educational Use
by: Fateen, Menna, et al.
Published: (2024)
by: Fateen, Menna, et al.
Published: (2024)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
by: Kim, Tae Soo, et al.
Published: (2025)
by: Kim, Tae Soo, et al.
Published: (2025)
Why is prompting hard? Understanding prompts on binary sequence predictors
by: Wenliang, Li Kevin, et al.
Published: (2025)
by: Wenliang, Li Kevin, et al.
Published: (2025)
Similar Items
-
I only read it for the plot! Maturity Ratings Affect Fanfiction Style and Community Engagement
by: Jacobsen, Mia, et al.
Published: (2025) -
Multilingual Embedding Probes Fail to Generalize Across Learner Corpora
by: Lyngbaek, Laurits, et al.
Published: (2026) -
Science is Exploration: Computational Frontiers for Conceptual Metaphor Theory
by: Hicke, Rebecca M. M., et al.
Published: (2024) -
Context is Key(NMF): Modelling Topical Information Dynamics in Chinese Diaspora Media
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2024) -
Are Chatbots Reliable Text Annotators? Sometimes
by: Kristensen-McLachlan, Ross Deans, et al.
Published: (2023)