Ace-CEFR -- A Dataset for Automated Evaluation of the Linguistic Difficulty of Conversational Texts for LLM Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Kogan, David, Schumacher, Max, Nguyen, Sam, Suzuki, Masanori, Smith, Melissa, Bellows, Chloe Sophia, Bernstein, Jared |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Retcon -- a Prompt-Based Technique for Precise Control of LLMs in Conversations
by: Kogan, David, et al.
Published: (2026)
by: Kogan, David, et al.
Published: (2026)
AceParse: A Comprehensive Dataset with Diverse Structured Texts for Academic Literature Parsing
by: Ji, Huawei, et al.
Published: (2024)
by: Ji, Huawei, et al.
Published: (2024)
Question Type, Cognitive Load, and CEFR Alignment: Evaluating LLM-Generated EFL Grammar Drill Exercises
by: Woollaston, Steve, et al.
Published: (2026)
by: Woollaston, Steve, et al.
Published: (2026)
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
by: Scaria, Nicy, et al.
Published: (2024)
by: Scaria, Nicy, et al.
Published: (2024)
UniversalCEFR: Enabling Open Multilingual Research on Language Proficiency Assessment
by: Imperial, Joseph Marvin, et al.
Published: (2025)
by: Imperial, Joseph Marvin, et al.
Published: (2025)
A CEFR-Inspired Classification Framework with Fuzzy C-Means To Automate Assessment of Programming Skills in Scratch
by: Hidalgo-Aragón, Ricardo, et al.
Published: (2026)
by: Hidalgo-Aragón, Ricardo, et al.
Published: (2026)
CEFR-Annotated WordNet: LLM-Based Proficiency-Guided Semantic Database for Language Learning
by: Kikuchi, Masato, et al.
Published: (2025)
by: Kikuchi, Masato, et al.
Published: (2025)
Linguistically Informed Multimodal Fusion for Vietnamese Scene-Text Image Captioning: Dataset, Graph Framework, and Phonological Attention
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
Testing the Relationship of Linguistic Complexity to Second Language Learners’ Comparative Judgment on Text Difficulty
by: Xiaopeng Zhang, et al.
Published: (2024)
by: Xiaopeng Zhang, et al.
Published: (2024)
Invariants: Computation and Applications
by: Kogan, Irina A.
Published: (2024)
by: Kogan, Irina A.
Published: (2024)
Controlling Language Difficulty in Dialogues with Linguistic Features
by: Xu, Shuyao, et al.
Published: (2025)
by: Xu, Shuyao, et al.
Published: (2025)
ShareChat: A Dataset of Chatbot Conversations in the Wild
by: Yan, Yueru, et al.
Published: (2025)
by: Yan, Yueru, et al.
Published: (2025)
Alignment Drift in CEFR-prompted LLMs for Interactive Spanish Tutoring
by: Almasi, Mina, et al.
Published: (2025)
by: Almasi, Mina, et al.
Published: (2025)
Reasoning Models Ace the CFA Exams
by: Patel, Jaisal, et al.
Published: (2025)
by: Patel, Jaisal, et al.
Published: (2025)
Efficiently Leveraging Linguistic Priors for Scene Text Spotting
by: Nguyen, Nguyen, et al.
Published: (2024)
by: Nguyen, Nguyen, et al.
Published: (2024)
Difficulties with Evaluating a Deception Detector for AIs
by: Smith, Lewis, et al.
Published: (2025)
by: Smith, Lewis, et al.
Published: (2025)
An Ace in the Hole for Soil Health and Nitrogen Tracking
by: Megan Sever
Published: (2024)
by: Megan Sever
Published: (2024)
mAceReason-Math: A Dataset of High-Quality Multilingual Math Problems Ready For RLVR
by: Dobler, Konstantin, et al.
Published: (2026)
by: Dobler, Konstantin, et al.
Published: (2026)
The Curious Decline of Linguistic Diversity: Training Language Models on Synthetic Text
by: Guo, Yanzhu, et al.
Published: (2023)
by: Guo, Yanzhu, et al.
Published: (2023)
Planificación local y diseño participativo en Chipilo, Puebla, México
by: Melissa Schumacher-González
Published: (2018)
by: Melissa Schumacher-González
Published: (2018)
AceWGS: An LLM-Aided Framework to Accelerate Catalyst Design for Water-Gas Shift Reactions
by: Chattoraj, Joyjit, et al.
Published: (2025)
by: Chattoraj, Joyjit, et al.
Published: (2025)
Analysing written production competence descriptors for academic and professional purposes and their calibration to the CEFR
by: Joana Pierce McMahon
Published: (2012)
by: Joana Pierce McMahon
Published: (2012)
A Lightweight Measure of Classification Difficulty from Application Dataset Characteristics
by: Cao, Bryan Bo, et al.
Published: (2024)
by: Cao, Bryan Bo, et al.
Published: (2024)
Kingfish Ace: Application of Gaming to Help Air Force Leaders Understand Agile Combat Employment
by: Troy B. Pierce, et al.
Published: (2024)
by: Troy B. Pierce, et al.
Published: (2024)
Scalable Hyperpolarized MRI Enabled by Ace‐SABRE of [1‐ 13 C]Pyruvate
by: Stephen J. McBride, et al.
Published: (2025)
by: Stephen J. McBride, et al.
Published: (2025)
Scalable Hyperpolarized MRI Enabled by Ace‐SABRE of [1‐ 13 C]Pyruvate
by: Stephen J. McBride, et al.
Published: (2025)
by: Stephen J. McBride, et al.
Published: (2025)
AceGPT, Localizing Large Language Models in Arabic
by: Huang, Huang, et al.
Published: (2023)
by: Huang, Huang, et al.
Published: (2023)
AceMap: Knowledge Discovery through Academic Graph
by: Wang, Xinbing, et al.
Published: (2024)
by: Wang, Xinbing, et al.
Published: (2024)
Revisiting Generalization Across Difficulty Levels: It's Not So Easy
by: Kordi, Yeganeh, et al.
Published: (2025)
by: Kordi, Yeganeh, et al.
Published: (2025)
A Judge-free LLM Open-ended Generation Benchmark Based on the Distributional Hypothesis
by: Imajo, Kentaro, et al.
Published: (2025)
by: Imajo, Kentaro, et al.
Published: (2025)
The Anatomy of Speech Persuasion: Linguistic Shifts in LLM-Modified Speeches
by: Barkar, Alisa, et al.
Published: (2025)
by: Barkar, Alisa, et al.
Published: (2025)
AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding
by: Oh, Gyutaek, et al.
Published: (2025)
by: Oh, Gyutaek, et al.
Published: (2025)
Identifying Hearing Difficulty Moments in Conversational Audio
by: Collins, Jack, et al.
Published: (2025)
by: Collins, Jack, et al.
Published: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
by: Muñoz-Ortiz, Alberto, et al.
Published: (2023)
by: Muñoz-Ortiz, Alberto, et al.
Published: (2023)
Telephone Surveys Meet Conversational AI: Evaluating a LLM-Based Telephone Survey System at Scale
by: Lang, Max M., et al.
Published: (2025)
by: Lang, Max M., et al.
Published: (2025)
AceTone: Bridging Words and Colors for Conditional Image Grading
by: Ma, Tianren, et al.
Published: (2026)
by: Ma, Tianren, et al.
Published: (2026)
Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution
by: Xiong, Feng, et al.
Published: (2026)
by: Xiong, Feng, et al.
Published: (2026)
Dataset Difficulty and the Role of Inductive Bias
by: Kwok, Devin, et al.
Published: (2024)
by: Kwok, Devin, et al.
Published: (2024)
Comparing International and National Language Assessments: Cambridge, Oxford, ÖSYM, and MEB within the CEFR Framework
by: Bilal BUDAK, et al.
Published: (2025)
by: Bilal BUDAK, et al.
Published: (2025)
Prompting ChatGPT for Chinese Learning as L2: A CEFR and EBCL Level Study
by: Lin-Zucker, Miao, et al.
Published: (2025)
by: Lin-Zucker, Miao, et al.
Published: (2025)
Similar Items
-
Retcon -- a Prompt-Based Technique for Precise Control of LLMs in Conversations
by: Kogan, David, et al.
Published: (2026) -
AceParse: A Comprehensive Dataset with Diverse Structured Texts for Academic Literature Parsing
by: Ji, Huawei, et al.
Published: (2024) -
Question Type, Cognitive Load, and CEFR Alignment: Evaluating LLM-Generated EFL Grammar Drill Exercises
by: Woollaston, Steve, et al.
Published: (2026) -
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
by: Scaria, Nicy, et al.
Published: (2024) -
UniversalCEFR: Enabling Open Multilingual Research on Language Proficiency Assessment
by: Imperial, Joseph Marvin, et al.
Published: (2025)