Codebook-Injected Dialogue Segmentation for Multi-Utterance Constructs Annotation: LLM-Assisted and Gold-Label-Free Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Jinsook, Vanacore, Kirk, Zhou, Zhuqian, Ahtisham, Bakhtawar, Grutter, Jeanine, Kizilcec, Rene F. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
LLM Reasoning Predicts When Models Are Right: Evidence from Coding Classroom Discourse
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
Optimizing LLM Annotation of Classroom Discourse through Multi-Agent Orchestration
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset
by: Zhou, Zhuqian, et al.
Published: (2026)
by: Zhou, Zhuqian, et al.
Published: (2026)
Sandpiper: Orchestrated AI-Annotation for Educational Discourse at Scale
by: Hedley, Daryl, et al.
Published: (2026)
by: Hedley, Daryl, et al.
Published: (2026)
Tutor Move Taxonomy: A Theory-Aligned Framework for Analyzing Instructional Moves in Tutoring
by: Zhou, Zhuqian, et al.
Published: (2026)
by: Zhou, Zhuqian, et al.
Published: (2026)
Million Tutoring Moves (MTM): An Open Multimodal Dataset for the Science of Tutoring
by: Kizilcec, René, et al.
Published: (2026)
by: Kizilcec, René, et al.
Published: (2026)
How well do Large Language Models Recognize Instructional Moves? Establishing Baselines for Foundation Models in Educational Discourse
by: Vanacore, Kirk, et al.
Published: (2025)
by: Vanacore, Kirk, et al.
Published: (2025)
Poor Alignment and Steerability of Large Language Models: Evidence from College Admission Essays
by: Lee, Jinsook, et al.
Published: (2025)
by: Lee, Jinsook, et al.
Published: (2025)
Does the TalkMoves Codebook Generalize to One-on-One Tutoring and Multimodal Interaction?
by: Focsan, Corina Luca, et al.
Published: (2026)
by: Focsan, Corina Luca, et al.
Published: (2026)
The Digital Divide in Generative AI: Evidence from Large Language Model Use in College Admissions Essays
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
The Life Cycle of Large Language Models: A Review of Biases in Education
by: Lee, Jinsook, et al.
Published: (2024)
by: Lee, Jinsook, et al.
Published: (2024)
The life cycle of large language models in education: A framework for understanding sources of bias
by: Jinsook Lee, et al.
Published: (2024)
by: Jinsook Lee, et al.
Published: (2024)
Modernizing Ground Truth: Four Shifts Toward Improving Reliability and Validity in AI in Education
by: Thomas, Danielle R., et al.
Published: (2026)
by: Thomas, Danielle R., et al.
Published: (2026)
Algorithms for College Admissions Decision Support: Impacts of Policy Change and Inherent Variability
by: Lee, Jinsook, et al.
Published: (2024)
by: Lee, Jinsook, et al.
Published: (2024)
Shiksha Copilot: Teacher-AI Collaboration for Curating and Customizing Lesson Plans in Low-Resource Schools
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
LUCID: LLM-Generated Utterances for Complex and Interesting Dialogues
by: Stacey, Joe, et al.
Published: (2024)
by: Stacey, Joe, et al.
Published: (2024)
Corporate cash holdings and industry risk
by: Jinsook Lee
Published: (2024)
by: Jinsook Lee
Published: (2024)
An Unsupervised Dialogue Topic Segmentation Model Based on Utterance Rewriting
by: Hou, Xia, et al.
Published: (2024)
by: Hou, Xia, et al.
Published: (2024)
Can Machines Learn the True Probabilities?
by: Kim, Jinsook
Published: (2024)
by: Kim, Jinsook
Published: (2024)
Learning LLM Preference over Intra-Dialogue Pairs: A Framework for Utterance-level Understandings
by: Liu, Xuanqing, et al.
Published: (2025)
by: Liu, Xuanqing, et al.
Published: (2025)
Multi-Gas analysis of ambient air using FTIR spectroscopy over Mexico City
by: Michel. Grutter
Published: (2003)
by: Michel. Grutter
Published: (2003)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
by: Park, Jeiyoon, et al.
Published: (2022)
by: Park, Jeiyoon, et al.
Published: (2022)
Student Engagement in AI Assisted Complex Problem Solving: A Pilot Study of Human AI Rubik's Cube Collaboration
by: Vanacore, Kirk, et al.
Published: (2025)
by: Vanacore, Kirk, et al.
Published: (2025)
Multi-Utterance Speech Separation and Association Trained on Short Segments
by: Wang, Yuzhu, et al.
Published: (2025)
by: Wang, Yuzhu, et al.
Published: (2025)
PicPersona-TOD : A Dataset for Personalizing Utterance Style in Task-Oriented Dialogue with Image Persona
by: Lee, Jihyun, et al.
Published: (2025)
by: Lee, Jihyun, et al.
Published: (2025)
LLM-Driven Preference Data Synthesis for Proactive Prediction of the Next User Utterance in Human-Machine Dialogue
by: Wang, Jinqiang, et al.
Published: (2025)
by: Wang, Jinqiang, et al.
Published: (2025)
Beyond ‘statistical significance’: A nontechnical primer of Bayesian statistics and Bayes factors for health researchers
by: Ahtisham Younas
Published: (2024)
by: Ahtisham Younas
Published: (2024)
Unpacking Policy Implementation Science: What Is It and Why Does It Matter in Nursing?
by: Ahtisham Younas
Published: (2025)
by: Ahtisham Younas
Published: (2025)
Intersectional Praxis for Addressing Health Inequalities and Promoting Population Health
by: Ahtisham Younas
Published: (2025)
by: Ahtisham Younas
Published: (2025)
Improving Dialogue Discourse Parsing through Discourse-aware Utterance Clarification
by: Fan, Yaxin, et al.
Published: (2025)
by: Fan, Yaxin, et al.
Published: (2025)
Where LLM Annotators Fail: Label-Free Learning on Graphs with LLMs
by: Thapaliya, Safal, et al.
Published: (2026)
by: Thapaliya, Safal, et al.
Published: (2026)
Can You Tell It's AI? Human Perception of Synthetic Voices in Vishing Scenarios
by: Bhatti, Zoha Hayat, et al.
Published: (2026)
by: Bhatti, Zoha Hayat, et al.
Published: (2026)
Injecting linguistic knowledge into BERT for Dialogue State Tracking
by: Feng, Xiaohan, et al.
Published: (2023)
by: Feng, Xiaohan, et al.
Published: (2023)
GEEPERs: Principal Stratification using Principal Scores and Stacked Estimating Equations
by: Sales, Adam C., et al.
Published: (2022)
by: Sales, Adam C., et al.
Published: (2022)
The Path to Conversational AI Tutors: Integrating Tutoring Best Practices and Targeted Technologies to Produce Scalable AI Agents
by: Vanacore, Kirk, et al.
Published: (2026)
by: Vanacore, Kirk, et al.
Published: (2026)
Does Algorithmic Uncertainty Sway Human Experts? Evidence from a Field Experiment in Selective College Admissions
by: Lee, Hansol, et al.
Published: (2026)
by: Lee, Hansol, et al.
Published: (2026)
CapTalk: Unified Voice Design for Single-Utterance and Dialogue Speech Generation
by: Su, Xiaosu, et al.
Published: (2026)
by: Su, Xiaosu, et al.
Published: (2026)
A Theory of Machine Learning
by: Kim, Jinsook, et al.
Published: (2024)
by: Kim, Jinsook, et al.
Published: (2024)
Similar Items
-
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
by: Lee, Jinsook, et al.
Published: (2026) -
AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
by: Ahtisham, Bakhtawar, et al.
Published: (2025) -
LLM Reasoning Predicts When Models Are Right: Evidence from Coding Classroom Discourse
by: Ahtisham, Bakhtawar, et al.
Published: (2026) -
Optimizing LLM Annotation of Classroom Discourse through Multi-Agent Orchestration
by: Ahtisham, Bakhtawar, et al.
Published: (2026) -
Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset
by: Zhou, Zhuqian, et al.
Published: (2026)