An Interpretable and Crosslingual Method for Evaluating Second-Language Dialogues
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Rena, Wu, Jingxuan, Wu, Xuetong, Roever, Carsten, Wu, Jing, Lv, Long, Lau, Jey Han |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Interaction Matters: An Evaluation Framework for Interactive Dialogue Assessment on English Second Language Conversations
por: Gao, Rena, et al.
Publicado: (2024)
por: Gao, Rena, et al.
Publicado: (2024)
Can LLMs Simulate L2-English Dialogue? An Information-Theoretic Analysis of L1-Dependent Biases
por: Gao, Rena, et al.
Publicado: (2025)
por: Gao, Rena, et al.
Publicado: (2025)
Moderation Matters:Measuring Conversational Moderation Impact in English as a Second Language Group Discussion
por: Gao, Rena, et al.
Publicado: (2025)
por: Gao, Rena, et al.
Publicado: (2025)
'No' Matters: Out-of-Distribution Detection in Multimodality Long Dialogue
por: Gao, Rena, et al.
Publicado: (2024)
por: Gao, Rena, et al.
Publicado: (2024)
Factual Dialogue Summarization via Learning from Large Language Models
por: Zhu, Rongxin, et al.
Publicado: (2024)
por: Zhu, Rongxin, et al.
Publicado: (2024)
Assessing Digital Interactional Competence for Second‐Language and First‐Language Chinese Speakers: Effects of Proficiency, Mode, and Setting
por: David Wei Dai, et al.
Publicado: (2026)
por: David Wei Dai, et al.
Publicado: (2026)
Evaluating Evidence Attribution in Generated Fact Checking Explanations
por: Xing, Rui, et al.
Publicado: (2024)
por: Xing, Rui, et al.
Publicado: (2024)
Beyond Seen Data: Improving KBQA Generalization Through Schema-Guided Logical Form Generation
por: Gao, Shengxiang, et al.
Publicado: (2025)
por: Gao, Shengxiang, et al.
Publicado: (2025)
CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics
por: Chen, Ming-Bin, et al.
Publicado: (2026)
por: Chen, Ming-Bin, et al.
Publicado: (2026)
MoDEM: Mixture of Domain Expert Models
por: Simonds, Toby, et al.
Publicado: (2024)
por: Simonds, Toby, et al.
Publicado: (2024)
Training and Evaluating with Human Label Variation: An Empirical Study
por: Kurniawan, Kemal, et al.
Publicado: (2025)
por: Kurniawan, Kemal, et al.
Publicado: (2025)
RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
por: Huang, Jing, et al.
Publicado: (2024)
por: Huang, Jing, et al.
Publicado: (2024)
CMA-R:Causal Mediation Analysis for Explaining Rumour Detection
por: Tian, Lin, et al.
Publicado: (2024)
por: Tian, Lin, et al.
Publicado: (2024)
A Sentiment Consolidation Framework for Meta-Review Generation
por: Li, Miao, et al.
Publicado: (2024)
por: Li, Miao, et al.
Publicado: (2024)
To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction
por: Kurniawan, Kemal, et al.
Publicado: (2024)
por: Kurniawan, Kemal, et al.
Publicado: (2024)
On the Interplay between Human Label Variation and Model Fairness
por: Kurniawan, Kemal, et al.
Publicado: (2025)
por: Kurniawan, Kemal, et al.
Publicado: (2025)
Who Wrote the Book? Detecting and Attributing LLM Ghostwriters
por: Shetty, Anudeex, et al.
Publicado: (2026)
por: Shetty, Anudeex, et al.
Publicado: (2026)
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
por: Shetty, Anudeex, et al.
Publicado: (2024)
por: Shetty, Anudeex, et al.
Publicado: (2024)
Context Volume Drives Performance: Tackling Domain Shift in Extremely Low-Resource Translation via RAG
por: Setiawan, David Samuel, et al.
Publicado: (2026)
por: Setiawan, David Samuel, et al.
Publicado: (2026)
WHoW: A Cross-domain Approach for Analysing Conversation Moderation
por: Chen, Ming-Bin, et al.
Publicado: (2024)
por: Chen, Ming-Bin, et al.
Publicado: (2024)
Crosslingual Optimized Metric for Translation Assessment of Indian Languages
por: Ahsan, Arafat, et al.
Publicado: (2025)
por: Ahsan, Arafat, et al.
Publicado: (2025)
Large Language Model based Situational Dialogues for Second Language Learning
por: Xu, Shuyao, et al.
Publicado: (2024)
por: Xu, Shuyao, et al.
Publicado: (2024)
Post-Training Language Models for Crosslingual Consistency
por: Liu, Tianyu, et al.
Publicado: (2026)
por: Liu, Tianyu, et al.
Publicado: (2026)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
por: Li, Miao, et al.
Publicado: (2025)
por: Li, Miao, et al.
Publicado: (2025)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
por: Xing, Rui, et al.
Publicado: (2025)
por: Xing, Rui, et al.
Publicado: (2025)
How Transliterations Improve Crosslingual Alignment
por: Liu, Yihong, et al.
Publicado: (2024)
por: Liu, Yihong, et al.
Publicado: (2024)
Explaining and Mitigating Crosslingual Tokenizer Inequities
por: Arnett, Catherine, et al.
Publicado: (2025)
por: Arnett, Catherine, et al.
Publicado: (2025)
On the Entity-Level Alignment in Crosslingual Consistency
por: Liu, Yihong, et al.
Publicado: (2025)
por: Liu, Yihong, et al.
Publicado: (2025)
Crosslingual Reasoning through Test-Time Scaling
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
Crosslingual Capabilities and Knowledge Barriers in Multilingual Large Language Models
por: Chua, Lynn, et al.
Publicado: (2024)
por: Chua, Lynn, et al.
Publicado: (2024)
Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space
por: Madani, Mohammad Reza Ghasemi, et al.
Publicado: (2026)
por: Madani, Mohammad Reza Ghasemi, et al.
Publicado: (2026)
Crosslingual On-Policy Self-Distillation for Multilingual Reasoning
por: Liu, Yihong, et al.
Publicado: (2026)
por: Liu, Yihong, et al.
Publicado: (2026)
Meta4XNLI: A Crosslingual Parallel Corpus for Metaphor Detection and Interpretation
por: Sanchez-Bayona, Elisa, et al.
Publicado: (2024)
por: Sanchez-Bayona, Elisa, et al.
Publicado: (2024)
Predicting Sentence Acceptability Judgments in Multimodal Contexts
por: Jang, Hyewon, et al.
Publicado: (2026)
por: Jang, Hyewon, et al.
Publicado: (2026)
Entropy2Vec: Crosslingual Language Modeling Entropy as End-to-End Learnable Language Representations
por: Irawan, Patrick Amadeus, et al.
Publicado: (2025)
por: Irawan, Patrick Amadeus, et al.
Publicado: (2025)
FLUKE: A Linguistically-Driven and Task-Agnostic Framework for Robustness Evaluation
por: Otmakhova, Yulia, et al.
Publicado: (2025)
por: Otmakhova, Yulia, et al.
Publicado: (2025)
Evaluating Large Language Models in Analysing Classroom Dialogue
por: Long, Yun, et al.
Publicado: (2024)
por: Long, Yun, et al.
Publicado: (2024)
Distilling Monolingual and Crosslingual Word-in-Context Representations
por: Arase, Yuki, et al.
Publicado: (2024)
por: Arase, Yuki, et al.
Publicado: (2024)
Controlling Language Difficulty in Dialogues with Linguistic Features
por: Xu, Shuyao, et al.
Publicado: (2025)
por: Xu, Shuyao, et al.
Publicado: (2025)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
por: Jiang, Yanbei, et al.
Publicado: (2026)
por: Jiang, Yanbei, et al.
Publicado: (2026)
Ejemplares similares
-
Interaction Matters: An Evaluation Framework for Interactive Dialogue Assessment on English Second Language Conversations
por: Gao, Rena, et al.
Publicado: (2024) -
Can LLMs Simulate L2-English Dialogue? An Information-Theoretic Analysis of L1-Dependent Biases
por: Gao, Rena, et al.
Publicado: (2025) -
Moderation Matters:Measuring Conversational Moderation Impact in English as a Second Language Group Discussion
por: Gao, Rena, et al.
Publicado: (2025) -
'No' Matters: Out-of-Distribution Detection in Multimodality Long Dialogue
por: Gao, Rena, et al.
Publicado: (2024) -
Factual Dialogue Summarization via Learning from Large Language Models
por: Zhu, Rongxin, et al.
Publicado: (2024)