Rethinking Response Evaluation from Interlocutor's Eye for Open-Domain Dialogue Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Tsuta, Yuma, Yoshinaga, Naoki, Sato, Shoetsu, Toyoda, Masashi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Tracing Multilingual Knowledge Acquisition Dynamics in Domain Adaptation: A Case Study of English-Japanese Biomedical Adaptation
por: Zhao, Xin, et al.
Publicado: (2025)
por: Zhao, Xin, et al.
Publicado: (2025)
When Harry Meets Superman: The Role of The Interlocutor in Persona-Based Dialogue Generation
por: Occhipinti, Daniela, et al.
Publicado: (2025)
por: Occhipinti, Daniela, et al.
Publicado: (2025)
Commentary Generation from Data Records of Multiplayer Strategy Esports Game
por: Wang, Zihan, et al.
Publicado: (2022)
por: Wang, Zihan, et al.
Publicado: (2022)
On the Benchmarking of LLMs for Open-Domain Dialogue Evaluation
por: Mendonça, John, et al.
Publicado: (2024)
por: Mendonça, John, et al.
Publicado: (2024)
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs
por: Mendonça, John, et al.
Publicado: (2024)
por: Mendonça, John, et al.
Publicado: (2024)
CausalScore: An Automatic Reference-Free Metric for Assessing Response Relevance in Open-Domain Dialogue Systems
por: Feng, Tao, et al.
Publicado: (2024)
por: Feng, Tao, et al.
Publicado: (2024)
What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
Is He Extroverted? Identifying Missing Relevant Personas for Faithful User Simulation
por: Su, Weiwen, et al.
Publicado: (2026)
por: Su, Weiwen, et al.
Publicado: (2026)
MEDAL: A Framework for Benchmarking LLMs as Multilingual Open-Domain Dialogue Evaluators
por: Mendonça, John, et al.
Publicado: (2025)
por: Mendonça, John, et al.
Publicado: (2025)
A Large Collection of Model-generated Contradictory Responses for Consistency-aware Dialogue Systems
por: Sato, Shiki, et al.
Publicado: (2024)
por: Sato, Shiki, et al.
Publicado: (2024)
Rethinking the Evaluation of Dialogue Systems: Effects of User Feedback on Crowdworkers and LLMs
por: Siro, Clemencia, et al.
Publicado: (2024)
por: Siro, Clemencia, et al.
Publicado: (2024)
Neuron Empirical Gradient: Discovering and Quantifying Neurons Global Linear Controllability
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation
por: Qiu, Huachuan, et al.
Publicado: (2023)
por: Qiu, Huachuan, et al.
Publicado: (2023)
SLIDE: A Framework Integrating Small and Large Language Models for Open-Domain Dialogues Evaluation
por: Zhao, Kun, et al.
Publicado: (2024)
por: Zhao, Kun, et al.
Publicado: (2024)
Emphasising Structured Information: Integrating Abstract Meaning Representation into LLMs for Enhanced Open-Domain Dialogue Evaluation
por: Yang, Bohao, et al.
Publicado: (2024)
por: Yang, Bohao, et al.
Publicado: (2024)
DRE: An Effective Dual-Refined Method for Integrating Small and Large Language Models in Open-Domain Dialogue Evaluation
por: Zhao, Kun, et al.
Publicado: (2025)
por: Zhao, Kun, et al.
Publicado: (2025)
Rethinking Evaluation in Retrieval-Augmented Personalized Dialogue: A Cognitive and Linguistic Perspective
por: Zhang, Tianyi, et al.
Publicado: (2026)
por: Zhang, Tianyi, et al.
Publicado: (2026)
Facilitating Pornographic Text Detection for Open-Domain Dialogue Systems via Knowledge Distillation of Large Language Models
por: Qiu, Huachuan, et al.
Publicado: (2024)
por: Qiu, Huachuan, et al.
Publicado: (2024)
Dynamic Stochastic Decoding Strategy for Open-Domain Dialogue Generation
por: Li, Yiwei, et al.
Publicado: (2024)
por: Li, Yiwei, et al.
Publicado: (2024)
Modeling the One-to-Many Property in Open-Domain Dialogue with LLMs
por: Lee, Jing Yang, et al.
Publicado: (2025)
por: Lee, Jing Yang, et al.
Publicado: (2025)
PairEval: Open-domain Dialogue Evaluation with Pairwise Comparison
por: Park, ChaeHun, et al.
Publicado: (2024)
por: Park, ChaeHun, et al.
Publicado: (2024)
SumRec: A Framework for Recommendation using Open-Domain Dialogue
por: Asahara, Ryutaro, et al.
Publicado: (2024)
por: Asahara, Ryutaro, et al.
Publicado: (2024)
Multi-Dimensional Prompt Chaining to Improve Open-Domain Dialogue Generation
por: Teng, Livia Leong Hui
Publicado: (2026)
por: Teng, Livia Leong Hui
Publicado: (2026)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
por: Choi, Younwoo, et al.
Publicado: (2025)
por: Choi, Younwoo, et al.
Publicado: (2025)
SHARE: Shared Memory-Aware Open-Domain Long-Term Dialogue Dataset Constructed from Movie Script
por: Kim, Eunwon, et al.
Publicado: (2024)
por: Kim, Eunwon, et al.
Publicado: (2024)
How Stylistic Similarity Shapes Preferences in Dialogue Dataset with User and Third Party Evaluations
por: Numaya, Ikumi, et al.
Publicado: (2025)
por: Numaya, Ikumi, et al.
Publicado: (2025)
RATFM: Retrieval-augmented Time Series Foundation Model for Anomaly Detection
por: Maru, Chihiro, et al.
Publicado: (2025)
por: Maru, Chihiro, et al.
Publicado: (2025)
Controllable and Diverse Data Augmentation with Large Language Model for Low-Resource Open-Domain Dialogue Generation
por: Liu, Zhenhua, et al.
Publicado: (2024)
por: Liu, Zhenhua, et al.
Publicado: (2024)
From What to Respond to When to Respond: Timely Response Generation for Open-domain Dialogue Agents
por: Jang, Seongbo, et al.
Publicado: (2025)
por: Jang, Seongbo, et al.
Publicado: (2025)
Rethinking Loss Functions for Fact Verification
por: Mukobara, Yuta, et al.
Publicado: (2024)
por: Mukobara, Yuta, et al.
Publicado: (2024)
Multi-Faceted Evaluation of Tool-Augmented Dialogue Systems
por: Hou, Zhaoyi Joey, et al.
Publicado: (2025)
por: Hou, Zhaoyi Joey, et al.
Publicado: (2025)
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems
por: Vasselli, Justin, et al.
Publicado: (2025)
por: Vasselli, Justin, et al.
Publicado: (2025)
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
por: Wang, Hongru, et al.
Publicado: (2025)
por: Wang, Hongru, et al.
Publicado: (2025)
Leveraging AI Graders for Missing Score Imputation to Achieve Accurate Ability Estimation in Constructed-Response Tests
por: Uto, Masaki, et al.
Publicado: (2025)
por: Uto, Masaki, et al.
Publicado: (2025)
Interpreting Answers to Yes-No Questions in Dialogues from Multiple Domains
por: Wang, Zijie, et al.
Publicado: (2024)
por: Wang, Zijie, et al.
Publicado: (2024)
Commonsense Generation and Evaluation for Dialogue Systems using Large Language Models
por: Estecha-Garitagoitia, Marcos, et al.
Publicado: (2025)
por: Estecha-Garitagoitia, Marcos, et al.
Publicado: (2025)
Discourse-Aware Dual-Track Streaming Response for Low-Latency Spoken Dialogue Systems
por: Liu, Siyuan, et al.
Publicado: (2026)
por: Liu, Siyuan, et al.
Publicado: (2026)
DiaCDM: Cognitive Diagnosis in Teacher-Student Dialogues using the Initiation-Response-Evaluation Framework
por: Jia, Rui, et al.
Publicado: (2025)
por: Jia, Rui, et al.
Publicado: (2025)
Rethinking the Alignment of Psychotherapy Dialogue Generation with Motivational Interviewing Strategies
por: Sun, Xin, et al.
Publicado: (2024)
por: Sun, Xin, et al.
Publicado: (2024)
Ejemplares similares
-
Tracing Multilingual Knowledge Acquisition Dynamics in Domain Adaptation: A Case Study of English-Japanese Biomedical Adaptation
por: Zhao, Xin, et al.
Publicado: (2025) -
When Harry Meets Superman: The Role of The Interlocutor in Persona-Based Dialogue Generation
por: Occhipinti, Daniela, et al.
Publicado: (2025) -
Commentary Generation from Data Records of Multiplayer Strategy Esports Game
por: Wang, Zihan, et al.
Publicado: (2022) -
On the Benchmarking of LLMs for Open-Domain Dialogue Evaluation
por: Mendonça, John, et al.
Publicado: (2024) -
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs
por: Mendonça, John, et al.
Publicado: (2024)