REALTALK: A 21-Day Real-World Dataset for Long-Term Conversation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Dong-Ho, Maharana, Adyasha, Pujara, Jay, Ren, Xiang, Barbieri, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Very Long-Term Conversational Memory of LLM Agents
by: Maharana, Adyasha, et al.
Published: (2024)
by: Maharana, Adyasha, et al.
Published: (2024)
Which Questions Improve Learning the Most? Utility Estimation of Questions with LM-based Simulations
by: Lee, Dong-Ho, et al.
Published: (2025)
by: Lee, Dong-Ho, et al.
Published: (2025)
Integrating Pre-Trained Language Model with Physical Layer Communications
by: Lee, Ju-Hyung, et al.
Published: (2024)
by: Lee, Ju-Hyung, et al.
Published: (2024)
Compress to Impress: Unleashing the Potential of Compressive Memory in Real-World Long-Term Conversations
by: Chen, Nuo, et al.
Published: (2024)
by: Chen, Nuo, et al.
Published: (2024)
Toward Better Temporal Structures for Geopolitical Events Forecasting
by: Ahrabian, Kian, et al.
Published: (2026)
by: Ahrabian, Kian, et al.
Published: (2026)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
by: Maharana, Adyasha, et al.
Published: (2023)
by: Maharana, Adyasha, et al.
Published: (2023)
Stark: Social Long-Term Multi-Modal Conversation with Persona Commonsense Knowledge
by: Lee, Young-Jun, et al.
Published: (2024)
by: Lee, Young-Jun, et al.
Published: (2024)
Generation-Based and Emotion-Reflected Memory Update: Creating the KEEM Dataset for Better Long-Term Conversation
by: Kang, Jeonghyun, et al.
Published: (2026)
by: Kang, Jeonghyun, et al.
Published: (2026)
Multilingual Topic Classification in X: Dataset and Analysis
by: Antypas, Dimosthenis, et al.
Published: (2024)
by: Antypas, Dimosthenis, et al.
Published: (2024)
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
by: Zheng, Lianmin, et al.
Published: (2023)
by: Zheng, Lianmin, et al.
Published: (2023)
EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory
by: Li, Yuyang, et al.
Published: (2026)
by: Li, Yuyang, et al.
Published: (2026)
PLUG: Leveraging Pivot Language in Cross-Lingual Instruction Tuning
by: Zhang, Zhihan, et al.
Published: (2023)
by: Zhang, Zhihan, et al.
Published: (2023)
A Systematic Analysis of Base Model Choice for Reward Modeling
by: Ahrabian, Kian, et al.
Published: (2025)
by: Ahrabian, Kian, et al.
Published: (2025)
HyperMem: Hypergraph Memory for Long-Term Conversations
by: Yue, Juwei, et al.
Published: (2026)
by: Yue, Juwei, et al.
Published: (2026)
A Simple Yet Strong Baseline for Long-Term Conversational Memory of LLM Agents
by: Zhou, Sizhe, et al.
Published: (2025)
by: Zhou, Sizhe, et al.
Published: (2025)
Multi-Domain ABSA Conversation Dataset Generation via LLMs for Real-World Evaluation and Model Comparison
by: Pandit, Tejul, et al.
Published: (2025)
by: Pandit, Tejul, et al.
Published: (2025)
Can LLMs Infer Personality from Real World Conversations?
by: Zhu, Jianfeng, et al.
Published: (2025)
by: Zhu, Jianfeng, et al.
Published: (2025)
HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents
by: Cao, Shuqi, et al.
Published: (2026)
by: Cao, Shuqi, et al.
Published: (2026)
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
by: Sen, Sahil, et al.
Published: (2026)
by: Sen, Sahil, et al.
Published: (2026)
DRInQ: Evaluating Conversational Implicature with Controlled Context Variation
by: Arai, Hirona Jacqueline, et al.
Published: (2026)
by: Arai, Hirona Jacqueline, et al.
Published: (2026)
The Curious Case of Nonverbal Abstract Reasoning with Multi-Modal Large Language Models
by: Ahrabian, Kian, et al.
Published: (2024)
by: Ahrabian, Kian, et al.
Published: (2024)
J-CRe3: A Japanese Conversation Dataset for Real-world Reference Resolution
by: Ueda, Nobuhiro, et al.
Published: (2024)
by: Ueda, Nobuhiro, et al.
Published: (2024)
On The Adaptation of Unlimiformer for Decoder-Only Transformers
by: Ahrabian, Kian, et al.
Published: (2024)
by: Ahrabian, Kian, et al.
Published: (2024)
MemRouter: Memory-as-Embedding Routing for Long-Term Conversational Agents
by: Hu, Tianyu, et al.
Published: (2026)
by: Hu, Tianyu, et al.
Published: (2026)
MTP: A Dataset for Multi-Modal Turning Points in Casual Conversations
by: Ho, Gia-Bao Dinh, et al.
Published: (2024)
by: Ho, Gia-Bao Dinh, et al.
Published: (2024)
A Practical Analysis of Human Alignment with *PO
by: Ahrabian, Kian, et al.
Published: (2024)
by: Ahrabian, Kian, et al.
Published: (2024)
From Single to Multi-Granularity: Toward Long-Term Memory Association and Selection of Conversational Agents
by: Xu, Derong, et al.
Published: (2025)
by: Xu, Derong, et al.
Published: (2025)
EpiCache: Episodic KV Cache Management for Long-Term Conversation on Resource-Constrained Environments
by: Kim, Minsoo, et al.
Published: (2025)
by: Kim, Minsoo, et al.
Published: (2025)
AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data
by: Rashidian, Sina, et al.
Published: (2025)
by: Rashidian, Sina, et al.
Published: (2025)
EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval
by: Acuna, Julian
Published: (2026)
by: Acuna, Julian
Published: (2026)
Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM Agents
by: Bei, Yuanchen, et al.
Published: (2026)
by: Bei, Yuanchen, et al.
Published: (2026)
An Empirical Evaluation of Encoder Architectures for Fast Real-Time Long Conversational Understanding
by: Senthilnathan, Annamalai, et al.
Published: (2025)
by: Senthilnathan, Annamalai, et al.
Published: (2025)
HateDay: Insights from a Global Hate Speech Dataset Representative of a Day on Twitter
by: Tonneau, Manuel, et al.
Published: (2024)
by: Tonneau, Manuel, et al.
Published: (2024)
LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
LEMONADE: A Large Multilingual Expert-Annotated Abstractive Event Dataset for the Real World
by: Semnani, Sina J., et al.
Published: (2025)
by: Semnani, Sina J., et al.
Published: (2025)
RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking
by: Jiang, Yifan, et al.
Published: (2024)
by: Jiang, Yifan, et al.
Published: (2024)
ES-MemEval: Benchmarking Conversational Agents on Personalized Long-Term Emotional Support
by: Chen, Tiantian, et al.
Published: (2026)
by: Chen, Tiantian, et al.
Published: (2026)
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
by: Xue, Haochen, et al.
Published: (2025)
by: Xue, Haochen, et al.
Published: (2025)
REALM: A Dataset of Real-World LLM Use Cases
by: Cheng, Jingwen, et al.
Published: (2025)
by: Cheng, Jingwen, et al.
Published: (2025)
Real World Conversational Entity Linking Requires More Than Zeroshots
by: Hoveyda, Mohanna, et al.
Published: (2024)
by: Hoveyda, Mohanna, et al.
Published: (2024)
Similar Items
-
Evaluating Very Long-Term Conversational Memory of LLM Agents
by: Maharana, Adyasha, et al.
Published: (2024) -
Which Questions Improve Learning the Most? Utility Estimation of Questions with LM-based Simulations
by: Lee, Dong-Ho, et al.
Published: (2025) -
Integrating Pre-Trained Language Model with Physical Layer Communications
by: Lee, Ju-Hyung, et al.
Published: (2024) -
Compress to Impress: Unleashing the Potential of Compressive Memory in Real-World Long-Term Conversations
by: Chen, Nuo, et al.
Published: (2024) -
Toward Better Temporal Structures for Geopolitical Events Forecasting
by: Ahrabian, Kian, et al.
Published: (2026)