CoSER: A Comprehensive Literary Dataset and Framework for Training and Evaluating LLM Role-Playing and Persona Simulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xintao, Wang, Heng, Zhang, Yifei, Yuan, Xinfeng, Xu, Rui, Huang, Jen-tse, Yuan, Siyu, Guo, Haoran, Chen, Jiangjie, Zhou, Shuchang, Wang, Wei, Xiao, Yanghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Character is Destiny: Can Role-Playing Language Agents Make Persona-Driven Decisions?
von: Xu, Rui, et al.
Veröffentlicht: (2024)
von: Xu, Rui, et al.
Veröffentlicht: (2024)
InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews
von: Wang, Xintao, et al.
Veröffentlicht: (2023)
von: Wang, Xintao, et al.
Veröffentlicht: (2023)
Capturing Minds, Not Just Words: Enhancing Role-Playing Language Models with Personality-Indicative Data
von: Ran, Yiting, et al.
Veröffentlicht: (2024)
von: Ran, Yiting, et al.
Veröffentlicht: (2024)
MINDECHO: Role-Playing Language Agents for Key Opinion Leaders
von: Xu, Rui, et al.
Veröffentlicht: (2024)
von: Xu, Rui, et al.
Veröffentlicht: (2024)
From Persona to Personalization: A Survey on Role-Playing Language Agents
von: Chen, Jiangjie, et al.
Veröffentlicht: (2024)
von: Chen, Jiangjie, et al.
Veröffentlicht: (2024)
ARIA: Training Language Agents with Intention-Driven Reward Aggregation
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works
von: Yuan, Xinfeng, et al.
Veröffentlicht: (2024)
von: Yuan, Xinfeng, et al.
Veröffentlicht: (2024)
Curse of Knowledge: When Complex Evaluation Context Benefits yet Biases LLM Judges
von: Li, Weiyuan, et al.
Veröffentlicht: (2025)
von: Li, Weiyuan, et al.
Veröffentlicht: (2025)
Enhancing Language Agent Strategic Reasoning through Self-Play in Adversarial Games
von: Zhang, Yikai, et al.
Veröffentlicht: (2025)
von: Zhang, Yikai, et al.
Veröffentlicht: (2025)
SurveyAgent: A Conversational System for Personalized and Efficient Research Survey
von: Wang, Xintao, et al.
Veröffentlicht: (2024)
von: Wang, Xintao, et al.
Veröffentlicht: (2024)
DEEPER Insight into Your User: Directed Persona Refinement for Dynamic Persona Modeling
von: Chen, Aili, et al.
Veröffentlicht: (2025)
von: Chen, Aili, et al.
Veröffentlicht: (2025)
TimeArena: Shaping Efficient Multitasking Language Agents in a Time-Aware Simulation
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
Light Up the Shadows: Enhance Long-Tailed Entity Grounding with Concept-Guided Vision-Language Models
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
ORIGAMISPACE: Benchmarking Multimodal LLMs in Multi-Step Spatial Reasoning with Mathematical Constraints
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
ANALOGYKB: Unlocking Analogical Reasoning of Language Models with A Million-scale Knowledge Base
von: Yuan, Siyu, et al.
Veröffentlicht: (2023)
von: Yuan, Siyu, et al.
Veröffentlicht: (2023)
Can LLMs Learn to Map the World from Local Descriptions?
von: Xia, Sirui, et al.
Veröffentlicht: (2025)
von: Xia, Sirui, et al.
Veröffentlicht: (2025)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
von: Ng, Man Tik, et al.
Veröffentlicht: (2024)
von: Ng, Man Tik, et al.
Veröffentlicht: (2024)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
von: Wang, Xintao, et al.
Veröffentlicht: (2026)
von: Wang, Xintao, et al.
Veröffentlicht: (2026)
ThinkDial: An Open Recipe for Controlling Reasoning Effort in Large Language Models
von: He, Qianyu, et al.
Veröffentlicht: (2025)
von: He, Qianyu, et al.
Veröffentlicht: (2025)
OpenCharacter: Training Customizable Role-Playing LLMs with Large-Scale Synthetic Personas
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2025)
ConcEPT: Concept-Enhanced Pre-Training for Language Models
von: Wang, Xintao, et al.
Veröffentlicht: (2024)
von: Wang, Xintao, et al.
Veröffentlicht: (2024)
A Unified Plug-and-Play Algorithm with Projected Landweber Operator for Split Convex Feasibility Problems
von: Zhang, Shuchang, et al.
Veröffentlicht: (2024)
von: Zhang, Shuchang, et al.
Veröffentlicht: (2024)
Past Meets Present: Creating Historical Analogy with Large Language Models
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
SelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
Revealing the Barriers of Language Agents in Planning
von: Xie, Jian, et al.
Veröffentlicht: (2024)
von: Xie, Jian, et al.
Veröffentlicht: (2024)
Chain-of-Knowledge: Integrating Knowledge Reasoning into Large Language Models by Learning from Knowledge Graphs
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
Memory-Driven Role-Playing: Evaluation and Enhancement of Persona Knowledge Utilization in LLMs
von: Wang, Kai, et al.
Veröffentlicht: (2026)
von: Wang, Kai, et al.
Veröffentlicht: (2026)
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models
von: Shi, Wenlong, et al.
Veröffentlicht: (2026)
von: Shi, Wenlong, et al.
Veröffentlicht: (2026)
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
von: Yuan, Youliang, et al.
Veröffentlicht: (2024)
von: Yuan, Youliang, et al.
Veröffentlicht: (2024)
The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models
von: Xiao, Yunze, et al.
Veröffentlicht: (2026)
von: Xiao, Yunze, et al.
Veröffentlicht: (2026)
PersonaEval: Are LLM Evaluators Human Enough to Judge Role-Play?
von: Zhou, Lingfeng, et al.
Veröffentlicht: (2025)
von: Zhou, Lingfeng, et al.
Veröffentlicht: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
von: Huang, Jen-tse, et al.
Veröffentlicht: (2025)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2025)
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies
von: Zhou, Jiaxu, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaxu, et al.
Veröffentlicht: (2025)
On the Failure of Latent State Persistence in Large Language Models
von: Huang, Jen-tse, et al.
Veröffentlicht: (2025)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2025)
PowerAttention: Exponentially Scaling of Receptive Fields for Effective Sparse Attention
von: Chen, Lida, et al.
Veröffentlicht: (2025)
von: Chen, Lida, et al.
Veröffentlicht: (2025)
ComboBench: Can LLMs Manipulate Physical Devices to Play Virtual Reality Games?
von: Li, Shuqing, et al.
Veröffentlicht: (2025)
von: Li, Shuqing, et al.
Veröffentlicht: (2025)
Crafting Customisable Characters with LLMs: A Persona-Driven Role-Playing Agent Framework
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
SMAT: Staged Multi-Agent Training for Co-Adaptive Exoskeleton Control
von: Yuan, Yifei, et al.
Veröffentlicht: (2026)
von: Yuan, Yifei, et al.
Veröffentlicht: (2026)
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment
von: Li, Jiajia, et al.
Veröffentlicht: (2026)
von: Li, Jiajia, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Character is Destiny: Can Role-Playing Language Agents Make Persona-Driven Decisions?
von: Xu, Rui, et al.
Veröffentlicht: (2024) -
InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews
von: Wang, Xintao, et al.
Veröffentlicht: (2023) -
Capturing Minds, Not Just Words: Enhancing Role-Playing Language Models with Personality-Indicative Data
von: Ran, Yiting, et al.
Veröffentlicht: (2024) -
MINDECHO: Role-Playing Language Agents for Key Opinion Leaders
von: Xu, Rui, et al.
Veröffentlicht: (2024) -
From Persona to Personalization: A Survey on Role-Playing Language Agents
von: Chen, Jiangjie, et al.
Veröffentlicht: (2024)