Gespeichert in:
| Hauptverfasser: | Suh, Joseph, Raj, Ayush, Kang, Minwoo, Chang, Serina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.09808 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
Graph-Based Alternatives to LLMs for Human Simulation
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
von: Suh, Joseph, et al.
Veröffentlicht: (2025)
Deep Binding of Language Model Virtual Personas: a Study on Approximating Political Partisan Misperceptions
von: Kang, Minwoo, et al.
Veröffentlicht: (2025)
von: Kang, Minwoo, et al.
Veröffentlicht: (2025)
Identity, Cooperation and Framing Effects within Groups of Real and Simulated Humans
von: Moon, Suhong, et al.
Veröffentlicht: (2026)
von: Moon, Suhong, et al.
Veröffentlicht: (2026)
Rediscovering the Latent Dimensions of Personality with Large Language Models as Trait Descriptors
von: Suh, Joseph, et al.
Veröffentlicht: (2024)
von: Suh, Joseph, et al.
Veröffentlicht: (2024)
Valid Survey Simulations with Limited Human Data: The Roles of Prompting, Fine-Tuning, and Rectification
von: Krsteski, Stefan, et al.
Veröffentlicht: (2025)
von: Krsteski, Stefan, et al.
Veröffentlicht: (2025)
ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant Simulation
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?
von: Dou, Yao, et al.
Veröffentlicht: (2025)
von: Dou, Yao, et al.
Veröffentlicht: (2025)
Quantifying the Persona Effect in LLM Simulations
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
ChatBench: From Static Benchmarks to Human-AI Evaluation
von: Chang, Serina, et al.
Veröffentlicht: (2025)
von: Chang, Serina, et al.
Veröffentlicht: (2025)
LifeSim: Long-Horizon User Life Simulator for Personalized Assistant Evaluation
von: Duan, Feiyu, et al.
Veröffentlicht: (2026)
von: Duan, Feiyu, et al.
Veröffentlicht: (2026)
Virtual Personas for Language Models via an Anthology of Backstories
von: Moon, Suhong, et al.
Veröffentlicht: (2024)
von: Moon, Suhong, et al.
Veröffentlicht: (2024)
Non-Collaborative User Simulators for Tool Agents
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Shim, Jeonghoon, et al.
Veröffentlicht: (2025)
SafeChat: A Framework for Building Trustworthy Collaborative Assistants and a Case Study of its Usefulness
von: Srivastava, Biplav, et al.
Veröffentlicht: (2025)
von: Srivastava, Biplav, et al.
Veröffentlicht: (2025)
When Wording Steers the Evaluation: Framing Bias in LLM judges
von: Hwang, Yerin, et al.
Veröffentlicht: (2026)
von: Hwang, Yerin, et al.
Veröffentlicht: (2026)
User-Assistant Bias in LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
Judging Against the Reference: Uncovering Knowledge-Driven Failures in LLM-Judges on QA Evaluation
von: Lee, Dongryeol, et al.
Veröffentlicht: (2026)
von: Lee, Dongryeol, et al.
Veröffentlicht: (2026)
Multi-trait User Simulation with Adaptive Decoding for Conversational Task Assistants
von: Ferreira, Rafael, et al.
Veröffentlicht: (2024)
von: Ferreira, Rafael, et al.
Veröffentlicht: (2024)
LLM-based Medical Assistant Personalization with Short- and Long-Term Memory Coordination
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
MathVC: An LLM-Simulated Multi-Character Virtual Classroom for Mathematics Education
von: Yue, Murong, et al.
Veröffentlicht: (2024)
von: Yue, Murong, et al.
Veröffentlicht: (2024)
An EcoSage Assistant: Towards Building A Multimodal Plant Care Dialogue Assistant
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
User Modeling Challenges in Interactive AI Assistant Systems
von: Su, Megan, et al.
Veröffentlicht: (2024)
von: Su, Megan, et al.
Veröffentlicht: (2024)
Human-AI Collaborative Taxonomy Construction: A Case Study in Profession-Specific Writing Assistants
von: Lee, Minhwa, et al.
Veröffentlicht: (2024)
von: Lee, Minhwa, et al.
Veröffentlicht: (2024)
SafeSearch: Do Not Trade Safety for Utility in LLM Search Agents
von: Zhan, Qiusi, et al.
Veröffentlicht: (2025)
von: Zhan, Qiusi, et al.
Veröffentlicht: (2025)
A Survey on LLM-based Conversational User Simulation
von: Ni, Bo, et al.
Veröffentlicht: (2026)
von: Ni, Bo, et al.
Veröffentlicht: (2026)
KAUCUS: Knowledge Augmented User Simulators for Training Language Model Assistants
von: Dhole, Kaustubh D.
Veröffentlicht: (2024)
von: Dhole, Kaustubh D.
Veröffentlicht: (2024)
MemSim: A Bayesian Simulator for Evaluating Memory of LLM-based Personal Assistants
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
Building A Coding Assistant via the Retrieval-Augmented Language Model
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning
von: White, Isadora, et al.
Veröffentlicht: (2025)
von: White, Isadora, et al.
Veröffentlicht: (2025)
Improving LLM-Powered EDA Assistants with RAFT
von: Shi, Luyao, et al.
Veröffentlicht: (2025)
von: Shi, Luyao, et al.
Veröffentlicht: (2025)
LLM Agent Meets Agentic AI: Can LLM Agents Simulate Customers to Evaluate Agentic-AI-based Shopping Assistants?
von: Sun, Lu, et al.
Veröffentlicht: (2025)
von: Sun, Lu, et al.
Veröffentlicht: (2025)
Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily Assistant
von: He, Gaole, et al.
Veröffentlicht: (2025)
von: He, Gaole, et al.
Veröffentlicht: (2025)
CARE: A Clue-guided Assistant for CSRs to Read User Manuals
von: Du, Weihong, et al.
Veröffentlicht: (2024)
von: Du, Weihong, et al.
Veröffentlicht: (2024)
Users Mispredict Their Own Preferences for AI Writing Assistance
von: Lai, Vivian, et al.
Veröffentlicht: (2026)
von: Lai, Vivian, et al.
Veröffentlicht: (2026)
Reliable LLM-based User Simulator for Task-Oriented Dialogue Systems
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
Personality Matters: User Traits Predict LLM Preferences in Multi-Turn Collaborative Tasks
von: Yunusov, Sarfaroz, et al.
Veröffentlicht: (2025)
von: Yunusov, Sarfaroz, et al.
Veröffentlicht: (2025)
Counterfactual Fairness Evaluation of LLM-Based Contact Center Agent Quality Assurance System
von: Mayilvaghanan, Kawin, et al.
Veröffentlicht: (2026)
von: Mayilvaghanan, Kawin, et al.
Veröffentlicht: (2026)
Goal Alignment in LLM-Based User Simulators for Conversational AI
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
Social Conjuring: Multi-User Runtime Collaboration with AI in Building Virtual 3D Worlds
von: Kobenova, Amina, et al.
Veröffentlicht: (2024)
von: Kobenova, Amina, et al.
Veröffentlicht: (2024)
DuetSim: Building User Simulator with Dual Large Language Models for Task-Oriented Dialogues
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions
von: Suh, Joseph, et al.
Veröffentlicht: (2025) -
Graph-Based Alternatives to LLMs for Human Simulation
von: Suh, Joseph, et al.
Veröffentlicht: (2025) -
Deep Binding of Language Model Virtual Personas: a Study on Approximating Political Partisan Misperceptions
von: Kang, Minwoo, et al.
Veröffentlicht: (2025) -
Identity, Cooperation and Framing Effects within Groups of Real and Simulated Humans
von: Moon, Suhong, et al.
Veröffentlicht: (2026) -
Rediscovering the Latent Dimensions of Personality with Large Language Models as Trait Descriptors
von: Suh, Joseph, et al.
Veröffentlicht: (2024)