Human Simulacra: Benchmarking the Personification of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xie, Qiuejie, Feng, Qiming, Zhang, Tianqi, Li, Qingqiu, Yang, Linyi, Zhang, Yuejie, Feng, Rui, He, Liang, Gao, Shang, Zhang, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Empirical Analysis of Uncertainty in Large Language Model Evaluations
von: Xie, Qiujie, et al.
Veröffentlicht: (2025)
von: Xie, Qiujie, et al.
Veröffentlicht: (2025)
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
von: Yuan, Runtian, et al.
Veröffentlicht: (2025)
von: Yuan, Runtian, et al.
Veröffentlicht: (2025)
Vision-Language Model Based Multi-Expert Fusion for CT Image Classification
von: Bai, Jianfa, et al.
Veröffentlicht: (2026)
von: Bai, Jianfa, et al.
Veröffentlicht: (2026)
Working with Large Language Models to Enhance Messaging Effectiveness for Vaccine Confidence
von: Gullison, Lucinda, et al.
Veröffentlicht: (2025)
von: Gullison, Lucinda, et al.
Veröffentlicht: (2025)
Anatomical Structure-Guided Medical Vision-Language Pre-training
von: Li, Qingqiu, et al.
Veröffentlicht: (2024)
von: Li, Qingqiu, et al.
Veröffentlicht: (2024)
Advancing Lung Disease Diagnosis in 3D CT Scans
von: Li, Qingqiu, et al.
Veröffentlicht: (2025)
von: Li, Qingqiu, et al.
Veröffentlicht: (2025)
Domain Adaptation Using Pseudo Labels for COVID-19 Detection
von: Yuan, Runtian, et al.
Veröffentlicht: (2024)
von: Yuan, Runtian, et al.
Veröffentlicht: (2024)
Advancing COVID-19 Detection in 3D CT Scans
von: Li, Qingqiu, et al.
Veröffentlicht: (2024)
von: Li, Qingqiu, et al.
Veröffentlicht: (2024)
Multi-Source COVID-19 Detection via Variance Risk Extrapolation
von: Yuan, Runtian, et al.
Veröffentlicht: (2025)
von: Yuan, Runtian, et al.
Veröffentlicht: (2025)
Beyond Private or Public: Large Language Models as Quasi-Public Goods in the AI Economy
von: Zhang, Yukun, et al.
Veröffentlicht: (2025)
von: Zhang, Yukun, et al.
Veröffentlicht: (2025)
Affective Computing in the Era of Large Language Models: A Survey from the NLP Perspective
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
Comparing Rationality Between Large Language Models and Humans: Insights and Open Questions
von: Alsagheer, Dana, et al.
Veröffentlicht: (2024)
von: Alsagheer, Dana, et al.
Veröffentlicht: (2024)
AOR: Anatomical Ontology-Guided Reasoning for Medical Large Multimodal Model in Chest X-Ray Interpretation
von: Li, Qingqiu, et al.
Veröffentlicht: (2025)
von: Li, Qingqiu, et al.
Veröffentlicht: (2025)
Collapsed Language Models Promote Fairness
von: Xu, Jingxuan, et al.
Veröffentlicht: (2024)
von: Xu, Jingxuan, et al.
Veröffentlicht: (2024)
Benchmarking Large Language Models on Homework Assessment in Circuit Analysis
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
Human-Level and Beyond: Benchmarking Large Language Models Against Clinical Pharmacists in Prescription Review
von: Yang, Yan, et al.
Veröffentlicht: (2025)
von: Yang, Yan, et al.
Veröffentlicht: (2025)
Can Large Language Models Be Trusted Paper Reviewers? A Feasibility Study
von: Li, Chuanlei, et al.
Veröffentlicht: (2025)
von: Li, Chuanlei, et al.
Veröffentlicht: (2025)
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
Personality Alignment of Large Language Models
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
Human-like Social Compliance in Large Language Models: Unifying Sycophancy and Conformity through Signal Competition Dynamics
von: Zhang, Long, et al.
Veröffentlicht: (2025)
von: Zhang, Long, et al.
Veröffentlicht: (2025)
Enhancing Systematic Interoperability: Convergences and Mismatches between Web 3.0 and the EU Data Act
von: Xu, Linyi, et al.
Veröffentlicht: (2025)
von: Xu, Linyi, et al.
Veröffentlicht: (2025)
Ask Again, Then Fail: Large Language Models' Vacillations in Judgment
von: Xie, Qiming, et al.
Veröffentlicht: (2023)
von: Xie, Qiming, et al.
Veröffentlicht: (2023)
Personification of a Virtue (?)
von: Scan-the-World
Veröffentlicht: (2026)
von: Scan-the-World
Veröffentlicht: (2026)
Evaluating Chinese Large Language Models: The Influence of Persona Assignment on Stereotypes and Safeguards
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
Mutual Wanting in Human--AI Interaction: Empirical Evidence from Large-Scale Analysis of GPT Model Transitions
von: Shang, HaoYang, et al.
Veröffentlicht: (2025)
von: Shang, HaoYang, et al.
Veröffentlicht: (2025)
The Future of Learning: Large Language Models through the Lens of Students
von: Zhang, He, et al.
Veröffentlicht: (2024)
von: Zhang, He, et al.
Veröffentlicht: (2024)
CycleResearcher: Improving Automated Research via Automated Review
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study
von: Xu, Liuchang, et al.
Veröffentlicht: (2024)
von: Xu, Liuchang, et al.
Veröffentlicht: (2024)
Is Self-knowledge and Action Consistent or Not: Investigating Large Language Model's Personality
von: Ai, Yiming, et al.
Veröffentlicht: (2024)
von: Ai, Yiming, et al.
Veröffentlicht: (2024)
JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
von: Xiao, Yunze, et al.
Veröffentlicht: (2025)
von: Xiao, Yunze, et al.
Veröffentlicht: (2025)
CDEval: A Benchmark for Measuring the Cultural Dimensions of Large Language Models
von: Wang, Yuhang, et al.
Veröffentlicht: (2023)
von: Wang, Yuhang, et al.
Veröffentlicht: (2023)
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
von: Xu, Jilan, et al.
Veröffentlicht: (2025)
von: Xu, Jilan, et al.
Veröffentlicht: (2025)
PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models
von: Bao, Han, et al.
Veröffentlicht: (2026)
von: Bao, Han, et al.
Veröffentlicht: (2026)
Guiding IoT-Based Healthcare Alert Systems with Large Language Models
von: Gao, Yulan, et al.
Veröffentlicht: (2024)
von: Gao, Yulan, et al.
Veröffentlicht: (2024)
Quantitative Insights into Large Language Model Usage and Trust in Academia: An Empirical Study
von: Jung, Minseok, et al.
Veröffentlicht: (2024)
von: Jung, Minseok, et al.
Veröffentlicht: (2024)
Exploring Large Language Model Agents for Piloting Social Experiments
von: Piao, Jinghua, et al.
Veröffentlicht: (2025)
von: Piao, Jinghua, et al.
Veröffentlicht: (2025)
Clinical Priors Guided Lung Disease Detection in 3D CT Scans
von: Lu, Kejin, et al.
Veröffentlicht: (2026)
von: Lu, Kejin, et al.
Veröffentlicht: (2026)
EducaSim: Interactive Simulacra for CS1 Instructional Practice
von: Mohne, Cameron, et al.
Veröffentlicht: (2026)
von: Mohne, Cameron, et al.
Veröffentlicht: (2026)
MDK12-Bench: A Comprehensive Evaluation of Multimodal Large Language Models on Multidisciplinary Exams
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Empirical Analysis of Uncertainty in Large Language Model Evaluations
von: Xie, Qiujie, et al.
Veröffentlicht: (2025) -
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024) -
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
von: Yuan, Runtian, et al.
Veröffentlicht: (2025) -
Vision-Language Model Based Multi-Expert Fusion for CT Image Classification
von: Bai, Jianfa, et al.
Veröffentlicht: (2026) -
Working with Large Language Models to Enhance Messaging Effectiveness for Vaccine Confidence
von: Gullison, Lucinda, et al.
Veröffentlicht: (2025)