HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xintao, Yang, Jian, Li, Weiyuan, Xie, Rui, Huang, Jen-tse, Gao, Jun, Huang, Shuai, Kang, Yueping, Gou, Yuanli, Feng, Hongwei, Xiao, Yanghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Applying Cognitive Design Patterns to General LLM Agents
von: Wray, Robert E., et al.
Veröffentlicht: (2025)
von: Wray, Robert E., et al.
Veröffentlicht: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
von: Karinshak, Elise, et al.
Veröffentlicht: (2024)
von: Karinshak, Elise, et al.
Veröffentlicht: (2024)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
von: Jia, Xiao
Veröffentlicht: (2026)
von: Jia, Xiao
Veröffentlicht: (2026)
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
von: Ketir, Si-Belkacem Yamine, et al.
Veröffentlicht: (2026)
von: Ketir, Si-Belkacem Yamine, et al.
Veröffentlicht: (2026)
Can LLM Graph Reasoning Generalize beyond Pattern Memorization?
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2024)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
von: Palit, Sayon, et al.
Veröffentlicht: (2025)
von: Palit, Sayon, et al.
Veröffentlicht: (2025)
LLMs Aren't Human: A Critical Perspective on LLM Personality
von: Zierahn, Kim, et al.
Veröffentlicht: (2026)
von: Zierahn, Kim, et al.
Veröffentlicht: (2026)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026)
von: Bachar, Or, et al.
Veröffentlicht: (2026)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
von: Figueiredo, Vanessa
Veröffentlicht: (2025)
von: Figueiredo, Vanessa
Veröffentlicht: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026)
von: Souza, Débora, et al.
Veröffentlicht: (2026)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
LLMs and the Human Condition
von: Wallis, Peter
Veröffentlicht: (2024)
von: Wallis, Peter
Veröffentlicht: (2024)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
von: Du, Bangde, et al.
Veröffentlicht: (2025)
von: Du, Bangde, et al.
Veröffentlicht: (2025)
Review of Case-Based Reasoning for LLM Agents: Theoretical Foundations, Architectural Components, and Cognitive Integration
von: Hatalis, Kostas, et al.
Veröffentlicht: (2025)
von: Hatalis, Kostas, et al.
Veröffentlicht: (2025)
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
Exploring Collatz Dynamics with Human-LLM Collaboration
von: Chang, Edward Y.
Veröffentlicht: (2026)
von: Chang, Edward Y.
Veröffentlicht: (2026)
Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning
von: Cazares, Manuel Israel
Veröffentlicht: (2026)
von: Cazares, Manuel Israel
Veröffentlicht: (2026)
FlexQuant: A Flexible and Efficient Dynamic Precision Switching Framework for LLM Quantization
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
von: Liu, Fangxin, et al.
Veröffentlicht: (2025)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
SUBLLM: A Novel Efficient Architecture with Token Sequence Subsampling for LLM
von: Wang, Quandong, et al.
Veröffentlicht: (2024)
von: Wang, Quandong, et al.
Veröffentlicht: (2024)
SagaLLM: Context Management, Validation, and Transaction Guarantees for Multi-Agent LLM Planning
von: Chang, Edward Y., et al.
Veröffentlicht: (2025)
von: Chang, Edward Y., et al.
Veröffentlicht: (2025)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
von: Yan, Lingyong, et al.
Veröffentlicht: (2026)
von: Yan, Lingyong, et al.
Veröffentlicht: (2026)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
Large Language Model (LLM) Bias Index -- LLMBI
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
Searching for the Most Human-like Emergent Language
von: Boldt, Brendon, et al.
Veröffentlicht: (2025)
von: Boldt, Brendon, et al.
Veröffentlicht: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
Revisiting Word Embeddings in the LLM Era
von: Mahajan, Yash, et al.
Veröffentlicht: (2024)
von: Mahajan, Yash, et al.
Veröffentlicht: (2024)
AsyncTLS: Efficient Generative LLM Inference with Asynchronous Two-level Sparse Attention
von: Hu, Yuxuan, et al.
Veröffentlicht: (2026)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2026)
Vibe-Creation: The Epistemology of Human-AI Emergent Cognition
von: Levin, Ilya
Veröffentlicht: (2026)
von: Levin, Ilya
Veröffentlicht: (2026)
Distinguishing Ignorance from Error in LLM Hallucinations
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
Decoding-Free Sampling Strategies for LLM Marginalization
von: Pohl, David, et al.
Veröffentlicht: (2025)
von: Pohl, David, et al.
Veröffentlicht: (2025)
Active Context Compression: Autonomous Memory Management in LLM Agents
von: Verma, Nikhil
Veröffentlicht: (2026)
von: Verma, Nikhil
Veröffentlicht: (2026)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Cognitive bias in LLM reasoning compromises interpretation of clinical oncology notes
von: Kenaston, Matthew W., et al.
Veröffentlicht: (2025)
von: Kenaston, Matthew W., et al.
Veröffentlicht: (2025)
Benchmarking Cognitive Biases in Large Language Models as Evaluators
von: Koo, Ryan, et al.
Veröffentlicht: (2023)
von: Koo, Ryan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Applying Cognitive Design Patterns to General LLM Agents
von: Wray, Robert E., et al.
Veröffentlicht: (2025) -
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023) -
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026) -
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
von: Karinshak, Elise, et al.
Veröffentlicht: (2024) -
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
von: Jia, Xiao
Veröffentlicht: (2026)