PsycoLLM: Enhancing LLM for Psychological Understanding and Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Hu, Jinpeng, Dong, Tengteng, Gang, Luo, Ma, Hui, Zou, Peng, Sun, Xiao, Guo, Dan, Yang, Xun, Wang, Meng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Psyche-R1: Towards Reliable Psychological LLMs through Unified Empathy, Expertise, and Reasoning
di: Dai, Chongyuan, et al.
Pubblicazione: (2025)
di: Dai, Chongyuan, et al.
Pubblicazione: (2025)
Traits Run Deep: Enhancing Personality Assessment via Psychology-Guided LLM Representations and Multimodal Apparent Behaviors
di: Li, Jia, et al.
Pubblicazione: (2025)
di: Li, Jia, et al.
Pubblicazione: (2025)
Think-Augmented Function Calling: Improving LLM Parameter Accuracy Through Embedded Reasoning
di: Wei, Lei, et al.
Pubblicazione: (2026)
di: Wei, Lei, et al.
Pubblicazione: (2026)
AgentMental: An Interactive Multi-Agent Framework for Explainable and Adaptive Mental Health Assessment
di: Hu, Jinpeng, et al.
Pubblicazione: (2025)
di: Hu, Jinpeng, et al.
Pubblicazione: (2025)
Unveiling LLM Evaluation Focused on Metrics: Challenges and Solutions
di: Hu, Taojun, et al.
Pubblicazione: (2024)
di: Hu, Taojun, et al.
Pubblicazione: (2024)
Understanding Layer Significance in LLM Alignment
di: Shi, Guangyuan, et al.
Pubblicazione: (2024)
di: Shi, Guangyuan, et al.
Pubblicazione: (2024)
Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text
di: Mohamed, Amr, et al.
Pubblicazione: (2025)
di: Mohamed, Amr, et al.
Pubblicazione: (2025)
DOCBENCH: A Benchmark for Evaluating LLM-based Document Reading Systems
di: Zou, Anni, et al.
Pubblicazione: (2024)
di: Zou, Anni, et al.
Pubblicazione: (2024)
Ψ-Arena: Interactive Assessment and Optimization of LLM-based Psychological Counselors with Tripartite Feedback
di: Zhu, Shijing, et al.
Pubblicazione: (2025)
di: Zhu, Shijing, et al.
Pubblicazione: (2025)
TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories
di: Dong, Honghua, et al.
Pubblicazione: (2025)
di: Dong, Honghua, et al.
Pubblicazione: (2025)
In-Context Examples Matter: Improving Emotion Recognition in Conversation with Instruction Tuning
di: Ma, Hui, et al.
Pubblicazione: (2025)
di: Ma, Hui, et al.
Pubblicazione: (2025)
CSCE: Boosting LLM Reasoning by Simultaneous Enhancing of Causal Significance and Consistency
di: Wang, Kangsheng, et al.
Pubblicazione: (2024)
di: Wang, Kangsheng, et al.
Pubblicazione: (2024)
LLM Hallucination Detection: HSAD
di: Li, JinXin, et al.
Pubblicazione: (2025)
di: Li, JinXin, et al.
Pubblicazione: (2025)
ScreenLLM: Stateful Screen Schema for Efficient Action Understanding and Prediction
di: Jin, Yiqiao, et al.
Pubblicazione: (2025)
di: Jin, Yiqiao, et al.
Pubblicazione: (2025)
LLM-based NLG Evaluation: Current Status and Challenges
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
Enhancing LLM Reasoning with Multi-Path Collaborative Reactive and Reflection agents
di: He, Chengbo, et al.
Pubblicazione: (2024)
di: He, Chengbo, et al.
Pubblicazione: (2024)
Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding
di: Yoo, Haneul, et al.
Pubblicazione: (2024)
di: Yoo, Haneul, et al.
Pubblicazione: (2024)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
Benchmarking LLM Guardrails in Handling Multilingual Toxicity
di: Yang, Yahan, et al.
Pubblicazione: (2024)
di: Yang, Yahan, et al.
Pubblicazione: (2024)
Explaining Length Bias in LLM-Based Preference Evaluations
di: Hu, Zhengyu, et al.
Pubblicazione: (2024)
di: Hu, Zhengyu, et al.
Pubblicazione: (2024)
WEST: LLM based Speech Toolkit for Speech Understanding, Generation, and Interaction
di: Zhang, Binbin, et al.
Pubblicazione: (2025)
di: Zhang, Binbin, et al.
Pubblicazione: (2025)
HEART-Bench: Do LLM Agents Exhibit Human-like Psychology?
di: Peng, Weihan, et al.
Pubblicazione: (2026)
di: Peng, Weihan, et al.
Pubblicazione: (2026)
MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors
di: Luo, Xiaotian, et al.
Pubblicazione: (2026)
di: Luo, Xiaotian, et al.
Pubblicazione: (2026)
Are LLM-based Evaluators Confusing NLG Quality Criteria?
di: Hu, Xinyu, et al.
Pubblicazione: (2024)
di: Hu, Xinyu, et al.
Pubblicazione: (2024)
DuanzAI: Slang-Enhanced LLM with Prompt for Humor Understanding
di: Rohn, Yesian
Pubblicazione: (2024)
di: Rohn, Yesian
Pubblicazione: (2024)
Code Fingerprints: Disentangled Attribution of LLM-Generated Code
di: Guo, Jiaxun, et al.
Pubblicazione: (2026)
di: Guo, Jiaxun, et al.
Pubblicazione: (2026)
Citation-Enhanced Generation for LLM-based Chatbots
di: Li, Weitao, et al.
Pubblicazione: (2024)
di: Li, Weitao, et al.
Pubblicazione: (2024)
LLM-Guided Strategy Synthesis for Scalable Equality Saturation
di: Yin, Chenyun, et al.
Pubblicazione: (2026)
di: Yin, Chenyun, et al.
Pubblicazione: (2026)
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
di: Cao, Yixin, et al.
Pubblicazione: (2025)
di: Cao, Yixin, et al.
Pubblicazione: (2025)
IDGen: Item Discrimination Induced Prompt Generation for LLM Evaluation
di: Lin, Fan, et al.
Pubblicazione: (2024)
di: Lin, Fan, et al.
Pubblicazione: (2024)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
di: Huang, Jiazhen, et al.
Pubblicazione: (2026)
di: Huang, Jiazhen, et al.
Pubblicazione: (2026)
BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation
di: Sun, Peng, et al.
Pubblicazione: (2026)
di: Sun, Peng, et al.
Pubblicazione: (2026)
Understanding LLM Embeddings for Regression
di: Tang, Eric, et al.
Pubblicazione: (2024)
di: Tang, Eric, et al.
Pubblicazione: (2024)
LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning
di: Meng, Silin, et al.
Pubblicazione: (2024)
di: Meng, Silin, et al.
Pubblicazione: (2024)
Exploring LLM Multi-Agents for ICD Coding
di: Li, Rumeng, et al.
Pubblicazione: (2024)
di: Li, Rumeng, et al.
Pubblicazione: (2024)
HuggingGraph: Understanding the Supply Chain of LLM Ecosystem
di: Rahman, Mohammad Shahedur, et al.
Pubblicazione: (2025)
di: Rahman, Mohammad Shahedur, et al.
Pubblicazione: (2025)
RocketEval: Efficient Automated LLM Evaluation via Grading Checklist
di: Wei, Tianjun, et al.
Pubblicazione: (2025)
di: Wei, Tianjun, et al.
Pubblicazione: (2025)
PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
di: Yang, Dongjie, et al.
Pubblicazione: (2024)
di: Yang, Dongjie, et al.
Pubblicazione: (2024)
HTAA: Enhancing LLM Planning via Hybrid Toolset Agentization & Adaptation
di: Huang, Chengrui, et al.
Pubblicazione: (2026)
di: Huang, Chengrui, et al.
Pubblicazione: (2026)
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
di: Yang, Hang, et al.
Pubblicazione: (2024)
di: Yang, Hang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Psyche-R1: Towards Reliable Psychological LLMs through Unified Empathy, Expertise, and Reasoning
di: Dai, Chongyuan, et al.
Pubblicazione: (2025) -
Traits Run Deep: Enhancing Personality Assessment via Psychology-Guided LLM Representations and Multimodal Apparent Behaviors
di: Li, Jia, et al.
Pubblicazione: (2025) -
Think-Augmented Function Calling: Improving LLM Parameter Accuracy Through Embedded Reasoning
di: Wei, Lei, et al.
Pubblicazione: (2026) -
AgentMental: An Interactive Multi-Agent Framework for Explainable and Adaptive Mental Health Assessment
di: Hu, Jinpeng, et al.
Pubblicazione: (2025) -
Unveiling LLM Evaluation Focused on Metrics: Challenges and Solutions
di: Hu, Taojun, et al.
Pubblicazione: (2024)