Can You Share Your Story? Modeling Clients' Metacognition and Openness for LLM Therapist Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Kim, Minju, Yoo, Dongje, Hwang, Yeonjun, Kang, Minseok, Kim, Namyoung, Gwak, Minju, Kwak, Beong-woo, Chae, Hyungjoo, Kim, Harim, Lee, Yunjoong, Kim, Min Hee, Jung, Dayi, Chung, Kyong-Mee, Yeo, Jinyoung |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ToolHaystack: Stress-Testing Tool-Augmented Language Models in Realistic Long-Term Interactions
por: Kwak, Beong-woo, et al.
Publicado: (2025)
por: Kwak, Beong-woo, et al.
Publicado: (2025)
Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web Navigation
por: Chae, Hyungjoo, et al.
Publicado: (2024)
por: Chae, Hyungjoo, et al.
Publicado: (2024)
Web-Shepherd: Advancing PRMs for Reinforcing Web Agents
por: Chae, Hyungjoo, et al.
Publicado: (2025)
por: Chae, Hyungjoo, et al.
Publicado: (2025)
PRINCIPLES: Synthetic Strategy Memory for Proactive Dialogue Agents
por: Kim, Namyoung, et al.
Publicado: (2025)
por: Kim, Namyoung, et al.
Publicado: (2025)
Pearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset
por: Kim, Minjin, et al.
Publicado: (2024)
por: Kim, Minjin, et al.
Publicado: (2024)
Towards Lifelong Dialogue Agents via Timeline-based Memory Management
por: Ong, Kai Tzu-iunn, et al.
Publicado: (2024)
por: Ong, Kai Tzu-iunn, et al.
Publicado: (2024)
CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space
por: Hwang, Yeonjun, et al.
Publicado: (2026)
por: Hwang, Yeonjun, et al.
Publicado: (2026)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
por: Lee, Seungbeen, et al.
Publicado: (2024)
por: Lee, Seungbeen, et al.
Publicado: (2024)
One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL
por: Chae, Hyungjoo, et al.
Publicado: (2025)
por: Chae, Hyungjoo, et al.
Publicado: (2025)
Cactus: Towards Psychological Counseling Conversations using Cognitive Behavioral Theory
por: Lee, Suyeon, et al.
Publicado: (2024)
por: Lee, Suyeon, et al.
Publicado: (2024)
Language Models as Compilers: Simulating Pseudocode Execution Improves Algorithmic Reasoning in Language Models
por: Chae, Hyungjoo, et al.
Publicado: (2024)
por: Chae, Hyungjoo, et al.
Publicado: (2024)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
Revisiting the UID Hypothesis in LLM Reasoning Traces
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
LLM Meets Scene Graph: Can Large Language Models Understand and Generate Scene Graphs? A Benchmark and Empirical Study
por: Yang, Dongil, et al.
Publicado: (2025)
por: Yang, Dongil, et al.
Publicado: (2025)
Embodied Agents Meet Personalization: Investigating Challenges and Solutions Through the Lens of Memory Utilization
por: Kwon, Taeyoon, et al.
Publicado: (2025)
por: Kwon, Taeyoon, et al.
Publicado: (2025)
EMBGuard: Constructing Hazard-Aware Guardrails for Safe Planning in Embodied Agents
por: Choi, Dongwook, et al.
Publicado: (2026)
por: Choi, Dongwook, et al.
Publicado: (2026)
VerifiNER: Verification-augmented NER via Knowledge-grounded Reasoning with Large Language Models
por: Kim, Seoyeon, et al.
Publicado: (2024)
por: Kim, Seoyeon, et al.
Publicado: (2024)
How Do Teachers Create Pedagogical Chatbots?: Current Practices and Challenges
por: Yoo, Minju, et al.
Publicado: (2025)
por: Yoo, Minju, et al.
Publicado: (2025)
LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training
por: Gwak, Minju, et al.
Publicado: (2026)
por: Gwak, Minju, et al.
Publicado: (2026)
On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length
por: Kim, Sunghwan, et al.
Publicado: (2026)
por: Kim, Sunghwan, et al.
Publicado: (2026)
Coffee-Gym: An Environment for Evaluating and Improving Natural Language Feedback on Erroneous Code
por: Chae, Hyungjoo, et al.
Publicado: (2024)
por: Chae, Hyungjoo, et al.
Publicado: (2024)
Rethinking Reward Model Evaluation Through the Lens of Reward Overoptimization
por: Kim, Sunghwan, et al.
Publicado: (2025)
por: Kim, Sunghwan, et al.
Publicado: (2025)
Convergence of orbital integrals on unitary groups in positive characteristic
por: Kim, Wansu, et al.
Publicado: (2026)
por: Kim, Wansu, et al.
Publicado: (2026)
PaP-NF: Probabilistic Long-Term Time Series Forecasting via Prefix-as-Prompt Reprogramming and Normalizing Flows
por: Kim, Minju, et al.
Publicado: (2026)
por: Kim, Minju, et al.
Publicado: (2026)
Sounds of Hidden Agents: The Development of Causal Reasoning About Musical Sounds
por: Minju Kim, et al.
Publicado: (2025)
por: Minju Kim, et al.
Publicado: (2025)
SecureMCP: A Policy-Enforced LLM Data Access Framework for AIoT Systems via Model Context Protocol
por: Kim, Wonbae, et al.
Publicado: (2026)
por: Kim, Wonbae, et al.
Publicado: (2026)
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt
por: Lee, Suyeon, et al.
Publicado: (2024)
por: Lee, Suyeon, et al.
Publicado: (2024)
Phase dependence of growth mechanisms in the daily energetics of the North Atlantic Oscillation
por: Minju Kim, et al.
Publicado: (2024)
por: Minju Kim, et al.
Publicado: (2024)
LEGO-Eval: Towards Fine-Grained Evaluation on Synthesizing 3D Embodied Environments with Tool Augmentation
por: Hwangbo, Gyeom, et al.
Publicado: (2025)
por: Hwangbo, Gyeom, et al.
Publicado: (2025)
Polymerization‐Induced Direct Photolithography of Quantum Dots
por: Taehyung Kim, et al.
Publicado: (2025)
por: Taehyung Kim, et al.
Publicado: (2025)
Fine-Grained and Thematic Evaluation of LLMs in Social Deduction Game
por: Kim, Byungjun, et al.
Publicado: (2024)
por: Kim, Byungjun, et al.
Publicado: (2024)
Leveraging Large Language Models for Active Merchant Non-player Characters
por: Kim, Byungjun, et al.
Publicado: (2024)
por: Kim, Byungjun, et al.
Publicado: (2024)
PHISH in MESH: Korean Adversarial Phonetic Substitution and Phonetic-Semantic Feature Integration Defense
por: Kim, Byungjun, et al.
Publicado: (2025)
por: Kim, Byungjun, et al.
Publicado: (2025)
PAC-BENCH: Evaluating Multi-Agent Collaboration under Privacy Constraints
por: Park, Minjun, et al.
Publicado: (2026)
por: Park, Minjun, et al.
Publicado: (2026)
Horocycles in hyperbolic 3-manifolds with round Sierpiński limit sets
por: Kim, Dongryul M., et al.
Publicado: (2025)
por: Kim, Dongryul M., et al.
Publicado: (2025)
Evaluating Robustness of Reward Models for Mathematical Reasoning
por: Kim, Sunghwan, et al.
Publicado: (2024)
por: Kim, Sunghwan, et al.
Publicado: (2024)
Multiple Photon Subtraction on Light with Tunable Intensity Correlations
por: Minju Kim, et al.
Publicado: (2026)
por: Minju Kim, et al.
Publicado: (2026)
A Dual-Prompting for Interpretable Mental Health Language Models
por: Jeon, Hyolim, et al.
Publicado: (2024)
por: Jeon, Hyolim, et al.
Publicado: (2024)
YA-TA: Towards Personalized Question-Answering Teaching Assistants using Instructor-Student Dual Retrieval-augmented Knowledge Fusion
por: Yang, Dongil, et al.
Publicado: (2024)
por: Yang, Dongil, et al.
Publicado: (2024)
Can LLMs and humans be friends? Uncovering factors affecting human-AI intimacy formation
por: Hong, Yeseon, et al.
Publicado: (2025)
por: Hong, Yeseon, et al.
Publicado: (2025)
Ejemplares similares
-
ToolHaystack: Stress-Testing Tool-Augmented Language Models in Realistic Long-Term Interactions
por: Kwak, Beong-woo, et al.
Publicado: (2025) -
Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web Navigation
por: Chae, Hyungjoo, et al.
Publicado: (2024) -
Web-Shepherd: Advancing PRMs for Reinforcing Web Agents
por: Chae, Hyungjoo, et al.
Publicado: (2025) -
PRINCIPLES: Synthetic Strategy Memory for Proactive Dialogue Agents
por: Kim, Namyoung, et al.
Publicado: (2025) -
Pearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset
por: Kim, Minjin, et al.
Publicado: (2024)