Benchmarking Motivational Interviewing Competence of Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Jha, Aishwariya, Shivaprakash, Prakrithi, Shukla, Lekhansh, Mukherjee, Animesh, Chand, Prabhat, Murthy, Pratima |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025)
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025)
Lost without translation -- Can transformer (language models) understand mood states?
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025)
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025)
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
por: Haritha Gireesh, et al.
Publicado: (2025)
por: Haritha Gireesh, et al.
Publicado: (2025)
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
por: Kumar, Subham, et al.
Publicado: (2025)
por: Kumar, Subham, et al.
Publicado: (2025)
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
por: Kumar, Subham, et al.
Publicado: (2025)
por: Kumar, Subham, et al.
Publicado: (2025)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
por: Banerjee, Somnath, et al.
Publicado: (2026)
por: Banerjee, Somnath, et al.
Publicado: (2026)
When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing
por: Kong, Haein, et al.
Publicado: (2025)
por: Kong, Haein, et al.
Publicado: (2025)
AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
por: Adak, Sayantan, et al.
Publicado: (2025)
por: Adak, Sayantan, et al.
Publicado: (2025)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
por: Banerjee, Somnath, et al.
Publicado: (2024)
por: Banerjee, Somnath, et al.
Publicado: (2024)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
por: Banerjee, Somnath, et al.
Publicado: (2026)
por: Banerjee, Somnath, et al.
Publicado: (2026)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
por: Brown, Andrew, et al.
Publicado: (2024)
por: Brown, Andrew, et al.
Publicado: (2024)
Motivation in Large Language Models
por: Nahum, Omer, et al.
Publicado: (2026)
por: Nahum, Omer, et al.
Publicado: (2026)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
por: Adak, Sayantan, et al.
Publicado: (2025)
por: Adak, Sayantan, et al.
Publicado: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
por: Adak, Sayantan, et al.
Publicado: (2024)
por: Adak, Sayantan, et al.
Publicado: (2024)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
por: Bohacek, Maty, et al.
Publicado: (2025)
por: Bohacek, Maty, et al.
Publicado: (2025)
Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation
por: Mukherjee, Arka, et al.
Publicado: (2025)
por: Mukherjee, Arka, et al.
Publicado: (2025)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
por: Nag, Arijit, et al.
Publicado: (2024)
por: Nag, Arijit, et al.
Publicado: (2024)
Attributional Safety Failures in Large Language Models under Code-Mixed Perturbations
por: Banerjee, Somnath, et al.
Publicado: (2025)
por: Banerjee, Somnath, et al.
Publicado: (2025)
UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
por: Adeeba, Farah, et al.
Publicado: (2025)
por: Adeeba, Farah, et al.
Publicado: (2025)
Evaluating Clinical Competencies of Large Language Models with a General Practice Benchmark
por: Li, Zheqing, et al.
Publicado: (2025)
por: Li, Zheqing, et al.
Publicado: (2025)
Consistent Client Simulation for Motivational Interviewing-based Counseling
por: Yang, Yizhe, et al.
Publicado: (2025)
por: Yang, Yizhe, et al.
Publicado: (2025)
PapersPlease: A Benchmark for Evaluating Motivational Values of Large Language Models Based on ERG Theory
por: Myung, Junho, et al.
Publicado: (2025)
por: Myung, Junho, et al.
Publicado: (2025)
Traces of Social Competence in Large Language Models
por: Kouwenhoven, Tom, et al.
Publicado: (2026)
por: Kouwenhoven, Tom, et al.
Publicado: (2026)
Evaluating the Deductive Competence of Large Language Models
por: Seals, Spencer M., et al.
Publicado: (2023)
por: Seals, Spencer M., et al.
Publicado: (2023)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
por: Waldis, Andreas, et al.
Publicado: (2024)
por: Waldis, Andreas, et al.
Publicado: (2024)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
por: Rai, Anand, et al.
Publicado: (2025)
por: Rai, Anand, et al.
Publicado: (2025)
ProSocialAlign: Preference Conditioned Test Time Alignment in Language Models
por: Banerjee, Somnath, et al.
Publicado: (2025)
por: Banerjee, Somnath, et al.
Publicado: (2025)
EMMI -- Empathic Multimodal Motivational Interviews Dataset: Analyses and Annotations
por: Galland, Lucie, et al.
Publicado: (2024)
por: Galland, Lucie, et al.
Publicado: (2024)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
por: Park, Dojun, et al.
Publicado: (2024)
por: Park, Dojun, et al.
Publicado: (2024)
Efficient Continual Pre-training of LLMs for Low-resource Languages
por: Nag, Arijit, et al.
Publicado: (2024)
por: Nag, Arijit, et al.
Publicado: (2024)
Extrinsic Evaluation of Cultural Competence in Large Language Models
por: Bhatt, Shaily, et al.
Publicado: (2024)
por: Bhatt, Shaily, et al.
Publicado: (2024)
A Novel Psychometrics-Based Approach to Developing Professional Competency Benchmark for Large Language Models
por: Kardanova, Elena, et al.
Publicado: (2024)
por: Kardanova, Elena, et al.
Publicado: (2024)
Examining Spanish Counseling with MIDAS: a Motivational Interviewing Dataset in Spanish
por: Gunal, Aylin, et al.
Publicado: (2025)
por: Gunal, Aylin, et al.
Publicado: (2025)
Exploring a New Competency Modeling Process with Large Language Models
por: Du, Silin, et al.
Publicado: (2026)
por: Du, Silin, et al.
Publicado: (2026)
Parallel Corpus Augmentation using Masked Language Models
por: Kumari, Vibhuti, et al.
Publicado: (2024)
por: Kumari, Vibhuti, et al.
Publicado: (2024)
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia
por: Ayash, Lama, et al.
Publicado: (2025)
por: Ayash, Lama, et al.
Publicado: (2025)
Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
por: Banerjee, Somnath, et al.
Publicado: (2024)
por: Banerjee, Somnath, et al.
Publicado: (2024)
Metadata Conditioned Large Language Models for Localization
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
por: Mukherjee, Anjishnu, et al.
Publicado: (2026)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
por: Acharya, Arkadeep, et al.
Publicado: (2024)
por: Acharya, Arkadeep, et al.
Publicado: (2024)
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning
por: Xie, Zhouhang, et al.
Publicado: (2024)
por: Xie, Zhouhang, et al.
Publicado: (2024)
Ejemplares similares
-
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025) -
Lost without translation -- Can transformer (language models) understand mood states?
por: Shivaprakash, Prakrithi, et al.
Publicado: (2025) -
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
por: Haritha Gireesh, et al.
Publicado: (2025) -
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
por: Kumar, Subham, et al.
Publicado: (2025) -
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
por: Kumar, Subham, et al.
Publicado: (2025)