Benchmarking Motivational Interviewing Competence of Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jha, Aishwariya, Shivaprakash, Prakrithi, Shukla, Lekhansh, Mukherjee, Animesh, Chand, Prabhat, Murthy, Pratima |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
par: Shivaprakash, Prakrithi, et autres
Publié: (2025)
par: Shivaprakash, Prakrithi, et autres
Publié: (2025)
Lost without translation -- Can transformer (language models) understand mood states?
par: Shivaprakash, Prakrithi, et autres
Publié: (2025)
par: Shivaprakash, Prakrithi, et autres
Publié: (2025)
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
par: Haritha Gireesh, et autres
Publié: (2025)
par: Haritha Gireesh, et autres
Publié: (2025)
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
par: Kumar, Subham, et autres
Publié: (2025)
par: Kumar, Subham, et autres
Publié: (2025)
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
par: Kumar, Subham, et autres
Publié: (2025)
par: Kumar, Subham, et autres
Publié: (2025)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
par: Banerjee, Somnath, et autres
Publié: (2026)
par: Banerjee, Somnath, et autres
Publié: (2026)
When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing
par: Kong, Haein, et autres
Publié: (2025)
par: Kong, Haein, et autres
Publié: (2025)
AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
par: Adak, Sayantan, et autres
Publié: (2025)
par: Adak, Sayantan, et autres
Publié: (2025)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
par: Banerjee, Somnath, et autres
Publié: (2024)
par: Banerjee, Somnath, et autres
Publié: (2024)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
par: Banerjee, Somnath, et autres
Publié: (2026)
par: Banerjee, Somnath, et autres
Publié: (2026)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
par: Brown, Andrew, et autres
Publié: (2024)
par: Brown, Andrew, et autres
Publié: (2024)
Motivation in Large Language Models
par: Nahum, Omer, et autres
Publié: (2026)
par: Nahum, Omer, et autres
Publié: (2026)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
par: Adak, Sayantan, et autres
Publié: (2025)
par: Adak, Sayantan, et autres
Publié: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
par: Adak, Sayantan, et autres
Publié: (2024)
par: Adak, Sayantan, et autres
Publié: (2024)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
par: Bohacek, Maty, et autres
Publié: (2025)
par: Bohacek, Maty, et autres
Publié: (2025)
Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation
par: Mukherjee, Arka, et autres
Publié: (2025)
par: Mukherjee, Arka, et autres
Publié: (2025)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
par: Nag, Arijit, et autres
Publié: (2024)
par: Nag, Arijit, et autres
Publié: (2024)
Attributional Safety Failures in Large Language Models under Code-Mixed Perturbations
par: Banerjee, Somnath, et autres
Publié: (2025)
par: Banerjee, Somnath, et autres
Publié: (2025)
UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
par: Adeeba, Farah, et autres
Publié: (2025)
par: Adeeba, Farah, et autres
Publié: (2025)
Evaluating Clinical Competencies of Large Language Models with a General Practice Benchmark
par: Li, Zheqing, et autres
Publié: (2025)
par: Li, Zheqing, et autres
Publié: (2025)
Consistent Client Simulation for Motivational Interviewing-based Counseling
par: Yang, Yizhe, et autres
Publié: (2025)
par: Yang, Yizhe, et autres
Publié: (2025)
PapersPlease: A Benchmark for Evaluating Motivational Values of Large Language Models Based on ERG Theory
par: Myung, Junho, et autres
Publié: (2025)
par: Myung, Junho, et autres
Publié: (2025)
Traces of Social Competence in Large Language Models
par: Kouwenhoven, Tom, et autres
Publié: (2026)
par: Kouwenhoven, Tom, et autres
Publié: (2026)
Evaluating the Deductive Competence of Large Language Models
par: Seals, Spencer M., et autres
Publié: (2023)
par: Seals, Spencer M., et autres
Publié: (2023)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
par: Waldis, Andreas, et autres
Publié: (2024)
par: Waldis, Andreas, et autres
Publié: (2024)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
par: Rai, Anand, et autres
Publié: (2025)
par: Rai, Anand, et autres
Publié: (2025)
ProSocialAlign: Preference Conditioned Test Time Alignment in Language Models
par: Banerjee, Somnath, et autres
Publié: (2025)
par: Banerjee, Somnath, et autres
Publié: (2025)
EMMI -- Empathic Multimodal Motivational Interviews Dataset: Analyses and Annotations
par: Galland, Lucie, et autres
Publié: (2024)
par: Galland, Lucie, et autres
Publié: (2024)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
par: Park, Dojun, et autres
Publié: (2024)
par: Park, Dojun, et autres
Publié: (2024)
Efficient Continual Pre-training of LLMs for Low-resource Languages
par: Nag, Arijit, et autres
Publié: (2024)
par: Nag, Arijit, et autres
Publié: (2024)
Extrinsic Evaluation of Cultural Competence in Large Language Models
par: Bhatt, Shaily, et autres
Publié: (2024)
par: Bhatt, Shaily, et autres
Publié: (2024)
A Novel Psychometrics-Based Approach to Developing Professional Competency Benchmark for Large Language Models
par: Kardanova, Elena, et autres
Publié: (2024)
par: Kardanova, Elena, et autres
Publié: (2024)
Examining Spanish Counseling with MIDAS: a Motivational Interviewing Dataset in Spanish
par: Gunal, Aylin, et autres
Publié: (2025)
par: Gunal, Aylin, et autres
Publié: (2025)
Exploring a New Competency Modeling Process with Large Language Models
par: Du, Silin, et autres
Publié: (2026)
par: Du, Silin, et autres
Publié: (2026)
Parallel Corpus Augmentation using Masked Language Models
par: Kumari, Vibhuti, et autres
Publié: (2024)
par: Kumari, Vibhuti, et autres
Publié: (2024)
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia
par: Ayash, Lama, et autres
Publié: (2025)
par: Ayash, Lama, et autres
Publié: (2025)
Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
par: Banerjee, Somnath, et autres
Publié: (2024)
par: Banerjee, Somnath, et autres
Publié: (2024)
Metadata Conditioned Large Language Models for Localization
par: Mukherjee, Anjishnu, et autres
Publié: (2026)
par: Mukherjee, Anjishnu, et autres
Publié: (2026)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
par: Acharya, Arkadeep, et autres
Publié: (2024)
par: Acharya, Arkadeep, et autres
Publié: (2024)
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning
par: Xie, Zhouhang, et autres
Publié: (2024)
par: Xie, Zhouhang, et autres
Publié: (2024)
Documents similaires
-
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
par: Shivaprakash, Prakrithi, et autres
Publié: (2025) -
Lost without translation -- Can transformer (language models) understand mood states?
par: Shivaprakash, Prakrithi, et autres
Publié: (2025) -
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
par: Haritha Gireesh, et autres
Publié: (2025) -
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
par: Kumar, Subham, et autres
Publié: (2025) -
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
par: Kumar, Subham, et autres
Publié: (2025)