Benchmarking Motivational Interviewing Competence of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jha, Aishwariya, Shivaprakash, Prakrithi, Shukla, Lekhansh, Mukherjee, Animesh, Chand, Prabhat, Murthy, Pratima |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
Lost without translation -- Can transformer (language models) understand mood states?
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025)
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
von: Haritha Gireesh, et al.
Veröffentlicht: (2025)
von: Haritha Gireesh, et al.
Veröffentlicht: (2025)
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing
von: Kong, Haein, et al.
Veröffentlicht: (2025)
von: Kong, Haein, et al.
Veröffentlicht: (2025)
AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
von: Adak, Sayantan, et al.
Veröffentlicht: (2025)
von: Adak, Sayantan, et al.
Veröffentlicht: (2025)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
Motivation in Large Language Models
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
von: Adak, Sayantan, et al.
Veröffentlicht: (2025)
von: Adak, Sayantan, et al.
Veröffentlicht: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
von: Adak, Sayantan, et al.
Veröffentlicht: (2024)
von: Adak, Sayantan, et al.
Veröffentlicht: (2024)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
von: Bohacek, Maty, et al.
Veröffentlicht: (2025)
von: Bohacek, Maty, et al.
Veröffentlicht: (2025)
Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation
von: Mukherjee, Arka, et al.
Veröffentlicht: (2025)
von: Mukherjee, Arka, et al.
Veröffentlicht: (2025)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
von: Nag, Arijit, et al.
Veröffentlicht: (2024)
von: Nag, Arijit, et al.
Veröffentlicht: (2024)
Attributional Safety Failures in Large Language Models under Code-Mixed Perturbations
von: Banerjee, Somnath, et al.
Veröffentlicht: (2025)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2025)
UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
von: Adeeba, Farah, et al.
Veröffentlicht: (2025)
von: Adeeba, Farah, et al.
Veröffentlicht: (2025)
Evaluating Clinical Competencies of Large Language Models with a General Practice Benchmark
von: Li, Zheqing, et al.
Veröffentlicht: (2025)
von: Li, Zheqing, et al.
Veröffentlicht: (2025)
Consistent Client Simulation for Motivational Interviewing-based Counseling
von: Yang, Yizhe, et al.
Veröffentlicht: (2025)
von: Yang, Yizhe, et al.
Veröffentlicht: (2025)
PapersPlease: A Benchmark for Evaluating Motivational Values of Large Language Models Based on ERG Theory
von: Myung, Junho, et al.
Veröffentlicht: (2025)
von: Myung, Junho, et al.
Veröffentlicht: (2025)
Traces of Social Competence in Large Language Models
von: Kouwenhoven, Tom, et al.
Veröffentlicht: (2026)
von: Kouwenhoven, Tom, et al.
Veröffentlicht: (2026)
Evaluating the Deductive Competence of Large Language Models
von: Seals, Spencer M., et al.
Veröffentlicht: (2023)
von: Seals, Spencer M., et al.
Veröffentlicht: (2023)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
von: Rai, Anand, et al.
Veröffentlicht: (2025)
von: Rai, Anand, et al.
Veröffentlicht: (2025)
ProSocialAlign: Preference Conditioned Test Time Alignment in Language Models
von: Banerjee, Somnath, et al.
Veröffentlicht: (2025)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2025)
EMMI -- Empathic Multimodal Motivational Interviews Dataset: Analyses and Annotations
von: Galland, Lucie, et al.
Veröffentlicht: (2024)
von: Galland, Lucie, et al.
Veröffentlicht: (2024)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
Efficient Continual Pre-training of LLMs for Low-resource Languages
von: Nag, Arijit, et al.
Veröffentlicht: (2024)
von: Nag, Arijit, et al.
Veröffentlicht: (2024)
Extrinsic Evaluation of Cultural Competence in Large Language Models
von: Bhatt, Shaily, et al.
Veröffentlicht: (2024)
von: Bhatt, Shaily, et al.
Veröffentlicht: (2024)
A Novel Psychometrics-Based Approach to Developing Professional Competency Benchmark for Large Language Models
von: Kardanova, Elena, et al.
Veröffentlicht: (2024)
von: Kardanova, Elena, et al.
Veröffentlicht: (2024)
Examining Spanish Counseling with MIDAS: a Motivational Interviewing Dataset in Spanish
von: Gunal, Aylin, et al.
Veröffentlicht: (2025)
von: Gunal, Aylin, et al.
Veröffentlicht: (2025)
Exploring a New Competency Modeling Process with Large Language Models
von: Du, Silin, et al.
Veröffentlicht: (2026)
von: Du, Silin, et al.
Veröffentlicht: (2026)
Parallel Corpus Augmentation using Masked Language Models
von: Kumari, Vibhuti, et al.
Veröffentlicht: (2024)
von: Kumari, Vibhuti, et al.
Veröffentlicht: (2024)
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia
von: Ayash, Lama, et al.
Veröffentlicht: (2025)
von: Ayash, Lama, et al.
Veröffentlicht: (2025)
Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
Metadata Conditioned Large Language Models for Localization
von: Mukherjee, Anjishnu, et al.
Veröffentlicht: (2026)
von: Mukherjee, Anjishnu, et al.
Veröffentlicht: (2026)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning
von: Xie, Zhouhang, et al.
Veröffentlicht: (2024)
von: Xie, Zhouhang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025) -
Lost without translation -- Can transformer (language models) understand mood states?
von: Shivaprakash, Prakrithi, et al.
Veröffentlicht: (2025) -
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
von: Haritha Gireesh, et al.
Veröffentlicht: (2025) -
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
von: Kumar, Subham, et al.
Veröffentlicht: (2025) -
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
von: Kumar, Subham, et al.
Veröffentlicht: (2025)