Assessing LLM Response Quality in the Context of Technology-Facilitated Abuse

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Prakash, Vijay, Almansoori, Majed, Hu, Donghan, Chatterjee, Rahul, Huang, Danny Yuxing
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914339531784192
author Prakash, Vijay
Almansoori, Majed
Hu, Donghan
Chatterjee, Rahul
Huang, Danny Yuxing
author_facet Prakash, Vijay
Almansoori, Majed
Hu, Donghan
Chatterjee, Rahul
Huang, Danny Yuxing
contents Technology-facilitated abuse (TFA) is a pervasive form of intimate partner violence (IPV) that leverages digital tools to control, surveil, or harm survivors. While tech clinics are one of the reliable sources of support for TFA survivors, they face limitations due to staffing constraints and logistical barriers. As a result, many survivors turn to online resources for assistance. With the growing accessibility and popularity of large language models (LLMs), and increasing interest from IPV organizations, survivors may begin to consult LLM-based chatbots before seeking help from tech clinics. In this work, we present the first expert-led manual evaluation of four LLMs - two widely used general-purpose non-reasoning models and two domain-specific models designed for IPV contexts - focused on their effectiveness in responding to TFA-related questions. Using real-world questions collected from literature and online forums, we assess the quality of zero-shot single-turn LLM responses generated with a survivor safety-centered prompt on criteria tailored to the TFA domain. Additionally, we conducted a user study to evaluate the perceived actionability of these responses from the perspective of individuals who have experienced TFA. Our findings, grounded in both expert assessment and user feedback, provide insights into the current capabilities and limitations of LLMs in the TFA context and may inform the design, development, and fine-tuning of future models for this domain. We conclude with concrete recommendations to improve LLM performance for survivor support.
format Preprint
id arxiv_https___arxiv_org_abs_2602_17672
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Assessing LLM Response Quality in the Context of Technology-Facilitated Abuse
Prakash, Vijay
Almansoori, Majed
Hu, Donghan
Chatterjee, Rahul
Huang, Danny Yuxing
Human-Computer Interaction
Artificial Intelligence
Computation and Language
Cryptography and Security
Computers and Society
Technology-facilitated abuse (TFA) is a pervasive form of intimate partner violence (IPV) that leverages digital tools to control, surveil, or harm survivors. While tech clinics are one of the reliable sources of support for TFA survivors, they face limitations due to staffing constraints and logistical barriers. As a result, many survivors turn to online resources for assistance. With the growing accessibility and popularity of large language models (LLMs), and increasing interest from IPV organizations, survivors may begin to consult LLM-based chatbots before seeking help from tech clinics. In this work, we present the first expert-led manual evaluation of four LLMs - two widely used general-purpose non-reasoning models and two domain-specific models designed for IPV contexts - focused on their effectiveness in responding to TFA-related questions. Using real-world questions collected from literature and online forums, we assess the quality of zero-shot single-turn LLM responses generated with a survivor safety-centered prompt on criteria tailored to the TFA domain. Additionally, we conducted a user study to evaluate the perceived actionability of these responses from the perspective of individuals who have experienced TFA. Our findings, grounded in both expert assessment and user feedback, provide insights into the current capabilities and limitations of LLMs in the TFA context and may inform the design, development, and fine-tuning of future models for this domain. We conclude with concrete recommendations to improve LLM performance for survivor support.
title Assessing LLM Response Quality in the Context of Technology-Facilitated Abuse
topic Human-Computer Interaction
Artificial Intelligence
Computation and Language
Cryptography and Security
Computers and Society
url https://arxiv.org/abs/2602.17672