Promoting Online Safety by Simulating Unsafe Conversations with LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Hoffman, Owen, Peng, Kangze, You, Zehua, Kamal, Sajid, Venkatagiri, Sukrit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ScamPilot: Simulating Conversations with LLMs to Protect Against Online Scams
by: Hoffman, Owen, et al.
Published: (2026)
by: Hoffman, Owen, et al.
Published: (2026)
OSINT Research Studios: A Flexible Crowdsourcing Framework to Scale Up Open Source Intelligence Investigations
by: Mukhopadhyay, Anirban, et al.
Published: (2024)
by: Mukhopadhyay, Anirban, et al.
Published: (2024)
Examining Solidarity Against AI-Enabled Surveillance at the Intersection of Workplace and Carceral Realities
by: McErlean, Morgan, et al.
Published: (2025)
by: McErlean, Morgan, et al.
Published: (2025)
An Experimental Study of Competitive Market Behavior Through LLMs
by: Jia, Jingru, et al.
Published: (2024)
by: Jia, Jingru, et al.
Published: (2024)
ChoiceMates: Supporting Unfamiliar Online Decision-Making with Multi-Agent Conversational Interactions
by: Park, Jeongeon, et al.
Published: (2023)
by: Park, Jeongeon, et al.
Published: (2023)
MeetMap: Real-Time Collaborative Dialogue Mapping with LLMs in Online Meetings
by: Chen, Xinyue, et al.
Published: (2025)
by: Chen, Xinyue, et al.
Published: (2025)
A Multi-Agent Conversational Bandit Approach to Online Evaluation and Selection of User-Aligned LLM Responses
by: Dai, Xiangxiang, et al.
Published: (2025)
by: Dai, Xiangxiang, et al.
Published: (2025)
Reasoning LLMs for User-Aware Multimodal Conversational Agents
by: Rahimi, Hamed, et al.
Published: (2025)
by: Rahimi, Hamed, et al.
Published: (2025)
Enhancing Patient-Centric Communication: Leveraging LLMs to Simulate Patient Perspectives
by: Ma, Xinyao, et al.
Published: (2025)
by: Ma, Xinyao, et al.
Published: (2025)
RecUserSim: A Realistic and Diverse User Simulator for Evaluating Conversational Recommender Systems
by: Chen, Luyu, et al.
Published: (2025)
by: Chen, Luyu, et al.
Published: (2025)
A LLM-based Controllable, Scalable, Human-Involved User Simulator Framework for Conversational Recommender Systems
by: Zhu, Lixi, et al.
Published: (2024)
by: Zhu, Lixi, et al.
Published: (2024)
AI persuading AI vs AI persuading Humans: LLMs' Differential Effectiveness in Promoting Pro-Environmental Behavior
by: Doudkin, Alexander, et al.
Published: (2025)
by: Doudkin, Alexander, et al.
Published: (2025)
Exploring Conversational Design Choices in LLMs for Pedagogical Purposes: Socratic and Narrative Approaches for Improving Instructor's Teaching Practice
by: Chen, Si, et al.
Published: (2025)
by: Chen, Si, et al.
Published: (2025)
Leveraging Large Language Models (LLMs) to Support Collaborative Human-AI Online Risk Data Annotation
by: Park, Jinkyung, et al.
Published: (2024)
by: Park, Jinkyung, et al.
Published: (2024)
AI Conversational Interviewing: Transforming Surveys with LLMs as Adaptive Interviewers
by: Wuttke, Alexander, et al.
Published: (2024)
by: Wuttke, Alexander, et al.
Published: (2024)
LLM-DaaS: LLM-driven Drone-as-a-Service Operations from Text User Requests
by: Wassim, Lillian, et al.
Published: (2024)
by: Wassim, Lillian, et al.
Published: (2024)
Evaluation of Pedestrian Safety in a High-Fidelity Simulation Environment Framework
by: Ma, Lin, et al.
Published: (2022)
by: Ma, Lin, et al.
Published: (2022)
Direct Advantage Regression: Aligning LLMs with Online AI Reward
by: He, Li, et al.
Published: (2025)
by: He, Li, et al.
Published: (2025)
The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs
by: Zhang, Xing, et al.
Published: (2026)
by: Zhang, Xing, et al.
Published: (2026)
Time Warp: The Gap Between Developers' Ideal vs Actual Workweeks in an AI-Driven Era
by: Kumar, Sukrit, et al.
Published: (2025)
by: Kumar, Sukrit, et al.
Published: (2025)
Is Conversational XAI All You Need? Human-AI Decision Making With a Conversational XAI Assistant
by: He, Gaole, et al.
Published: (2025)
by: He, Gaole, et al.
Published: (2025)
Proactive Conversational Agents with Inner Thoughts
by: Liu, Xingyu Bruce, et al.
Published: (2024)
by: Liu, Xingyu Bruce, et al.
Published: (2024)
SycoEval-EM: Sycophancy Evaluation of Large Language Models in Simulated Clinical Encounters for Emergency Care
by: Peng, Dongshen, et al.
Published: (2026)
by: Peng, Dongshen, et al.
Published: (2026)
ScriptDoctor: Automatic Generation of PuzzleScript Games via Large Language Models and Tree Search
by: Earle, Sam, et al.
Published: (2025)
by: Earle, Sam, et al.
Published: (2025)
Exploring Emotion-Sensitive LLM-Based Conversational AI
by: Brun, Antonin, et al.
Published: (2025)
by: Brun, Antonin, et al.
Published: (2025)
User Misconceptions of LLM-Based Conversational Programming Assistants
by: O'Brien, Gabrielle, et al.
Published: (2025)
by: O'Brien, Gabrielle, et al.
Published: (2025)
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
by: Taneja, Karan, et al.
Published: (2026)
by: Taneja, Karan, et al.
Published: (2026)
The Value-Sensitive Conversational Agent Co-Design Framework
by: Sadek, Malak, et al.
Published: (2023)
by: Sadek, Malak, et al.
Published: (2023)
TherapyProbe: Generating Design Knowledge for Relational Safety in Mental Health Chatbots Through Adversarial Simulation
by: Chandra, Joydeep, et al.
Published: (2026)
by: Chandra, Joydeep, et al.
Published: (2026)
A Dynamic and High-Precision Method for Scenario-Based HRA Synthetic Data Collection in Multi-Agent Collaborative Environments Driven by LLMs
by: Xiao, Xingyu, et al.
Published: (2025)
by: Xiao, Xingyu, et al.
Published: (2025)
Promoting Critical Thinking With Domain-Specific Generative AI Provocations
by: von Davier, Thomas Şerban, et al.
Published: (2026)
by: von Davier, Thomas Şerban, et al.
Published: (2026)
Conversational Explanations: Discussing Explainable AI with Non-AI Experts
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
The Illusion of Empathy: How AI Chatbots Shape Conversation Perception
by: Liu, Tingting, et al.
Published: (2024)
by: Liu, Tingting, et al.
Published: (2024)
Exploring Human-LLM Conversations: Mental Models and the Originator of Toxicity
by: Schneider, Johannes, et al.
Published: (2024)
by: Schneider, Johannes, et al.
Published: (2024)
Automated Visualization Makeovers with LLMs
by: Gangwar, Siddharth, et al.
Published: (2025)
by: Gangwar, Siddharth, et al.
Published: (2025)
Conversational Agents for Older Adults' Health: A Systematic Literature Review
by: An, Jiaxin, et al.
Published: (2025)
by: An, Jiaxin, et al.
Published: (2025)
AI-Enabled Conversational Journaling for Advancing Parkinson's Disease Symptom Tracking
by: Rashik, Mashrur, et al.
Published: (2025)
by: Rashik, Mashrur, et al.
Published: (2025)
AI Standardized Patient Improves Human Conversations in Advanced Cancer Care
by: Haut, Kurtis, et al.
Published: (2025)
by: Haut, Kurtis, et al.
Published: (2025)
How Managers Perceive AI-Assisted Conversational Training for Workplace Communication
by: Wilhelm, Lance T., et al.
Published: (2025)
by: Wilhelm, Lance T., et al.
Published: (2025)
Personalized to Persuade: The Effects of Contextualization and Warmth on Trust and Reliance in Conversational AI
by: Yazan, Mert, et al.
Published: (2026)
by: Yazan, Mert, et al.
Published: (2026)
Similar Items
-
ScamPilot: Simulating Conversations with LLMs to Protect Against Online Scams
by: Hoffman, Owen, et al.
Published: (2026) -
OSINT Research Studios: A Flexible Crowdsourcing Framework to Scale Up Open Source Intelligence Investigations
by: Mukhopadhyay, Anirban, et al.
Published: (2024) -
Examining Solidarity Against AI-Enabled Surveillance at the Intersection of Workplace and Carceral Realities
by: McErlean, Morgan, et al.
Published: (2025) -
An Experimental Study of Competitive Market Behavior Through LLMs
by: Jia, Jingru, et al.
Published: (2024) -
ChoiceMates: Supporting Unfamiliar Online Decision-Making with Multi-Agent Conversational Interactions
by: Park, Jeongeon, et al.
Published: (2023)