Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
Fuente:
arXiv
Guardado en:
| Autores principales: | Park, Jihyung, Afroogh, Saleh, Atkinson, David, Jiao, Junfeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
por: Park, Jihyung, et al.
Publicado: (2026)
por: Park, Jihyung, et al.
Publicado: (2026)
The global landscape of academic guidelines for generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2024)
por: Jiao, Junfeng, et al.
Publicado: (2024)
Navigating LLM Ethics: Advancements, Challenges, and Future Directions
por: Jiao, Junfeng, et al.
Publicado: (2024)
por: Jiao, Junfeng, et al.
Publicado: (2024)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Evaluating the Effectiveness of OpenAI's Parental Control System
por: Ersoz, Kerem, et al.
Publicado: (2026)
por: Ersoz, Kerem, et al.
Publicado: (2026)
SafeMate: A Modular RAG-Based Agent for Context-Aware Emergency Guidance
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
por: Yuan, Youliang, et al.
Publicado: (2024)
por: Yuan, Youliang, et al.
Publicado: (2024)
Balancing Knowledge Delivery and Emotional Comfort in Healthcare Conversational Systems
por: Tsai, Shang-Chi, et al.
Publicado: (2025)
por: Tsai, Shang-Chi, et al.
Publicado: (2025)
Do LLMs "Feel"? Emotion Circuits Discovery and Control
por: Wang, Chenxi, et al.
Publicado: (2025)
por: Wang, Chenxi, et al.
Publicado: (2025)
Dialogue You Can Trust: Human and AI Perspectives on Generated Conversations
por: Ebubechukwu, Ike, et al.
Publicado: (2024)
por: Ebubechukwu, Ike, et al.
Publicado: (2024)
A Survey of Personality, Persona, and Profile in Conversational Agents and Chatbots
por: Sutcliffe, Richard
Publicado: (2023)
por: Sutcliffe, Richard
Publicado: (2023)
Persona-Based Conversational AI: State of the Art and Challenges
por: Liu, Junfeng, et al.
Publicado: (2022)
por: Liu, Junfeng, et al.
Publicado: (2022)
Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR
por: Lee, Chanuk, et al.
Publicado: (2026)
por: Lee, Chanuk, et al.
Publicado: (2026)
Generative AI and LLMs in Industry: A text-mining Analysis and Critical Evaluation of Guidelines and Policy Statements Across Fourteen Industrial Sectors
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
How News Feels: Understanding Affective Bias in Multilingual Headlines for Human-Centered Media Design
por: Ameen, Mohd Ruhul, et al.
Publicado: (2025)
por: Ameen, Mohd Ruhul, et al.
Publicado: (2025)
Mapping out AI Functions in Intelligent Disaster (Mis)Management and AI-Caused Disasters
por: Pouresmaeil, Yasser, et al.
Publicado: (2025)
por: Pouresmaeil, Yasser, et al.
Publicado: (2025)
AI Empathy Erodes Cognitive Autonomy in Younger Users
por: Jiao, Junfeng, et al.
Publicado: (2026)
por: Jiao, Junfeng, et al.
Publicado: (2026)
Towards Supporting Legal Argumentation with NLP: Is More Data Really All You Need?
por: Santosh, T. Y. S. S, et al.
Publicado: (2024)
por: Santosh, T. Y. S. S, et al.
Publicado: (2024)
ShareChat: A Dataset of Chatbot Conversations in the Wild
por: Yan, Yueru, et al.
Publicado: (2025)
por: Yan, Yueru, et al.
Publicado: (2025)
A Complete Survey on LLM-based AI Chatbots
por: Dam, Sumit Kumar, et al.
Publicado: (2024)
por: Dam, Sumit Kumar, et al.
Publicado: (2024)
Towards Proactive Information Probing: Customer Service Chatbots Harvesting Value from Conversation
por: Huang, Chen, et al.
Publicado: (2026)
por: Huang, Chen, et al.
Publicado: (2026)
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
por: Byun, Ju-Seung, et al.
Publicado: (2024)
por: Byun, Ju-Seung, et al.
Publicado: (2024)
Trust in AI: Progress, Challenges, and Future Directions
por: Afroogh, Saleh, et al.
Publicado: (2024)
por: Afroogh, Saleh, et al.
Publicado: (2024)
Mind the Goal: Data-Efficient Goal-Oriented Evaluation of Conversational Agents and Chatbots using Teacher Models
por: Piskala, Deepak Babu, et al.
Publicado: (2025)
por: Piskala, Deepak Babu, et al.
Publicado: (2025)
How Reliable AI Chatbots are for Disease Prediction from Patient Complaints?
por: Nipu, Ayesha Siddika, et al.
Publicado: (2024)
por: Nipu, Ayesha Siddika, et al.
Publicado: (2024)
When PCOS Meets Eating Disorders: An Explainable AI Approach to Detecting the Hidden Triple Burden
por: Prasad, Apoorv, et al.
Publicado: (2026)
por: Prasad, Apoorv, et al.
Publicado: (2026)
IGGA: A Dataset of Industrial Guidelines and Policy Statements for Generative AIs
por: Jiao, Junfeng, et al.
Publicado: (2025)
por: Jiao, Junfeng, et al.
Publicado: (2025)
JU-NLP at Touché: Covert Advertisement in Conversational AI-Generation and Detection Strategies
por: Dutta, Arka, et al.
Publicado: (2025)
por: Dutta, Arka, et al.
Publicado: (2025)
On Overcoming Miscalibrated Conversational Priors in LLM-based Chatbots
por: Herlihy, Christine, et al.
Publicado: (2024)
por: Herlihy, Christine, et al.
Publicado: (2024)
Cash or Comfort? How LLMs Value Your Inconvenience
por: Cedro, Mateusz, et al.
Publicado: (2025)
por: Cedro, Mateusz, et al.
Publicado: (2025)
First-Person Fairness in Chatbots
por: Eloundou, Tyna, et al.
Publicado: (2024)
por: Eloundou, Tyna, et al.
Publicado: (2024)
Does This Look Familiar to You? Knowledge Analysis via Model Internal Representations
por: Park, Sihyun
Publicado: (2025)
por: Park, Sihyun
Publicado: (2025)
Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses
por: Hussain, Khizar, et al.
Publicado: (2026)
por: Hussain, Khizar, et al.
Publicado: (2026)
How Different AI Chatbots Behave? Benchmarking Large Language Models in Behavioral Economics Games
por: Xie, Yutong, et al.
Publicado: (2024)
por: Xie, Yutong, et al.
Publicado: (2024)
Attention Smoothing Is All You Need For Unlearning
por: Zade, Saleh Zare, et al.
Publicado: (2026)
por: Zade, Saleh Zare, et al.
Publicado: (2026)
Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic Interactions
por: Jang, Jihyoung, et al.
Publicado: (2025)
por: Jang, Jihyoung, et al.
Publicado: (2025)
Mitigating the Negative Impact of Over-association for Conversational Query Production
por: Wang, Ante, et al.
Publicado: (2024)
por: Wang, Ante, et al.
Publicado: (2024)
Meanings and Feelings of Large Language Models: Observability of Latent States in Generative AI
por: Liu, Tian Yu, et al.
Publicado: (2024)
por: Liu, Tian Yu, et al.
Publicado: (2024)
ChatEL: Entity Linking with Chatbots
por: Ding, Yifan, et al.
Publicado: (2024)
por: Ding, Yifan, et al.
Publicado: (2024)
Are Chatbots Reliable Text Annotators? Sometimes
por: Kristensen-McLachlan, Ross Deans, et al.
Publicado: (2023)
por: Kristensen-McLachlan, Ross Deans, et al.
Publicado: (2023)
Ejemplares similares
-
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
por: Park, Jihyung, et al.
Publicado: (2026) -
The global landscape of academic guidelines for generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2024) -
Navigating LLM Ethics: Advancements, Challenges, and Future Directions
por: Jiao, Junfeng, et al.
Publicado: (2024) -
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
por: Jiao, Junfeng, et al.
Publicado: (2025) -
Evaluating the Effectiveness of OpenAI's Parental Control System
por: Ersoz, Kerem, et al.
Publicado: (2026)