Let Them Down Easy! Contextual Effects of LLM Guardrails on User Perceptions and Preferences
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Mingqian, Hu, Wenjia, Zhao, Patrick, Eslami, Motahhare, Hwang, Jena D., Brahman, Faeze, Rose, Carolyn, Sap, Maarten |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations
by: Park, Eunkyu, et al.
Published: (2025)
by: Park, Eunkyu, et al.
Published: (2025)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
by: Zhou, Kaitlyn, et al.
Published: (2024)
by: Zhou, Kaitlyn, et al.
Published: (2024)
Strategies for Designing Responsibly within a Capitalist Enterprise
by: Xie, Shixian, et al.
Published: (2026)
by: Xie, Shixian, et al.
Published: (2026)
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
by: Fan, Xianzhe, et al.
Published: (2024)
by: Fan, Xianzhe, et al.
Published: (2024)
Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
by: Zhou, Kaitlyn, et al.
Published: (2024)
by: Zhou, Kaitlyn, et al.
Published: (2024)
Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations
by: Zheng, Mingqian, et al.
Published: (2026)
by: Zheng, Mingqian, et al.
Published: (2026)
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
by: Chakrabarty, Tuhin, et al.
Published: (2023)
by: Chakrabarty, Tuhin, et al.
Published: (2023)
Breaking Political Filter Bubbles via Social Comparison
by: Soliman, Nouran, et al.
Published: (2024)
by: Soliman, Nouran, et al.
Published: (2024)
WeAudit: Scaffolding User Auditors and AI Practitioners in Auditing Generative AI
by: Deng, Wesley Hanwen, et al.
Published: (2025)
by: Deng, Wesley Hanwen, et al.
Published: (2025)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
by: Cohen, Myke C., et al.
Published: (2026)
by: Cohen, Myke C., et al.
Published: (2026)
Examining the Effect of Explanations of AI Privacy Redaction in AI-mediated Interactions
by: Kaushik, Roshni, et al.
Published: (2026)
by: Kaushik, Roshni, et al.
Published: (2026)
Guardrail Selection in Line Charts to Contextualize Persuasive Visualizations
by: Nadib, Khandaker Abrar, et al.
Published: (2026)
by: Nadib, Khandaker Abrar, et al.
Published: (2026)
Vipera: Blending Visual and LLM-Driven Guidance for Systematic Auditing of Text-to-Image Generative AI
by: Huang, Yanwei, et al.
Published: (2025)
by: Huang, Yanwei, et al.
Published: (2025)
'Simulacrum of Stories': Examining Large Language Models as Qualitative Research Participants
by: Kapania, Shivani, et al.
Published: (2024)
by: Kapania, Shivani, et al.
Published: (2024)
Let's Talk About It: Making Scientific Computational Reproducibility Easy
by: Costa, Lázaro, et al.
Published: (2025)
by: Costa, Lázaro, et al.
Published: (2025)
Rethinking Theory of Mind Benchmarks for LLMs: Towards A User-Centered Perspective
by: Wang, Qiaosi, et al.
Published: (2025)
by: Wang, Qiaosi, et al.
Published: (2025)
MIRAGE: Multi-model Interface for Reviewing and Auditing Generative Text-to-Image AI
by: Maldaner, Matheus Kunzler, et al.
Published: (2025)
by: Maldaner, Matheus Kunzler, et al.
Published: (2025)
Children's Overtrust and Shifting Perspectives of Generative AI
by: Solyst, Jaemarie, et al.
Published: (2024)
by: Solyst, Jaemarie, et al.
Published: (2024)
The Fall of an Algorithm: Characterizing the Dynamics Toward Abandonment
by: Johnson, Nari, et al.
Published: (2024)
by: Johnson, Nari, et al.
Published: (2024)
Seeing Twice: How Side-by-Side T2I Comparison Changes Auditing Strategies
by: Maldaner, Matheus Kunzler, et al.
Published: (2025)
by: Maldaner, Matheus Kunzler, et al.
Published: (2025)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
by: Park, Eunkyu, et al.
Published: (2025)
by: Park, Eunkyu, et al.
Published: (2025)
The Art of Saying No: Contextual Noncompliance in Language Models
by: Brahman, Faeze, et al.
Published: (2024)
by: Brahman, Faeze, et al.
Published: (2024)
Vipera: Towards systematic auditing of generative text-to-image models at scale
by: Huang, Yanwei, et al.
Published: (2025)
by: Huang, Yanwei, et al.
Published: (2025)
Not Like Us, Hunty: Measuring Perceptions and Behavioral Effects of Minoritized Anthropomorphic Cues in LLMs
by: Basoah, Jeffrey, et al.
Published: (2025)
by: Basoah, Jeffrey, et al.
Published: (2025)
Framing an AI with Values Reduces AI Reliance in AI-supported Writing Tasks
by: Gao, Alice, et al.
Published: (2026)
by: Gao, Alice, et al.
Published: (2026)
What Comes After Harm? Mapping Reparative Actions in AI through Justice Frameworks
by: Xiao, Sijia, et al.
Published: (2025)
by: Xiao, Sijia, et al.
Published: (2025)
"I Don't Think RAI Applies to My Model'' -- Engaging Non-champions with Sticky Stories for Responsible AI Work
by: Nahar, Nadia, et al.
Published: (2025)
by: Nahar, Nadia, et al.
Published: (2025)
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
by: Fan, Xianzhe, et al.
Published: (2024)
by: Fan, Xianzhe, et al.
Published: (2024)
"Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking Partner
by: Chan, Szeyi, et al.
Published: (2023)
by: Chan, Szeyi, et al.
Published: (2023)
"It Might be Technically Impressive, But It's Practically Useless to us": Motivations, Practices, Challenges, and Opportunities for Cross-Functional Collaboration around AI within the News Industry
by: Xiao, Qing, et al.
Published: (2024)
by: Xiao, Qing, et al.
Published: (2024)
AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
by: Su, Zhe, et al.
Published: (2024)
by: Su, Zhe, et al.
Published: (2024)
Counterspeakers' Perspectives: Unveiling Barriers and AI Needs in the Fight against Online Hate
by: Mun, Jimin, et al.
Published: (2024)
by: Mun, Jimin, et al.
Published: (2024)
Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees' Practices, Challenges, and Needs
by: Johnson, Nari, et al.
Published: (2024)
by: Johnson, Nari, et al.
Published: (2024)
Investigating Youth AI Auditing
by: Solyst, Jaemarie, et al.
Published: (2025)
by: Solyst, Jaemarie, et al.
Published: (2025)
Public Technologies Transforming Work of the Public and the Public Sector
by: Kim, Seyun, et al.
Published: (2024)
by: Kim, Seyun, et al.
Published: (2024)
"It Is Easy Using My Apps:" Understanding Technology Use and Needs of Adults with Down Syndrome
by: Johnson, Hailey L., et al.
Published: (2024)
by: Johnson, Hailey L., et al.
Published: (2024)
Leftover Lunch: Advantage-based Offline Reinforcement Learning for Language Models
by: Baheti, Ashutosh, et al.
Published: (2023)
by: Baheti, Ashutosh, et al.
Published: (2023)
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
by: Cohen, Myke C., et al.
Published: (2025)
by: Cohen, Myke C., et al.
Published: (2025)
Understanding Users' Privacy Perceptions Towards LLM's RAG-based Memory
by: Zhang, Shuning, et al.
Published: (2025)
by: Zhang, Shuning, et al.
Published: (2025)
To Embody or Not: The Effect Of Embodiment On User Perception Of LLM-based Conversational Agents
by: Wang, Kyra, et al.
Published: (2025)
by: Wang, Kyra, et al.
Published: (2025)
Similar Items
-
Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations
by: Park, Eunkyu, et al.
Published: (2025) -
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
by: Zhou, Kaitlyn, et al.
Published: (2024) -
Strategies for Designing Responsibly within a Capitalist Enterprise
by: Xie, Shixian, et al.
Published: (2026) -
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
by: Fan, Xianzhe, et al.
Published: (2024) -
Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
by: Zhou, Kaitlyn, et al.
Published: (2024)