Can Interpretability Layouts Influence Human Perception of Offensive Sentences?
Fuente:
arXiv
Saved in:
| Main Authors: | Santos, Thiago Freitas dos, Osman, Nardine, Schorlemmer, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
Understanding Impact of Human Feedback via Influence Functions
by: Min, Taywon, et al.
Published: (2025)
by: Min, Taywon, et al.
Published: (2025)
Predictive AI Can Support Human Learning while Preserving Error Diversity
by: He, Vivianna Fang, et al.
Published: (2025)
by: He, Vivianna Fang, et al.
Published: (2025)
Vi(E)va LLM! A Conceptual Stack for Evaluating and Interpreting Generative AI-based Visualizations
by: Podo, Luca, et al.
Published: (2024)
by: Podo, Luca, et al.
Published: (2024)
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
by: Xiao, Chang, et al.
Published: (2024)
by: Xiao, Chang, et al.
Published: (2024)
Towards Human Haptic Gesture Interpretation for Robotic Systems
by: Bianchini, Bibit, et al.
Published: (2020)
by: Bianchini, Bibit, et al.
Published: (2020)
Optimizing Delegation in Collaborative Human-AI Hybrid Teams
by: Fuchs, Andrew, et al.
Published: (2024)
by: Fuchs, Andrew, et al.
Published: (2024)
A look under the hood of the Interactive Deep Learning Enterprise (No-IDLE)
by: Sonntag, Daniel, et al.
Published: (2024)
by: Sonntag, Daniel, et al.
Published: (2024)
Learning for Detecting Norm Violation in Online Communities
by: Santos, Thiago Freitas dos, et al.
Published: (2021)
by: Santos, Thiago Freitas dos, et al.
Published: (2021)
Beyond Accuracy: Robustness, Interpretability and Expressiveness of EEG Foundation Models
by: Širca, Urban, et al.
Published: (2026)
by: Širca, Urban, et al.
Published: (2026)
Explainability for Machine Learning Models: From Data Adaptability to User Perception
by: Delaunay, julien
Published: (2024)
by: Delaunay, julien
Published: (2024)
Care for the Mind Amid Chronic Diseases: An Interpretable AI Approach Using IoT
by: Xie, Jiaheng, et al.
Published: (2022)
by: Xie, Jiaheng, et al.
Published: (2022)
Diagrammatization and Abduction to Improve AI Interpretability With Domain-Aligned Explanations for Medical Diagnosis
by: Lim, Brian Y., et al.
Published: (2023)
by: Lim, Brian Y., et al.
Published: (2023)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
by: Yu, Guanghui, et al.
Published: (2024)
by: Yu, Guanghui, et al.
Published: (2024)
Human Expertise in Algorithmic Prediction
by: Alur, Rohan, et al.
Published: (2024)
by: Alur, Rohan, et al.
Published: (2024)
Human-Computer Interaction and Human-AI Collaboration in Advanced Air Mobility: A Comprehensive Review
by: Sagirli, Fatma Yamac, et al.
Published: (2024)
by: Sagirli, Fatma Yamac, et al.
Published: (2024)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
by: Chaudhari, Shravan, et al.
Published: (2025)
by: Chaudhari, Shravan, et al.
Published: (2025)
Align When They Want, Complement When They Need! Human-Centered Ensembles for Adaptive Human-AI Collaboration
by: Amin, Hasan, et al.
Published: (2026)
by: Amin, Hasan, et al.
Published: (2026)
Does Calibration Affect Human Actions?
by: Nizri, Meir, et al.
Published: (2025)
by: Nizri, Meir, et al.
Published: (2025)
Human-AI Collaborative Uncertainty Quantification
by: Noorani, Sima, et al.
Published: (2025)
by: Noorani, Sima, et al.
Published: (2025)
CREW: Facilitating Human-AI Teaming Research
by: Zhang, Lingyu, et al.
Published: (2024)
by: Zhang, Lingyu, et al.
Published: (2024)
Fully Data-driven but Interpretable Human Behavioural Modelling with Differentiable Discrete Choice Model
by: Makinoshima, Fumiyasu, et al.
Published: (2024)
by: Makinoshima, Fumiyasu, et al.
Published: (2024)
Do It For Me vs. Do It With Me: Investigating User Perceptions of Different Paradigms of Automation in Copilots for Feature-Rich Software
by: Khurana, Anjali, et al.
Published: (2025)
by: Khurana, Anjali, et al.
Published: (2025)
Can we use LLMs to bootstrap reinforcement learning? -- A case study in digital health behavior change
by: Albers, Nele, et al.
Published: (2025)
by: Albers, Nele, et al.
Published: (2025)
Designing Interpretable ML System to Enhance Trust in Healthcare: A Systematic Review to Proposed Responsible Clinician-AI-Collaboration Framework
by: Nasarian, Elham, et al.
Published: (2023)
by: Nasarian, Elham, et al.
Published: (2023)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
A No Free Lunch Theorem for Human-AI Collaboration
by: Peng, Kenny, et al.
Published: (2024)
by: Peng, Kenny, et al.
Published: (2024)
Strength Estimation and Human-Like Strength Adjustment in Games
by: Chen, Chun Jung, et al.
Published: (2025)
by: Chen, Chun Jung, et al.
Published: (2025)
Advancing Human-Machine Teaming: Concepts, Challenges, and Applications
by: Chen, Dian, et al.
Published: (2025)
by: Chen, Dian, et al.
Published: (2025)
AI Agents for Inventory Control: Human-LLM-OR Complementarity
by: Baek, Jackie, et al.
Published: (2026)
by: Baek, Jackie, et al.
Published: (2026)
Learning to Decide with AI Assistance under Human-Alignment
by: Benz, Nina Corvelo, et al.
Published: (2026)
by: Benz, Nina Corvelo, et al.
Published: (2026)
The Human-Data-Model Interaction Canvas for Visual Analytics
by: Bernard, Jürgen
Published: (2025)
by: Bernard, Jürgen
Published: (2025)
Rationalize: Shared Semantic Reasoning for Human-AI Alignment
by: Dasgupta, Aritra, et al.
Published: (2026)
by: Dasgupta, Aritra, et al.
Published: (2026)
Toward Human-AI Complementarity Across Diverse Tasks
by: Xu, Yuzheng, et al.
Published: (2026)
by: Xu, Yuzheng, et al.
Published: (2026)
MedSyn: Enhancing Diagnostics with Human-AI Collaboration
by: Sayin, Burcu, et al.
Published: (2025)
by: Sayin, Burcu, et al.
Published: (2025)
Affective and Conversational Predictors of Re-Engagement in Human-Robot Interactions -- A Student-Centered Study with A Humanoid Social Robot
by: Kang, Hangyeol, et al.
Published: (2025)
by: Kang, Hangyeol, et al.
Published: (2025)
Esports Debut as a Medal Event at 2023 Asian Games: Exploring Public Perceptions with BERTopic and GPT-4 Topic Fine-Tuning
by: Qian, Tyreal Yizhou, et al.
Published: (2024)
by: Qian, Tyreal Yizhou, et al.
Published: (2024)
Why and When LLM-Based Assistants Can Go Wrong: Investigating the Effectiveness of Prompt-Based Interactions for Software Help-Seeking
by: Khurana, Anjali, et al.
Published: (2024)
by: Khurana, Anjali, et al.
Published: (2024)
LookALike: Human Mimicry based collaborative decision making
by: Karanjai, Rabimba, et al.
Published: (2024)
by: Karanjai, Rabimba, et al.
Published: (2024)
Off-Policy Selection for Initiating Human-Centric Experimental Design
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
Similar Items
-
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025) -
Understanding Impact of Human Feedback via Influence Functions
by: Min, Taywon, et al.
Published: (2025) -
Predictive AI Can Support Human Learning while Preserving Error Diversity
by: He, Vivianna Fang, et al.
Published: (2025) -
Vi(E)va LLM! A Conceptual Stack for Evaluating and Interpreting Generative AI-based Visualizations
by: Podo, Luca, et al.
Published: (2024) -
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
by: Xiao, Chang, et al.
Published: (2024)