Multi-Turn Neural Transparency: Surfacing Neural Activations Improves User Calibration to LLM Behavioral Drift
Fuente:
arXiv
Saved in:
| Main Authors: | Karny, Sheer, Baez, Anthony, Pataranutaporn, Pat |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural Transparency: Mechanistic Interpretability Interfaces for Anticipating Model Behaviors for Personalized AI
by: Karny, Sheer, et al.
Published: (2025)
by: Karny, Sheer, et al.
Published: (2025)
AI-Wrapped: Participatory, Privacy-Preserving Measurement of Longitudinal LLM Use In-the-Wild
by: Fang, Cathy Mengying, et al.
Published: (2026)
by: Fang, Cathy Mengying, et al.
Published: (2026)
Atlas of Human-AI Interaction (v1): An Interactive Meta-Science Platform for Large-Scale Research Literature Sensemaking
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
"My Boyfriend is AI": A Computational Analysis of Human-AI Companionship in Reddit's AI Community
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
Chatbot Companionship: A Mixed-Methods Study of Companion Chatbot Usage Patterns and Their Relationship to Loneliness in Active Users
by: Liu, Auren R., et al.
Published: (2024)
by: Liu, Auren R., et al.
Published: (2024)
OceanChat: The Effect of Virtual Conversational AI Agents on Sustainable Attitude and Behavior Change
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
AI persuading AI vs AI persuading Humans: LLMs' Differential Effectiveness in Promoting Pro-Environmental Behavior
by: Doudkin, Alexander, et al.
Published: (2025)
by: Doudkin, Alexander, et al.
Published: (2025)
Exploring the Impact of AI Value Alignment in Collaborative Ideation: Effects on Perception, Ownership, and Output
by: Guo, Alicia, et al.
Published: (2024)
by: Guo, Alicia, et al.
Published: (2024)
Super-intelligence or Superstition? Exploring Psychological Factors Influencing Belief in AI Predictions about Personal Behavior
by: Lee, Eunhae, et al.
Published: (2024)
by: Lee, Eunhae, et al.
Published: (2024)
Large Language Models Predict Human Well-being -- But Not Equally Everywhere
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
"Death" of a Chatbot: Investigating and Designing Toward Psychologically Safe Endings for Human-AI Relationships
by: Poonsiriwong, Rachel, et al.
Published: (2026)
by: Poonsiriwong, Rachel, et al.
Published: (2026)
Simulating Psychological Risks in Human-AI Interactions: Real-Case Informed Modeling of AI-Induced Addiction, Anorexia, Depression, Homicide, Psychosis, and Suicide
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
Text2Tradition: From Epistemological Tensions to AI-Mediated Cross-Cultural Co-Creation
by: Pataranutaporn, Pat, et al.
Published: (2024)
by: Pataranutaporn, Pat, et al.
Published: (2024)
Synthetic Human Memories: AI-Edited Images and Videos Can Implant False Memories and Distort Recollection
by: Pataranutaporn, Pat, et al.
Published: (2024)
by: Pataranutaporn, Pat, et al.
Published: (2024)
Personality Matters: User Traits Predict LLM Preferences in Multi-Turn Collaborative Tasks
by: Yunusov, Sarfaroz, et al.
Published: (2025)
by: Yunusov, Sarfaroz, et al.
Published: (2025)
Human-AI Collaboration: Trade-offs Between Performance and Preferences
by: Mayer, Lukas William, et al.
Published: (2025)
by: Mayer, Lukas William, et al.
Published: (2025)
How AI and Human Behaviors Shape Psychosocial Effects of Extended Chatbot Use: A Longitudinal Randomized Controlled Study
by: Fang, Cathy Mengying, et al.
Published: (2025)
by: Fang, Cathy Mengying, et al.
Published: (2025)
People over trust AI-generated medical responses and view them to be as valid as doctors, despite low accuracy
by: Shekar, Shruthi, et al.
Published: (2024)
by: Shekar, Shruthi, et al.
Published: (2024)
What Large Language Models Know and What People Think They Know
by: Steyvers, Mark, et al.
Published: (2024)
by: Steyvers, Mark, et al.
Published: (2024)
Designing an LLM-Based Behavioral Activation Chatbot for Young People with Depression: Insights from an Evaluation with Artificial Users and Clinical Experts
by: Kuhlmeier, Florian Onur, et al.
Published: (2025)
by: Kuhlmeier, Florian Onur, et al.
Published: (2025)
Predicting User Behavior in Smart Spaces with LLM-Enhanced Logs and Personalized Prompts
by: Song, Yunpeng, et al.
Published: (2024)
by: Song, Yunpeng, et al.
Published: (2024)
Trust in Transparency: How Explainable AI Shapes User Perceptions
by: Sunny, Allen Daniel
Published: (2025)
by: Sunny, Allen Daniel
Published: (2025)
Future You: A Conversation with an AI-Generated Future Self Reduces Anxiety, Negative Emotions, and Increases Future Self-Continuity
by: Pataranutaporn, Pat, et al.
Published: (2024)
by: Pataranutaporn, Pat, et al.
Published: (2024)
Model Behavior Specification by Leveraging LLM Self-Playing and Self-Improving
by: Park, Soya, et al.
Published: (2025)
by: Park, Soya, et al.
Published: (2025)
Graph Neural Networks for User Satisfaction Classification in Human-Computer Interaction
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
Human and LLM-Based Voice Assistant Interaction: An Analytical Framework for User Verbal and Nonverbal Behaviors
by: Chan, Szeyi, et al.
Published: (2024)
by: Chan, Szeyi, et al.
Published: (2024)
Conversational AI Powered by Large Language Models Amplifies False Memories in Witness Interviews
by: Chan, Samantha, et al.
Published: (2024)
by: Chan, Samantha, et al.
Published: (2024)
Rethinking User Empowerment in AI Recommender System: Innovating Transparent and Controllable Interfaces
by: Wu, Mengke, et al.
Published: (2025)
by: Wu, Mengke, et al.
Published: (2025)
CoPrompter: User-Centric Evaluation of LLM Instruction Alignment for Improved Prompt Engineering
by: Joshi, Ishika, et al.
Published: (2024)
by: Joshi, Ishika, et al.
Published: (2024)
Simulating Life Paths with Digital Twins: AI-Generated Future Selves Influence Decision-Making and Expand Human Choice
by: Poonsiriwong, Rachel, et al.
Published: (2025)
by: Poonsiriwong, Rachel, et al.
Published: (2025)
Temporal Drift in Privacy Recall: Users Misremember From Verbatim Loss to Gist-Based Overexposure
by: Guo, Haoze, et al.
Published: (2025)
by: Guo, Haoze, et al.
Published: (2025)
Behind the Feed: A Taxonomy of User-Facing Cues for Algorithmic Transparency in Social Media
by: Guo, Haoze, et al.
Published: (2026)
by: Guo, Haoze, et al.
Published: (2026)
Exploring Similarity between Neural and LLM Trajectories in Language Processing
by: Xiao, Xin, et al.
Published: (2025)
by: Xiao, Xin, et al.
Published: (2025)
Mind Drifts, Data Shifts: Utilizing Mind Wandering to Track the Evolution of User Experience with Data Visualizations
by: Arunkumar, Anjana, et al.
Published: (2024)
by: Arunkumar, Anjana, et al.
Published: (2024)
QueryGenie: Making LLM-Based Database Querying Transparent and Controllable
by: Chen, Longfei, et al.
Published: (2025)
by: Chen, Longfei, et al.
Published: (2025)
Future You: Designing and Evaluating Multimodal AI-generated Digital Twins for Strengthening Future Self-Continuity
by: Albrecht, Constanze, et al.
Published: (2025)
by: Albrecht, Constanze, et al.
Published: (2025)
Authorship Drift: How Self-Efficacy and Trust Evolve During LLM-Assisted Writing
by: Park, Yeon Su, et al.
Published: (2026)
by: Park, Yeon Su, et al.
Published: (2026)
Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions
by: Chen, Meng, et al.
Published: (2025)
by: Chen, Meng, et al.
Published: (2025)
Explainable XR: Understanding User Behaviors of XR Environments using LLM-assisted Analytics Framework
by: Kim, Yoonsang, et al.
Published: (2025)
by: Kim, Yoonsang, et al.
Published: (2025)
Neural and Cognitive Impacts of AI: The Influence of Task Subjectivity on Human-LLM Collaboration
by: Russell, Matthew, et al.
Published: (2025)
by: Russell, Matthew, et al.
Published: (2025)
Similar Items
-
Neural Transparency: Mechanistic Interpretability Interfaces for Anticipating Model Behaviors for Personalized AI
by: Karny, Sheer, et al.
Published: (2025) -
AI-Wrapped: Participatory, Privacy-Preserving Measurement of Longitudinal LLM Use In-the-Wild
by: Fang, Cathy Mengying, et al.
Published: (2026) -
Atlas of Human-AI Interaction (v1): An Interactive Meta-Science Platform for Large-Scale Research Literature Sensemaking
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025) -
"My Boyfriend is AI": A Computational Analysis of Human-AI Companionship in Reddit's AI Community
by: Pataranutaporn, Pat, et al.
Published: (2025) -
Chatbot Companionship: A Mixed-Methods Study of Companion Chatbot Usage Patterns and Their Relationship to Loneliness in Active Users
by: Liu, Auren R., et al.
Published: (2024)