Neural Transparency: Mechanistic Interpretability Interfaces for Anticipating Model Behaviors for Personalized AI
Fuente:
arXiv
Saved in:
| Main Authors: | Karny, Sheer, Baez, Anthony, Pataranutaporn, Pat |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Turn Neural Transparency: Surfacing Neural Activations Improves User Calibration to LLM Behavioral Drift
by: Karny, Sheer, et al.
Published: (2026)
by: Karny, Sheer, et al.
Published: (2026)
Super-intelligence or Superstition? Exploring Psychological Factors Influencing Belief in AI Predictions about Personal Behavior
by: Lee, Eunhae, et al.
Published: (2024)
by: Lee, Eunhae, et al.
Published: (2024)
AI persuading AI vs AI persuading Humans: LLMs' Differential Effectiveness in Promoting Pro-Environmental Behavior
by: Doudkin, Alexander, et al.
Published: (2025)
by: Doudkin, Alexander, et al.
Published: (2025)
Atlas of Human-AI Interaction (v1): An Interactive Meta-Science Platform for Large-Scale Research Literature Sensemaking
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
OceanChat: The Effect of Virtual Conversational AI Agents on Sustainable Attitude and Behavior Change
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
AI-Wrapped: Participatory, Privacy-Preserving Measurement of Longitudinal LLM Use In-the-Wild
by: Fang, Cathy Mengying, et al.
Published: (2026)
by: Fang, Cathy Mengying, et al.
Published: (2026)
"My Boyfriend is AI": A Computational Analysis of Human-AI Companionship in Reddit's AI Community
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
"Death" of a Chatbot: Investigating and Designing Toward Psychologically Safe Endings for Human-AI Relationships
by: Poonsiriwong, Rachel, et al.
Published: (2026)
by: Poonsiriwong, Rachel, et al.
Published: (2026)
Synthetic Human Memories: AI-Edited Images and Videos Can Implant False Memories and Distort Recollection
by: Pataranutaporn, Pat, et al.
Published: (2024)
by: Pataranutaporn, Pat, et al.
Published: (2024)
Human-AI Collaboration: Trade-offs Between Performance and Preferences
by: Mayer, Lukas William, et al.
Published: (2025)
by: Mayer, Lukas William, et al.
Published: (2025)
People over trust AI-generated medical responses and view them to be as valid as doctors, despite low accuracy
by: Shekar, Shruthi, et al.
Published: (2024)
by: Shekar, Shruthi, et al.
Published: (2024)
Future You: A Conversation with an AI-Generated Future Self Reduces Anxiety, Negative Emotions, and Increases Future Self-Continuity
by: Pataranutaporn, Pat, et al.
Published: (2024)
by: Pataranutaporn, Pat, et al.
Published: (2024)
What Large Language Models Know and What People Think They Know
by: Steyvers, Mark, et al.
Published: (2024)
by: Steyvers, Mark, et al.
Published: (2024)
Conversational AI Powered by Large Language Models Amplifies False Memories in Witness Interviews
by: Chan, Samantha, et al.
Published: (2024)
by: Chan, Samantha, et al.
Published: (2024)
Simulating Life Paths with Digital Twins: AI-Generated Future Selves Influence Decision-Making and Expand Human Choice
by: Poonsiriwong, Rachel, et al.
Published: (2025)
by: Poonsiriwong, Rachel, et al.
Published: (2025)
Future You: Designing and Evaluating Multimodal AI-generated Digital Twins for Strengthening Future Self-Continuity
by: Albrecht, Constanze, et al.
Published: (2025)
by: Albrecht, Constanze, et al.
Published: (2025)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
Exploring the Impact of AI Value Alignment in Collaborative Ideation: Effects on Perception, Ownership, and Output
by: Guo, Alicia, et al.
Published: (2024)
by: Guo, Alicia, et al.
Published: (2024)
An Epistemic Human-Aware Task Planner which Anticipates Human Beliefs and Decisions
by: Shekhar, Shashank, et al.
Published: (2024)
by: Shekhar, Shashank, et al.
Published: (2024)
Investigating Affective Use and Emotional Well-being on ChatGPT
by: Phang, Jason, et al.
Published: (2025)
by: Phang, Jason, et al.
Published: (2025)
Simulating Psychological Risks in Human-AI Interactions: Real-Case Informed Modeling of AI-Induced Addiction, Anorexia, Depression, Homicide, Psychosis, and Suicide
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025)
Optimizing Generative AI's Accuracy and Transparency in Inductive Thematic Analysis: A Human-AI Comparison
by: Nyaaba, Matthew, et al.
Published: (2025)
by: Nyaaba, Matthew, et al.
Published: (2025)
Large Language Models Predict Human Well-being -- But Not Equally Everywhere
by: Pataranutaporn, Pat, et al.
Published: (2025)
by: Pataranutaporn, Pat, et al.
Published: (2025)
Chatbot Companionship: A Mixed-Methods Study of Companion Chatbot Usage Patterns and Their Relationship to Loneliness in Active Users
by: Liu, Auren R., et al.
Published: (2024)
by: Liu, Auren R., et al.
Published: (2024)
SayNext-Bench: Why Do LLMs Struggle with Next-Utterance Anticipation?
by: Yang, Yueyi, et al.
Published: (2026)
by: Yang, Yueyi, et al.
Published: (2026)
TRAFA: Anticipating User Actions to Reduce Errors in Procedural Tasks with Predictive Feedback
by: Mokhtar, Sassan, et al.
Published: (2026)
by: Mokhtar, Sassan, et al.
Published: (2026)
Towards Personalized Brain-Computer Interface Application Based on Endogenous EEG Paradigms
by: Kwak, Heon-Gyu, et al.
Published: (2024)
by: Kwak, Heon-Gyu, et al.
Published: (2024)
Beyond Autocomplete: Designing CopilotLens Towards Transparent and Explainable AI Coding Agents
by: Ye, Runlong, et al.
Published: (2025)
by: Ye, Runlong, et al.
Published: (2025)
Towards Meaningful Transparency in Civic AI Systems
by: Murray-Rust, Dave, et al.
Published: (2025)
by: Murray-Rust, Dave, et al.
Published: (2025)
Contrastive Explanations That Anticipate Human Misconceptions Can Improve Human Decision-Making Skills
by: Buçinca, Zana, et al.
Published: (2024)
by: Buçinca, Zana, et al.
Published: (2024)
Anticipating User Needs: Insights from Design Fiction on Conversational Agents for Computational Thinking
by: Penney, Jacob, et al.
Published: (2023)
by: Penney, Jacob, et al.
Published: (2023)
More Than "Means to an End": Supporting Reasoning with Transparently Designed AI Data Science Processes
by: Sivaraman, Venkatesh, et al.
Published: (2026)
by: Sivaraman, Venkatesh, et al.
Published: (2026)
Understanding Mode Switching in Human-AI Collaboration: Behavioral Insights and Predictive Modeling
by: Nargund, Avinash Ajit, et al.
Published: (2025)
by: Nargund, Avinash Ajit, et al.
Published: (2025)
When Should an AI Act? A Human-Centered Model of Scene, Context, and Behavior for Agentic AI Design
by: Jung, Soyoung, et al.
Published: (2026)
by: Jung, Soyoung, et al.
Published: (2026)
Designing a Dashboard for Transparency and Control of Conversational AI
by: Chen, Yida, et al.
Published: (2024)
by: Chen, Yida, et al.
Published: (2024)
Envisioning Possibilities and Challenges of AI for Personalized Cancer Care
by: Kong, Elaine, et al.
Published: (2024)
by: Kong, Elaine, et al.
Published: (2024)
Investigating Personalized Driving Behaviors in Dilemma Zones: Analysis and Prediction of Stop-or-Go Decisions
by: Qin, Ziye, et al.
Published: (2024)
by: Qin, Ziye, et al.
Published: (2024)
Human-AI Teaming Using Large Language Models: Boosting Brain-Computer Interfacing (BCI) and Brain Research
by: Kapitonova, Maryna, et al.
Published: (2024)
by: Kapitonova, Maryna, et al.
Published: (2024)
Ensuring Fairness with Transparent Auditing of Quantitative Bias in AI Systems
by: Yuan, Chih-Cheng Rex, et al.
Published: (2024)
by: Yuan, Chih-Cheng Rex, et al.
Published: (2024)
Steering AI-Driven Personalization of Scientific Text for General Audiences
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Similar Items
-
Multi-Turn Neural Transparency: Surfacing Neural Activations Improves User Calibration to LLM Behavioral Drift
by: Karny, Sheer, et al.
Published: (2026) -
Super-intelligence or Superstition? Exploring Psychological Factors Influencing Belief in AI Predictions about Personal Behavior
by: Lee, Eunhae, et al.
Published: (2024) -
AI persuading AI vs AI persuading Humans: LLMs' Differential Effectiveness in Promoting Pro-Environmental Behavior
by: Doudkin, Alexander, et al.
Published: (2025) -
Atlas of Human-AI Interaction (v1): An Interactive Meta-Science Platform for Large-Scale Research Literature Sensemaking
by: Archiwaranguprok, Chayapatr, et al.
Published: (2025) -
OceanChat: The Effect of Virtual Conversational AI Agents on Sustainable Attitude and Behavior Change
by: Pataranutaporn, Pat, et al.
Published: (2025)