Conceptors for Semantic Steering
Fuente:
arXiv
Saved in:
| Main Authors: | Triantafyllopoulos, Ilias, Cho, Young-Min, Tao, Ren, Miao, Miranda Muqing, Rai, Sunny, Ungar, Lyle, Guntuku, Sharath Chandra, Ryant, Neville, Sedoc, João |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Concise Agent is Less Expert: Revealing Side Effects of Using Style Features on Conversational Agents
by: Cho, Young-Min, et al.
Published: (2026)
by: Cho, Young-Min, et al.
Published: (2026)
Building Knowledge-Guided Lexica to Model Cultural Variation
by: Havaldar, Shreya, et al.
Published: (2024)
by: Havaldar, Shreya, et al.
Published: (2024)
Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems
by: Cho, Young-Min, et al.
Published: (2025)
by: Cho, Young-Min, et al.
Published: (2025)
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Correctness-Optimized Residual Activation Lens (CORAL): Transferrable and Calibration-Aware Inference-Time Steering
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG
by: Triantafyllopoulos, Ilias, et al.
Published: (2025)
by: Triantafyllopoulos, Ilias, et al.
Published: (2025)
Culturally-Aware Conversations: A Framework & Benchmark for LLMs
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Closing the Confidence-Faithfulness Gap in Large Language Models
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Language-based Valence and Arousal Expressions between the United States and China: a Cross-Cultural Examination
by: Cho, Young-Min, et al.
Published: (2024)
by: Cho, Young-Min, et al.
Published: (2024)
Social Norms in Cinema: A Cross-Cultural Analysis of Shame, Pride and Prejudice
by: Rai, Sunny, et al.
Published: (2024)
by: Rai, Sunny, et al.
Published: (2024)
Real-Time Deadlines Reveal Temporal Awareness Failures in LLM Strategic Dialogues
by: Sehgal, Neil K. R., et al.
Published: (2026)
by: Sehgal, Neil K. R., et al.
Published: (2026)
Cross-Cultural Differences in Mental Health Expressions on Social Media
by: Rai, Sunny, et al.
Published: (2024)
by: Rai, Sunny, et al.
Published: (2024)
The Impact of Language Mixing on Bilingual LLM Reasoning
by: Li, Yihao, et al.
Published: (2025)
by: Li, Yihao, et al.
Published: (2025)
Self-Reported Side Effects of Semaglutide and Tirzepatide in Online Communities
by: Sehgal, Neil K. R., et al.
Published: (2026)
by: Sehgal, Neil K. R., et al.
Published: (2026)
Optimized but Unowned: How AI-Authored Goals Undermine the Motivation They Are Meant to Drive
by: Chi, Vivienne Bihe, et al.
Published: (2026)
by: Chi, Vivienne Bihe, et al.
Published: (2026)
Evaluating Speech-to-Text Systems with PennSound
by: Wright, Jonathan, et al.
Published: (2025)
by: Wright, Jonathan, et al.
Published: (2025)
Voice-Based Chatbots for English Speaking Practice in Multilingual Low-Resource Indian Schools: A Multi-Stakeholder Study
by: Shashidhara, Sneha, et al.
Published: (2026)
by: Shashidhara, Sneha, et al.
Published: (2026)
Conversations with AI Chatbots Increase Short-Term Vaccine Intentions But Do Not Outperform Standard Public Health Messaging
by: Sehgal, Neil K. R., et al.
Published: (2025)
by: Sehgal, Neil K. R., et al.
Published: (2025)
Designing Mental-Health Chatbots for Indian Adolescents: Mixed-Methods Evidence, a Boundary-Object Lens, and a Design-Tensions Framework
by: Sehgal, Neil K. R., et al.
Published: (2025)
by: Sehgal, Neil K. R., et al.
Published: (2025)
Exploring Socio-Cultural Challenges and Opportunities in Designing Mental Health Chatbots for Adolescents in India
by: Sehgal, Neil K. R., et al.
Published: (2025)
by: Sehgal, Neil K. R., et al.
Published: (2025)
When Support Escalates Distress: Regulation and Escalation in LLM Responses to Venting and Advice-Seeking
by: Chi, Vivienne Bihe, et al.
Published: (2026)
by: Chi, Vivienne Bihe, et al.
Published: (2026)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
by: Salecha, Aadesh, et al.
Published: (2024)
by: Salecha, Aadesh, et al.
Published: (2024)
PAL: Designing Conversational Agents as Scalable, Cooperative Patient Simulators for Palliative-Care Training
by: Sehgal, Neil K. R., et al.
Published: (2025)
by: Sehgal, Neil K. R., et al.
Published: (2025)
Hallucination, Monofacts, and Miscalibration: An Empirical Investigation
by: Miao, Miranda Muqing, et al.
Published: (2025)
by: Miao, Miranda Muqing, et al.
Published: (2025)
Modeling Human Subjectivity in LLMs Using Explicit and Implicit Human Factors in Personas
by: Giorgi, Salvatore, et al.
Published: (2024)
by: Giorgi, Salvatore, et al.
Published: (2024)
DBOT: Artificial Intelligence for Systematic Long-Term Investing
by: Dhar, Vasant, et al.
Published: (2025)
by: Dhar, Vasant, et al.
Published: (2025)
To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation
by: Cheng, Xiang, et al.
Published: (2024)
by: Cheng, Xiang, et al.
Published: (2024)
Interpreting and Mitigating Unwanted Uncertainty in LLMs
by: Roy, Tiasa Singha, et al.
Published: (2025)
by: Roy, Tiasa Singha, et al.
Published: (2025)
Reasoning and the Trusting Behavior of DeepSeek and GPT: An Experiment Revealing Hidden Fault Lines in Large Language Models
by: Li, Rubing, et al.
Published: (2025)
by: Li, Rubing, et al.
Published: (2025)
Prompt-Counterfactual Explanations for Generative AI System Behavior
by: Goethals, Sofie, et al.
Published: (2026)
by: Goethals, Sofie, et al.
Published: (2026)
Multiscale Encoder and Omni-Dimensional Dynamic Convolution Enrichment in nnU-Net for Brain Tumor Segmentation
by: Mistry, Sahaj K., et al.
Published: (2024)
by: Mistry, Sahaj K., et al.
Published: (2024)
Know Me, Respond to Me: Benchmarking LLMs for Dynamic User Profiling and Personalized Responses at Scale
by: Jiang, Bowen, et al.
Published: (2025)
by: Jiang, Bowen, et al.
Published: (2025)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
by: Gupta, Ashray, et al.
Published: (2025)
by: Gupta, Ashray, et al.
Published: (2025)
Towards Style Alignment in Cross-Cultural Translation
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Comparing Styles across Languages: A Cross-Cultural Exploration of Politeness
by: Havaldar, Shreya, et al.
Published: (2023)
by: Havaldar, Shreya, et al.
Published: (2023)
Who speaks like a style of Vitamin: Towards Syntax-Aware DialogueSummarization using Multi-task Learning
by: Lee, Seolhwa, et al.
Published: (2021)
by: Lee, Seolhwa, et al.
Published: (2021)
How to Choose How to Choose Your Chatbot: A Massively Multi-System MultiReference Data Set for Dialog Metric Evaluation
by: Khayrallah, Huda, et al.
Published: (2023)
by: Khayrallah, Huda, et al.
Published: (2023)
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Effect of Static vs. Conversational AI-Generated Messages on Colorectal Cancer Screening Intent: a Randomized Controlled Trial
by: Sehgal, Neil K. R., et al.
Published: (2025)
by: Sehgal, Neil K. R., et al.
Published: (2025)
Large Human Language Models: A Need and the Challenges
by: Soni, Nikita, et al.
Published: (2023)
by: Soni, Nikita, et al.
Published: (2023)
Similar Items
-
A Concise Agent is Less Expert: Revealing Side Effects of Using Style Features on Conversational Agents
by: Cho, Young-Min, et al.
Published: (2026) -
Building Knowledge-Guided Lexica to Model Cultural Variation
by: Havaldar, Shreya, et al.
Published: (2024) -
Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems
by: Cho, Young-Min, et al.
Published: (2025) -
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
by: Miao, Miranda Muqing, et al.
Published: (2026) -
Correctness-Optimized Residual Activation Lens (CORAL): Transferrable and Calibration-Aware Inference-Time Steering
by: Miao, Miranda Muqing, et al.
Published: (2026)