Incongruent Positivity: When Miscalibrated Positivity Undermines Online Supportive Conversations
Fuente:
arXiv
Salvato in:
| Autori principali: | Almajed, Leen, ALdayel, Abeer |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Overcoming Miscalibrated Conversational Priors in LLM-based Chatbots
di: Herlihy, Christine, et al.
Pubblicazione: (2024)
di: Herlihy, Christine, et al.
Pubblicazione: (2024)
Extreme Miscalibration and the Illusion of Adversarial Robustness
di: Raina, Vyas, et al.
Pubblicazione: (2024)
di: Raina, Vyas, et al.
Pubblicazione: (2024)
Hallucination, Monofacts, and Miscalibration: An Empirical Investigation
di: Miao, Miranda Muqing, et al.
Pubblicazione: (2025)
di: Miao, Miranda Muqing, et al.
Pubblicazione: (2025)
EMBRACE: Shaping Inclusive Opinion Representation by Aligning Implicit Conversations with Social Norms
di: Aldayel, Abeer, et al.
Pubblicazione: (2025)
di: Aldayel, Abeer, et al.
Pubblicazione: (2025)
Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech
di: Corrêa, Pedro, et al.
Pubblicazione: (2025)
di: Corrêa, Pedro, et al.
Pubblicazione: (2025)
The Confidence Dichotomy: Analyzing and Mitigating Miscalibration in Tool-Use Agents
di: Xuan, Weihao, et al.
Pubblicazione: (2026)
di: Xuan, Weihao, et al.
Pubblicazione: (2026)
Large Language Models are Miscalibrated In-Context Learners
di: Li, Chengzu, et al.
Pubblicazione: (2023)
di: Li, Chengzu, et al.
Pubblicazione: (2023)
Position Bias Mitigates Position Bias:Mitigate Position Bias Through Inter-Position Knowledge Distillation
di: Wang, Yifei, et al.
Pubblicazione: (2025)
di: Wang, Yifei, et al.
Pubblicazione: (2025)
Hatred Stems from Ignorance! Distillation of the Persuasion Modes in Countering Conversational Hate Speech
di: Alyahya, Ghadi, et al.
Pubblicazione: (2024)
di: Alyahya, Ghadi, et al.
Pubblicazione: (2024)
Mapping from Meaning: Addressing the Miscalibration of Prompt-Sensitive Language Models
di: Cox, Kyle, et al.
Pubblicazione: (2025)
di: Cox, Kyle, et al.
Pubblicazione: (2025)
When Bias Pretends to Be Truth: How Spurious Correlations Undermine Hallucination Detection in LLMs
di: Wang, Shaowen, et al.
Pubblicazione: (2025)
di: Wang, Shaowen, et al.
Pubblicazione: (2025)
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
di: Chang, Trenton, et al.
Pubblicazione: (2025)
di: Chang, Trenton, et al.
Pubblicazione: (2025)
Detecting Hope Across Languages: Multiclass Classification for Positive Online Discourse
di: Abiola, T. O., et al.
Pubblicazione: (2025)
di: Abiola, T. O., et al.
Pubblicazione: (2025)
Hallucinations Undermine Trust; Metacognition is a Way Forward
di: Yona, Gal, et al.
Pubblicazione: (2026)
di: Yona, Gal, et al.
Pubblicazione: (2026)
When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training
di: Wang, Haonan, et al.
Pubblicazione: (2024)
di: Wang, Haonan, et al.
Pubblicazione: (2024)
Beyond Position Bias: Shifting Context Compression from Position-Driven to Semantic-Driven
di: Tang, Jiwei, et al.
Pubblicazione: (2026)
di: Tang, Jiwei, et al.
Pubblicazione: (2026)
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
di: Okutomi, Akira
Pubblicazione: (2025)
di: Okutomi, Akira
Pubblicazione: (2025)
Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders
di: Sun, Xin, et al.
Pubblicazione: (2026)
di: Sun, Xin, et al.
Pubblicazione: (2026)
Positional Description for Numerical Normalization
di: Gupta, Deepanshu, et al.
Pubblicazione: (2024)
di: Gupta, Deepanshu, et al.
Pubblicazione: (2024)
Position IDs Matter: An Enhanced Position Layout for Efficient Context Compression in Large Language Models
di: Zhao, Runsong, et al.
Pubblicazione: (2024)
di: Zhao, Runsong, et al.
Pubblicazione: (2024)
Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models
di: Bakhshi, Asim D.
Pubblicazione: (2026)
di: Bakhshi, Asim D.
Pubblicazione: (2026)
Do Large Language Models Plan Answer Positions? Position Bias in Multiple-Choice Question Generation
di: Tang, Xuemei, et al.
Pubblicazione: (2026)
di: Tang, Xuemei, et al.
Pubblicazione: (2026)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
di: Plaut, Benjamin, et al.
Pubblicazione: (2024)
di: Plaut, Benjamin, et al.
Pubblicazione: (2024)
Context-aware Rotary Position Embedding
di: Veisi, Ali, et al.
Pubblicazione: (2025)
di: Veisi, Ali, et al.
Pubblicazione: (2025)
Engagement Undermines Safety: How Stereotypes and Toxicity Shape Humor in Language Models
di: Dogra, Atharvan, et al.
Pubblicazione: (2025)
di: Dogra, Atharvan, et al.
Pubblicazione: (2025)
Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models?
di: Jha, Basab, et al.
Pubblicazione: (2025)
di: Jha, Basab, et al.
Pubblicazione: (2025)
Compromesso! Italian Many-Shot Jailbreaks Undermine the Safety of Large Language Models
di: Pernisi, Fabio, et al.
Pubblicazione: (2024)
di: Pernisi, Fabio, et al.
Pubblicazione: (2024)
Steering Conversational Large Language Models for Long Emotional Support Conversations
di: Madani, Navid, et al.
Pubblicazione: (2024)
di: Madani, Navid, et al.
Pubblicazione: (2024)
PositionID: LLMs can Control Lengths, Copy and Paste with Explicit Positional Awareness
di: Wang, Zekun, et al.
Pubblicazione: (2024)
di: Wang, Zekun, et al.
Pubblicazione: (2024)
Learning When to Quit in Sales Conversations
di: Manzoor, Emaad, et al.
Pubblicazione: (2025)
di: Manzoor, Emaad, et al.
Pubblicazione: (2025)
Wavelet-based Positional Representation for Long Context
di: Oka, Yui, et al.
Pubblicazione: (2025)
di: Oka, Yui, et al.
Pubblicazione: (2025)
DoPE: Denoising Rotary Position Embedding
di: Xiong, Jing, et al.
Pubblicazione: (2025)
di: Xiong, Jing, et al.
Pubblicazione: (2025)
Socratic Reasoning Improves Positive Text Rewriting
di: Goel, Anmol, et al.
Pubblicazione: (2024)
di: Goel, Anmol, et al.
Pubblicazione: (2024)
A Morphology-Based Investigation of Positional Encodings
di: Ghosh, Poulami, et al.
Pubblicazione: (2024)
di: Ghosh, Poulami, et al.
Pubblicazione: (2024)
On Positional Bias of Faithfulness for Long-form Summarization
di: Wan, David, et al.
Pubblicazione: (2024)
di: Wan, David, et al.
Pubblicazione: (2024)
Serial Position Effects of Large Language Models
di: Guo, Xiaobo, et al.
Pubblicazione: (2024)
di: Guo, Xiaobo, et al.
Pubblicazione: (2024)
Evaluating, Synthesizing, and Enhancing for Customer Support Conversation
di: Zhu, Jie, et al.
Pubblicazione: (2025)
di: Zhu, Jie, et al.
Pubblicazione: (2025)
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks)
di: Priyanshu, Aman, et al.
Pubblicazione: (2024)
di: Priyanshu, Aman, et al.
Pubblicazione: (2024)
Selective Rotary Position Embedding
di: Movahedi, Sajad, et al.
Pubblicazione: (2025)
di: Movahedi, Sajad, et al.
Pubblicazione: (2025)
On the Geometry of Positional Encodings in Transformers
di: Cirrincione, Giansalvo
Pubblicazione: (2026)
di: Cirrincione, Giansalvo
Pubblicazione: (2026)
Documenti analoghi
-
On Overcoming Miscalibrated Conversational Priors in LLM-based Chatbots
di: Herlihy, Christine, et al.
Pubblicazione: (2024) -
Extreme Miscalibration and the Illusion of Adversarial Robustness
di: Raina, Vyas, et al.
Pubblicazione: (2024) -
Hallucination, Monofacts, and Miscalibration: An Empirical Investigation
di: Miao, Miranda Muqing, et al.
Pubblicazione: (2025) -
EMBRACE: Shaping Inclusive Opinion Representation by Aligning Implicit Conversations with Social Norms
di: Aldayel, Abeer, et al.
Pubblicazione: (2025) -
Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech
di: Corrêa, Pedro, et al.
Pubblicazione: (2025)