Conversations: Love Them, Hate Them, Steer Them
Fuente:
arXiv
Saved in:
| Main Authors: | Chebrolu, Niranjan, Yeo, Gerard Christopher, Jaidka, Kokil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Limitations of Steering in Language Model Alignment
by: Niranjan, Chebrolu, et al.
Published: (2025)
by: Niranjan, Chebrolu, et al.
Published: (2025)
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
by: Chebrolu, Niranjan, et al.
Published: (2025)
by: Chebrolu, Niranjan, et al.
Published: (2025)
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
by: Yeo, Gerard Christopher, et al.
Published: (2025)
by: Yeo, Gerard Christopher, et al.
Published: (2025)
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
by: Yeo, Gerard Christopher, et al.
Published: (2024)
by: Yeo, Gerard Christopher, et al.
Published: (2024)
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
by: Furniturewala, Shaz, et al.
Published: (2026)
by: Furniturewala, Shaz, et al.
Published: (2026)
Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
by: Yeo, Gerard, et al.
Published: (2025)
by: Yeo, Gerard, et al.
Published: (2025)
Layer of Truth: Probing Belief Shifts under Continual Pre-Training Poisoning
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
by: Singh, Sahajpreet, et al.
Published: (2025)
by: Singh, Sahajpreet, et al.
Published: (2025)
Impact of Decoding Methods on Human Alignment of Conversational LLMs
by: Furniturewala, Shaz, et al.
Published: (2024)
by: Furniturewala, Shaz, et al.
Published: (2024)
Turn-Level Empathy Prediction Using Psychological Indicators
by: Furniturewala, Shaz, et al.
Published: (2024)
by: Furniturewala, Shaz, et al.
Published: (2024)
The MediaSpin Dataset: Post-Publication News Headline Edits Annotated for Media Bias
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
Reading Between the Lines: How Electronic Nonverbal Cues shape Emotion Decoding
by: Kumar, Taara, et al.
Published: (2026)
by: Kumar, Taara, et al.
Published: (2026)
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
by: Churina, Svetlana, et al.
Published: (2024)
by: Churina, Svetlana, et al.
Published: (2024)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
by: Khalifa, Muhammad, et al.
Published: (2024)
by: Khalifa, Muhammad, et al.
Published: (2024)
One Whisper to Grade Them All
by: Phan, Nhan, et al.
Published: (2025)
by: Phan, Nhan, et al.
Published: (2025)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
Researching Skills: They Need Them When They Need Them.
by: Burgess, Barbara J.
Published: (1987)
by: Burgess, Barbara J.
Published: (1987)
Hateful Meme Detection through Context-Sensitive Prompting and Fine-Grained Labeling
by: Ouyang, Rongxin, et al.
Published: (2024)
by: Ouyang, Rongxin, et al.
Published: (2024)
Fantastic Biases (What are They) and Where to Find Them
by: Barriere, Valentin
Published: (2024)
by: Barriere, Valentin
Published: (2024)
GitSearch: Enhancing Community Notes Generation with Gap-Informed Targeted Search
by: Singh, Sahajpreet, et al.
Published: (2026)
by: Singh, Sahajpreet, et al.
Published: (2026)
OmniCaptioner: One Captioner to Rule Them All
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
Mirror, Mirror on the Wall -- Which is the Best Model of Them All?
by: Sayed, Dina, et al.
Published: (2025)
by: Sayed, Dina, et al.
Published: (2025)
An Annotated Dataset of Errors in Premodern Greek and Baselines for Detecting Them
by: Brooks, Creston, et al.
Published: (2024)
by: Brooks, Creston, et al.
Published: (2024)
LLMs Can Plan Only If We Tell Them
by: Sel, Bilgehan, et al.
Published: (2025)
by: Sel, Bilgehan, et al.
Published: (2025)
Barriers to Universal Reasoning With Transformers (And How to Overcome Them)
by: Kraus, Oliver, et al.
Published: (2026)
by: Kraus, Oliver, et al.
Published: (2026)
LTD-Bench: Evaluating Large Language Models by Letting Them Draw
by: Lin, Liuhao, et al.
Published: (2025)
by: Lin, Liuhao, et al.
Published: (2025)
One Prompt To Rule Them All: LLMs for Opinion Summary Evaluation
by: Siledar, Tejpalsingh, et al.
Published: (2024)
by: Siledar, Tejpalsingh, et al.
Published: (2024)
LLMs for Doctors: Leveraging Medical LLMs to Assist Doctors, Not Replace Them
by: Xie, Wenya, et al.
Published: (2024)
by: Xie, Wenya, et al.
Published: (2024)
Fantastic Bugs and Where to Find Them in AI Benchmarks
by: Truong, Sang, et al.
Published: (2025)
by: Truong, Sang, et al.
Published: (2025)
Reasoning Inconsistencies and How to Mitigate Them in Deep Learning
by: Arakelyan, Erik
Published: (2025)
by: Arakelyan, Erik
Published: (2025)
DiscoverLLM: From Executing Intents to Discovering Them
by: Kim, Tae Soo, et al.
Published: (2026)
by: Kim, Tae Soo, et al.
Published: (2026)
LLMs and Finetuning: Benchmarking cross-domain performance for hate speech detection
by: Nasir, Ahmad, et al.
Published: (2023)
by: Nasir, Ahmad, et al.
Published: (2023)
Low-Perplexity LLM-Generated Sequences and Where To Find Them
by: Wuhrmann, Arthur, et al.
Published: (2025)
by: Wuhrmann, Arthur, et al.
Published: (2025)
HALoGEN: Fantastic LLM Hallucinations and Where to Find Them
by: Ravichander, Abhilasha, et al.
Published: (2025)
by: Ravichander, Abhilasha, et al.
Published: (2025)
TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
by: Wang, Yidong, et al.
Published: (2025)
by: Wang, Yidong, et al.
Published: (2025)
Facts and People: Where to Find Them.
by: Hirigoyen, Maria
Published: (1971)
by: Hirigoyen, Maria
Published: (1971)
Predicting Sentence-Level Factuality of News and Bias of Media Outlets
by: Vargas, Francielle, et al.
Published: (2023)
by: Vargas, Francielle, et al.
Published: (2023)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
by: Choi, Alexander S., et al.
Published: (2024)
by: Choi, Alexander S., et al.
Published: (2024)
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
by: Tan, Fiona Anting, et al.
Published: (2024)
by: Tan, Fiona Anting, et al.
Published: (2024)
Disentangling Codemixing in Chats: The NUS ABC Codemixed Corpus
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
Similar Items
-
On the Limitations of Steering in Language Model Alignment
by: Niranjan, Chebrolu, et al.
Published: (2025) -
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
by: Chebrolu, Niranjan, et al.
Published: (2025) -
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
by: Yeo, Gerard Christopher, et al.
Published: (2025) -
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
by: Yeo, Gerard Christopher, et al.
Published: (2024) -
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
by: Furniturewala, Shaz, et al.
Published: (2026)