Impact of Decoding Methods on Human Alignment of Conversational LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Furniturewala, Shaz, Jaidka, Kokil, Sharma, Yashvardhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Turn-Level Empathy Prediction Using Psychological Indicators
by: Furniturewala, Shaz, et al.
Published: (2024)
by: Furniturewala, Shaz, et al.
Published: (2024)
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
by: Yeo, Gerard Christopher, et al.
Published: (2024)
by: Yeo, Gerard Christopher, et al.
Published: (2024)
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
by: Furniturewala, Shaz, et al.
Published: (2026)
by: Furniturewala, Shaz, et al.
Published: (2026)
Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
by: Furniturewala, Shaz, et al.
Published: (2024)
by: Furniturewala, Shaz, et al.
Published: (2024)
Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content
by: Furniturewala, Shaz, et al.
Published: (2025)
by: Furniturewala, Shaz, et al.
Published: (2025)
Reading Between the Lines: How Electronic Nonverbal Cues shape Emotion Decoding
by: Kumar, Taara, et al.
Published: (2026)
by: Kumar, Taara, et al.
Published: (2026)
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
by: Churina, Svetlana, et al.
Published: (2024)
by: Churina, Svetlana, et al.
Published: (2024)
LLMs and Finetuning: Benchmarking cross-domain performance for hate speech detection
by: Nasir, Ahmad, et al.
Published: (2023)
by: Nasir, Ahmad, et al.
Published: (2023)
The MediaSpin Dataset: Post-Publication News Headline Edits Annotated for Media Bias
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
Conversations: Love Them, Hate Them, Steer Them
by: Chebrolu, Niranjan, et al.
Published: (2025)
by: Chebrolu, Niranjan, et al.
Published: (2025)
On the Limitations of Steering in Language Model Alignment
by: Niranjan, Chebrolu, et al.
Published: (2025)
by: Niranjan, Chebrolu, et al.
Published: (2025)
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
by: Yeo, Gerard Christopher, et al.
Published: (2025)
by: Yeo, Gerard Christopher, et al.
Published: (2025)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
by: Yeo, Gerard, et al.
Published: (2025)
by: Yeo, Gerard, et al.
Published: (2025)
GitSearch: Enhancing Community Notes Generation with Gap-Informed Targeted Search
by: Singh, Sahajpreet, et al.
Published: (2026)
by: Singh, Sahajpreet, et al.
Published: (2026)
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
by: Chebrolu, Niranjan, et al.
Published: (2025)
by: Chebrolu, Niranjan, et al.
Published: (2025)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
by: Singh, Sahajpreet, et al.
Published: (2025)
by: Singh, Sahajpreet, et al.
Published: (2025)
Predicting Sentence-Level Factuality of News and Bias of Media Outlets
by: Vargas, Francielle, et al.
Published: (2023)
by: Vargas, Francielle, et al.
Published: (2023)
Disentangling Codemixing in Chats: The NUS ABC Codemixed Corpus
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild
by: Singh, Sahajpreet, et al.
Published: (2026)
by: Singh, Sahajpreet, et al.
Published: (2026)
Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
Hateful Meme Detection through Context-Sensitive Prompting and Fine-Grained Labeling
by: Ouyang, Rongxin, et al.
Published: (2024)
by: Ouyang, Rongxin, et al.
Published: (2024)
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
by: Tan, Fiona Anting, et al.
Published: (2024)
by: Tan, Fiona Anting, et al.
Published: (2024)
PAD: Personalized Alignment of LLMs at Decoding-Time
by: Chen, Ruizhe, et al.
Published: (2024)
by: Chen, Ruizhe, et al.
Published: (2024)
A Thorough Examination of Decoding Methods in the Era of LLMs
by: Shi, Chufan, et al.
Published: (2024)
by: Shi, Chufan, et al.
Published: (2024)
LLMs Can Infer Political Alignment from Online Conversations
by: Lee, Byunghwee, et al.
Published: (2026)
by: Lee, Byunghwee, et al.
Published: (2026)
Robustness and Confounders in the Demographic Alignment of LLMs with Human Perceptions of Offensiveness
by: Alipour, Shayan, et al.
Published: (2024)
by: Alipour, Shayan, et al.
Published: (2024)
Nudging: Inference-time Alignment of LLMs via Guided Decoding
by: Fei, Yu, et al.
Published: (2024)
by: Fei, Yu, et al.
Published: (2024)
Layer of Truth: Probing Belief Shifts under Continual Pre-Training Poisoning
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
Alignment-Augmented Speculative Decoding with Alignment Sampling and Conditional Verification
by: Wang, Jikai, et al.
Published: (2025)
by: Wang, Jikai, et al.
Published: (2025)
Modeling Empathetic Alignment in Conversation
by: Yang, Jiamin, et al.
Published: (2024)
by: Yang, Jiamin, et al.
Published: (2024)
Conversational Agents and the Understanding of Human Language: Reflections on AI, LLMs, and Cognitive Science
by: Popescu-Belis, Andrei
Published: (2026)
by: Popescu-Belis, Andrei
Published: (2026)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
by: Huang, Dun-Ming, et al.
Published: (2024)
by: Huang, Dun-Ming, et al.
Published: (2024)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
by: Nie, Shangrui, et al.
Published: (2025)
by: Nie, Shangrui, et al.
Published: (2025)
HAL: Inducing Human-likeness in LLMs with Alignment
by: Hasan, Masum, et al.
Published: (2026)
by: Hasan, Masum, et al.
Published: (2026)
MAD: Multi-Alignment MEG-to-Text Decoding
by: Yang, Yiqian, et al.
Published: (2024)
by: Yang, Yiqian, et al.
Published: (2024)
Multi-Drafter Speculative Decoding with Alignment Feedback
by: Kim, Taehyeon, et al.
Published: (2026)
by: Kim, Taehyeon, et al.
Published: (2026)
Conversational Alignment with Artificial Intelligence in Context
by: Sterken, Rachel Katharine, et al.
Published: (2025)
by: Sterken, Rachel Katharine, et al.
Published: (2025)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
by: Goel, Raghavv, et al.
Published: (2024)
by: Goel, Raghavv, et al.
Published: (2024)
TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs
by: Xiao, Sibo, et al.
Published: (2025)
by: Xiao, Sibo, et al.
Published: (2025)
Similar Items
-
Turn-Level Empathy Prediction Using Psychological Indicators
by: Furniturewala, Shaz, et al.
Published: (2024) -
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
by: Yeo, Gerard Christopher, et al.
Published: (2024) -
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
by: Furniturewala, Shaz, et al.
Published: (2026) -
Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
by: Furniturewala, Shaz, et al.
Published: (2024) -
Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content
by: Furniturewala, Shaz, et al.
Published: (2025)