What if you said that differently?: How Explanation Formats Affect Human Feedback Efficacy and User Perception
Fuente:
arXiv
Salvato in:
| Autori principali: | Malaviya, Chaitanya, Lee, Subin, Roth, Dan, Yatskar, Mark |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ExpertQA: Expert-Curated Questions and Attributed Answers
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2023)
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2023)
Flattery, Fluff, and Fog: Diagnosing and Mitigating Idiosyncratic Biases in Preference Models
di: Bharadwaj, Anirudh, et al.
Pubblicazione: (2025)
di: Bharadwaj, Anirudh, et al.
Pubblicazione: (2025)
Contextualized Evaluations: Judging Language Model Responses to Underspecified Queries
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2024)
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2024)
ResearchQA: Evaluating Scholarly Question Answering at Scale Across 75 Fields with Survey-Mined Questions and Rubrics
di: Yifei, Li S., et al.
Pubblicazione: (2025)
di: Yifei, Li S., et al.
Pubblicazione: (2025)
A Simple Joint Model for Improved Contextual Neural Lemmatization
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)
On Reference (In-)Determinacy in Natural Language Inference
di: Chen, Sihao, et al.
Pubblicazione: (2025)
di: Chen, Sihao, et al.
Pubblicazione: (2025)
DOLOMITES: Domain-Specific Long-Form Methodical Tasks
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2024)
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2024)
Dynamic Clue Bottlenecks: Towards Interpretable-by-Design Visual Question Answering
di: Fu, Xingyu, et al.
Pubblicazione: (2023)
di: Fu, Xingyu, et al.
Pubblicazione: (2023)
LLM-based Hierarchical Concept Decomposition for Interpretable Fine-Grained Image Classification
di: Qu, Renyi, et al.
Pubblicazione: (2024)
di: Qu, Renyi, et al.
Pubblicazione: (2024)
Fakes of Varying Shades: How Warning Affects Human Perception and Engagement Regarding LLM Hallucinations
di: Nahar, Mahjabin, et al.
Pubblicazione: (2024)
di: Nahar, Mahjabin, et al.
Pubblicazione: (2024)
Do LLM Self-Explanations Help Users Predict Model Behavior? Evaluating Counterfactual Simulatability with Pragmatic Perturbations
di: Hong, Pingjun, et al.
Pubblicazione: (2026)
di: Hong, Pingjun, et al.
Pubblicazione: (2026)
EvalAgent: Discovering Implicit Evaluation Criteria from the Web
di: Wadhwa, Manya, et al.
Pubblicazione: (2025)
di: Wadhwa, Manya, et al.
Pubblicazione: (2025)
AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?
di: Yoran, Ori, et al.
Pubblicazione: (2024)
di: Yoran, Ori, et al.
Pubblicazione: (2024)
How Does the Disclosure of AI Assistance Affect the Perceptions of Writing?
di: Li, Zhuoyan, et al.
Pubblicazione: (2024)
di: Li, Zhuoyan, et al.
Pubblicazione: (2024)
What to Format and How: A Benchmark and Workflow Approach for Document Formatting
di: Rao, Shihao, et al.
Pubblicazione: (2026)
di: Rao, Shihao, et al.
Pubblicazione: (2026)
Comparing How a Chatbot References User Utterances from Previous Chatting Sessions: An Investigation of Users' Privacy Concerns and Perceptions
di: Cox, Samuel Rhys, et al.
Pubblicazione: (2023)
di: Cox, Samuel Rhys, et al.
Pubblicazione: (2023)
From Lists to Emojis: How Format Bias Affects Model Alignment
di: Zhang, Xuanchang, et al.
Pubblicazione: (2024)
di: Zhang, Xuanchang, et al.
Pubblicazione: (2024)
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
User Feedback in Human-LLM Dialogues: A Lens to Understand Users But Noisy as a Learning Signal
di: Liu, Yuhan, et al.
Pubblicazione: (2025)
di: Liu, Yuhan, et al.
Pubblicazione: (2025)
Benchmarking LLM Guardrails in Handling Multilingual Toxicity
di: Yang, Yahan, et al.
Pubblicazione: (2024)
di: Yang, Yahan, et al.
Pubblicazione: (2024)
Not All Explanations Simulate Equally: Comparing Verbalized Feature Attributions and Self-Generated Rationales
di: Hong, Pingjun, et al.
Pubblicazione: (2026)
di: Hong, Pingjun, et al.
Pubblicazione: (2026)
Calibrating Large Language Models with Sample Consistency
di: Lyu, Qing, et al.
Pubblicazione: (2024)
di: Lyu, Qing, et al.
Pubblicazione: (2024)
Prompting Science Report 3: I'll pay you or I'll kill you -- but will you care?
di: Meincke, Lennart, et al.
Pubblicazione: (2025)
di: Meincke, Lennart, et al.
Pubblicazione: (2025)
What Language(s) Does Aya-23 Think In? How Multilinguality Affects Internal Language Representations
di: Trinley, Katharina, et al.
Pubblicazione: (2025)
di: Trinley, Katharina, et al.
Pubblicazione: (2025)
User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums
di: Kulyabin, Mikhail, et al.
Pubblicazione: (2025)
di: Kulyabin, Mikhail, et al.
Pubblicazione: (2025)
HumanOmni-Speaker: Identifying Who said What and When
di: Bai, Detao, et al.
Pubblicazione: (2026)
di: Bai, Detao, et al.
Pubblicazione: (2026)
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck
di: Ludan, Josh Magnus, et al.
Pubblicazione: (2023)
di: Ludan, Josh Magnus, et al.
Pubblicazione: (2023)
WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback
di: Shi, Taiwei, et al.
Pubblicazione: (2024)
di: Shi, Taiwei, et al.
Pubblicazione: (2024)
MrGuard: A Multilingual Reasoning Guardrail for Universal LLM Safety
di: Yang, Yahan, et al.
Pubblicazione: (2025)
di: Yang, Yahan, et al.
Pubblicazione: (2025)
On the Calibration of Multilingual Question Answering LLMs
di: Yang, Yahan, et al.
Pubblicazione: (2023)
di: Yang, Yahan, et al.
Pubblicazione: (2023)
Reasoning is about giving reasons
di: Shah, Krunal, et al.
Pubblicazione: (2025)
di: Shah, Krunal, et al.
Pubblicazione: (2025)
Conflicts in Texts: Data, Implications and Challenges
di: Liu, Siyi, et al.
Pubblicazione: (2025)
di: Liu, Siyi, et al.
Pubblicazione: (2025)
What talking you?: Translating Code-Mixed Messaging Texts to English
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2024)
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2024)
What's In My Human Feedback? Learning Interpretable Descriptions of Preference Data
di: Movva, Rajiv, et al.
Pubblicazione: (2025)
di: Movva, Rajiv, et al.
Pubblicazione: (2025)
Compact Example-Based Explanations for Language Models
di: Schoenegger, Loris, et al.
Pubblicazione: (2026)
di: Schoenegger, Loris, et al.
Pubblicazione: (2026)
BankMathBench: A Benchmark for Numerical Reasoning in Banking Scenarios
di: Lee, Yunseung, et al.
Pubblicazione: (2026)
di: Lee, Yunseung, et al.
Pubblicazione: (2026)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
What Affects the Effective Depth of Large Language Models?
di: Hu, Yi, et al.
Pubblicazione: (2025)
di: Hu, Yi, et al.
Pubblicazione: (2025)
ConSiDERS-The-Human Evaluation Framework: Rethinking Human Evaluation for Generative Large Language Models
di: Elangovan, Aparna, et al.
Pubblicazione: (2024)
di: Elangovan, Aparna, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ExpertQA: Expert-Curated Questions and Attributed Answers
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2023) -
Flattery, Fluff, and Fog: Diagnosing and Mitigating Idiosyncratic Biases in Preference Models
di: Bharadwaj, Anirudh, et al.
Pubblicazione: (2025) -
Contextualized Evaluations: Judging Language Model Responses to Underspecified Queries
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2024) -
ResearchQA: Evaluating Scholarly Question Answering at Scale Across 75 Fields with Survey-Mined Questions and Rubrics
di: Yifei, Li S., et al.
Pubblicazione: (2025) -
A Simple Joint Model for Improved Contextual Neural Lemmatization
di: Malaviya, Chaitanya, et al.
Pubblicazione: (2019)