Evaluation of an LLM in Identifying Logical Fallacies: A Call for Rigor When Adopting LLMs in HCI Research
Fuente:
arXiv
Saved in:
| Main Authors: | Lim, Gionnieve, Perrault, Simon T. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Iffy-Or-Not: Extending the Web to Support the Critical Evaluation of Fallacious Texts
by: Lim, Gionnieve, et al.
Published: (2025)
by: Lim, Gionnieve, et al.
Published: (2025)
Fact Checking Chatbot: A Misinformation Intervention for Instant Messaging Apps and an Analysis of Trust in the Fact Checkers
by: Lim, Gionnieve, et al.
Published: (2024)
by: Lim, Gionnieve, et al.
Published: (2024)
Rapid AIdeation: Generating Ideas With the Self and in Collaboration With Large Language Models
by: Lim, Gionnieve, et al.
Published: (2024)
by: Lim, Gionnieve, et al.
Published: (2024)
Effects of Automated Misinformation Warning Labels on the Intents to Like, Comment and Share Posts
by: Lim, Gionnieve, et al.
Published: (2024)
by: Lim, Gionnieve, et al.
Published: (2024)
XAI in Automated Fact-Checking? The Benefits Are Modest and There's No One-Explanation-Fits-All
by: Lim, Gionnieve, et al.
Published: (2023)
by: Lim, Gionnieve, et al.
Published: (2023)
Local Perceptions and Practices of News Sharing and Fake News
by: Lim, Gionnieve, et al.
Published: (2020)
by: Lim, Gionnieve, et al.
Published: (2020)
Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation
by: Lim, Gionnieve, et al.
Published: (2025)
by: Lim, Gionnieve, et al.
Published: (2025)
Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs
by: Nahar, Mahjabin, et al.
Published: (2026)
by: Nahar, Mahjabin, et al.
Published: (2026)
Expanding Horizons in HCI Research Through LLM-Driven Qualitative Analysis
by: Torii, Maya Grace, et al.
Published: (2024)
by: Torii, Maya Grace, et al.
Published: (2024)
CollabCoder: A Lower-barrier, Rigorous Workflow for Inductive Collaborative Qualitative Analysis with Large Language Models
by: Gao, Jie, et al.
Published: (2023)
by: Gao, Jie, et al.
Published: (2023)
Help Me Reflect: Leveraging Self-Reflection Interface Nudges to Enhance Deliberativeness on Online Deliberation Platforms
by: Yeo, Shun Yi, et al.
Published: (2024)
by: Yeo, Shun Yi, et al.
Published: (2024)
The Potential and Implications of Generative AI on HCI Education
by: Kharrufa, Ahmed, et al.
Published: (2024)
by: Kharrufa, Ahmed, et al.
Published: (2024)
Avatar Visual Similarity for Social HCI: Increasing Self-Awareness
by: Hilpert, Bernhard, et al.
Published: (2024)
by: Hilpert, Bernhard, et al.
Published: (2024)
AdaptoML-UX: An Adaptive User-centered GUI-based AutoML Toolkit for Non-AI Experts and HCI Researchers
by: Gomaa, Amr, et al.
Published: (2024)
by: Gomaa, Amr, et al.
Published: (2024)
MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning
by: Bhattarai, Ankit, et al.
Published: (2026)
by: Bhattarai, Ankit, et al.
Published: (2026)
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
by: Li, Karen Jia-Hui, et al.
Published: (2025)
by: Li, Karen Jia-Hui, et al.
Published: (2025)
The HCI GenAI CO2ST Calculator: A Tool for Calculating the Carbon Footprint of Generative AI Use in Human-Computer Interaction Research
by: Inie, Nanna, et al.
Published: (2025)
by: Inie, Nanna, et al.
Published: (2025)
The European Commitment to Human-Centered Technology: The Integral Role of HCI in the EU AI Act's Success
by: Valdez, André Calero, et al.
Published: (2024)
by: Valdez, André Calero, et al.
Published: (2024)
Experience Paper: Adopting Activity Recognition in On-demand Food Delivery Business
by: Xu, Huatao, et al.
Published: (2025)
by: Xu, Huatao, et al.
Published: (2025)
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
by: Wei, Kevin L., et al.
Published: (2025)
by: Wei, Kevin L., et al.
Published: (2025)
When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis
by: Najera, Aisha, et al.
Published: (2026)
by: Najera, Aisha, et al.
Published: (2026)
Establishing Rigorous and Cost-effective Clinical Trials for Artificial Intelligence Models
by: Gao, Wanling, et al.
Published: (2024)
by: Gao, Wanling, et al.
Published: (2024)
Reflexis: Supporting Reflexivity and Rigor in Collaborative Qualitative Analysis through Design for Deliberation
by: Ye, Runlong, et al.
Published: (2026)
by: Ye, Runlong, et al.
Published: (2026)
Evaluating LLMs for Visualization Generation and Understanding
by: Khan, Saadiq Rauf, et al.
Published: (2025)
by: Khan, Saadiq Rauf, et al.
Published: (2025)
Addressing the Ecological Fallacy in Larger LMs with Human Context
by: Soni, Nikita, et al.
Published: (2026)
by: Soni, Nikita, et al.
Published: (2026)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
by: Jiang, Guanxuan, et al.
Published: (2025)
by: Jiang, Guanxuan, et al.
Published: (2025)
Bridging Cultural Distance Between Models Default and Local Classroom Demands: How Global Teachers Adopt GenAI to Support Everyday Teaching Practices
by: Xiao, Ruiwei, et al.
Published: (2025)
by: Xiao, Ruiwei, et al.
Published: (2025)
"Oops! ChatGPT is Temporarily Unavailable!": A Diary Study on Knowledge Workers' Experiences of LLM Withdrawal
by: Oh, Eunseo, et al.
Published: (2026)
by: Oh, Eunseo, et al.
Published: (2026)
ComplLLM: Fine-tuning LLMs to Discover Complementary Signals for Decision-making
by: Guo, Ziyang, et al.
Published: (2026)
by: Guo, Ziyang, et al.
Published: (2026)
Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
Logic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMs
by: Rahdari, Behnam, et al.
Published: (2023)
by: Rahdari, Behnam, et al.
Published: (2023)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?
by: Yin, Xiaoyun, et al.
Published: (2025)
by: Yin, Xiaoyun, et al.
Published: (2025)
Leveraging LLMs to Predict Affective States via Smartphone Sensor Features
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Evaluating the Semantic Profiling Abilities of LLMs for Natural Language Utterances in Data Visualization
by: Bako, Hannah K., et al.
Published: (2024)
by: Bako, Hannah K., et al.
Published: (2024)
Talk Less, Call Right: Enhancing Role-Play LLM Agents with Automatic Prompt Optimization and Role Prompting
by: Ruangtanusak, Saksorn, et al.
Published: (2025)
by: Ruangtanusak, Saksorn, et al.
Published: (2025)
H is for Human and How (Not) To Evaluate Qualitative Research in HCI
by: Crabtree, Andy
Published: (2024)
by: Crabtree, Andy
Published: (2024)
Evaluating graph-based explanations for AI-based recommender systems
by: Delarue, Simon, et al.
Published: (2024)
by: Delarue, Simon, et al.
Published: (2024)
ConsistencyAI: A Benchmark to Assess LLMs' Factual Consistency When Responding to Different Demographic Groups
by: Banyas, Peter, et al.
Published: (2025)
by: Banyas, Peter, et al.
Published: (2025)
It's Not Just Labeling -- A Research on LLM Generated Feedback Interpretability and Image Labeling Sketch Features
by: Li, Baichuan, et al.
Published: (2025)
by: Li, Baichuan, et al.
Published: (2025)
Similar Items
-
Iffy-Or-Not: Extending the Web to Support the Critical Evaluation of Fallacious Texts
by: Lim, Gionnieve, et al.
Published: (2025) -
Fact Checking Chatbot: A Misinformation Intervention for Instant Messaging Apps and an Analysis of Trust in the Fact Checkers
by: Lim, Gionnieve, et al.
Published: (2024) -
Rapid AIdeation: Generating Ideas With the Self and in Collaboration With Large Language Models
by: Lim, Gionnieve, et al.
Published: (2024) -
Effects of Automated Misinformation Warning Labels on the Intents to Like, Comment and Share Posts
by: Lim, Gionnieve, et al.
Published: (2024) -
XAI in Automated Fact-Checking? The Benefits Are Modest and There's No One-Explanation-Fits-All
by: Lim, Gionnieve, et al.
Published: (2023)