Comparing Apples to Oranges: A Dataset & Analysis of LLM Humour Understanding from Traditional Puns to Topical Jokes
Fuente:
arXiv
Saved in:
| Main Authors: | Loakman, Tyler, Thorne, William, Lin, Chenghua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Who's Laughing Now? An Overview of Computational Humour Generation and Explanation
by: Loakman, Tyler, et al.
Published: (2025)
by: Loakman, Tyler, et al.
Published: (2025)
Train & Constrain: Phonologically Informed Tongue-Twister Generation from Topics and Paraphrases
by: Loakman, Tyler, et al.
Published: (2024)
by: Loakman, Tyler, et al.
Published: (2024)
ReproHum #0087-01: Human Evaluation Reproduction Report for Generating Fact Checking Explanations
by: Loakman, Tyler, et al.
Published: (2024)
by: Loakman, Tyler, et al.
Published: (2024)
Seeing isn't Hearing: Benchmarking Vision Language Models at Interpreting Spectrograms
by: Loakman, Tyler, et al.
Published: (2025)
by: Loakman, Tyler, et al.
Published: (2025)
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
by: Loakman, Tyler, et al.
Published: (2024)
by: Loakman, Tyler, et al.
Published: (2024)
CADGE: Context-Aware Dialogue Generation Enhanced with Graph-Structured Knowledge Aggregation
by: Zhang, Hongbo, et al.
Published: (2023)
by: Zhang, Hongbo, et al.
Published: (2023)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
by: Wang, Shun, et al.
Published: (2024)
by: Wang, Shun, et al.
Published: (2024)
Explaining Humour Style Classifications: An XAI Approach to Understanding Computational Humour Analysis
by: Kenneth, Mary Ogbuka, et al.
Published: (2025)
by: Kenneth, Mary Ogbuka, et al.
Published: (2025)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
by: Bansal, Hritik, et al.
Published: (2024)
by: Bansal, Hritik, et al.
Published: (2024)
A Survey of Pun Generation: Datasets, Evaluations and Methodologies
by: Su, Yuchen, et al.
Published: (2025)
by: Su, Yuchen, et al.
Published: (2025)
LongEval: A Comprehensive Analysis of Long-Text Generation Through a Plan-based Paradigm
by: Wu, Siwei, et al.
Published: (2025)
by: Wu, Siwei, et al.
Published: (2025)
Drivel-ology: Challenging LLMs with Interpreting Nonsense with Depth
by: Wang, Yang, et al.
Published: (2025)
by: Wang, Yang, et al.
Published: (2025)
Exploring Task Performance with Interpretable Models via Sparse Auto-Encoders
by: Wang, Shun, et al.
Published: (2025)
by: Wang, Shun, et al.
Published: (2025)
Visual Puns from Idioms: An Iterative LLM-T2IM-MLLM Framework
by: Xiao, Kelaiti, et al.
Published: (2025)
by: Xiao, Kelaiti, et al.
Published: (2025)
Evaluating LLM-Based Grant Proposal Review via Structured Perturbations
by: Thorne, William, et al.
Published: (2026)
by: Thorne, William, et al.
Published: (2026)
Pun Unintended: LLMs and the Illusion of Humor Understanding
by: Zangari, Alessandro, et al.
Published: (2025)
by: Zangari, Alessandro, et al.
Published: (2025)
Not All Jokes Land: Evaluating Large Language Models Understanding of Workplace Humor
by: Shafiei, Mohammadamin, et al.
Published: (2025)
by: Shafiei, Mohammadamin, et al.
Published: (2025)
Words at Play: Benchmarking Audio Pun Understanding in Large Audio-Language Models
by: Su, Yuchen, et al.
Published: (2026)
by: Su, Yuchen, et al.
Published: (2026)
"A good pun is its own reword": Can Large Language Models Understand Puns?
by: Xu, Zhijun, et al.
Published: (2024)
by: Xu, Zhijun, et al.
Published: (2024)
Two Ways to Set Up Wireless Hotspot: Comparing Apples and Oranges
by: Mutch, Andrew, et al.
Published: (2006)
by: Mutch, Andrew, et al.
Published: (2006)
Psychology-Driven Enhancement of Humour Translation
by: Su, Yuchen, et al.
Published: (2025)
by: Su, Yuchen, et al.
Published: (2025)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
"I See What You Did There": Can Large Vision-Language Models Understand Multimodal Puns?
by: Xu, Naen, et al.
Published: (2026)
by: Xu, Naen, et al.
Published: (2026)
Multilingual Topic Classification in X: Dataset and Analysis
by: Antypas, Dimosthenis, et al.
Published: (2024)
by: Antypas, Dimosthenis, et al.
Published: (2024)
Comparing Apples to Oranges: A Taxonomy for Navigating the Global Landscape of AI Regulation
by: Alanoca, Sacha, et al.
Published: (2025)
by: Alanoca, Sacha, et al.
Published: (2025)
Identifying Imaging Follow-Up in Radiology Reports: A Comparative Analysis of Traditional ML and LLM Approaches
by: Park, Namu, et al.
Published: (2025)
by: Park, Namu, et al.
Published: (2025)
A Two-Model Approach for Humour Style Recognition
by: Kenneth, Mary Ogbuka, et al.
Published: (2024)
by: Kenneth, Mary Ogbuka, et al.
Published: (2024)
CAST: Corpus-Aware Self-similarity Enhanced Topic modelling
by: Ma, Yanan, et al.
Published: (2024)
by: Ma, Yanan, et al.
Published: (2024)
Towards Multimodal Prediction of Spontaneous Humour: A Novel Dataset and First Results
by: Christ, Lukas, et al.
Published: (2022)
by: Christ, Lukas, et al.
Published: (2022)
No Joke: An Embodied Conversational Agent Greeting Older Adults with Humour or a Smile Unrelated to Initial Acceptance
by: Li, Ge "Rikaku", et al.
Published: (2024)
by: Li, Ge "Rikaku", et al.
Published: (2024)
URLs Help, Topics Guide: Understanding Metadata Utility in LLM Training
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
ReceiptSense: Beyond Traditional OCR -- A Dataset for Receipt Understanding
by: Abdallah, Abdelrahman, et al.
Published: (2024)
by: Abdallah, Abdelrahman, et al.
Published: (2024)
One Joke to Rule them All? On the (Im)possibility of Generalizing Humor
by: Turgeman, Mor, et al.
Published: (2025)
by: Turgeman, Mor, et al.
Published: (2025)
On the Rigour of Scientific Writing: Criteria, Analysis, and Insights
by: James, Joseph, et al.
Published: (2024)
by: James, Joseph, et al.
Published: (2024)
Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and Reasoning
by: Yang, Bohao, et al.
Published: (2025)
by: Yang, Bohao, et al.
Published: (2025)
Large Language Models Offer an Alternative to the Traditional Approach of Topic Modelling
by: Mu, Yida, et al.
Published: (2024)
by: Mu, Yida, et al.
Published: (2024)
BioMNER: A Dataset for Biomedical Method Entity Recognition
by: Tang, Chen, et al.
Published: (2024)
by: Tang, Chen, et al.
Published: (2024)
Increasing the Difficulty of Automatically Generated Questions via Reinforcement Learning with Synthetic Preference
by: Thorne, William, et al.
Published: (2024)
by: Thorne, William, et al.
Published: (2024)
Natural Language Generation
by: van Miltenburg, Emiel, et al.
Published: (2025)
by: van Miltenburg, Emiel, et al.
Published: (2025)
They Said Memes Were Harmless-We Found the Ones That Hurt: Decoding Jokes, Symbols, and Cultural References
by: Tripathi, Sahil, et al.
Published: (2026)
by: Tripathi, Sahil, et al.
Published: (2026)
Similar Items
-
Who's Laughing Now? An Overview of Computational Humour Generation and Explanation
by: Loakman, Tyler, et al.
Published: (2025) -
Train & Constrain: Phonologically Informed Tongue-Twister Generation from Topics and Paraphrases
by: Loakman, Tyler, et al.
Published: (2024) -
ReproHum #0087-01: Human Evaluation Reproduction Report for Generating Fact Checking Explanations
by: Loakman, Tyler, et al.
Published: (2024) -
Seeing isn't Hearing: Benchmarking Vision Language Models at Interpreting Spectrograms
by: Loakman, Tyler, et al.
Published: (2025) -
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
by: Loakman, Tyler, et al.
Published: (2024)