Drivel-ology: Challenging LLMs with Interpreting Nonsense with Depth
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yang, Xiao, Chenghao, Hsiao, Chia-Yi, Chang, Zi Yan, Chen, Chi-Li, Loakman, Tyler, Lin, Chenghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Seeing isn't Hearing: Benchmarking Vision Language Models at Interpreting Spectrograms
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
ReproHum #0087-01: Human Evaluation Reproduction Report for Generating Fact Checking Explanations
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
Train & Constrain: Phonologically Informed Tongue-Twister Generation from Topics and Paraphrases
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
Who's Laughing Now? An Overview of Computational Humour Generation and Explanation
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
Comparing Apples to Oranges: A Dataset & Analysis of LLM Humour Understanding from Traditional Puns to Topical Jokes
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
Exploring Task Performance with Interpretable Models via Sparse Auto-Encoders
von: Wang, Shun, et al.
Veröffentlicht: (2025)
von: Wang, Shun, et al.
Veröffentlicht: (2025)
CADGE: Context-Aware Dialogue Generation Enhanced with Graph-Structured Knowledge Aggregation
von: Zhang, Hongbo, et al.
Veröffentlicht: (2023)
von: Zhang, Hongbo, et al.
Veröffentlicht: (2023)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
von: Wang, Shun, et al.
Veröffentlicht: (2024)
von: Wang, Shun, et al.
Veröffentlicht: (2024)
Effective Distillation of Table-based Reasoning Ability from LLMs
von: Yang, Bohao, et al.
Veröffentlicht: (2023)
von: Yang, Bohao, et al.
Veröffentlicht: (2023)
On the Rigour of Scientific Writing: Criteria, Analysis, and Insights
von: James, Joseph, et al.
Veröffentlicht: (2024)
von: James, Joseph, et al.
Veröffentlicht: (2024)
Crafting Customisable Characters with LLMs: A Persona-Driven Role-Playing Agent Framework
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
Beyond One-Size-Fits-All: Inversion Learning for Highly Effective NLG Evaluation Prompts
von: Hong, Hanhua, et al.
Veröffentlicht: (2025)
von: Hong, Hanhua, et al.
Veröffentlicht: (2025)
LongEval: A Comprehensive Analysis of Long-Text Generation Through a Plan-based Paradigm
von: Wu, Siwei, et al.
Veröffentlicht: (2025)
von: Wu, Siwei, et al.
Veröffentlicht: (2025)
RIGOURATE: Quantifying Scientific Exaggeration with Evidence-Aligned Claim Evaluation
von: James, Joseph, et al.
Veröffentlicht: (2026)
von: James, Joseph, et al.
Veröffentlicht: (2026)
Adversarial Defence without Adversarial Defence: Enhancing Language Model Robustness via Instance-level Principal Component Removal
von: Wang, Yang, et al.
Veröffentlicht: (2025)
von: Wang, Yang, et al.
Veröffentlicht: (2025)
BioMNER: A Dataset for Biomedical Method Entity Recognition
von: Tang, Chen, et al.
Veröffentlicht: (2024)
von: Tang, Chen, et al.
Veröffentlicht: (2024)
Finding Challenging Metaphors that Confuse Pretrained Language Models
von: Li, Yucheng, et al.
Veröffentlicht: (2024)
von: Li, Yucheng, et al.
Veröffentlicht: (2024)
From Facts to Insights: A Study on the Generation and Evaluation of Analytical Reports for Deciphering Earnings Calls
von: Goldsack, Tomas, et al.
Veröffentlicht: (2024)
von: Goldsack, Tomas, et al.
Veröffentlicht: (2024)
Audio Contrastive-based Fine-tuning: Decoupling Representation Learning and Classification
von: Wang, Yang, et al.
Veröffentlicht: (2023)
von: Wang, Yang, et al.
Veröffentlicht: (2023)
Tougher Text, Smarter Models: Raising the Bar for Adversarial Defence Benchmarks
von: Wang, Yang, et al.
Veröffentlicht: (2025)
von: Wang, Yang, et al.
Veröffentlicht: (2025)
In Defense of "Ignorant Drivel".
von: Boyce, Bert R., et al.
Veröffentlicht: (1987)
von: Boyce, Bert R., et al.
Veröffentlicht: (1987)
Overview of the NLPCC 2025 Shared Task: Gender Bias Mitigation Challenge
von: Li, Yizhi, et al.
Veröffentlicht: (2025)
von: Li, Yizhi, et al.
Veröffentlicht: (2025)
Equipping Transformer with Random-Access Reading for Long-Context Understanding
von: Yang, Chenghao, et al.
Veröffentlicht: (2024)
von: Yang, Chenghao, et al.
Veröffentlicht: (2024)
LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores
von: Liu, Yiqi, et al.
Veröffentlicht: (2023)
von: Liu, Yiqi, et al.
Veröffentlicht: (2023)
Unlearning in LLMs: Methods, Evaluation, and Open Challenges
von: Lizzo, Tyler, et al.
Veröffentlicht: (2026)
von: Lizzo, Tyler, et al.
Veröffentlicht: (2026)
X-ray Made Simple: Lay Radiology Report Generation and Robust Evaluation
von: Zhao, Kun, et al.
Veröffentlicht: (2024)
von: Zhao, Kun, et al.
Veröffentlicht: (2024)
Emphasising Structured Information: Integrating Abstract Meaning Representation into LLMs for Enhanced Open-Domain Dialogue Evaluation
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
von: Yang, Bohao, et al.
Veröffentlicht: (2024)
Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration
von: Huang, Langlin, et al.
Veröffentlicht: (2026)
von: Huang, Langlin, et al.
Veröffentlicht: (2026)
The(y)ology
von: Brumberg-Kraus, Max Yeshaye
Veröffentlicht: (2023)
von: Brumberg-Kraus, Max Yeshaye
Veröffentlicht: (2023)
Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
CAST: Corpus-Aware Self-similarity Enhanced Topic modelling
von: Ma, Yanan, et al.
Veröffentlicht: (2024)
von: Ma, Yanan, et al.
Veröffentlicht: (2024)
Quantifier Scope Interpretation in Language Learners and LLMs
von: Fang, Shaohua, et al.
Veröffentlicht: (2025)
von: Fang, Shaohua, et al.
Veröffentlicht: (2025)
Prompt-Induced Linguistic Fingerprints for LLM-Generated Fake News Detection
von: Wang, Chi, et al.
Veröffentlicht: (2025)
von: Wang, Chi, et al.
Veröffentlicht: (2025)
The Achilles' Heel of Angular Margins: A Chebyshev Polynomial Fix for Speaker Verification
von: Wang, Yang, et al.
Veröffentlicht: (2026)
von: Wang, Yang, et al.
Veröffentlicht: (2026)
Natural Language Generation
von: van Miltenburg, Emiel, et al.
Veröffentlicht: (2025)
von: van Miltenburg, Emiel, et al.
Veröffentlicht: (2025)
Self-Evolved Reward Learning for LLMs
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
von: Huang, Chenghua, et al.
Veröffentlicht: (2024)
An Open Source Data Contamination Report for Large Language Models
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
LatestEval: Addressing Data Contamination in Language Model Evaluation through Dynamic and Time-Sensitive Test Construction
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
Paraphrase-Aligned Machine Translation
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Seeing isn't Hearing: Benchmarking Vision Language Models at Interpreting Spectrograms
von: Loakman, Tyler, et al.
Veröffentlicht: (2025) -
ReproHum #0087-01: Human Evaluation Reproduction Report for Generating Fact Checking Explanations
von: Loakman, Tyler, et al.
Veröffentlicht: (2024) -
Train & Constrain: Phonologically Informed Tongue-Twister Generation from Topics and Paraphrases
von: Loakman, Tyler, et al.
Veröffentlicht: (2024) -
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
von: Loakman, Tyler, et al.
Veröffentlicht: (2024) -
Who's Laughing Now? An Overview of Computational Humour Generation and Explanation
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)