Finding Challenging Metaphors that Confuse Pretrained Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yucheng, Guerin, Frank, Lin, Chenghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Open Source Data Contamination Report for Large Language Models
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
LatestEval: Addressing Data Contamination in Language Model Evaluation through Dynamic and Time-Sensitive Test Construction
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
von: Li, Yucheng, et al.
Veröffentlicht: (2023)
Evaluating Large Language Models for Generalization and Robustness via Data Compression
von: Li, Yucheng, et al.
Veröffentlicht: (2024)
von: Li, Yucheng, et al.
Veröffentlicht: (2024)
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
von: Wang, Shun, et al.
Veröffentlicht: (2024)
von: Wang, Shun, et al.
Veröffentlicht: (2024)
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
von: Qu, Xingwei, et al.
Veröffentlicht: (2024)
von: Qu, Xingwei, et al.
Veröffentlicht: (2024)
On the Rigour of Scientific Writing: Criteria, Analysis, and Insights
von: James, Joseph, et al.
Veröffentlicht: (2024)
von: James, Joseph, et al.
Veröffentlicht: (2024)
CMDAG: A Chinese Metaphor Dataset with Annotated Grounds as CoT for Boosting Metaphor Generation
von: Shao, Yujie, et al.
Veröffentlicht: (2024)
von: Shao, Yujie, et al.
Veröffentlicht: (2024)
BioMNER: A Dataset for Biomedical Method Entity Recognition
von: Tang, Chen, et al.
Veröffentlicht: (2024)
von: Tang, Chen, et al.
Veröffentlicht: (2024)
Seeing isn't Hearing: Benchmarking Vision Language Models at Interpreting Spectrograms
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
von: Loakman, Tyler, et al.
Veröffentlicht: (2025)
Natural Language Generation
von: van Miltenburg, Emiel, et al.
Veröffentlicht: (2025)
von: van Miltenburg, Emiel, et al.
Veröffentlicht: (2025)
Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation
von: Zhang, Collin, et al.
Veröffentlicht: (2025)
von: Zhang, Collin, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Zero-shot Lay Summarisation in Biomedicine and Beyond
von: Goldsack, Tomas, et al.
Veröffentlicht: (2025)
von: Goldsack, Tomas, et al.
Veröffentlicht: (2025)
RIGOURATE: Quantifying Scientific Exaggeration with Evidence-Aligned Claim Evaluation
von: James, Joseph, et al.
Veröffentlicht: (2026)
von: James, Joseph, et al.
Veröffentlicht: (2026)
Metaphor Understanding Challenge Dataset for LLMs
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
Vision Language Models are Confused Tourists
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
von: Warstadt, Alex, et al.
Veröffentlicht: (2025)
von: Warstadt, Alex, et al.
Veröffentlicht: (2025)
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
Overview of the NLPCC 2025 Shared Task: Gender Bias Mitigation Challenge
von: Li, Yizhi, et al.
Veröffentlicht: (2025)
von: Li, Yizhi, et al.
Veröffentlicht: (2025)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
von: Hu, Michael Y., et al.
Veröffentlicht: (2024)
von: Hu, Michael Y., et al.
Veröffentlicht: (2024)
A Dual-Perspective Metaphor Detection Framework Using Large Language Models
von: Lin, Yujie, et al.
Veröffentlicht: (2024)
von: Lin, Yujie, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Language Confusion in LLMs
von: Marchisio, Kelly, et al.
Veröffentlicht: (2024)
von: Marchisio, Kelly, et al.
Veröffentlicht: (2024)
Controlling Language Confusion in Multilingual LLMs
von: Lee, Nahyun, et al.
Veröffentlicht: (2025)
von: Lee, Nahyun, et al.
Veröffentlicht: (2025)
Tougher Text, Smarter Models: Raising the Bar for Adversarial Defence Benchmarks
von: Wang, Yang, et al.
Veröffentlicht: (2025)
von: Wang, Yang, et al.
Veröffentlicht: (2025)
Adversarial Confusion Attack: Disrupting Multimodal Large Language Models
von: Hoscilowicz, Jakub, et al.
Veröffentlicht: (2025)
von: Hoscilowicz, Jakub, et al.
Veröffentlicht: (2025)
Confusion-Aware Rubric Optimization for LLM-based Automated Grading
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
The Frequency Confound in Language-Model Surprisal and Metaphor Novelty
von: Momen, Omar, et al.
Veröffentlicht: (2026)
von: Momen, Omar, et al.
Veröffentlicht: (2026)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
von: Nie, Ercong, et al.
Veröffentlicht: (2025)
von: Nie, Ercong, et al.
Veröffentlicht: (2025)
ReproHum #0087-01: Human Evaluation Reproduction Report for Generating Fact Checking Explanations
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
von: Loakman, Tyler, et al.
Veröffentlicht: (2024)
Towards Multimodal Metaphor Understanding: A Chinese Dataset and Model for Metaphor Mapping Identification
von: Zhang, Dongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Dongyu, et al.
Veröffentlicht: (2025)
LVPruning: An Effective yet Simple Language-Guided Vision Token Pruning Approach for Multi-modal Large Language Models
von: Sun, Yizheng, et al.
Veröffentlicht: (2025)
von: Sun, Yizheng, et al.
Veröffentlicht: (2025)
SLIDE: A Framework Integrating Small and Large Language Models for Open-Domain Dialogues Evaluation
von: Zhao, Kun, et al.
Veröffentlicht: (2024)
von: Zhao, Kun, et al.
Veröffentlicht: (2024)
Automated Multiple Mini Interview (MMI) Scoring
von: Huynh, Ryan, et al.
Veröffentlicht: (2026)
von: Huynh, Ryan, et al.
Veröffentlicht: (2026)
Conceptual Metaphor Theory as a Prompting Paradigm for Large Language Models
von: Kramer, Oliver
Veröffentlicht: (2025)
von: Kramer, Oliver
Veröffentlicht: (2025)
Na'vi or Knave: Jailbreaking Language Models via Metaphorical Avatars
von: Yan, Yu, et al.
Veröffentlicht: (2024)
von: Yan, Yu, et al.
Veröffentlicht: (2024)
Harnessing the Intrinsic Knowledge of Pretrained Language Models for Challenging Text Classification Settings
von: Gao, Lingyu
Veröffentlicht: (2024)
von: Gao, Lingyu
Veröffentlicht: (2024)
Large Language Model Displays Emergent Ability to Interpret Novel Literary Metaphors
von: Ichien, Nicholas, et al.
Veröffentlicht: (2023)
von: Ichien, Nicholas, et al.
Veröffentlicht: (2023)
Pretraining Language Models Using Translationese
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
Geographic Adaptation of Pretrained Language Models
von: Hofmann, Valentin, et al.
Veröffentlicht: (2022)
von: Hofmann, Valentin, et al.
Veröffentlicht: (2022)
Pretraining Language Models for Diachronic Linguistic Change Discovery
von: Fittschen, Elisabeth, et al.
Veröffentlicht: (2025)
von: Fittschen, Elisabeth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Open Source Data Contamination Report for Large Language Models
von: Li, Yucheng, et al.
Veröffentlicht: (2023) -
LatestEval: Addressing Data Contamination in Language Model Evaluation through Dynamic and Time-Sensitive Test Construction
von: Li, Yucheng, et al.
Veröffentlicht: (2023) -
Evaluating Large Language Models for Generalization and Robustness via Data Compression
von: Li, Yucheng, et al.
Veröffentlicht: (2024) -
With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language Models
von: Loakman, Tyler, et al.
Veröffentlicht: (2024) -
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
von: Wang, Shun, et al.
Veröffentlicht: (2024)