When LLMs Meet Cunning Texts: A Fallacy Understanding Benchmark for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yinghui, Zhou, Qingyu, Luo, Yuanzhen, Ma, Shirong, Li, Yangning, Zheng, Hai-Tao, Hu, Xuming, Yu, Philip S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the (In)Effectiveness of Large Language Models for Chinese Text Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
Rethinking the Roles of Large Language Models in Chinese Grammatical Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
Correct Like Humans: Progressive Learning Framework for Chinese Text Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
LatEval: An Interactive LLMs Evaluation Benchmark with Incomplete Information from Lateral Thinking Puzzles
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding
von: Li, Yinghui, et al.
Veröffentlicht: (2026)
von: Li, Yinghui, et al.
Veröffentlicht: (2026)
Enhancing Phrase Representation by Information Bottleneck Guided Text Diffusion Process for Keyphrase Extraction
von: Luo, Yuanzhen, et al.
Veröffentlicht: (2023)
von: Luo, Yuanzhen, et al.
Veröffentlicht: (2023)
UltraWiki: Ultra-fine-grained Entity Set Expansion with Negative Seed Entities
von: Li, Yangning, et al.
Veröffentlicht: (2024)
von: Li, Yangning, et al.
Veröffentlicht: (2024)
Exploring the Implicit Semantic Ability of Multimodal Large Language Models: A Pilot Study on Entity Set Expansion
von: Wang, Hebin, et al.
Veröffentlicht: (2024)
von: Wang, Hebin, et al.
Veröffentlicht: (2024)
Mitigating Catastrophic Forgetting in Multi-domain Chinese Spelling Correction by Multi-stage Knowledge Transfer Framework
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Automatic Context Pattern Generation for Entity Set Expansion
von: Li, Yinghui, et al.
Veröffentlicht: (2022)
von: Li, Yinghui, et al.
Veröffentlicht: (2022)
ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions
von: Ye, Jingheng, et al.
Veröffentlicht: (2024)
von: Ye, Jingheng, et al.
Veröffentlicht: (2024)
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
von: Li, Yangning, et al.
Veröffentlicht: (2022)
von: Li, Yangning, et al.
Veröffentlicht: (2022)
Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
von: Li, Yangning, et al.
Veröffentlicht: (2024)
von: Li, Yangning, et al.
Veröffentlicht: (2024)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
von: Chen, Shaoshen, et al.
Veröffentlicht: (2025)
von: Chen, Shaoshen, et al.
Veröffentlicht: (2025)
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
RAISE: Reinforced Adaptive Instruction Selection For Large Language Models
von: Lv, Qingsong, et al.
Veröffentlicht: (2025)
von: Lv, Qingsong, et al.
Veröffentlicht: (2025)
Reason from Fallacy: Enhancing Large Language Models' Logical Reasoning through Logical Fallacy Understanding
von: Li, Yanda, et al.
Veröffentlicht: (2024)
von: Li, Yanda, et al.
Veröffentlicht: (2024)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Bidirectional End-to-End Learning of Retriever-Reader Paradigm for Entity Linking
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
von: Hu, Xuming, et al.
Veröffentlicht: (2024)
von: Hu, Xuming, et al.
Veröffentlicht: (2024)
Large Language Models Meet NLP: A Survey
von: Qin, Libo, et al.
Veröffentlicht: (2024)
von: Qin, Libo, et al.
Veröffentlicht: (2024)
One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
When Text Embedding Meets Large Language Model: A Comprehensive Survey
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction
von: Ye, Jingheng, et al.
Veröffentlicht: (2025)
von: Ye, Jingheng, et al.
Veröffentlicht: (2025)
Finetuning LLMs for EvaCun 2025 token prediction shared task
von: Jon, Josef, et al.
Veröffentlicht: (2025)
von: Jon, Josef, et al.
Veröffentlicht: (2025)
MER 2025: When Affective Computing Meets Large Language Models
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
When Fuzzing Meets LLMs: Challenges and Opportunities
von: Jiang, Yu, et al.
Veröffentlicht: (2024)
von: Jiang, Yu, et al.
Veröffentlicht: (2024)
EvoConfig: Self-Evolving Multi-Agent Systems for Efficient Autonomous Environment Configuration
von: Guo, Xinshuai, et al.
Veröffentlicht: (2026)
von: Guo, Xinshuai, et al.
Veröffentlicht: (2026)
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
Are LLMs Good Zero-Shot Fallacy Classifiers?
von: Pan, Fengjun, et al.
Veröffentlicht: (2024)
von: Pan, Fengjun, et al.
Veröffentlicht: (2024)
When Reasoning Meets Compression: Understanding the Effects of LLMs Compression on Large Reasoning Models
von: Zhang, Nan, et al.
Veröffentlicht: (2025)
von: Zhang, Nan, et al.
Veröffentlicht: (2025)
From Token to Line: Enhancing Code Generation with a Long-Term Perspective
von: Lu, Tingwei, et al.
Veröffentlicht: (2025)
von: Lu, Tingwei, et al.
Veröffentlicht: (2025)
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
von: Zheng, Kening, et al.
Veröffentlicht: (2024)
von: Zheng, Kening, et al.
Veröffentlicht: (2024)
SpatialText: A Pure-Text Cognitive Benchmark for Spatial Understanding in Large Language Models
von: Jiang, Peiyao, et al.
Veröffentlicht: (2026)
von: Jiang, Peiyao, et al.
Veröffentlicht: (2026)
A Survey of Text Watermarking in the Era of Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
On the (In)Effectiveness of Large Language Models for Chinese Text Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023) -
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
von: Huang, Shulin, et al.
Veröffentlicht: (2023) -
Rethinking the Roles of Large Language Models in Chinese Grammatical Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2024) -
Correct Like Humans: Progressive Learning Framework for Chinese Text Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023) -
LatEval: An Interactive LLMs Evaluation Benchmark with Incomplete Information from Lateral Thinking Puzzles
von: Huang, Shulin, et al.
Veröffentlicht: (2023)