Zero-Shot Commonsense Validation and Reasoning with Large Language Models: An Evaluation on SemEval-2020 Task 4 Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alfugaha, Rawand, AL-Smadi, Mohammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IntegrityAI at GenAI Detection Task 2: Detecting Machine-Generated Academic Essays in English and Arabic Using ELECTRA and Stylometry
von: AL-Smadi, Mohammad
Veröffentlicht: (2025)
von: AL-Smadi, Mohammad
Veröffentlicht: (2025)
LyS at SemEval 2025 Task 8: Zero-Shot Code Generation for Tabular QA
von: Gude, Adrián, et al.
Veröffentlicht: (2025)
von: Gude, Adrián, et al.
Veröffentlicht: (2025)
QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment
von: AL-Smadi, Mohammad
Veröffentlicht: (2026)
von: AL-Smadi, Mohammad
Veröffentlicht: (2026)
SHROOM-INDElab at SemEval-2024 Task 6: Zero- and Few-Shot LLM-Based Classification for Hallucination Detection
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
Automated Detection of Dosing Errors in Clinical Trial Narratives: A Multi-Modal Feature Engineering Approach with LightGBM
von: AL-Smadi, Mohammad
Veröffentlicht: (2026)
von: AL-Smadi, Mohammad
Veröffentlicht: (2026)
SemEval-2026 Task 12: Abductive Event Reasoning: Towards Real-World Event Causal Inference for Large Language Models
von: Cao, Pengfei, et al.
Veröffentlicht: (2026)
von: Cao, Pengfei, et al.
Veröffentlicht: (2026)
REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models
von: Lee, DongGeon, et al.
Veröffentlicht: (2025)
von: Lee, DongGeon, et al.
Veröffentlicht: (2025)
Cyber for AI at SemEval-2025 Task 4: Forgotten but Not Lost: The Balancing Act of Selective Unlearning in Large Language Models
von: P, Dinesh Srivasthav, et al.
Veröffentlicht: (2025)
von: P, Dinesh Srivasthav, et al.
Veröffentlicht: (2025)
ITLC at SemEval-2026 Task 11: Normalization and Deterministic Parsing for Formal Reasoning in LLMs
von: Muhamad, Wicaksono Leksono, et al.
Veröffentlicht: (2026)
von: Muhamad, Wicaksono Leksono, et al.
Veröffentlicht: (2026)
SHA256 at SemEval-2025 Task 4: Selective Amnesia -- Constrained Unlearning for Large Language Models via Knowledge Isolation
von: Agrawal, Saransh, et al.
Veröffentlicht: (2025)
von: Agrawal, Saransh, et al.
Veröffentlicht: (2025)
NLP at UC Santa Cruz at SemEval-2024 Task 5: Legal Answer Validation using Few-Shot Multi-Choice QA
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
Team QUST at SemEval-2025 Task 10: Evaluating Large Language Models in Multiclass Multi-label Classification of News Entity Framing
von: Liu, Jiyan, et al.
Veröffentlicht: (2025)
von: Liu, Jiyan, et al.
Veröffentlicht: (2025)
SemEval-2024 Task 2: Safe Biomedical Natural Language Inference for Clinical Trials
von: Jullien, Mael, et al.
Veröffentlicht: (2024)
von: Jullien, Mael, et al.
Veröffentlicht: (2024)
SemEval 2024 -- Task 10: Emotion Discovery and Reasoning its Flip in Conversation (EDiReF)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
von: Kumar, Shivani, et al.
Veröffentlicht: (2024)
SemEval-2024 Task 3: Multimodal Emotion Cause Analysis in Conversations
von: Wang, Fanfan, et al.
Veröffentlicht: (2024)
von: Wang, Fanfan, et al.
Veröffentlicht: (2024)
mdok-style at SemEval-2026 Task 10: Finetuning LLMs for Conspiracy Detection
von: Macko, Dominik
Veröffentlicht: (2026)
von: Macko, Dominik
Veröffentlicht: (2026)
ZJUKLAB at SemEval-2025 Task 4: Unlearning via Model Merging
von: Xu, Haoming, et al.
Veröffentlicht: (2025)
von: Xu, Haoming, et al.
Veröffentlicht: (2025)
SLPL SHROOM at SemEval2024 Task 06: A comprehensive study on models ability to detect hallucination
von: Fallah, Pouya, et al.
Veröffentlicht: (2024)
von: Fallah, Pouya, et al.
Veröffentlicht: (2024)
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning
von: Mathur, Suyash Vardhan, et al.
Veröffentlicht: (2024)
von: Mathur, Suyash Vardhan, et al.
Veröffentlicht: (2024)
Lacuna Inc. at SemEval-2025 Task 4: LoRA-Enhanced Influence-Based Unlearning for LLMs
von: Kudelya, Aleksey, et al.
Veröffentlicht: (2025)
von: Kudelya, Aleksey, et al.
Veröffentlicht: (2025)
mdok-style at SemEval-2026 Task 9: Finetuning LLMs for Multilingual Polarization Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
Empaths at SemEval-2025 Task 11: Retrieval-Augmented Approach to Perceived Emotions Prediction
von: Morozov, Lev, et al.
Veröffentlicht: (2025)
von: Morozov, Lev, et al.
Veröffentlicht: (2025)
SmurfCat at SemEval-2024 Task 6: Leveraging Synthetic Data for Hallucination Detection
von: Rykov, Elisei, et al.
Veröffentlicht: (2024)
von: Rykov, Elisei, et al.
Veröffentlicht: (2024)
IITK at SemEval-2024 Task 4: Hierarchical Embeddings for Detection of Persuasion Techniques in Memes
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
Abdelhak at SemEval-2024 Task 9 : Decoding Brainteasers, The Efficacy of Dedicated Models Versus ChatGPT
von: Kelious, Abdelhak, et al.
Veröffentlicht: (2024)
von: Kelious, Abdelhak, et al.
Veröffentlicht: (2024)
AIMA at SemEval-2024 Task 3: Simple Yet Powerful Emotion Cause Pair Analysis
von: Kure, Alireza Ghahramani, et al.
Veröffentlicht: (2025)
von: Kure, Alireza Ghahramani, et al.
Veröffentlicht: (2025)
SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense
von: Jiang, Yifan, et al.
Veröffentlicht: (2024)
von: Jiang, Yifan, et al.
Veröffentlicht: (2024)
Homa at SemEval-2025 Task 5: Aligning Librarian Records with OntoAligner for Subject Tagging
von: Tekanlou, Hadi Bayrami Asl, et al.
Veröffentlicht: (2025)
von: Tekanlou, Hadi Bayrami Asl, et al.
Veröffentlicht: (2025)
HausaNLP at SemEval-2025 Task 3: Towards a Fine-Grained Model-Aware Hallucination Detection
von: Bala, Maryam, et al.
Veröffentlicht: (2025)
von: Bala, Maryam, et al.
Veröffentlicht: (2025)
Ustnlp16 at SemEval-2025 Task 9: Improving Model Performance through Imbalance Handling and Focal Loss
von: Cai, Zhuoang, et al.
Veröffentlicht: (2025)
von: Cai, Zhuoang, et al.
Veröffentlicht: (2025)
AILS-NTUA at SemEval-2024 Task 9: Cracking Brain Teasers: Transformer Models for Lateral Thinking Puzzles
von: Panagiotopoulos, Ioannis, et al.
Veröffentlicht: (2024)
von: Panagiotopoulos, Ioannis, et al.
Veröffentlicht: (2024)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
von: Vemula, Saketh Reddy, et al.
Veröffentlicht: (2025)
von: Vemula, Saketh Reddy, et al.
Veröffentlicht: (2025)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
von: Spiegel, Michal, et al.
Veröffentlicht: (2024)
von: Spiegel, Michal, et al.
Veröffentlicht: (2024)
FZI-WIM at SemEval-2024 Task 2: Self-Consistent CoT for Complex NLI in Biomedical Domain
von: Liu, Jin, et al.
Veröffentlicht: (2024)
von: Liu, Jin, et al.
Veröffentlicht: (2024)
HalluSearch at SemEval-2025 Task 3: A Search-Enhanced RAG Pipeline for Hallucination Detection
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
MRT at SemEval-2025 Task 8: Maximizing Recovery from Tables with Multiple Steps
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
von: Fane, Enfa, et al.
Veröffentlicht: (2025)
von: Fane, Enfa, et al.
Veröffentlicht: (2025)
FLANS at SemEval-2026 Task 7: RAG with Open-Sourced Smaller LLMs for Everyday Knowledge Across Diverse Languages and Cultures
von: Bogdanova, Liliia, et al.
Veröffentlicht: (2026)
von: Bogdanova, Liliia, et al.
Veröffentlicht: (2026)
AIMA at SemEval-2024 Task 10: History-Based Emotion Recognition in Hindi-English Code-Mixed Conversations
von: Abootorabi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
von: Abootorabi, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
RFBES at SemEval-2024 Task 8: Investigating Syntactic and Semantic Features for Distinguishing AI-Generated and Human-Written Texts
von: Rad, Mohammad Heydari, et al.
Veröffentlicht: (2024)
von: Rad, Mohammad Heydari, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
IntegrityAI at GenAI Detection Task 2: Detecting Machine-Generated Academic Essays in English and Arabic Using ELECTRA and Stylometry
von: AL-Smadi, Mohammad
Veröffentlicht: (2025) -
LyS at SemEval 2025 Task 8: Zero-Shot Code Generation for Tabular QA
von: Gude, Adrián, et al.
Veröffentlicht: (2025) -
QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment
von: AL-Smadi, Mohammad
Veröffentlicht: (2026) -
SHROOM-INDElab at SemEval-2024 Task 6: Zero- and Few-Shot LLM-Based Classification for Hallucination Detection
von: Allen, Bradley P., et al.
Veröffentlicht: (2024) -
Automated Detection of Dosing Errors in Clinical Trial Narratives: A Multi-Modal Feature Engineering Approach with LightGBM
von: AL-Smadi, Mohammad
Veröffentlicht: (2026)