Addressing Data Scarcity in Bangla Fake News Detection: An LLM-Based Dataset Augmentation Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Sani, Ahmed Alfey, Zaoad, Kazi Akib, Adib, Shefayat E Shams, Muqtadir, Md Abdul, Abrar, Ajwad |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification
by: Adib, Shefayat E Shams, et al.
Published: (2026)
by: Adib, Shefayat E Shams, et al.
Published: (2026)
Assessing Large Language Models for Medical QA: Zero-Shot and LLM-as-a-Judge Evaluation
by: Adib, Shefayat E Shams, et al.
Published: (2026)
by: Adib, Shefayat E Shams, et al.
Published: (2026)
BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali
by: Adib, Shefayat E Shams, et al.
Published: (2026)
by: Adib, Shefayat E Shams, et al.
Published: (2026)
BanglaSummEval: Reference-Free Factual Consistency Evaluation for Bangla Summarization
by: Rafid, Ahmed, et al.
Published: (2026)
by: Rafid, Ahmed, et al.
Published: (2026)
MixSarc: A Bangla-English Code-Mixed Corpus for Implicit Meaning Identification
by: Alam, Kazi Samin Yasar, et al.
Published: (2026)
by: Alam, Kazi Samin Yasar, et al.
Published: (2026)
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization
by: Abrar, Ajwad, et al.
Published: (2025)
by: Abrar, Ajwad, et al.
Published: (2025)
BanglaMedQA and BanglaMMedBench: Evaluating Retrieval-Augmented Generation Strategies for Bangla Biomedical Question Answering
by: Sultana, Sadia, et al.
Published: (2025)
by: Sultana, Sadia, et al.
Published: (2025)
Generalization Gaps in Political Fake News Detection: An Empirical Study on the LIAR Dataset
by: Hasan, S Mahmudul, et al.
Published: (2025)
by: Hasan, S Mahmudul, et al.
Published: (2025)
Breaking the Fake News Barrier: Deep Learning Approaches in Bangla Language
by: Mondal, Pronoy Kumar, et al.
Published: (2025)
by: Mondal, Pronoy Kumar, et al.
Published: (2025)
Bangla Fake News Detection Based On Multichannel Combined CNN-LSTM
by: George, Md. Zahin Hossain, et al.
Published: (2025)
by: George, Md. Zahin Hossain, et al.
Published: (2025)
Hybrid Feature Combinations with CNN for Bangla Fake News Classification
by: Hussain, Md Gulzar, et al.
Published: (2026)
by: Hussain, Md Gulzar, et al.
Published: (2026)
Vashantor: A Large-scale Multilingual Benchmark Dataset for Automated Translation of Bangla Regional Dialects to Bangla Language
by: Faria, Fatema Tuj Johora, et al.
Published: (2023)
by: Faria, Fatema Tuj Johora, et al.
Published: (2023)
From Scarcity to Capability: Empowering Fake News Detection in Low-Resource Languages with LLMs
by: Shibu, Hrithik Majumdar, et al.
Published: (2025)
by: Shibu, Hrithik Majumdar, et al.
Published: (2025)
Distinguishing Repetition Disfluency from Morphological Reduplication in Bangla ASR Transcripts: A Novel Corpus and Benchmarking Analysis
by: Arpa, Zaara Zabeen, et al.
Published: (2025)
by: Arpa, Zaara Zabeen, et al.
Published: (2025)
Strengthening False Information Propagation Detection: Leveraging SVM and Sophisticated Text Vectorization Techniques in comparison to BERT
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance
by: Eyasir, Abrar, et al.
Published: (2026)
by: Eyasir, Abrar, et al.
Published: (2026)
eGov-Lens: Code and Dataset
by: Fuad, Md. Muqtadir, et al.
Published: (2026)
by: Fuad, Md. Muqtadir, et al.
Published: (2026)
BanglaFake: Constructing and Evaluating a Specialized Bengali Deepfake Audio Dataset
by: Fahad, Istiaq Ahmed, et al.
Published: (2025)
by: Fahad, Istiaq Ahmed, et al.
Published: (2025)
Enhancing Bangla Fake News Detection Using Bidirectional Gated Recurrent Units and Deep Learning Techniques
by: Roy, Utsha, et al.
Published: (2024)
by: Roy, Utsha, et al.
Published: (2024)
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
by: Dihan, Mahir Labib, et al.
Published: (2025)
by: Dihan, Mahir Labib, et al.
Published: (2025)
Bangla BERT for Hyperpartisan News Detection: A Semi-Supervised and Explainable AI Approach
by: Hasan, Mohammad Mehadi, et al.
Published: (2025)
by: Hasan, Mohammad Mehadi, et al.
Published: (2025)
BDA: Bangla Text Data Augmentation Framework
by: Tariquzzaman, Md., et al.
Published: (2024)
by: Tariquzzaman, Md., et al.
Published: (2024)
Towards Real-Time Fake News Detection under Evidence Scarcity
by: Wei, Guangyu, et al.
Published: (2025)
by: Wei, Guangyu, et al.
Published: (2025)
ChitroJera: A Regionally Relevant Visual Question Answering Dataset for Bangla
by: Barua, Deeparghya Dutta, et al.
Published: (2024)
by: Barua, Deeparghya Dutta, et al.
Published: (2024)
Break the Checkbox: Challenging Closed-Style Evaluations of Cultural Alignment in LLMs
by: Kabir, Mohsinul, et al.
Published: (2025)
by: Kabir, Mohsinul, et al.
Published: (2025)
Larger models yield better results? Streamlined severity classification of ADHD-related concerns using BERT-based knowledge distillation
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
by: Hasan, Md Arid, et al.
Published: (2025)
by: Hasan, Md Arid, et al.
Published: (2025)
Religious Bias Landscape in Language and Text-to-Image Models: Analysis, Detection, and Debiasing Strategies
by: Abrar, Ajwad, et al.
Published: (2025)
by: Abrar, Ajwad, et al.
Published: (2025)
From Chat to Checkup: Can Large Language Models Assist in Diabetes Prediction?
by: Sakib, Shadman, et al.
Published: (2025)
by: Sakib, Shadman, et al.
Published: (2025)
Bangla-WhisperDiar: Fine-Tuning Whisper and PyAnnote for Bangla Long-Form Speech Recognition and Speaker Diarization
by: Bhuiyan, Mohammed Aman, et al.
Published: (2026)
by: Bhuiyan, Mohammed Aman, et al.
Published: (2026)
AI- Enhanced Stethoscope in Remote Diagnostics for Cardiopulmonary Diseases
by: Ghouse, Hania, et al.
Published: (2025)
by: Ghouse, Hania, et al.
Published: (2025)
FirstAidQA: A Synthetic Dataset for First Aid and Emergency Response in Low-Connectivity Settings
by: Muna, Saiyma Sittul, et al.
Published: (2025)
by: Muna, Saiyma Sittul, et al.
Published: (2025)
Bangla MedER: Multi-BERT Ensemble Approach for the Recognition of Bangla Medical Entity
by: Aurpa, Tanjim Taharat, et al.
Published: (2025)
by: Aurpa, Tanjim Taharat, et al.
Published: (2025)
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
by: Haider, Fabiha, et al.
Published: (2024)
by: Haider, Fabiha, et al.
Published: (2024)
Bangla Sign Language Recognition With Multimodal Deep Learning Fusion
by: Adib Hasan, et al.
Published: (2025)
by: Adib Hasan, et al.
Published: (2025)
Enhancing Traffic Sign Recognition with Tailored Data Augmentation: Addressing Class Imbalance and Instance Scarcity
by: Alsiyeu, Ulan, et al.
Published: (2024)
by: Alsiyeu, Ulan, et al.
Published: (2024)
Official-NV: An LLM-Generated News Video Dataset for Multimodal Fake News Detection
by: Wang, Yihao, et al.
Published: (2024)
by: Wang, Yihao, et al.
Published: (2024)
How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking
by: Ahmed, Rafid, et al.
Published: (2026)
by: Ahmed, Rafid, et al.
Published: (2026)
CogniAlign: Survivability-Grounded Multi-Agent Moral Reasoning for Safe and Transparent AI
by: Ali, Hasin Jawad, et al.
Published: (2025)
by: Ali, Hasin Jawad, et al.
Published: (2025)
FLea: Addressing Data Scarcity and Label Skew in Federated Learning via Privacy-preserving Feature Augmentation
by: Xia, Tong, et al.
Published: (2023)
by: Xia, Tong, et al.
Published: (2023)
Similar Items
-
LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification
by: Adib, Shefayat E Shams, et al.
Published: (2026) -
Assessing Large Language Models for Medical QA: Zero-Shot and LLM-as-a-Judge Evaluation
by: Adib, Shefayat E Shams, et al.
Published: (2026) -
BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali
by: Adib, Shefayat E Shams, et al.
Published: (2026) -
BanglaSummEval: Reference-Free Factual Consistency Evaluation for Bangla Summarization
by: Rafid, Ahmed, et al.
Published: (2026) -
MixSarc: A Bangla-English Code-Mixed Corpus for Implicit Meaning Identification
by: Alam, Kazi Samin Yasar, et al.
Published: (2026)