Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition
Fuente:
arXiv
Saved in:
| Main Author: | Rayhan, Md. Sultan Al |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion State-Guided Projected Gradient for Inverse Problems
by: Zirvi, Rayhan, et al.
Published: (2024)
by: Zirvi, Rayhan, et al.
Published: (2024)
LOCR: Location-Guided Transformer for Optical Character Recognition
by: Sun, Yu, et al.
Published: (2024)
by: Sun, Yu, et al.
Published: (2024)
BanglaNet: Bangla Handwritten Character Recognition using Ensembling of Convolutional Neural Network
by: Saha, Chandrika, et al.
Published: (2024)
by: Saha, Chandrika, et al.
Published: (2024)
Performance Analysis of Few-Shot Learning Approaches for Bangla Handwritten Character and Digit Recognition
by: Ahamed, Mehedi, et al.
Published: (2025)
by: Ahamed, Mehedi, et al.
Published: (2025)
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
by: Ma, Yuhang, et al.
Published: (2024)
by: Ma, Yuhang, et al.
Published: (2024)
Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition
by: Zhou, Yuxi, et al.
Published: (2026)
by: Zhou, Yuxi, et al.
Published: (2026)
SGCCNet: Single-Stage 3D Object Detector With Saliency-Guided Data Augmentation and Confidence Correction Mechanism
by: Liang, Ao, et al.
Published: (2024)
by: Liang, Ao, et al.
Published: (2024)
Multi-Modal Character Localization and Extraction for Chinese Text Recognition
by: Li, Qilong, et al.
Published: (2026)
by: Li, Qilong, et al.
Published: (2026)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
by: Papadimitriou, Christos, et al.
Published: (2024)
by: Papadimitriou, Christos, et al.
Published: (2024)
SASG-DA: Sparse-Aware Semantic-Guided Diffusion Augmentation For Myoelectric Gesture Recognition
by: Liu, Chen, et al.
Published: (2025)
by: Liu, Chen, et al.
Published: (2025)
A Comprehensive Literature Review on Sweet Orange Leaf Diseases
by: Emon, Yousuf Rayhan, et al.
Published: (2023)
by: Emon, Yousuf Rayhan, et al.
Published: (2023)
Nepali Sign Language Characters Recognition: Dataset Development and Deep Learning Approaches
by: Poudel, Birat, et al.
Published: (2025)
by: Poudel, Birat, et al.
Published: (2025)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
by: Kaliosis, Panagiotis, et al.
Published: (2025)
by: Kaliosis, Panagiotis, et al.
Published: (2025)
Compound Expression Recognition via Multi Model Ensemble
by: Yu, Jun, et al.
Published: (2024)
by: Yu, Jun, et al.
Published: (2024)
Temporal Label Hierachical Network for Compound Emotion Recognition
by: Li, Sunan, et al.
Published: (2024)
by: Li, Sunan, et al.
Published: (2024)
Taming Diffusion Probabilistic Models for Character Control
by: Chen, Rui, et al.
Published: (2024)
by: Chen, Rui, et al.
Published: (2024)
TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision
by: Gillani, Syeda Anshrah, et al.
Published: (2025)
by: Gillani, Syeda Anshrah, et al.
Published: (2025)
Compound Expression Recognition via Large Vision-Language Models
by: Yu, Jun, et al.
Published: (2025)
by: Yu, Jun, et al.
Published: (2025)
DTGen: Generative Diffusion-Based Few-Shot Data Augmentation for Fine-Grained Dirty Tableware Recognition
by: Hao, Lifei, et al.
Published: (2025)
by: Hao, Lifei, et al.
Published: (2025)
ContextAnyone: Context-Aware Diffusion for Character-Consistent Text-to-Video Generation
by: Mai, Ziyang, et al.
Published: (2025)
by: Mai, Ziyang, et al.
Published: (2025)
Enhancing Traffic Sign Recognition with Tailored Data Augmentation: Addressing Class Imbalance and Instance Scarcity
by: Alsiyeu, Ulan, et al.
Published: (2024)
by: Alsiyeu, Ulan, et al.
Published: (2024)
AnyTop: Character Animation Diffusion with Any Topology
by: Gat, Inbar, et al.
Published: (2025)
by: Gat, Inbar, et al.
Published: (2025)
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
by: Rubaiyeat, Husne Ara, et al.
Published: (2025)
Enhancing Table Recognition with Vision LLMs: A Benchmark and Neighbor-Guided Toolchain Reasoner
by: Zhou, Yitong, et al.
Published: (2024)
by: Zhou, Yitong, et al.
Published: (2024)
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition
by: Bhatia, Gagan, et al.
Published: (2024)
by: Bhatia, Gagan, et al.
Published: (2024)
Forecasting When to Forecast: Accelerating Diffusion Models with Confidence-Gated Taylor
by: Guan, Xiaoliu, et al.
Published: (2025)
by: Guan, Xiaoliu, et al.
Published: (2025)
PsOCR: Benchmarking Large Multimodal Models for Optical Character Recognition in Low-resource Pashto Language
by: Haq, Ijazul, et al.
Published: (2025)
by: Haq, Ijazul, et al.
Published: (2025)
Interactive Character Control with Auto-Regressive Motion Diffusion Models
by: Shi, Yi, et al.
Published: (2023)
by: Shi, Yi, et al.
Published: (2023)
DANet: Enhancing Small Object Detection through an Efficient Deformable Attention Network
by: Mia, Md Sohag, et al.
Published: (2023)
by: Mia, Md Sohag, et al.
Published: (2023)
Toward Reliable Tea Leaf Disease Diagnosis Using Deep Learning Model: Enhancing Robustness With Explainable AI and Adversarial Training
by: Ghosh, Samanta, et al.
Published: (2026)
by: Ghosh, Samanta, et al.
Published: (2026)
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
by: Na, Kihyun, et al.
Published: (2025)
by: Na, Kihyun, et al.
Published: (2025)
Effective Data Augmentation With Diffusion Models
by: Trabucco, Brandon, et al.
Published: (2023)
by: Trabucco, Brandon, et al.
Published: (2023)
Remote Sensing Image Classification Using Deep Ensemble Learning
by: Islam, Niful, et al.
Published: (2026)
by: Islam, Niful, et al.
Published: (2026)
ZAYAN: Disentangled Contrastive Transformer for Tabular Remote Sensing Data
by: Habib, Al Zadid Sultan Bin, et al.
Published: (2026)
by: Habib, Al Zadid Sultan Bin, et al.
Published: (2026)
Semantic Data Augmentation for Long-tailed Facial Expression Recognition
by: Li, Zijian, et al.
Published: (2024)
by: Li, Zijian, et al.
Published: (2024)
Paired Image Generation with Diffusion-Guided Diffusion Models
by: Zhang, Haoxuan, et al.
Published: (2025)
by: Zhang, Haoxuan, et al.
Published: (2025)
Surgical Triplet Recognition via Diffusion Model
by: Liu, Daochang, et al.
Published: (2024)
by: Liu, Daochang, et al.
Published: (2024)
Rethinking Genomic Modeling Through Optical Character Recognition
by: Xiang, Hongxin, et al.
Published: (2026)
by: Xiang, Hongxin, et al.
Published: (2026)
Enhancing Object Detection Robustness: Detecting and Restoring Confidence in the Presence of Adversarial Patch Attacks
by: Kazoom, Roie, et al.
Published: (2024)
by: Kazoom, Roie, et al.
Published: (2024)
Similar Items
-
Diffusion State-Guided Projected Gradient for Inverse Problems
by: Zirvi, Rayhan, et al.
Published: (2024) -
LOCR: Location-Guided Transformer for Optical Character Recognition
by: Sun, Yu, et al.
Published: (2024) -
BanglaNet: Bangla Handwritten Character Recognition using Ensembling of Convolutional Neural Network
by: Saha, Chandrika, et al.
Published: (2024) -
Performance Analysis of Few-Shot Learning Approaches for Bangla Handwritten Character and Digit Recognition
by: Ahamed, Mehedi, et al.
Published: (2025) -
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
by: Ma, Yuhang, et al.
Published: (2024)