SegDT: A Diffusion Transformer-Based Segmentation Model for Medical Imaging
Fuente:
arXiv
Saved in:
| Main Authors: | Bekhouche, Salah Eddine, Maroun, Gaby, Dornaika, Fadi, Hadid, Abdenour |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Integrating ConvNeXt and Vision Transformers for Enhancing Facial Age Estimation
by: Maroun, Gaby, et al.
Published: (2025)
by: Maroun, Gaby, et al.
Published: (2025)
Decoding Matters: Efficient Mamba-Based Decoder with Distribution-Aware Deep Supervision for Medical Image Segmentation
by: Bougourzi, Fares, et al.
Published: (2026)
by: Bougourzi, Fares, et al.
Published: (2026)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
by: Djouama, Ahmed Marouane, et al.
Published: (2026)
by: Djouama, Ahmed Marouane, et al.
Published: (2026)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
by: Sellam, Abdellah Zakaria, et al.
Published: (2026)
by: Sellam, Abdellah Zakaria, et al.
Published: (2026)
Conflict-Aware Multimodal Fusion for Ambivalence and Hesitancy Recognition
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
Recent Advances in Medical Imaging Segmentation: A Survey
by: Bougourzi, Fares, et al.
Published: (2025)
by: Bougourzi, Fares, et al.
Published: (2025)
C-DiffDet+: Fusing Global Scene Context with Generative Denoising for High-Fidelity Car Damage Detection
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
Enhanced Arabic Text Retrieval with Attentive Relevance Scoring
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
HSMix: Hard and Soft Mixing Data Augmentation for Medical Image Segmentation
by: Sun, Danyang, et al.
Published: (2025)
by: Sun, Danyang, et al.
Published: (2025)
MambaCAFU: Hybrid Multi-Scale and Multi-Attention Model with Mamba-Based Fusion for Medical Image Segmentation
by: Bui, T-Mai, et al.
Published: (2025)
by: Bui, T-Mai, et al.
Published: (2025)
LoLA-SpecViT: Local Attention SwiGLU Vision Transformer with LoRA for Hyperspectral Imaging
by: Zidi, Fadi Abdeladhim, et al.
Published: (2025)
by: Zidi, Fadi Abdeladhim, et al.
Published: (2025)
D-TrAttUnet: Toward Hybrid CNN-Transformer Architecture for Generic and Subtle Segmentation in Medical Images
by: Bougourzi, Fares, et al.
Published: (2024)
by: Bougourzi, Fares, et al.
Published: (2024)
Rethinking Attention Gated with Hybrid Dual Pyramid Transformer-CNN for Generalized Segmentation in Medical Imaging
by: Bougourzi, Fares, et al.
Published: (2024)
by: Bougourzi, Fares, et al.
Published: (2024)
SAViL-Det: Semantic-Aware Vision-Language Model for Multi-Script Text Detection
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
by: Zighem, Mohammed-En-Nadhir, et al.
Published: (2025)
Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
A Re-node Self-training Approach for Deep Graph-based Semi-supervised Classification on Multi-view Image Data
by: Bi, Jingjun, et al.
Published: (2025)
by: Bi, Jingjun, et al.
Published: (2025)
Local and Global Context-and-Object-part-Aware Superpixel-based Data Augmentation for Deep Visual Recognition
by: Dornaika, Fadi, et al.
Published: (2025)
by: Dornaika, Fadi, et al.
Published: (2025)
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
by: Keita, Mamadou, et al.
Published: (2025)
by: Keita, Mamadou, et al.
Published: (2025)
SCS-SupCon: Sigmoid-based Common and Style Supervised Contrastive Learning with Adaptive Decision Boundaries
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
Lung Infection Severity Prediction Using Transformers with Conditional TransMix Augmentation and Cross-Attention
by: Slika, Bouthaina, et al.
Published: (2025)
by: Slika, Bouthaina, et al.
Published: (2025)
SegStitch: Multidimensional Transformer for Robust and Efficient Medical Imaging Segmentation
by: Tan, Shengbo, et al.
Published: (2024)
by: Tan, Shengbo, et al.
Published: (2024)
Extremely Fine-Grained Visual Classification over Resembling Glyphs in the Wild
by: Bougourzi, Fares, et al.
Published: (2024)
by: Bougourzi, Fares, et al.
Published: (2024)
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation
by: Perera, Shehan, et al.
Published: (2024)
by: Perera, Shehan, et al.
Published: (2024)
Shuffle Vision Transformer: Lightweight, Fast and Efficient Recognition of Driver Facial Expression
by: Saadi, Ibtissam, et al.
Published: (2024)
by: Saadi, Ibtissam, et al.
Published: (2024)
FIDAVL: Fake Image Detection and Attribution using Vision-Language Model
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
by: Kim, Chaehyun, et al.
Published: (2025)
by: Kim, Chaehyun, et al.
Published: (2025)
MCFCN: Multi-View Clustering via a Fusion-Consensus Graph Convolutional Network
by: Pei, Chenping, et al.
Published: (2025)
by: Pei, Chenping, et al.
Published: (2025)
Enhancing Semi-Supervised Multi-View Graph Convolutional Networks via Supervised Contrastive Learning and Self-Training
by: Xiao, Huaiyuan, et al.
Published: (2025)
by: Xiao, Huaiyuan, et al.
Published: (2025)
MedSeg-R: Medical Image Segmentation with Clinical Reasoning
by: Shao, Hao, et al.
Published: (2025)
by: Shao, Hao, et al.
Published: (2025)
SegVol: Universal and Interactive Volumetric Medical Image Segmentation
by: Du, Yuxin, et al.
Published: (2023)
by: Du, Yuxin, et al.
Published: (2023)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
PE-CLIP: A Parameter-Efficient Fine-Tuning of Vision Language Models for Dynamic Facial Expression Recognition
by: Saadi, Ibtissam, et al.
Published: (2025)
by: Saadi, Ibtissam, et al.
Published: (2025)
MedSeg-R: Reasoning Segmentation in Medical Images with Multimodal Large Language Models
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation
by: Ahmed, Salma J., et al.
Published: (2026)
by: Ahmed, Salma J., et al.
Published: (2026)
RefineSeg: Dual Coarse-to-Fine Learning for Medical Image Segmentation
by: Du, Anghong, et al.
Published: (2025)
by: Du, Anghong, et al.
Published: (2025)
SPARK-IL: Spectral Retrieval-Augmented RAG for Knowledge-driven Deepfake Detection via Incremental Learning
by: Eutamene, Hessen Bougueffa, et al.
Published: (2026)
by: Eutamene, Hessen Bougueffa, et al.
Published: (2026)
ConvTransSeg: A Multi-resolution Convolution-Transformer Network for Medical Image Segmentation
by: Gong, Zhendi, et al.
Published: (2022)
by: Gong, Zhendi, et al.
Published: (2022)
SegMamba: Long-range Sequential Modeling Mamba For 3D Medical Image Segmentation
by: Xing, Zhaohu, et al.
Published: (2024)
by: Xing, Zhaohu, et al.
Published: (2024)
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
by: Bekhouche, Salah Eddine, et al.
Published: (2025)
Similar Items
-
Integrating ConvNeXt and Vision Transformers for Enhancing Facial Age Estimation
by: Maroun, Gaby, et al.
Published: (2025) -
Decoding Matters: Efficient Mamba-Based Decoder with Distribution-Aware Deep Supervision for Medical Image Segmentation
by: Bougourzi, Fares, et al.
Published: (2026) -
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025) -
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
by: Djouama, Ahmed Marouane, et al.
Published: (2026) -
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
by: Sellam, Abdellah Zakaria, et al.
Published: (2026)