C-DiffDet+: Fusing Global Scene Context with Generative Denoising for High-Fidelity Car Damage Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Sellam, Abdellah Zakaria, Benaissa, Ilyes, Bekhouche, Salah Eddine, Hadid, Abdenour, Renó, Vito, Distante, Cosimo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
por: Djouama, Ahmed Marouane, et al.
Publicado: (2026)
por: Djouama, Ahmed Marouane, et al.
Publicado: (2026)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
Mamba Adaptive Anomaly Transformer with association discrepancy for time series
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
LoLA-SpecViT: Local Attention SwiGLU Vision Transformer with LoRA for Hyperspectral Imaging
por: Zidi, Fadi Abdeladhim, et al.
Publicado: (2025)
por: Zidi, Fadi Abdeladhim, et al.
Publicado: (2025)
SegDT: A Diffusion Transformer-Based Segmentation Model for Medical Imaging
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
Enhanced Arabic Text Retrieval with Attentive Relevance Scoring
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
Conflict-Aware Multimodal Fusion for Ambivalence and Hesitancy Recognition
por: Bekhouche, Salah Eddine, et al.
Publicado: (2026)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2026)
SPARK-IL: Spectral Retrieval-Augmented RAG for Knowledge-driven Deepfake Detection via Incremental Learning
por: Eutamene, Hessen Bougueffa, et al.
Publicado: (2026)
por: Eutamene, Hessen Bougueffa, et al.
Publicado: (2026)
Boosting House Price Estimations with Multi-Head Gated Attention
por: Sellam, Zakaria Abdellah, et al.
Publicado: (2024)
por: Sellam, Zakaria Abdellah, et al.
Publicado: (2024)
CVPD at QIAS 2026: RAG-Guided LLM Reasoning for Al-Mawarith Share Computation and Heir Allocation
por: Swaileh, Wassim, et al.
Publicado: (2026)
por: Swaileh, Wassim, et al.
Publicado: (2026)
SAViL-Det: Semantic-Aware Vision-Language Model for Multi-Script Text Detection
por: Zighem, Mohammed-En-Nadhir, et al.
Publicado: (2025)
por: Zighem, Mohammed-En-Nadhir, et al.
Publicado: (2025)
Knowledge-Based Convolutional Neural Network for the Simulation and Prediction of Two-Phase Darcy Flows
por: Elabid, Zakaria, et al.
Publicado: (2024)
por: Elabid, Zakaria, et al.
Publicado: (2024)
DiffDet4SAR: Diffusion-based Aircraft Target Detection Network for SAR Images
por: Jie, Zhou, et al.
Publicado: (2024)
por: Jie, Zhou, et al.
Publicado: (2024)
Integrating ConvNeXt and Vision Transformers for Enhancing Facial Age Estimation
por: Maroun, Gaby, et al.
Publicado: (2025)
por: Maroun, Gaby, et al.
Publicado: (2025)
TG-PhyNN: An Enhanced Physically-Aware Graph Neural Network framework for forecasting Spatio-Temporal Data
por: Elabid, Zakaria, et al.
Publicado: (2024)
por: Elabid, Zakaria, et al.
Publicado: (2024)
Cross-Modal Mapping and Dual-Branch Reconstruction for 2D-3D Multimodal Industrial Anomaly Detection
por: Daci, Radia, et al.
Publicado: (2026)
por: Daci, Radia, et al.
Publicado: (2026)
Recent Advances in Medical Imaging Segmentation: A Survey
por: Bougourzi, Fares, et al.
Publicado: (2025)
por: Bougourzi, Fares, et al.
Publicado: (2025)
When Geoscience Meets Generative AI and Large Language Models: Foundations, Trends, and Future Challenges
por: Hadid, Abdenour, et al.
Publicado: (2024)
por: Hadid, Abdenour, et al.
Publicado: (2024)
TempoKGAT: A Novel Graph Attention Network Approach for Temporal Graph Analysis
por: Sasal, Lena, et al.
Publicado: (2024)
por: Sasal, Lena, et al.
Publicado: (2024)
Decoding Matters: Efficient Mamba-Based Decoder with Distribution-Aware Deep Supervision for Medical Image Segmentation
por: Bougourzi, Fares, et al.
Publicado: (2026)
por: Bougourzi, Fares, et al.
Publicado: (2026)
When geoscience meets generative AI and large language models: Foundations, trends, and future challenges
por: Abdenour Hadid, et al.
Publicado: (2024)
por: Abdenour Hadid, et al.
Publicado: (2024)
Boosting Hyperspectral Image Classification with Gate-Shift-Fuse Mechanisms in a Novel CNN-Transformer Approach
por: Guerri, Mohamed Fadhlallah, et al.
Publicado: (2024)
por: Guerri, Mohamed Fadhlallah, et al.
Publicado: (2024)
Can Vision Transformers with ResNet's Global Features Fairly Authenticate Demographic Faces?
por: Sufian, Abu, et al.
Publicado: (2025)
por: Sufian, Abu, et al.
Publicado: (2025)
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
por: Skorupski, Patryk, et al.
Publicado: (2025)
por: Skorupski, Patryk, et al.
Publicado: (2025)
CineInfini: Adaptive Multi-Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi‑Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi‑Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi-Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
Face to Cartoon Incremental Super-Resolution using Knowledge Distillation
por: Devkatte, Trinetra, et al.
Publicado: (2024)
por: Devkatte, Trinetra, et al.
Publicado: (2024)
T2IBias: Uncovering Societal Bias Encoded in the Latent Space of Text-to-Image Generative Models
por: Sufian, Abu, et al.
Publicado: (2025)
por: Sufian, Abu, et al.
Publicado: (2025)
AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions
por: Zeggai, Abdellah, et al.
Publicado: (2025)
por: Zeggai, Abdellah, et al.
Publicado: (2025)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
por: Keita, Mamadou, et al.
Publicado: (2024)
por: Keita, Mamadou, et al.
Publicado: (2024)
Skew-Probabilistic Neural Networks for Learning from Imbalanced Data
por: Naik, Shraddha M., et al.
Publicado: (2023)
por: Naik, Shraddha M., et al.
Publicado: (2023)
D-TrAttUnet: Toward Hybrid CNN-Transformer Architecture for Generic and Subtle Segmentation in Medical Images
por: Bougourzi, Fares, et al.
Publicado: (2024)
por: Bougourzi, Fares, et al.
Publicado: (2024)
Note on Boundary Stabilization of Degenerate Schrödinger Equations
por: Benaissa, Abdelkader, et al.
Publicado: (2026)
por: Benaissa, Abdelkader, et al.
Publicado: (2026)
DATASHI: A Parallel English-Tashlhiyt Corpus for Orthography Normalization and Low-Resource Language Processing
por: Monir, Nasser-Eddine, et al.
Publicado: (2026)
por: Monir, Nasser-Eddine, et al.
Publicado: (2026)
Bi-LORA: A Vision-Language Approach for Synthetic Image Detection
por: Keita, Mamadou, et al.
Publicado: (2024)
por: Keita, Mamadou, et al.
Publicado: (2024)
Ejemplares similares
-
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025) -
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025) -
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
por: Djouama, Ahmed Marouane, et al.
Publicado: (2026) -
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026) -
Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)