Saved in:
| Main Authors: | Luo, Zhe, Fu, Weina, Liu, Shuai, Anwar, Saeed, Saqib, Muhammad, Bakshi, Sambit, Muhammad, Khan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.05771 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HazeSpace2M: A Dataset for Haze Aware Single Image Dehazing
by: Islam, Md Tanvir, et al.
Published: (2024)
by: Islam, Md Tanvir, et al.
Published: (2024)
RDD4D: 4D Attention-Guided Road Damage Detection And Classification
by: Alkalbani, Asma, et al.
Published: (2025)
by: Alkalbani, Asma, et al.
Published: (2025)
From Satellite to Street: A Hybrid Framework Integrating Stable Diffusion and PanoGAN for Consistent Cross-View Synthesis
by: Bajbaa, Khawlah, et al.
Published: (2025)
by: Bajbaa, Khawlah, et al.
Published: (2025)
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
by: Alghamdi, Leena, et al.
Published: (2025)
by: Alghamdi, Leena, et al.
Published: (2025)
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
by: Ibrahim, Muhammad, et al.
Published: (2025)
by: Ibrahim, Muhammad, et al.
Published: (2025)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
by: Asad, Muhammad Hamza, et al.
Published: (2023)
by: Asad, Muhammad Hamza, et al.
Published: (2023)
Detecting Severity of Diabetic Retinopathy from Fundus Images: A Transformer Network-based Review
by: Karkera, Tejas, et al.
Published: (2023)
by: Karkera, Tejas, et al.
Published: (2023)
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025)
by: Abid, Hassan, et al.
Published: (2025)
Transformer-based Spatial Grounding: A Comprehensive Survey
by: Haq, Ijazul, et al.
Published: (2025)
by: Haq, Ijazul, et al.
Published: (2025)
MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking
by: Nadeem, Numair, et al.
Published: (2025)
by: Nadeem, Numair, et al.
Published: (2025)
Segmenting Visuals With Querying Words: Language Anchors For Semi-Supervised Image Segmentation
by: Nadeem, Numair, et al.
Published: (2025)
by: Nadeem, Numair, et al.
Published: (2025)
Underwater Object Detection Enhancement via Channel Stabilization
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition
by: Ullah, Hayat, et al.
Published: (2025)
by: Ullah, Hayat, et al.
Published: (2025)
Exploring Convolutional Neural Networks for Rice Grain Classification: An Explainable AI Approach
by: Asif, Muhammad Junaid, et al.
Published: (2025)
by: Asif, Muhammad Junaid, et al.
Published: (2025)
A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification
by: Khan, Muhammad Kaleem Ullah
Published: (2026)
by: Khan, Muhammad Kaleem Ullah
Published: (2026)
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Early Detection of Late Blight Tomato Disease using Histogram Oriented Gradient based Support Vector Machine
by: Alhwaiti, Yousef, et al.
Published: (2023)
by: Alhwaiti, Yousef, et al.
Published: (2023)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
Rethinking Transformers Pre-training for Multi-Spectral Satellite Imagery
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment
by: Alsaafin, Mohammed, et al.
Published: (2024)
by: Alsaafin, Mohammed, et al.
Published: (2024)
Deep Models for Multi-View 3D Object Recognition: A Review
by: Alzahrani, Mona, et al.
Published: (2024)
by: Alzahrani, Mona, et al.
Published: (2024)
Bird Eye-View to Street-View: A Survey
by: Bajbaa, Khawlah, et al.
Published: (2024)
by: Bajbaa, Khawlah, et al.
Published: (2024)
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026)
by: George, Franky, et al.
Published: (2026)
A Review on Coarse to Fine-Grained Animal Action Recognition
by: Zia, Ali, et al.
Published: (2025)
by: Zia, Ali, et al.
Published: (2025)
Multi-Granularity Hand Action Detection
by: Zhe, Ting, et al.
Published: (2023)
by: Zhe, Ting, et al.
Published: (2023)
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks
by: Mohsin, Muhammad Ahmed, et al.
Published: (2025)
by: Mohsin, Muhammad Ahmed, et al.
Published: (2025)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
FSBI: Deepfakes Detection with Frequency Enhanced Self-Blended Images
by: Hasanaath, Ahmed Abul, et al.
Published: (2024)
by: Hasanaath, Ahmed Abul, et al.
Published: (2024)
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
LoLI-Street: Benchmarking Low-Light Image Enhancement and Beyond
by: Islam, Md Tanvir, et al.
Published: (2024)
by: Islam, Md Tanvir, et al.
Published: (2024)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
Context-Aware Detection of Mixed Critical Events using Video Classification
by: Akhlaq, Filza, et al.
Published: (2024)
by: Akhlaq, Filza, et al.
Published: (2024)
FSKD: Monocular Forest Structure Inference via LiDAR-to-RGBI Knowledge Distillation
by: Khan, Taimur, et al.
Published: (2026)
by: Khan, Taimur, et al.
Published: (2026)
High-Performance Inference Graph Convolutional Networks for Skeleton-Based Action Recognition
by: Wang, Junyi, et al.
Published: (2023)
by: Wang, Junyi, et al.
Published: (2023)
How Effective are Self-Supervised Models for Contact Identification in Videos
by: Gunawardhana, Malitha, et al.
Published: (2024)
by: Gunawardhana, Malitha, et al.
Published: (2024)
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
Confidence-Weighted Semi-Supervised Learning for Skin Lesion Segmentation Using Hybrid CNN-Transformer Networks
by: Qamar, Saqib
Published: (2025)
by: Qamar, Saqib
Published: (2025)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
Similar Items
-
HazeSpace2M: A Dataset for Haze Aware Single Image Dehazing
by: Islam, Md Tanvir, et al.
Published: (2024) -
RDD4D: 4D Attention-Guided Road Damage Detection And Classification
by: Alkalbani, Asma, et al.
Published: (2025) -
From Satellite to Street: A Hybrid Framework Integrating Stable Diffusion and PanoGAN for Consistent Cross-View Synthesis
by: Bajbaa, Khawlah, et al.
Published: (2025) -
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
by: Alghamdi, Leena, et al.
Published: (2025) -
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
by: Ibrahim, Muhammad, et al.
Published: (2025)