Robust and Label-Efficient Deep Waste Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Abid, Hassan, Muhammad, Khan, Khan, Muhammad Haris |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024)
by: Khan, Adnan, et al.
Published: (2024)
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
Towards Generalizing to Unseen Domains with Few Labels
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
Real-Time Object Detection in Occluded Environment with Background Cluttering Effects Using Deep Learning
by: Aamir, Syed Muhammad, et al.
Published: (2024)
by: Aamir, Syed Muhammad, et al.
Published: (2024)
Depth Attention for Robust RGB Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
by: Ali, Eman, et al.
Published: (2023)
by: Ali, Eman, et al.
Published: (2023)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
by: Sharifdeen, Ashshak, et al.
Published: (2025)
by: Sharifdeen, Ashshak, et al.
Published: (2025)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
by: Ali, Eman, et al.
Published: (2024)
by: Ali, Eman, et al.
Published: (2024)
Towards Multimodal Domain Generalization with Few Labels
by: Li, Hongzhao, et al.
Published: (2026)
by: Li, Hongzhao, et al.
Published: (2026)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
by: Chokuwa, Sharon, et al.
Published: (2024)
by: Chokuwa, Sharon, et al.
Published: (2024)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026)
by: Kuckreja, Kartik, et al.
Published: (2026)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
AURORA: Adaptive Unified Representation for Robust Ultrasound Analysis
by: Khan, Ufaq, et al.
Published: (2026)
by: Khan, Ufaq, et al.
Published: (2026)
Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments
by: Ali, Muhammad, et al.
Published: (2025)
by: Ali, Muhammad, et al.
Published: (2025)
Unsupervised Deep Graph Matching Based on Cycle Consistency
by: Tourani, Siddharth, et al.
Published: (2023)
by: Tourani, Siddharth, et al.
Published: (2023)
Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models
by: Baliah, Sanoojan, et al.
Published: (2024)
by: Baliah, Sanoojan, et al.
Published: (2024)
Agentic AI for Remote Sensing: Technical Challenges and Research Directions
by: Munir, Muhammad Akhtar, et al.
Published: (2026)
by: Munir, Muhammad Akhtar, et al.
Published: (2026)
ReConText3D: Replay-based Continual Text-to-3D Generation
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
by: Danish, Muhammad Sohail, et al.
Published: (2025)
by: Danish, Muhammad Sohail, et al.
Published: (2025)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
Underwater Object Detection Enhancement via Channel Stabilization
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite for Video-LMMs
by: Khattak, Muhammad Uzair, et al.
Published: (2024)
by: Khattak, Muhammad Uzair, et al.
Published: (2024)
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
by: Ali, Eman, et al.
Published: (2025)
by: Ali, Eman, et al.
Published: (2025)
Judging from Support-set: A New Way to Utilize Few-Shot Segmentation for Segmentation Refinement Process
by: Moon, Seonghyeon, et al.
Published: (2024)
by: Moon, Seonghyeon, et al.
Published: (2024)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
by: Tourani, Siddharth, et al.
Published: (2024)
by: Tourani, Siddharth, et al.
Published: (2024)
ThinkGeo: Evaluating Tool-Augmented Agents for Remote Sensing Tasks
by: Shabbir, Akashah, et al.
Published: (2025)
by: Shabbir, Akashah, et al.
Published: (2025)
GlobalWasteData: A Large-Scale, Integrated Dataset for Robust Waste Classification and Environmental Monitoring
by: Ijaz, Misbah, et al.
Published: (2026)
by: Ijaz, Misbah, et al.
Published: (2026)
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
by: Khan, Ufaq, et al.
Published: (2025)
by: Khan, Ufaq, et al.
Published: (2025)
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
by: Nawaz, Umair, et al.
Published: (2025)
by: Nawaz, Umair, et al.
Published: (2025)
TLAC: Two-stage LMM Augmented CLIP for Zero-Shot Classification
by: Munir, Ans, et al.
Published: (2025)
by: Munir, Ans, et al.
Published: (2025)
Compositional Zero-Shot Learning: A Survey
by: Munir, Ans, et al.
Published: (2025)
by: Munir, Ans, et al.
Published: (2025)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
OpenEarthAgent: A Unified Framework for Tool-Augmented Geospatial Agents
by: Shabbir, Akashah, et al.
Published: (2026)
by: Shabbir, Akashah, et al.
Published: (2026)
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
Multi-modal Medical Image Fusion For Non-Small Cell Lung Cancer Classification
by: Hassan, Salma, et al.
Published: (2024)
by: Hassan, Salma, et al.
Published: (2024)
Similar Items
-
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024) -
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025) -
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024) -
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025) -
Towards Generalizing to Unseen Domains with Few Labels
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)