Saved in:
| Main Authors: | Islam, Khawar, Zaheer, Muhammad Zaigham, Mahmood, Arif, Nandakumar, Karthik |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.14881 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenMix: Effective Data Augmentation with Generative Diffusion Model Image Editing
by: Islam, Khawar, et al.
Published: (2024)
by: Islam, Khawar, et al.
Published: (2024)
Face Pyramid Vision Transformer
by: Islam, Khawar, et al.
Published: (2022)
by: Islam, Khawar, et al.
Published: (2022)
Context-guided Responsible Data Augmentation with Diffusion Models
by: Islam, Khawar, et al.
Published: (2025)
by: Islam, Khawar, et al.
Published: (2025)
Clustering Aided Weakly Supervised Training to Detect Anomalous Events in Surveillance Videos
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
Collaborative Learning of Anomalies with Privacy (CLAP) for Unsupervised Video Anomaly Detection: A New Baseline
by: Al-lahham, Anas, et al.
Published: (2024)
by: Al-lahham, Anas, et al.
Published: (2024)
Stabilizing Adversarially Learned One-Class Novelty Detection Using Pseudo Anomalies
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
by: Demidov, Dmitry, et al.
Published: (2025)
by: Demidov, Dmitry, et al.
Published: (2025)
Constricting Normal Latent Space for Anomaly Detection with Normal-only Training Data
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
Deep Learning for Video-based Person Re-Identification: A Survey
by: Islam, Khawar
Published: (2023)
by: Islam, Khawar
Published: (2023)
Intra-finger Variability of Diffusion-based Latent Fingerprint Generation
by: Hussein, Noor, et al.
Published: (2026)
by: Hussein, Noor, et al.
Published: (2026)
Exploiting Autoencoder's Weakness to Generate Pseudo Anomalies
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
Vocabulary-free Fine-grained Visual Recognition via Enriched Contextually Grounded Vision-Language Model
by: Demidov, Dmitry, et al.
Published: (2025)
by: Demidov, Dmitry, et al.
Published: (2025)
SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training
by: Kumar, Komal, et al.
Published: (2026)
by: Kumar, Komal, et al.
Published: (2026)
Linking Faces and Voices Across Languages: Insights from the FAME 2026 Challenge
by: Moscati, Marta, et al.
Published: (2025)
by: Moscati, Marta, et al.
Published: (2025)
OpenEarthAgent: A Unified Framework for Tool-Augmented Geospatial Agents
by: Shabbir, Akashah, et al.
Published: (2026)
by: Shabbir, Akashah, et al.
Published: (2026)
STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
by: Srivatsan, Koushik, et al.
Published: (2024)
by: Srivatsan, Koushik, et al.
Published: (2024)
SB-BEVFusion: Enhancing the Robustness against Sensor Malfunction and Corruptions
by: Essl, Markus, et al.
Published: (2026)
by: Essl, Markus, et al.
Published: (2026)
Face-Voice Association with Inductive Bias for Maximum Class Separation
by: Moscati, Marta, et al.
Published: (2026)
by: Moscati, Marta, et al.
Published: (2026)
Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models
by: Imam, Raza, et al.
Published: (2024)
by: Imam, Raza, et al.
Published: (2024)
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
by: Nawaz, Umair, et al.
Published: (2025)
by: Nawaz, Umair, et al.
Published: (2025)
NoiseCutMix: A Novel Data Augmentation Approach by Mixing Estimated Noise in Diffusion Models
by: Takezaki, Shumpei, et al.
Published: (2025)
by: Takezaki, Shumpei, et al.
Published: (2025)
Implicit to Explicit Entropy Regularization: Benchmarking ViT Fine-tuning under Noisy Labels
by: Marrium, Maria, et al.
Published: (2024)
by: Marrium, Maria, et al.
Published: (2024)
FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing
by: Alam, Mohammed Talha, et al.
Published: (2025)
by: Alam, Mohammed Talha, et al.
Published: (2025)
SGD-Mix: Enhancing Domain-Specific Image Classification with Label-Preserving Data Augmentation
by: Dong, Yixuan, et al.
Published: (2025)
by: Dong, Yixuan, et al.
Published: (2025)
How Good is my Histopathology Vision-Language Foundation Model? A Holistic Benchmark
by: Majzoub, Roba Al, et al.
Published: (2025)
by: Majzoub, Roba Al, et al.
Published: (2025)
Enhancing 3D Human Pose Estimation Amidst Severe Occlusion with Dual Transformer Fusion
by: Ghafoor, Mehwish, et al.
Published: (2024)
by: Ghafoor, Mehwish, et al.
Published: (2024)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
by: Basu, Abhishek, et al.
Published: (2025)
by: Basu, Abhishek, et al.
Published: (2025)
ChildDiffusion: Unlocking the Potential of Generative AI and Controllable Augmentations for Child Facial Data using Stable Diffusion and Large Language Models
by: Farooq, Muhammad Ali, et al.
Published: (2024)
by: Farooq, Muhammad Ali, et al.
Published: (2024)
EdgeDAM: Real-time Object Tracking for Mobile Devices
by: Raza, Syed Muhammad, et al.
Published: (2026)
by: Raza, Syed Muhammad, et al.
Published: (2026)
Label-Efficient Data Augmentation with Video Diffusion Models for Guidewire Segmentation in Cardiac Fluoroscopy
by: Pan, Shaoyan, et al.
Published: (2024)
by: Pan, Shaoyan, et al.
Published: (2024)
Diffusion-Based Data Augmentation for Medical Image Segmentation
by: Nazir, Maham, et al.
Published: (2025)
by: Nazir, Maham, et al.
Published: (2025)
APPLE: Attribute-Preserving Pseudo-Labeling for Diffusion-Based Face Swapping
by: Kang, Jiwon, et al.
Published: (2026)
by: Kang, Jiwon, et al.
Published: (2026)
IPMix: Label-Preserving Data Augmentation Method for Training Robust Classifiers
by: Huang, Zhenglin, et al.
Published: (2023)
by: Huang, Zhenglin, et al.
Published: (2023)
Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan
by: Moscati, Marta, et al.
Published: (2025)
by: Moscati, Marta, et al.
Published: (2025)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
by: Tourani, Siddharth, et al.
Published: (2024)
by: Tourani, Siddharth, et al.
Published: (2024)
Similar Items
-
GenMix: Effective Data Augmentation with Generative Diffusion Model Image Editing
by: Islam, Khawar, et al.
Published: (2024) -
Face Pyramid Vision Transformer
by: Islam, Khawar, et al.
Published: (2022) -
Context-guided Responsible Data Augmentation with Diffusion Models
by: Islam, Khawar, et al.
Published: (2025) -
Clustering Aided Weakly Supervised Training to Detect Anomalous Events in Surveillance Videos
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022) -
Collaborative Learning of Anomalies with Privacy (CLAP) for Unsupervised Video Anomaly Detection: A New Baseline
by: Al-lahham, Anas, et al.
Published: (2024)