Autonomous AI Surveillance: Multimodal Deep Learning for Cognitive and Behavioral Monitoring
Fuente:
arXiv
Saved in:
| Main Authors: | Hamza, Ameer, But, Zuhaib Hussain, Arif, Umar, Samiya, Asad, M. Abdullah, Naeem, Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking
by: Nadeem, Numair, et al.
Published: (2025)
by: Nadeem, Numair, et al.
Published: (2025)
Resource-Efficient Medical Report Generation using Large Language Models
by: Abdullah, et al.
Published: (2024)
by: Abdullah, et al.
Published: (2024)
Advancements in Crop Analysis through Deep Learning and Explainable AI
by: Khan, Hamza
Published: (2025)
by: Khan, Hamza
Published: (2025)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
by: Asad, Muhammad Hamza, et al.
Published: (2023)
by: Asad, Muhammad Hamza, et al.
Published: (2023)
Segmenting Visuals With Querying Words: Language Anchors For Semi-Supervised Image Segmentation
by: Nadeem, Numair, et al.
Published: (2025)
by: Nadeem, Numair, et al.
Published: (2025)
A Comprehensive Survey on Deep Learning Solutions for 3D Flood Mapping
by: Jia, Wenfeng, et al.
Published: (2025)
by: Jia, Wenfeng, et al.
Published: (2025)
GLOFNet -- A Multimodal Dataset for GLOF Monitoring and Prediction
by: Fatima, Zuha, et al.
Published: (2025)
by: Fatima, Zuha, et al.
Published: (2025)
Multimodal and Multiview Deep Fusion for Autonomous Marine Navigation
by: Dagdilelis, Dimitrios, et al.
Published: (2025)
by: Dagdilelis, Dimitrios, et al.
Published: (2025)
Evaluating Deep Learning Models for African Wildlife Image Classification: From DenseNet to Vision Transformers
by: Aliyu, Lukman Jibril, et al.
Published: (2025)
by: Aliyu, Lukman Jibril, et al.
Published: (2025)
Streamlining Forest Wildfire Surveillance: AI-Enhanced UAVs Utilizing the FLAME Aerial Video Dataset for Lightweight and Efficient Monitoring
by: Zhao, Lemeng, et al.
Published: (2024)
by: Zhao, Lemeng, et al.
Published: (2024)
Ensemble Deep Learning and LLM-Assisted Reporting for Automated Skin Lesion Diagnosis
by: Khan, Sher, et al.
Published: (2025)
by: Khan, Sher, et al.
Published: (2025)
Video Forgery Detection for Surveillance Cameras: A Review
by: Tayfor, Noor B., et al.
Published: (2025)
by: Tayfor, Noor B., et al.
Published: (2025)
Dual-Encoder Transformer-Based Multimodal Learning for Ischemic Stroke Lesion Segmentation Using Diffusion MRI
by: Usman, Muhammad, et al.
Published: (2025)
by: Usman, Muhammad, et al.
Published: (2025)
Efficient Transformer for High Resolution Image Motion Deblurring
by: Akmaral, Amanturdieva, et al.
Published: (2025)
by: Akmaral, Amanturdieva, et al.
Published: (2025)
MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance
by: Zhao, Kaikai, et al.
Published: (2025)
by: Zhao, Kaikai, et al.
Published: (2025)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Anatomy-Guided Representation Learning Using a Transformer-Based Network for Thyroid Nodule Segmentation in Ultrasound Images
by: Farooq, Muhammad Umar, et al.
Published: (2025)
by: Farooq, Muhammad Umar, et al.
Published: (2025)
A Lightweight and Interpretable Deepfakes Detection Framework
by: Farooq, Muhammad Umar, et al.
Published: (2025)
by: Farooq, Muhammad Umar, et al.
Published: (2025)
Crowd Scene Analysis using Deep Learning Techniques
by: Asif, Muhammad Junaid
Published: (2025)
by: Asif, Muhammad Junaid
Published: (2025)
Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild
by: Yao, Shanle, et al.
Published: (2026)
by: Yao, Shanle, et al.
Published: (2026)
Crop Pest Classification Using Deep Learning Techniques: A Review
by: Ejaz, Muhammad Hassam, et al.
Published: (2025)
by: Ejaz, Muhammad Hassam, et al.
Published: (2025)
Hybrid CNN-ViT Framework for Motion-Blurred Scene Text Restoration
by: Rashid, Umar, et al.
Published: (2025)
by: Rashid, Umar, et al.
Published: (2025)
Rethinking RGB-D Fusion for Semantic Segmentation in Surgical Datasets
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
Adaptive Image Restoration for Video Surveillance: A Real-Time Approach
by: Amin, Muhammad Awais, et al.
Published: (2025)
by: Amin, Muhammad Awais, et al.
Published: (2025)
A Scalable and Generalized Deep Learning Framework for Anomaly Detection in Surveillance Videos
by: Jebur, Sabah Abdulazeez, et al.
Published: (2024)
by: Jebur, Sabah Abdulazeez, et al.
Published: (2024)
Physical Adversarial Attacks on AI Surveillance Systems:Detection, Tracking, and Visible--Infrared Evasion
by: DelaCruz, Miguel A., et al.
Published: (2026)
by: DelaCruz, Miguel A., et al.
Published: (2026)
DeepJIVE: Learning Joint and Individual Variation Explained from Multimodal Data Using Deep Learning
by: Drexler, Matthew, et al.
Published: (2025)
by: Drexler, Matthew, et al.
Published: (2025)
Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward
by: Islam, Muhammad, et al.
Published: (2025)
by: Islam, Muhammad, et al.
Published: (2025)
Using Deep Learning for Morphological Classification in Pigs with a Focus on Sanitary Monitoring
by: Bedin, Eduardo, et al.
Published: (2024)
by: Bedin, Eduardo, et al.
Published: (2024)
Enhancing AI Diagnostics: Autonomous Lesion Masking via Semi-Supervised Deep Learning
by: Wei, Ting-Ruen, et al.
Published: (2024)
by: Wei, Ting-Ruen, et al.
Published: (2024)
Research on the Application of Computer Vision Based on Deep Learning in Autonomous Driving Technology
by: Zhang, Jingyu, et al.
Published: (2024)
by: Zhang, Jingyu, et al.
Published: (2024)
Application of Multimodal Fusion Deep Learning Model in Disease Recognition
by: Liu, Xiaoyi, et al.
Published: (2024)
by: Liu, Xiaoyi, et al.
Published: (2024)
AI Meets Brain: Memory Systems from Cognitive Neuroscience to Autonomous Agents
by: Liang, Jiafeng, et al.
Published: (2025)
by: Liang, Jiafeng, et al.
Published: (2025)
YOLOv12: A Breakdown of the Key Architectural Features
by: Alif, Mujadded Al Rabbani, et al.
Published: (2025)
by: Alif, Mujadded Al Rabbani, et al.
Published: (2025)
AgriChat: A Multimodal Large Language Model for Agriculture Image Understanding
by: Boudiaf, Abderrahmene, et al.
Published: (2026)
by: Boudiaf, Abderrahmene, et al.
Published: (2026)
Automated Landfill Detection Using Deep Learning: A Comparative Study of Lightweight and Custom Architectures with the AerialWaste Dataset
by: Sharmily, Nowshin, et al.
Published: (2025)
by: Sharmily, Nowshin, et al.
Published: (2025)
Exploring Personalized Federated Learning Architectures for Violence Detection in Surveillance Videos
by: Kassir, Mohammad, et al.
Published: (2025)
by: Kassir, Mohammad, et al.
Published: (2025)
Addressing Image Authenticity When Cameras Use Generative AI
by: Masud, Umar, et al.
Published: (2026)
by: Masud, Umar, et al.
Published: (2026)
Evaluation of Safety Cognition Capability in Vision-Language Models for Autonomous Driving
by: Zhang, Enming, et al.
Published: (2025)
by: Zhang, Enming, et al.
Published: (2025)
Toward Cognitive Supersensing in Multimodal Large Language Model
by: Li, Boyi, et al.
Published: (2026)
by: Li, Boyi, et al.
Published: (2026)
Similar Items
-
MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking
by: Nadeem, Numair, et al.
Published: (2025) -
Resource-Efficient Medical Report Generation using Large Language Models
by: Abdullah, et al.
Published: (2024) -
Advancements in Crop Analysis through Deep Learning and Explainable AI
by: Khan, Hamza
Published: (2025) -
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
by: Asad, Muhammad Hamza, et al.
Published: (2023) -
Segmenting Visuals With Querying Words: Language Anchors For Semi-Supervised Image Segmentation
by: Nadeem, Numair, et al.
Published: (2025)