Guardado en:
| Autores principales: | Pandat, Ami, Rajasekhar, Punna, Aravamuthan, G., Vinod, Gopika, Shukla, Rohit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2512.17784 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
YolovN-CBi: A Lightweight and Efficient Architecture for Real-Time Detection of Small UAVs
por: Pandat, Ami, et al.
Publicado: (2025)
por: Pandat, Ami, et al.
Publicado: (2025)
SimD3: A Synthetic drone Dataset with Payload and Bird Distractor Modeling for Robust Detection
por: Pandat, Ami, et al.
Publicado: (2026)
por: Pandat, Ami, et al.
Publicado: (2026)
SecurePose: Automated Face Blurring and Human Movement Kinematics Extraction from Videos Recorded in Clinical Settings
por: Bajpai, Rishabh, et al.
Publicado: (2024)
por: Bajpai, Rishabh, et al.
Publicado: (2024)
Inconsistency-Aware Cross-Attention for Audio-Visual Fusion in Dimensional Emotion Recognition
por: Rajasekhar, G, et al.
Publicado: (2024)
por: Rajasekhar, G, et al.
Publicado: (2024)
Hybrid approach for face boundary marking and recognition in dense environments using deep‐learning techniques
por: Jayasree Marimuthu, et al.
Publicado: (2025)
por: Jayasree Marimuthu, et al.
Publicado: (2025)
Application of 2D Homography for High Resolution Traffic Data Collection using CCTV Cameras
por: Zhang, Linlin, et al.
Publicado: (2024)
por: Zhang, Linlin, et al.
Publicado: (2024)
SPOT!: Map-Guided LLM Agent for Unsupervised Multi-CCTV Dynamic Object Tracking
por: Roh, Yujin, et al.
Publicado: (2025)
por: Roh, Yujin, et al.
Publicado: (2025)
Table tennis ball spin estimation with an event camera
por: Gossard, Thomas, et al.
Publicado: (2024)
por: Gossard, Thomas, et al.
Publicado: (2024)
Recovering Origin Destination Flows from Bus CCTV: Early Results from Nairobi and Kigali
por: Kyatha, Nthenya, et al.
Publicado: (2025)
por: Kyatha, Nthenya, et al.
Publicado: (2025)
SSAVSV: Towards Unified Model for Self-Supervised Audio-Visual Speaker Verification
por: Rajasekhar, Gnana Praveen, et al.
Publicado: (2025)
por: Rajasekhar, Gnana Praveen, et al.
Publicado: (2025)
BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities
por: Sharma, Akash, et al.
Publicado: (2026)
por: Sharma, Akash, et al.
Publicado: (2026)
Automatic camera orientation estimation for a partially calibrated camera above a plane with a line at known planar distance
por: Dinya, Gergely, et al.
Publicado: (2025)
por: Dinya, Gergely, et al.
Publicado: (2025)
Hybrid Long and Short Range Flows for Point Cloud Filtering
por: Edirimuni, Dasith de Silva, et al.
Publicado: (2025)
por: Edirimuni, Dasith de Silva, et al.
Publicado: (2025)
Estimation of Psychosocial Work Environment Exposures Through Video Object Detection. Proof of Concept Using CCTV Footage
por: Hansen, Claus D., et al.
Publicado: (2024)
por: Hansen, Claus D., et al.
Publicado: (2024)
Active headrest combined with a depth camera-based ear-positioning system
por: Liu, Yuteng, et al.
Publicado: (2023)
por: Liu, Yuteng, et al.
Publicado: (2023)
Deep learning-based ecological analysis of camera trap images is impacted by training data quality and quantity
por: Bevan, Peggy A., et al.
Publicado: (2024)
por: Bevan, Peggy A., et al.
Publicado: (2024)
Enhancing Single-Image Facial Demorphing using Multimodal Large Language Models
por: Shukla, Nitish, et al.
Publicado: (2026)
por: Shukla, Nitish, et al.
Publicado: (2026)
Fast Wrong-way Cycling Detection in CCTV Videos: Sparse Sampling is All You Need
por: Xu, Jing, et al.
Publicado: (2024)
por: Xu, Jing, et al.
Publicado: (2024)
A Deep Learning-Based CCTV System for Automatic Smoking Detection in Fire Exit Zones
por: Sadat, Sami, et al.
Publicado: (2025)
por: Sadat, Sami, et al.
Publicado: (2025)
BrainMT: A Hybrid Mamba-Transformer Architecture for Modeling Long-Range Dependencies in Functional MRI Data
por: Kannan, Arunkumar, et al.
Publicado: (2025)
por: Kannan, Arunkumar, et al.
Publicado: (2025)
iiANET: Inception Inspired Attention Hybrid Network for efficient Long-Range Dependency
por: Yunusa, Haruna, et al.
Publicado: (2024)
por: Yunusa, Haruna, et al.
Publicado: (2024)
Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models
por: Baid, Ami, et al.
Publicado: (2026)
por: Baid, Ami, et al.
Publicado: (2026)
Intelligent CCTV for Urban Design: AI-Based Analysis of Soft Infrastructure at Intersections
por: Katariya, Vinit, et al.
Publicado: (2026)
por: Katariya, Vinit, et al.
Publicado: (2026)
XAI-based gait analysis of patients walking with Knee-Ankle-Foot orthosis using video cameras
por: Mishra, Arnav, et al.
Publicado: (2024)
por: Mishra, Arnav, et al.
Publicado: (2024)
A model-agnostic active learning approach for animal detection from camera traps
por: Nguyen, Thi Thu Thuy, et al.
Publicado: (2025)
por: Nguyen, Thi Thu Thuy, et al.
Publicado: (2025)
HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
por: Ben-Ami, Dan, et al.
Publicado: (2026)
por: Ben-Ami, Dan, et al.
Publicado: (2026)
On depth prediction for autonomous driving using self-supervised learning
por: Boulahbal, Houssem
Publicado: (2024)
por: Boulahbal, Houssem
Publicado: (2024)
Is Long Range Sequential Modeling Necessary For Colorectal Tumor Segmentation?
por: Srivastava, Abhishek, et al.
Publicado: (2025)
por: Srivastava, Abhishek, et al.
Publicado: (2025)
Multi-camera calibration with pattern rigs, including for non-overlapping cameras: CALICO
por: Tabb, Amy, et al.
Publicado: (2019)
por: Tabb, Amy, et al.
Publicado: (2019)
Digging into contrastive learning for robust depth estimation with diffusion models
por: Wang, Jiyuan, et al.
Publicado: (2024)
por: Wang, Jiyuan, et al.
Publicado: (2024)
When the Small-Loss Trick is Not Enough: Multi-Label Image Classification with Noisy Labels Applied to CCTV Sewer Inspections
por: Chelouche, Keryan, et al.
Publicado: (2024)
por: Chelouche, Keryan, et al.
Publicado: (2024)
Traversing Distortion-Perception Tradeoff using a Single Score-Based Generative Model
por: Wang, Yuhan, et al.
Publicado: (2025)
por: Wang, Yuhan, et al.
Publicado: (2025)
Monocular absolute depth estimation from endoscopy via domain-invariant feature learning and latent consistency
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
Efficient Few-shot Identity Preserving Attribute Editing for 3D-aware Deep Generative Models
por: Vinod, Vishal
Publicado: (2025)
por: Vinod, Vishal
Publicado: (2025)
Uncertainty and Energy based Loss Guided Semi-Supervised Semantic Segmentation
por: Thakur, Rini Smita, et al.
Publicado: (2025)
por: Thakur, Rini Smita, et al.
Publicado: (2025)
HCR-Net: A deep learning based script independent handwritten character recognition network
por: Chauhan, Vinod Kumar, et al.
Publicado: (2021)
por: Chauhan, Vinod Kumar, et al.
Publicado: (2021)
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
por: Lingenberg, Tobias, et al.
Publicado: (2024)
por: Lingenberg, Tobias, et al.
Publicado: (2024)
A Deep Ordinal Distortion Estimation Approach for Distortion Rectification
por: Liao, Kang, et al.
Publicado: (2020)
por: Liao, Kang, et al.
Publicado: (2020)
Multi-perspective monitoring of wildlife and human activities from camera traps and drones with deep learning models
por: Chen, Hao, et al.
Publicado: (2025)
por: Chen, Hao, et al.
Publicado: (2025)
Experimental Demonstration of Event-based Optical Camera Communication in Long-Range Outdoor Environment
por: Sumino, Miu, et al.
Publicado: (2025)
por: Sumino, Miu, et al.
Publicado: (2025)
Ejemplares similares
-
YolovN-CBi: A Lightweight and Efficient Architecture for Real-Time Detection of Small UAVs
por: Pandat, Ami, et al.
Publicado: (2025) -
SimD3: A Synthetic drone Dataset with Payload and Bird Distractor Modeling for Robust Detection
por: Pandat, Ami, et al.
Publicado: (2026) -
SecurePose: Automated Face Blurring and Human Movement Kinematics Extraction from Videos Recorded in Clinical Settings
por: Bajpai, Rishabh, et al.
Publicado: (2024) -
Inconsistency-Aware Cross-Attention for Audio-Visual Fusion in Dimensional Emotion Recognition
por: Rajasekhar, G, et al.
Publicado: (2024) -
Hybrid approach for face boundary marking and recognition in dense environments using deep‐learning techniques
por: Jayasree Marimuthu, et al.
Publicado: (2025)