VARS: Vision-based Assessment of Risk in Security Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Pranav, Gohil, Pratham, S, Sridhar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViDAS: Vision-based Danger Assessment and Scoring
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
Time Series, Vision, and Language: Exploring the Limits of Alignment in Contrastive Representation Spaces
by: Yashwante, Pratham, et al.
Published: (2026)
by: Yashwante, Pratham, et al.
Published: (2026)
How Do Inpainting Artifacts Propagate to Language?
by: Yashwante, Pratham, et al.
Published: (2026)
by: Yashwante, Pratham, et al.
Published: (2026)
An Explainable Two Stage Deep Learning Framework for Pericoronitis Assessment in Panoramic Radiographs Using YOLOv8 and ResNet-50
by: George, Ajo Babu, et al.
Published: (2026)
by: George, Ajo Babu, et al.
Published: (2026)
Gen-LangSplat: Generalized Language Gaussian Splatting with Pre-Trained Feature Compression
by: Saxena, Pranav
Published: (2025)
by: Saxena, Pranav
Published: (2025)
Evaluating Contextual Intelligence in Recyclability: A Comprehensive Study of Image-Based Reasoning Systems
by: Park, Eliot, et al.
Published: (2025)
by: Park, Eliot, et al.
Published: (2025)
WebSight: A Vision-First Architecture for Robust Web Agents
by: Bhathal, Tanvir, et al.
Published: (2025)
by: Bhathal, Tanvir, et al.
Published: (2025)
Symbolic Rule Extraction from Attention-Guided Sparse Representations in Vision Transformers
by: Padalkar, Parth, et al.
Published: (2025)
by: Padalkar, Parth, et al.
Published: (2025)
Give me a hint: Can LLMs take a hint to solve math problems?
by: Agrawal, Vansh, et al.
Published: (2024)
by: Agrawal, Vansh, et al.
Published: (2024)
Exploring Intrinsic Properties of Medical Images for Self-Supervised Binary Semantic Segmentation
by: Singh, Pranav, et al.
Published: (2024)
by: Singh, Pranav, et al.
Published: (2024)
A Survey of Adversarial Defenses in Vision-based Systems: Categorization, Methods and Challenges
by: Chattopadhyay, Nandish, et al.
Published: (2025)
by: Chattopadhyay, Nandish, et al.
Published: (2025)
Rice-VL: Evaluating Vision-Language Models for Cultural Understanding Across ASEAN Countries
by: Pranav, Tushar, et al.
Published: (2025)
by: Pranav, Tushar, et al.
Published: (2025)
Beta Distribution Learning for Reliable Roadway Crash Risk Assessment
by: Elallaf, Ahmad, et al.
Published: (2025)
by: Elallaf, Ahmad, et al.
Published: (2025)
TRACE: Transformer-based Risk Assessment for Clinical Evaluation
by: Christopoulos, Dionysis, et al.
Published: (2024)
by: Christopoulos, Dionysis, et al.
Published: (2024)
3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language Models
by: Sambara, Sraavya, et al.
Published: (2025)
by: Sambara, Sraavya, et al.
Published: (2025)
Two Steps Are All You Need: Efficient 3D Point Cloud Anomaly Detection with Consistency Models
by: A, Pranav, et al.
Published: (2026)
by: A, Pranav, et al.
Published: (2026)
LG-Traj: LLM Guided Pedestrian Trajectory Prediction
by: Chib, Pranav Singh, et al.
Published: (2024)
by: Chib, Pranav Singh, et al.
Published: (2024)
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
by: Chen, Zhihao, et al.
Published: (2023)
by: Chen, Zhihao, et al.
Published: (2023)
Process Integrated Computer Vision for Real-Time Failure Prediction in Steel Rolling Mill
by: Kurrey, Vaibhav, et al.
Published: (2025)
by: Kurrey, Vaibhav, et al.
Published: (2025)
Wildfire Detection Using Vision Transformer with the Wildfire Dataset
by: Vuppari, Gowtham Raj, et al.
Published: (2025)
by: Vuppari, Gowtham Raj, et al.
Published: (2025)
Self-Aug: Query and Entropy Adaptive Decoding for Large Vision-Language Models
by: Im, Eun Woo, et al.
Published: (2025)
by: Im, Eun Woo, et al.
Published: (2025)
Security Tensors as a Cross-Modal Bridge: Extending Text-Aligned Safety to Vision in LVLM
by: Li, Shen, et al.
Published: (2025)
by: Li, Shen, et al.
Published: (2025)
VITAL: Vision-Encoder-centered Pre-training for LMMs in Visual Quality Assessment
by: Jia, Ziheng, et al.
Published: (2025)
by: Jia, Ziheng, et al.
Published: (2025)
Cerebra: A Multidisciplinary AI Board for Multimodal Dementia Characterization and Risk Assessment
by: Liu, Sheng, et al.
Published: (2026)
by: Liu, Sheng, et al.
Published: (2026)
NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning
by: Rawal, Ishaan, et al.
Published: (2026)
by: Rawal, Ishaan, et al.
Published: (2026)
On Inherent Adversarial Robustness of Active Vision Systems
by: Mukherjee, Amitangshu, et al.
Published: (2024)
by: Mukherjee, Amitangshu, et al.
Published: (2024)
A Deep Multi-Modal Method for Patient Wound Healing Assessment
by: Oota, Subba Reddy, et al.
Published: (2026)
by: Oota, Subba Reddy, et al.
Published: (2026)
Synthetic Thermal and RGB Videos for Automatic Pain Assessment utilizing a Vision-MLP Architecture
by: Gkikas, Stefanos, et al.
Published: (2024)
by: Gkikas, Stefanos, et al.
Published: (2024)
A Computer Vision-Based Quality Assessment Technique for the automatic control of consumables for analytical laboratories
by: Zribi, Meriam, et al.
Published: (2024)
by: Zribi, Meriam, et al.
Published: (2024)
Solar PV Installation Potential Assessment on Building Facades Based on Vision and Language Foundation Models
by: Liu, Ruyu, et al.
Published: (2025)
by: Liu, Ruyu, et al.
Published: (2025)
Is it safe to cross? Interpretable Risk Assessment with GPT-4V for Safety-Aware Street Crossing
by: Hwang, Hochul, et al.
Published: (2024)
by: Hwang, Hochul, et al.
Published: (2024)
Real-Time Posture Monitoring and Risk Assessment for Manual Lifting Tasks Using MediaPipe and LSTM
by: Bagga, Ereena, et al.
Published: (2024)
by: Bagga, Ereena, et al.
Published: (2024)
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026)
by: Naimi, Safwen, et al.
Published: (2026)
Towards AI-Powered Video Assistant Referee System (VARS) for Association Football
by: Held, Jan, et al.
Published: (2024)
by: Held, Jan, et al.
Published: (2024)
ErgoChat: a Visual Query System for the Ergonomic Risk Assessment of Construction Workers
by: Fan, Chao, et al.
Published: (2024)
by: Fan, Chao, et al.
Published: (2024)
Multimodal Carotid Risk Stratification with Large Vision-Language Models: Benchmarking, Fine-Tuning, and Clinical Insights
by: Tsolissou, Daphne, et al.
Published: (2025)
by: Tsolissou, Daphne, et al.
Published: (2025)
REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction
by: Leem, Seowung, et al.
Published: (2026)
by: Leem, Seowung, et al.
Published: (2026)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
by: Kundu, Souvik, et al.
Published: (2025)
by: Kundu, Souvik, et al.
Published: (2025)
Aesthetic Assessment of Chinese Handwritings Based on Vision Language Models
by: Zheng, Chen, et al.
Published: (2026)
by: Zheng, Chen, et al.
Published: (2026)
Similar Items
-
ViDAS: Vision-based Danger Assessment and Scoring
by: Gupta, Pranav, et al.
Published: (2024) -
Time Series, Vision, and Language: Exploring the Limits of Alignment in Contrastive Representation Spaces
by: Yashwante, Pratham, et al.
Published: (2026) -
How Do Inpainting Artifacts Propagate to Language?
by: Yashwante, Pratham, et al.
Published: (2026) -
An Explainable Two Stage Deep Learning Framework for Pericoronitis Assessment in Panoramic Radiographs Using YOLOv8 and ResNet-50
by: George, Ajo Babu, et al.
Published: (2026) -
Gen-LangSplat: Generalized Language Gaussian Splatting with Pre-Trained Feature Compression
by: Saxena, Pranav
Published: (2025)