An Analysis Focused on Womens Safety: Can VAD Models Be Enhanced by a Multi-modal Dataset?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sangeeta, Prajwal, Maddikuntla Sai, Dogra, Debi Prosad, Thakare, Kamalakar Vijay, Jung, Hyungjoo, Kim, Ig-Jae, Choi, Heeseung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
"Human gaze behavior in small conversational groups"
von: NAYAK, KAMAKSHYA PRASAD, et al.
Veröffentlicht: (2024)
von: NAYAK, KAMAKSHYA PRASAD, et al.
Veröffentlicht: (2024)
Feature Reweighting for EEG-based Motor Imagery Classification
von: Lotey, Taveena, et al.
Veröffentlicht: (2023)
von: Lotey, Taveena, et al.
Veröffentlicht: (2023)
K-FACE: A Large-Scale KIST Face Database in Consideration with Unconstrained Environments
von: Choi, Yeji, et al.
Veröffentlicht: (2021)
von: Choi, Yeji, et al.
Veröffentlicht: (2021)
Dual Prototype Attention for Unsupervised Video Object Segmentation
von: Cho, Suhwan, et al.
Veröffentlicht: (2022)
von: Cho, Suhwan, et al.
Veröffentlicht: (2022)
Effective SAM Combination for Open-Vocabulary Semantic Segmentation
von: Lee, Minhyeok, et al.
Veröffentlicht: (2024)
von: Lee, Minhyeok, et al.
Veröffentlicht: (2024)
Flashback: Memory-Driven Zero-shot, Real-time Video Anomaly Detection
von: Lee, Hyogun, et al.
Veröffentlicht: (2025)
von: Lee, Hyogun, et al.
Veröffentlicht: (2025)
PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation
von: Yang, Yu, et al.
Veröffentlicht: (2026)
von: Yang, Yu, et al.
Veröffentlicht: (2026)
DUAL-VAD: Dual Benchmarks and Anomaly-Focused Sampling for Video Anomaly Detection
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
Navigating Label Ambiguity for Facial Expression Recognition in the Wild
von: Lee, JunGyu, et al.
Veröffentlicht: (2025)
von: Lee, JunGyu, et al.
Veröffentlicht: (2025)
PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
von: Lee, Seokyeong, et al.
Veröffentlicht: (2025)
von: Lee, Seokyeong, et al.
Veröffentlicht: (2025)
A.Glimpse : R.K.Narayan's Novel 'Swami and Friend's
von: Nitin S. Thakare
Veröffentlicht: (2025)
von: Nitin S. Thakare
Veröffentlicht: (2025)
MAIR++: Improving Multi-view Attention Inverse Rendering with Implicit Lighting Representation
von: Choi, JunYong, et al.
Veröffentlicht: (2024)
von: Choi, JunYong, et al.
Veröffentlicht: (2024)
Safety Alignment Can Be Not Superficial With Explicit Safety Signals
von: Li, Jianwei, et al.
Veröffentlicht: (2025)
von: Li, Jianwei, et al.
Veröffentlicht: (2025)
Channel-wise Noise Scheduled Diffusion for Inverse Rendering in Indoor Scenes
von: Choi, JunYong, et al.
Veröffentlicht: (2025)
von: Choi, JunYong, et al.
Veröffentlicht: (2025)
IG-FIQA: Improving Face Image Quality Assessment through Intra-class Variance Guidance robust to Inaccurate Pseudo-Labels
von: Kim, Minsoo, et al.
Veröffentlicht: (2024)
von: Kim, Minsoo, et al.
Veröffentlicht: (2024)
Late Cretaceous palm stem Palmoxylon lametaei sp. nov. from Bhisi Village, Maharashtra, India
von: Debi Dutta
Veröffentlicht: (2011)
von: Debi Dutta
Veröffentlicht: (2011)
Development of IPM package with safe pesticide residue: 1. Cabbage
von: Debi Sharma
Veröffentlicht: (2006)
von: Debi Sharma
Veröffentlicht: (2006)
Comparative study of pesticide residue pattern in vegetables grown using IPM and non-IPM practices
von: Debi Sharma
Veröffentlicht: (2009)
von: Debi Sharma
Veröffentlicht: (2009)
Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
von: Shin, Chaehun, et al.
Veröffentlicht: (2024)
Evidence-Focused Fact Summarization for Knowledge-Augmented Zero-Shot Question Answering
von: Ko, Sungho, et al.
Veröffentlicht: (2024)
von: Ko, Sungho, et al.
Veröffentlicht: (2024)
Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
VIGFace: Virtual Identity Generation for Privacy-Free Face Recognition
von: Kim, Minsoo, et al.
Veröffentlicht: (2024)
von: Kim, Minsoo, et al.
Veröffentlicht: (2024)
Social Media Clones: Exploring the Impact of Social Delegation with AI Clones through a Design Workbook Study
von: Liu, Jackie, et al.
Veröffentlicht: (2025)
von: Liu, Jackie, et al.
Veröffentlicht: (2025)
The AI Genie Phenomenon and Three Types of AI Chatbot Addiction: Escapist Roleplays, Pseudosocial Companions, and Epistemic Rabbit Holes
von: Shen, M. Karen, et al.
Veröffentlicht: (2026)
von: Shen, M. Karen, et al.
Veröffentlicht: (2026)
Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support
von: Shirvani, Mehrnoosh Sadat, et al.
Veröffentlicht: (2025)
von: Shirvani, Mehrnoosh Sadat, et al.
Veröffentlicht: (2025)
Iti-Validator: A Guardrail Framework for Validating and Correcting LLM-Generated Itineraries
von: Gadbail, Shravan, et al.
Veröffentlicht: (2025)
von: Gadbail, Shravan, et al.
Veröffentlicht: (2025)
Who Sits Where? Automated Detection of Director Interlocks in Indian Companies
von: Sancheti, Prateek, et al.
Veröffentlicht: (2026)
von: Sancheti, Prateek, et al.
Veröffentlicht: (2026)
ViSAGe: Video-to-Spatial Audio Generation
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
Can Paracorporeal Bi‐VAD Improve the Outcome of a Patient Candidate for Heart Transplant?
von: Giuseppe Fischetti, et al.
Veröffentlicht: (2025)
von: Giuseppe Fischetti, et al.
Veröffentlicht: (2025)
VoiceGuider: Enhancing Out-of-Domain Performance in Parameter-Efficient Speaker-Adaptive Text-to-Speech via Autoguidance
von: Yeom, Jiheum, et al.
Veröffentlicht: (2024)
von: Yeom, Jiheum, et al.
Veröffentlicht: (2024)
Diagnostic implication of thyroid spherules for cytological diagnosis of thyroid nodules
von: Heeseung Sohn, et al.
Veröffentlicht: (2024)
von: Heeseung Sohn, et al.
Veröffentlicht: (2024)
V-NAW: Video-based Noise-aware Adaptive Weighting for Facial Expression Recognition
von: Lee, JunGyu, et al.
Veröffentlicht: (2025)
von: Lee, JunGyu, et al.
Veröffentlicht: (2025)
Manifold Decoders: A Framework for Generative Modeling from Nonlinear Embeddings
von: Thakare, Riddhish, et al.
Veröffentlicht: (2025)
von: Thakare, Riddhish, et al.
Veröffentlicht: (2025)
Hydrodynamic cavitation for efficient removal of ammoniacal nitrogen: Optimization of process parameters and geometric configurations
von: Neha Thakare, et al.
Veröffentlicht: (2026)
von: Neha Thakare, et al.
Veröffentlicht: (2026)
Antagonistic Bowden-Cable Actuation of a Lightweight Robotic Hand: Toward Dexterous Manipulation for Payload Constrained Humanoids
von: Min, Sungjae, et al.
Veröffentlicht: (2025)
von: Min, Sungjae, et al.
Veröffentlicht: (2025)
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
von: Moon, Junwon, et al.
Veröffentlicht: (2026)
von: Moon, Junwon, et al.
Veröffentlicht: (2026)
Maintaining the Level of a Payload carried by Multi-Robot System on Irregular Surface
von: Yadav, Rishabh Dev, et al.
Veröffentlicht: (2025)
von: Yadav, Rishabh Dev, et al.
Veröffentlicht: (2025)
Peak ground motions generated by earthquakes (magnitude 2.5+, September 2010 - June 2023, southern California) that are larger or smaller than theoretically expected
von: Kilb, Debi, et al.
Veröffentlicht: (2025)
von: Kilb, Debi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
"Human gaze behavior in small conversational groups"
von: NAYAK, KAMAKSHYA PRASAD, et al.
Veröffentlicht: (2024) -
Feature Reweighting for EEG-based Motor Imagery Classification
von: Lotey, Taveena, et al.
Veröffentlicht: (2023) -
K-FACE: A Large-Scale KIST Face Database in Consideration with Unconstrained Environments
von: Choi, Yeji, et al.
Veröffentlicht: (2021) -
Dual Prototype Attention for Unsupervised Video Object Segmentation
von: Cho, Suhwan, et al.
Veröffentlicht: (2022) -
Effective SAM Combination for Open-Vocabulary Semantic Segmentation
von: Lee, Minhyeok, et al.
Veröffentlicht: (2024)