Detecting Content Rating Violations in Android Applications: A Vision-Language Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Denipitiyage, D., Silva, B., Seneviratne, S., Seneviratne, A., Chawla, S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs
von: Denipitiyage, Dishanika, et al.
Veröffentlicht: (2026)
von: Denipitiyage, Dishanika, et al.
Veröffentlicht: (2026)
Long-Tail Learning with Rebalanced Contrastive Loss
von: De Alvis, Charika, et al.
Veröffentlicht: (2023)
von: De Alvis, Charika, et al.
Veröffentlicht: (2023)
Detecting and Characterising Mobile App Metamorphosis in Google Play Store
von: Denipitiyage, D., et al.
Veröffentlicht: (2024)
von: Denipitiyage, D., et al.
Veröffentlicht: (2024)
Unveiling the Visual Counting Bottleneck in Vision-Language Models
von: Pang, Xingzhou, et al.
Veröffentlicht: (2026)
von: Pang, Xingzhou, et al.
Veröffentlicht: (2026)
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages
von: Farina, Matteo, et al.
Veröffentlicht: (2025)
von: Farina, Matteo, et al.
Veröffentlicht: (2025)
Diversity Matters: Revisiting Test-Time Compute in Vision-Language Models
von: Tong, Yijie, et al.
Veröffentlicht: (2026)
von: Tong, Yijie, et al.
Veröffentlicht: (2026)
Dual Memory Networks: A Versatile Adaptation Approach for Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
Competitive Learning for Achieving Content-specific Filters in Video Coding for Machines
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
Language Models as Black-Box Optimizers for Vision-Language Models
von: Liu, Shihong, et al.
Veröffentlicht: (2023)
von: Liu, Shihong, et al.
Veröffentlicht: (2023)
MCE: Towards a General Framework for Handling Missing Modalities under Imbalanced Missing Rates
von: Zhao, Binyu, et al.
Veröffentlicht: (2025)
von: Zhao, Binyu, et al.
Veröffentlicht: (2025)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
Latent Space Probing for Adult Content Detection in Video Generative Models
von: Khatri, Alizishaan, et al.
Veröffentlicht: (2026)
von: Khatri, Alizishaan, et al.
Veröffentlicht: (2026)
Detection of Cyberbullying in GIF using AI
von: Dave, Pal, et al.
Veröffentlicht: (2025)
von: Dave, Pal, et al.
Veröffentlicht: (2025)
Regularized Contrastive Partial Multi-view Outlier Detection
von: Wang, Yijia, et al.
Veröffentlicht: (2024)
von: Wang, Yijia, et al.
Veröffentlicht: (2024)
Relating CNN-Transformer Fusion Network for Change Detection
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
BRep Boundary and Junction Detection for CAD Reverse Engineering
von: Ali, Sk Aziz, et al.
Veröffentlicht: (2024)
von: Ali, Sk Aziz, et al.
Veröffentlicht: (2024)
EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE
von: Chen, Junyi, et al.
Veröffentlicht: (2023)
von: Chen, Junyi, et al.
Veröffentlicht: (2023)
Adversarially Robust Deepfake Detection via Adversarial Feature Similarity Learning
von: Khan, Sarwar
Veröffentlicht: (2024)
von: Khan, Sarwar
Veröffentlicht: (2024)
CreativeVR: Diffusion-Prior-Guided Approach for Structure and Motion Restoration in Generative and Real Videos
von: Panambur, Tejas, et al.
Veröffentlicht: (2025)
von: Panambur, Tejas, et al.
Veröffentlicht: (2025)
Bridging Compressed Image Latents and Multimodal Large Language Models
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
Understanding the Fine-Grained Knowledge Capabilities of Vision-Language Models
von: Ghosh, Dhruba, et al.
Veröffentlicht: (2026)
von: Ghosh, Dhruba, et al.
Veröffentlicht: (2026)
LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
RMAdapter: Reconstruction-based Multi-Modal Adapter for Vision-Language Models
von: Lin, Xiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiang, et al.
Veröffentlicht: (2025)
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
LinVT: Empower Your Image-level Large Language Model to Understand Videos
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
3DTV: A Feedforward Interpolation Network for Real-Time View Synthesis
von: Schulz, Stefan, et al.
Veröffentlicht: (2026)
von: Schulz, Stefan, et al.
Veröffentlicht: (2026)
X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
von: Sun, Zeyi, et al.
Veröffentlicht: (2024)
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
Parallel Backpropagation for Inverse of a Convolution with Application to Normalizing Flows
von: Nagar, Sandeep, et al.
Veröffentlicht: (2024)
von: Nagar, Sandeep, et al.
Veröffentlicht: (2024)
Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis
von: Huang, Po-Hsuan, et al.
Veröffentlicht: (2024)
von: Huang, Po-Hsuan, et al.
Veröffentlicht: (2024)
Vision-Language Meets the Skeleton: Progressively Distillation with Cross-Modal Knowledge for 3D Action Representation Learning
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Words or Vision: Do Vision-Language Models Have Blind Faith in Text?
von: Deng, Ailin, et al.
Veröffentlicht: (2025)
von: Deng, Ailin, et al.
Veröffentlicht: (2025)
4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer's Diagnosis
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching
von: Taniguchi, Takara, et al.
Veröffentlicht: (2026)
von: Taniguchi, Takara, et al.
Veröffentlicht: (2026)
Multimodal Real-Time Anomaly Detection and Industrial Applications
von: Verma, Aman, et al.
Veröffentlicht: (2025)
von: Verma, Aman, et al.
Veröffentlicht: (2025)
CinePile: A Long Video Question Answering Dataset and Benchmark
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies
von: Smith, Megan, et al.
Veröffentlicht: (2026)
von: Smith, Megan, et al.
Veröffentlicht: (2026)
Omnidirectional Video Super-Resolution using Deep Learning
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs
von: Denipitiyage, Dishanika, et al.
Veröffentlicht: (2026) -
Long-Tail Learning with Rebalanced Contrastive Loss
von: De Alvis, Charika, et al.
Veröffentlicht: (2023) -
Detecting and Characterising Mobile App Metamorphosis in Google Play Store
von: Denipitiyage, D., et al.
Veröffentlicht: (2024) -
Unveiling the Visual Counting Bottleneck in Vision-Language Models
von: Pang, Xingzhou, et al.
Veröffentlicht: (2026) -
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)