STING-BEE: Towards Vision-Language Model for Real-World X-ray Baggage Security Inspection
Fuente:
arXiv
Salvato in:
| Autori principali: | Velayudhan, Divya, Ahmed, Abdelfatah, Alansari, Mohamad, Gour, Neha, Behouch, Abderaouf, Hassan, Taimur, Wasim, Syed Talal, Maalej, Nabil, Naseer, Muzammal, Gall, Juergen, Bennamoun, Mohammed, Damiani, Ernesto, Werghi, Naoufel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs
di: Alansari, Mohamad, et al.
Pubblicazione: (2026)
di: Alansari, Mohamad, et al.
Pubblicazione: (2026)
StableMamba: Distillation-free Scaling of Large SSMs for Images and Videos
di: Suleman, Hamid, et al.
Pubblicazione: (2024)
di: Suleman, Hamid, et al.
Pubblicazione: (2024)
Rethinking Memory Design in SAM-Based Visual Object Tracking
di: Alansari, Mohamad, et al.
Pubblicazione: (2025)
di: Alansari, Mohamad, et al.
Pubblicazione: (2025)
Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
di: Yi, Jinhui, et al.
Pubblicazione: (2024)
MixANT: Observation-dependent Memory Propagation for Stochastic Dense Action Anticipation
di: Wasim, Syed Talal, et al.
Pubblicazione: (2025)
di: Wasim, Syed Talal, et al.
Pubblicazione: (2025)
RedSage: A Cybersecurity Generalist LLM
di: Suryanto, Naufal, et al.
Pubblicazione: (2026)
di: Suryanto, Naufal, et al.
Pubblicazione: (2026)
Cytoplasmic Strings Analysis in Human Embryo Time-Lapse Videos using Deep Learning Framework
di: Sohail, Anabia, et al.
Pubblicazione: (2025)
di: Sohail, Anabia, et al.
Pubblicazione: (2025)
CLDTracker: A Comprehensive Language Description for Visual Tracking
di: Alansari, Mohamad, et al.
Pubblicazione: (2025)
di: Alansari, Mohamad, et al.
Pubblicazione: (2025)
Advancing Histopathology with Deep Learning Under Data Scarcity: A Decade in Review
di: Obeid, Ahmad, et al.
Pubblicazione: (2024)
di: Obeid, Ahmad, et al.
Pubblicazione: (2024)
DyCON: Dynamic Uncertainty-aware Consistency and Contrastive Learning for Semi-supervised Medical Image Segmentation
di: Assefa, Maregu, et al.
Pubblicazione: (2025)
di: Assefa, Maregu, et al.
Pubblicazione: (2025)
Video-GroundingDINO: Towards Open-Vocabulary Spatio-Temporal Video Grounding
di: Wasim, Syed Talal, et al.
Pubblicazione: (2023)
di: Wasim, Syed Talal, et al.
Pubblicazione: (2023)
GroupMamba: Efficient Group-Based Visual State Space Model
di: Shaker, Abdelrahman, et al.
Pubblicazione: (2024)
di: Shaker, Abdelrahman, et al.
Pubblicazione: (2024)
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
di: Alawode, Basit, et al.
Pubblicazione: (2025)
di: Alawode, Basit, et al.
Pubblicazione: (2025)
Tomato Maturity Recognition with Convolutional Transformers
di: Khan, Asim, et al.
Pubblicazione: (2023)
di: Khan, Asim, et al.
Pubblicazione: (2023)
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
di: Javed, Sajid, et al.
Pubblicazione: (2024)
di: Javed, Sajid, et al.
Pubblicazione: (2024)
AdaRD-key: Adaptive Relevance-Diversity Keyframe Sampling for Long-form Video understanding
di: Zhang, Xian, et al.
Pubblicazione: (2025)
di: Zhang, Xian, et al.
Pubblicazione: (2025)
BENet: A Cross-domain Robust Network for Detecting Face Forgeries via Bias Expansion and Latent-space Attention
di: Liu, Weihua, et al.
Pubblicazione: (2024)
di: Liu, Weihua, et al.
Pubblicazione: (2024)
Multi-Modal Attention Networks for Enhanced Segmentation and Depth Estimation of Subsurface Defects in Pulse Thermography
di: Salah, Mohammed, et al.
Pubblicazione: (2025)
di: Salah, Mohammed, et al.
Pubblicazione: (2025)
Makeup-Guided Facial Privacy Protection via Untrained Neural Network Priors
di: Shamshad, Fahad, et al.
Pubblicazione: (2024)
di: Shamshad, Fahad, et al.
Pubblicazione: (2024)
Multi-Resolution Pathology-Language Pre-training Model with Text-Guided Visual Representation
di: Albastaki, Shahad, et al.
Pubblicazione: (2025)
di: Albastaki, Shahad, et al.
Pubblicazione: (2025)
RobMOT: Robust 3D Multi-Object Tracking by Observational Noise and State Estimation Drift Mitigation on LiDAR PointCloud
di: Nagy, Mohamed, et al.
Pubblicazione: (2024)
di: Nagy, Mohamed, et al.
Pubblicazione: (2024)
Towards Accurate State Estimation: Kalman Filter Incorporating Motion Dynamics for 3D Multi-Object Tracking
di: Nagy, Mohamed, et al.
Pubblicazione: (2025)
di: Nagy, Mohamed, et al.
Pubblicazione: (2025)
MUOT_3M: A 3 Million Frame Multimodal Underwater Benchmark and the MUTrack Tracking Method
di: Bakht, Ahsan Baidar, et al.
Pubblicazione: (2026)
di: Bakht, Ahsan Baidar, et al.
Pubblicazione: (2026)
LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection
di: Hossain, Akram, et al.
Pubblicazione: (2026)
di: Hossain, Akram, et al.
Pubblicazione: (2026)
SHREC'25 Track on Multiple Relief Patterns: Report and Analysis
di: Paolini, Gabriele, et al.
Pubblicazione: (2025)
di: Paolini, Gabriele, et al.
Pubblicazione: (2025)
Video Anomaly Detection in 10 Years: A Survey and Outlook
di: Abdalla, Moshira, et al.
Pubblicazione: (2024)
di: Abdalla, Moshira, et al.
Pubblicazione: (2024)
PromptSmooth: Certifying Robustness of Medical Vision-Language Models via Prompt Learning
di: Hussein, Noor, et al.
Pubblicazione: (2024)
di: Hussein, Noor, et al.
Pubblicazione: (2024)
MedContext: Learning Contextual Cues for Efficient Volumetric Medical Segmentation
di: Gani, Hanan, et al.
Pubblicazione: (2024)
di: Gani, Hanan, et al.
Pubblicazione: (2024)
Baggage Fees and Airline Performance
di: Wenyi Kuang, et al.
Pubblicazione: (2025)
di: Wenyi Kuang, et al.
Pubblicazione: (2025)
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification
di: Alawode, Basit, et al.
Pubblicazione: (2024)
di: Alawode, Basit, et al.
Pubblicazione: (2024)
To BEE or not to BEE: Estimating more than Entropy with Biased Entropy Estimators
di: la Torre, Ilaria Pia, et al.
Pubblicazione: (2025)
di: la Torre, Ilaria Pia, et al.
Pubblicazione: (2025)
A Robust Adversary Detection-Deactivation Method for Metaverse-oriented Collaborative Deep Learning
di: Li, Pengfei, et al.
Pubblicazione: (2023)
di: Li, Pengfei, et al.
Pubblicazione: (2023)
On Fun for Teaching Large Programming Courses
di: Maalej, Walid
Pubblicazione: (2026)
di: Maalej, Walid
Pubblicazione: (2026)
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models
di: Belal, Mohammad, et al.
Pubblicazione: (2024)
di: Belal, Mohammad, et al.
Pubblicazione: (2024)
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures
di: Belal, Mohammad, et al.
Pubblicazione: (2024)
di: Belal, Mohammad, et al.
Pubblicazione: (2024)
Language Guided Domain Generalized Medical Image Segmentation
di: Kunhimon, Shahina, et al.
Pubblicazione: (2024)
di: Kunhimon, Shahina, et al.
Pubblicazione: (2024)
Cross-Modal Self-Training: Aligning Images and Pointclouds to Learn Classification without Labels
di: Dharmasiri, Amaya, et al.
Pubblicazione: (2024)
di: Dharmasiri, Amaya, et al.
Pubblicazione: (2024)
Enhancing Novel Object Detection via Cooperative Foundational Models
di: Bharadwaj, Rohit, et al.
Pubblicazione: (2023)
di: Bharadwaj, Rohit, et al.
Pubblicazione: (2023)
NaviGNN: Multi-Agent Reinforcement Learning and Graph Neural Network for Sustainable Mobility in Futuristic Smart Cities
di: Bahi, Abderaouf, et al.
Pubblicazione: (2025)
di: Bahi, Abderaouf, et al.
Pubblicazione: (2025)
Crypto Technology -- Impact on Global Economy
di: Pillai, Arunkumar Velayudhan
Pubblicazione: (2024)
di: Pillai, Arunkumar Velayudhan
Pubblicazione: (2024)
Documenti analoghi
-
SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs
di: Alansari, Mohamad, et al.
Pubblicazione: (2026) -
StableMamba: Distillation-free Scaling of Large SSMs for Images and Videos
di: Suleman, Hamid, et al.
Pubblicazione: (2024) -
Rethinking Memory Design in SAM-Based Visual Object Tracking
di: Alansari, Mohamad, et al.
Pubblicazione: (2025) -
Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models
di: Yi, Jinhui, et al.
Pubblicazione: (2024) -
MixANT: Observation-dependent Memory Propagation for Stochastic Dense Action Anticipation
di: Wasim, Syed Talal, et al.
Pubblicazione: (2025)