STORM: Strategic Orchestration of Modalities for Rare Event Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kamboj, Payal, Banerjee, Ayan, Gupta, Sandeep K. S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generating customized prompts for Zero-Shot Rare Event Medical Image Classification using LLM
von: Kamboj, Payal, et al.
Veröffentlicht: (2025)
von: Kamboj, Payal, et al.
Veröffentlicht: (2025)
Framework for developing and evaluating ethical collaboration between expert and machine
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
The Expert Knowledge combined with AI outperforms AI Alone in Seizure Onset Zone Localization using resting state fMRI
von: Kamboj, Payal, et al.
Veröffentlicht: (2023)
von: Kamboj, Payal, et al.
Veröffentlicht: (2023)
XAI-MeD: Explainable Knowledge Guided Neuro-Symbolic Framework for Domain Generalization and Rare Class Detection in Medical Imaging
von: Urooj, Midhat, et al.
Veröffentlicht: (2026)
von: Urooj, Midhat, et al.
Veröffentlicht: (2026)
Detection of Deployment Operational Deviations for Safety and Security of AI-Enabled Human-Centric Cyber Physical Systems
von: Ngabonziza, Bernard, et al.
Veröffentlicht: (2026)
von: Ngabonziza, Bernard, et al.
Veröffentlicht: (2026)
Human Knowledge Integrated Multi-modal Learning for Single Source Domain Generalization
von: Banerjee, Ayan, et al.
Veröffentlicht: (2026)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2026)
EMMA: Extracting Multiple physical parameters from Multimodal Data
von: Shaikh, Farhat, et al.
Veröffentlicht: (2026)
von: Shaikh, Farhat, et al.
Veröffentlicht: (2026)
NEURO-GUARD: Neuro-Symbolic Generalization and Unbiased Adaptive Routing for Diagnostics -- Explainable Medical AI
von: Urooj, Midhat, et al.
Veröffentlicht: (2025)
von: Urooj, Midhat, et al.
Veröffentlicht: (2025)
Experience with Single Domain Generalization in Real World Medical Imaging Deployments
von: Banerjee, Ayan, et al.
Veröffentlicht: (2026)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2026)
Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach
von: Urooj, Midhat, et al.
Veröffentlicht: (2025)
von: Urooj, Midhat, et al.
Veröffentlicht: (2025)
The Progression of Transformers from Language to Vision to MOT: A Literature Review on Multi-Object Tracking with Transformers
von: Kamboj, Abhi
Veröffentlicht: (2024)
von: Kamboj, Abhi
Veröffentlicht: (2024)
Stylistic-STORM (ST-STORM) : Perceiving the Semantic Nature of Appearance
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering
von: Jia, Yiduo, et al.
Veröffentlicht: (2026)
von: Jia, Yiduo, et al.
Veröffentlicht: (2026)
CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models
von: Banerjee, Ayan, et al.
Veröffentlicht: (2025)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2025)
STORM: Token-Efficient Long Video Understanding for Multimodal LLMs
von: Jiang, Jindong, et al.
Veröffentlicht: (2025)
von: Jiang, Jindong, et al.
Veröffentlicht: (2025)
Image Classification using Fuzzy Pooling in Convolutional Kolmogorov-Arnold Networks
von: Igali, Ayan, et al.
Veröffentlicht: (2024)
von: Igali, Ayan, et al.
Veröffentlicht: (2024)
Robust Vision Systems for Connected and Autonomous Vehicles: Security Challenges and Attack Vectors
von: Gupta, Sandeep, et al.
Veröffentlicht: (2026)
von: Gupta, Sandeep, et al.
Veröffentlicht: (2026)
STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
von: Wang, Jinhong, et al.
Veröffentlicht: (2025)
STORM: Segment, Track, and Object Re-Localization from a Single Image
von: Deng, Yu, et al.
Veröffentlicht: (2025)
von: Deng, Yu, et al.
Veröffentlicht: (2025)
STORM: Search-Guided Generative World Models for Robotic Manipulation
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs
von: Luo, Ziyang, et al.
Veröffentlicht: (2026)
von: Luo, Ziyang, et al.
Veröffentlicht: (2026)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
von: Banerjee, Ayan, et al.
Veröffentlicht: (2025)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2025)
DocRevive: A Unified Pipeline for Document Text Restoration
von: Purkayastha, Kunal, et al.
Veröffentlicht: (2026)
von: Purkayastha, Kunal, et al.
Veröffentlicht: (2026)
STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models
von: Liang, Yiming, et al.
Veröffentlicht: (2026)
von: Liang, Yiming, et al.
Veröffentlicht: (2026)
Robult: Leveraging Redundancy and Modality Specific Features for Robust Multimodal Learning
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
STORM: End-to-End Referring Multi-Object Tracking in Videos
von: Lu, Zijia, et al.
Veröffentlicht: (2026)
von: Lu, Zijia, et al.
Veröffentlicht: (2026)
SuperEvent: Cross-Modal Learning of Event-based Keypoint Detection for SLAM
von: Burkhardt, Yannick, et al.
Veröffentlicht: (2025)
von: Burkhardt, Yannick, et al.
Veröffentlicht: (2025)
A Survey of IMU Based Cross-Modal Transfer Learning in Human Activity Recognition
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes
von: Yang, Jiawei, et al.
Veröffentlicht: (2024)
von: Yang, Jiawei, et al.
Veröffentlicht: (2024)
CPS-LLM: Large Language Model based Safe Usage Plan Generator for Human-in-the-Loop Human-in-the-Plant Cyber-Physical System
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
EventVGGT: Exploring Cross-Modal Distillation for Consistent Event-based Depth Estimation
von: Ren, Yinrui, et al.
Veröffentlicht: (2026)
von: Ren, Yinrui, et al.
Veröffentlicht: (2026)
TACO-Net: Topological Signatures Triumph in 3D Object Classification
von: Ghosh, Anirban, et al.
Veröffentlicht: (2025)
von: Ghosh, Anirban, et al.
Veröffentlicht: (2025)
Sparsity-based background removal for STORM super-resolution images
von: Valera, Patris, et al.
Veröffentlicht: (2024)
von: Valera, Patris, et al.
Veröffentlicht: (2024)
Your Interest, Your Summaries: Query-Focused Long Video Summarization
von: Patel, Nirav, et al.
Veröffentlicht: (2024)
von: Patel, Nirav, et al.
Veröffentlicht: (2024)
Cross-Modal Mapping: Mitigating the Modality Gap for Few-Shot Image Classification
von: Yang, Xi, et al.
Veröffentlicht: (2024)
von: Yang, Xi, et al.
Veröffentlicht: (2024)
Bidirectional Cross-Modal Prompting for Event-Frame Asymmetric Stereo
von: Xu, Ninghui, et al.
Veröffentlicht: (2026)
von: Xu, Ninghui, et al.
Veröffentlicht: (2026)
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Generating customized prompts for Zero-Shot Rare Event Medical Image Classification using LLM
von: Kamboj, Payal, et al.
Veröffentlicht: (2025) -
Framework for developing and evaluating ethical collaboration between expert and machine
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024) -
The Expert Knowledge combined with AI outperforms AI Alone in Seizure Onset Zone Localization using resting state fMRI
von: Kamboj, Payal, et al.
Veröffentlicht: (2023) -
XAI-MeD: Explainable Knowledge Guided Neuro-Symbolic Framework for Domain Generalization and Rare Class Detection in Medical Imaging
von: Urooj, Midhat, et al.
Veröffentlicht: (2026) -
Detection of Deployment Operational Deviations for Safety and Security of AI-Enabled Human-Centric Cyber Physical Systems
von: Ngabonziza, Bernard, et al.
Veröffentlicht: (2026)