Actor-agnostic Multi-label Action Recognition with Multi-modal Query
Fuente:
arXiv
Saved in:
| Main Authors: | Mondal, Anindya, Nag, Sauradip, Prada, Joaquin M, Zhu, Xiatian, Dutta, Anjan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OmniCount: Multi-label Object Counting with Semantic-Geometric Priors
by: Mondal, Anindya, et al.
Published: (2024)
by: Mondal, Anindya, et al.
Published: (2024)
Multi-modal Deep Learning
by: Yuhua, Chen
Published: (2024)
by: Yuhua, Chen
Published: (2024)
BreastSegNet: Multi-label Segmentation of Breast MRI
by: Li, Qihang, et al.
Published: (2025)
by: Li, Qihang, et al.
Published: (2025)
CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance
by: Mondal, Anindya, et al.
Published: (2025)
by: Mondal, Anindya, et al.
Published: (2025)
Multi-modal AI for comprehensive breast cancer prognostication
by: Witowski, Jan, et al.
Published: (2024)
by: Witowski, Jan, et al.
Published: (2024)
Disentangled Multi-modal Learning of Histology and Transcriptomics for Cancer Characterization
by: Zhang, Yupei, et al.
Published: (2025)
by: Zhang, Yupei, et al.
Published: (2025)
MMM-RS: A Multi-modal, Multi-GSD, Multi-scene Remote Sensing Dataset and Benchmark for Text-to-Image Generation
by: Luo, Jialin, et al.
Published: (2024)
by: Luo, Jialin, et al.
Published: (2024)
Multi-modal Contrastive Learning for Tumor-specific Missing Modality Synthesis
by: Lim, Minjoo, et al.
Published: (2025)
by: Lim, Minjoo, et al.
Published: (2025)
Cross-Fundus Transformer for Multi-modal Diabetic Retinopathy Grading with Cataract
by: Xiao, Fan, et al.
Published: (2024)
by: Xiao, Fan, et al.
Published: (2024)
Improving Multi-label Recognition using Class Co-Occurrence Probabilities
by: Rawlekar, Samyak, et al.
Published: (2024)
by: Rawlekar, Samyak, et al.
Published: (2024)
Multi-sensor Learning Enables Information Transfer across Different Sensory Data and Augments Multi-modality Imaging
by: Zhu, Lingting, et al.
Published: (2024)
by: Zhu, Lingting, et al.
Published: (2024)
Enhanced Masked Image Modeling to Avoid Model Collapse on Multi-modal MRI Datasets
by: Han, Linxuan, et al.
Published: (2024)
by: Han, Linxuan, et al.
Published: (2024)
Multi-modal Medical Image Fusion For Non-Small Cell Lung Cancer Classification
by: Hassan, Salma, et al.
Published: (2024)
by: Hassan, Salma, et al.
Published: (2024)
Unified Domain Adaptive Semantic Segmentation
by: Zhang, Zhe, et al.
Published: (2023)
by: Zhang, Zhe, et al.
Published: (2023)
CostFilter-AD: Enhancing Anomaly Detection through Matching Cost Filtering
by: Zhang, Zhe, et al.
Published: (2025)
by: Zhang, Zhe, et al.
Published: (2025)
Dual Thinking and Logical Processing -- Are Multi-modal Large Language Models Closing the Gap with Human Vision ?
by: Dayanandan, Kailas, et al.
Published: (2024)
by: Dayanandan, Kailas, et al.
Published: (2024)
Exploring the Efficacy of Partial Denoising Using Bit Plane Slicing for Enhanced Fracture Identification: A Comparative Study of Deep Learning-Based Approaches and Handcrafted Feature Extraction Techniques
by: Paul, Snigdha, et al.
Published: (2025)
by: Paul, Snigdha, et al.
Published: (2025)
A Multi-annotated and Multi-modal Dataset for Wide-angle Video Quality Assessment
by: Hu, Bo, et al.
Published: (2025)
by: Hu, Bo, et al.
Published: (2025)
impuTMAE: Multi-modal Transformer with Masked Pre-training for Missing Modalities Imputation in Cancer Survival Prediction
by: Boyko, Maria, et al.
Published: (2025)
by: Boyko, Maria, et al.
Published: (2025)
Dr. Tongue: Sign-Oriented Multi-label Detection for Remote Tongue Diagnosis
by: Chen, Yiliang, et al.
Published: (2025)
by: Chen, Yiliang, et al.
Published: (2025)
Adapting Frozen Mono-modal Backbones for Multi-modal Registration via Contrast-Agnostic Instance Optimization
by: Zhang, Yi, et al.
Published: (2026)
by: Zhang, Yi, et al.
Published: (2026)
Multi-modal brain MRI synthesis based on SwinUNETR
by: Pang, Haowen, et al.
Published: (2025)
by: Pang, Haowen, et al.
Published: (2025)
Unified Unsupervised Anomaly Detection via Matching Cost Filtering
by: Zhang, Zhe, et al.
Published: (2025)
by: Zhang, Zhe, et al.
Published: (2025)
Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities
by: Das, Anindya Bijoy, et al.
Published: (2025)
by: Das, Anindya Bijoy, et al.
Published: (2025)
Can Large Language Models Challenge CNNs in Medical Image Analysis?
by: Ahmed, Shibbir, et al.
Published: (2025)
by: Ahmed, Shibbir, et al.
Published: (2025)
UniMIC: Towards Universal Multi-modality Perceptual Image Compression
by: Gao, Yixin, et al.
Published: (2024)
by: Gao, Yixin, et al.
Published: (2024)
Multi-modal data generation with a deep metric variational autoencoder
by: Sundgaard, Josefine Vilsbøll, et al.
Published: (2022)
by: Sundgaard, Josefine Vilsbøll, et al.
Published: (2022)
Fed-MUnet: Multi-modal Federated Unet for Brain Tumor Segmentation
by: Zhou, Ruojun, et al.
Published: (2024)
by: Zhou, Ruojun, et al.
Published: (2024)
Multi-modal MRI Translation via Evidential Regression and Distribution Calibration
by: Liu, Jiyao, et al.
Published: (2024)
by: Liu, Jiyao, et al.
Published: (2024)
Generating Print-Ready Personalized AI Art Products from Minimal User Inputs
by: Pursell, Noah, et al.
Published: (2024)
by: Pursell, Noah, et al.
Published: (2024)
Modality-agnostic Domain Generalizable Medical Image Segmentation by Multi-Frequency in Multi-Scale Attention
by: Nam, Ju-Hyeon, et al.
Published: (2024)
by: Nam, Ju-Hyeon, et al.
Published: (2024)
Macro2Micro: A Rapid and Precise Cross-modal Magnetic Resonance Imaging Synthesis using Multi-scale Structural Brain Similarity
by: Kim, Sooyoung, et al.
Published: (2024)
by: Kim, Sooyoung, et al.
Published: (2024)
ICFNet: Integrated Cross-modal Fusion Network for Survival Prediction
by: Zhang, Binyu, et al.
Published: (2025)
by: Zhang, Binyu, et al.
Published: (2025)
An Interpretable Cross-Attentive Multi-modal MRI Fusion Framework for Schizophrenia Diagnosis
by: Zhou, Ziyu, et al.
Published: (2024)
by: Zhou, Ziyu, et al.
Published: (2024)
Multi-modal Learning with Missing Modality in Predicting Axillary Lymph Node Metastasis
by: Zhang, Shichuan, et al.
Published: (2024)
by: Zhang, Shichuan, et al.
Published: (2024)
CC-DCNet: Dynamic Convolutional Neural Network with Contrastive Constraints for Identifying Lung Cancer Subtypes on Multi-modality Images
by: Jin, Yuan, et al.
Published: (2024)
by: Jin, Yuan, et al.
Published: (2024)
SynthFM: Training Modality-agnostic Foundation Models for Medical Image Segmentation without Real Medical Data
by: Sengupta, Sourya, et al.
Published: (2025)
by: Sengupta, Sourya, et al.
Published: (2025)
Progressive Multi-Level Alignments for Semi-Supervised Domain Adaptation SAR Target Recognition Using Simulated Data
by: Zhang, Xinzheng, et al.
Published: (2024)
by: Zhang, Xinzheng, et al.
Published: (2024)
CLPIPS: A Personalized Metric for AI-Generated Image Similarity
by: Trinh, Khoi, et al.
Published: (2026)
by: Trinh, Khoi, et al.
Published: (2026)
Multi-modal Liver Segmentation and Fibrosis Staging Using Real-world MRI Images
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
Similar Items
-
OmniCount: Multi-label Object Counting with Semantic-Geometric Priors
by: Mondal, Anindya, et al.
Published: (2024) -
Multi-modal Deep Learning
by: Yuhua, Chen
Published: (2024) -
BreastSegNet: Multi-label Segmentation of Breast MRI
by: Li, Qihang, et al.
Published: (2025) -
CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance
by: Mondal, Anindya, et al.
Published: (2025) -
Multi-modal AI for comprehensive breast cancer prognostication
by: Witowski, Jan, et al.
Published: (2024)