Towards Adaptive Fusion of Multimodal Deep Networks for Human Action Recognition
Fuente:
arXiv
Saved in:
| Main Author: | Yudistira, Novanto |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Conformal Prediction for Reliable and Explainable Medical Image Classification
by: Octadion, One, et al.
Published: (2026)
by: Octadion, One, et al.
Published: (2026)
Efficient Object Detection of Marine Debris using Pruned YOLO Model
by: Aryaza, Abi, et al.
Published: (2025)
by: Aryaza, Abi, et al.
Published: (2025)
IndoHerb: Indonesia Medicinal Plants Recognition using Transfer Learning and Deep Learning
by: Musyaffa, Muhammad Salman Ikrar, et al.
Published: (2023)
by: Musyaffa, Muhammad Salman Ikrar, et al.
Published: (2023)
Input-Adaptive Visual Preprocessing for Efficient Fast Vision-Language Model Inference
by: Cahyani, Putu Indah Githa, et al.
Published: (2025)
by: Cahyani, Putu Indah Githa, et al.
Published: (2025)
Training-Free Disentangled Text-Guided Image Editing via Sparse Latent Constraints
by: Shabrina, Mutiara, et al.
Published: (2025)
by: Shabrina, Mutiara, et al.
Published: (2025)
Hybrid of DiffStride and Spectral Pooling in Convolutional Neural Networks
by: Rafif, Sulthan, et al.
Published: (2024)
by: Rafif, Sulthan, et al.
Published: (2024)
MAMI: Multi-Attentional Mutual-Information for Long Sequence Neuron Captioning
by: Fauzulhaq, Alfirsa Damasyifa, et al.
Published: (2024)
by: Fauzulhaq, Alfirsa Damasyifa, et al.
Published: (2024)
SNN-Driven Multimodal Human Action Recognition via Sparse Spatial-Temporal Data Fusion
by: Zheng, Naichuan, et al.
Published: (2025)
by: Zheng, Naichuan, et al.
Published: (2025)
MultiFuser: Multimodal Fusion Transformer for Enhanced Driver Action Recognition
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
Human-Centric Transformer for Domain Adaptive Action Recognition
by: Lin, Kun-Yu, et al.
Published: (2024)
by: Lin, Kun-Yu, et al.
Published: (2024)
SparseSwin: Swin Transformer with Sparse Transformer Block
by: Pinasthika, Krisna, et al.
Published: (2023)
by: Pinasthika, Krisna, et al.
Published: (2023)
MK-SGN: A Spiking Graph Convolutional Network with Multimodal Fusion and Knowledge Distillation for Skeleton-based Action Recognition
by: Zheng, Naichuan, et al.
Published: (2024)
by: Zheng, Naichuan, et al.
Published: (2024)
Active Generation Network of Human Skeleton for Action Recognition
by: Liu, Long, et al.
Published: (2024)
by: Liu, Long, et al.
Published: (2024)
FusionAgent: A Multimodal Agent with Dynamic Model Selection for Human Recognition
by: Zhu, Jie, et al.
Published: (2026)
by: Zhu, Jie, et al.
Published: (2026)
Patch as Node: Human-Centric Graph Representation Learning for Multimodal Action Recognition
by: Liang, Zeyu, et al.
Published: (2025)
by: Liang, Zeyu, et al.
Published: (2025)
TACFN: Transformer-based Adaptive Cross-modal Fusion Network for Multimodal Emotion Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Application of Multimodal Fusion Deep Learning Model in Disease Recognition
by: Liu, Xiaoyi, et al.
Published: (2024)
by: Liu, Xiaoyi, et al.
Published: (2024)
Human Action Recognition without Human
by: Kataoka, Hirokatsu, et al.
Published: (2016)
by: Kataoka, Hirokatsu, et al.
Published: (2016)
Adaptive Hyper-Graph Convolution Network for Skeleton-based Human Action Recognition with Virtual Connections
by: Zhou, Youwei, et al.
Published: (2024)
by: Zhou, Youwei, et al.
Published: (2024)
iPay: Integrated Payment Action Recognition via Multimodal Networks and Adaptive Spatial Prior Learning
by: Huang, Kaicong, et al.
Published: (2026)
by: Huang, Kaicong, et al.
Published: (2026)
Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?
by: Xie, Jianyang, et al.
Published: (2025)
by: Xie, Jianyang, et al.
Published: (2025)
A Multimodal Fusion Network For Student Emotion Recognition Based on Transformer and Tensor Product
by: Xiang, Ao, et al.
Published: (2024)
by: Xiang, Ao, et al.
Published: (2024)
Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion
by: Luo, Yabo, et al.
Published: (2026)
by: Luo, Yabo, et al.
Published: (2026)
MultiTSF: Transformer-based Sensor Fusion for Human-Centric Multi-view and Multi-modal Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
An Evolutionary Network Architecture Search Framework with Adaptive Multimodal Fusion for Hand Gesture Recognition
by: Xia, Yizhang, et al.
Published: (2024)
by: Xia, Yizhang, et al.
Published: (2024)
Explore Human Parsing Modality for Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Conflict-Aware Multimodal Fusion for Ambivalence and Hesitancy Recognition
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
by: Bekhouche, Salah Eddine, et al.
Published: (2026)
An Effective End-to-End Solution for Multimodal Action Recognition
by: Wang, Songping, et al.
Published: (2025)
by: Wang, Songping, et al.
Published: (2025)
Multimodal Prototype-Enhanced Network for Few-Shot Action Recognition
by: Ni, Xinzhe, et al.
Published: (2022)
by: Ni, Xinzhe, et al.
Published: (2022)
Towards Universal Skeleton-Based Action Recognition
by: Kuang, Jidong, et al.
Published: (2026)
by: Kuang, Jidong, et al.
Published: (2026)
Sample-level Adaptive Knowledge Distillation for Action Recognition
by: Li, Ping, et al.
Published: (2025)
by: Li, Ping, et al.
Published: (2025)
Adaptive Fusion Network with Temporal-Ranked and Motion-Intensity Dynamic Images for Micro-expression Recognition
by: Man, Thi Bich Phuong, et al.
Published: (2025)
by: Man, Thi Bich Phuong, et al.
Published: (2025)
HFGCN:Hypergraph Fusion Graph Convolutional Networks for Skeleton-Based Action Recognition
by: Dong, Pengcheng, et al.
Published: (2025)
by: Dong, Pengcheng, et al.
Published: (2025)
Leveraging Foundation Models for Multimodal Graph-Based Action Recognition
by: Ziaeetabar, Fatemeh, et al.
Published: (2025)
by: Ziaeetabar, Fatemeh, et al.
Published: (2025)
LS-HAR: Language Supervised Human Action Recognition with Salient Fusion, Construction Sites as a Use-Case
by: Mahdavian, Mohammad, et al.
Published: (2024)
by: Mahdavian, Mohammad, et al.
Published: (2024)
Enhancing Adaptive Deep Networks for Image Classification via Uncertainty-aware Decision Fusion
by: Zhang, Xu, et al.
Published: (2024)
by: Zhang, Xu, et al.
Published: (2024)
Adaptive Fusion of Radiomics and Deep Features for Lung Adenocarcinoma Subtype Recognition
by: Zhou, Jing, et al.
Published: (2023)
by: Zhou, Jing, et al.
Published: (2023)
AdaptiveFusion: Adaptive Multi-Modal Multi-View Fusion for 3D Human Body Reconstruction
by: Chen, Anjun, et al.
Published: (2024)
by: Chen, Anjun, et al.
Published: (2024)
Towards an Effective Action-Region Tracking Framework for Fine-grained Video Action Recognition
by: Sun, Baoli, et al.
Published: (2025)
by: Sun, Baoli, et al.
Published: (2025)
Similar Items
-
Adaptive Conformal Prediction for Reliable and Explainable Medical Image Classification
by: Octadion, One, et al.
Published: (2026) -
Efficient Object Detection of Marine Debris using Pruned YOLO Model
by: Aryaza, Abi, et al.
Published: (2025) -
IndoHerb: Indonesia Medicinal Plants Recognition using Transfer Learning and Deep Learning
by: Musyaffa, Muhammad Salman Ikrar, et al.
Published: (2023) -
Input-Adaptive Visual Preprocessing for Efficient Fast Vision-Language Model Inference
by: Cahyani, Putu Indah Githa, et al.
Published: (2025) -
Training-Free Disentangled Text-Guided Image Editing via Sparse Latent Constraints
by: Shabrina, Mutiara, et al.
Published: (2025)