Efficient Egocentric Action Recognition with Multimodal Data
Fuente:
arXiv
Saved in:
| Main Authors: | Calzavara, Marco, Kastrati, Ard, Macchini, Matteo, Vasilevski, Dushan, Wattenhofer, Roger |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AEye: A Visualization Tool for Image Datasets
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
by: Mosconi, Matteo, et al.
Published: (2024)
by: Mosconi, Matteo, et al.
Published: (2024)
Continual Multimodal Egocentric Activity Recognition via Modality-Aware Novel Detection
by: Lim, Wonseon, et al.
Published: (2026)
by: Lim, Wonseon, et al.
Published: (2026)
Object Aware Egocentric Online Action Detection
by: An, Joungbin, et al.
Published: (2024)
by: An, Joungbin, et al.
Published: (2024)
The Impact of Scaling Training Data on Adversarial Robustness
by: Zimmerli, Marco, et al.
Published: (2025)
by: Zimmerli, Marco, et al.
Published: (2025)
EPAM-Net: An Efficient Pose-driven Attention-guided Multimodal Network for Video Action Recognition
by: Abdelkawy, Ahmed, et al.
Published: (2024)
by: Abdelkawy, Ahmed, et al.
Published: (2024)
Active Multimodal Distillation for Few-shot Action Recognition
by: Feng, Weijia, et al.
Published: (2025)
by: Feng, Weijia, et al.
Published: (2025)
Seeing Through the Mask: Rethinking Adversarial Examples for CAPTCHAs
by: Jabary, Yahya, et al.
Published: (2024)
by: Jabary, Yahya, et al.
Published: (2024)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
by: Tse, Tze Ho Elden, et al.
Published: (2025)
by: Tse, Tze Ho Elden, et al.
Published: (2025)
Multimodal Prototype-Enhanced Network for Few-Shot Action Recognition
by: Ni, Xinzhe, et al.
Published: (2022)
by: Ni, Xinzhe, et al.
Published: (2022)
Aria-NeRF: Multimodal Egocentric View Synthesis
by: Sun, Jiankai, et al.
Published: (2023)
by: Sun, Jiankai, et al.
Published: (2023)
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
by: Chu, Qiaohui, et al.
Published: (2025)
by: Chu, Qiaohui, et al.
Published: (2025)
FLIP Reasoning Challenge
by: Plesner, Andreas, et al.
Published: (2025)
by: Plesner, Andreas, et al.
Published: (2025)
Beyond Perfect Scores: Proof-by-Contradiction for Trustworthy Machine Learning
by: Wadduwage, Dushan N., et al.
Published: (2026)
by: Wadduwage, Dushan N., et al.
Published: (2026)
EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
by: Li, Yuan-Ming, et al.
Published: (2024)
by: Li, Yuan-Ming, et al.
Published: (2024)
EgoGen: An Egocentric Synthetic Data Generator
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Interpretable Action Recognition on Hard to Classify Actions
by: Anichenko, Anastasia, et al.
Published: (2024)
by: Anichenko, Anastasia, et al.
Published: (2024)
Fire on Motion: Optimizing Video Pass-bands for Efficient Spiking Action Recognition
by: Ye, Shuhan, et al.
Published: (2026)
by: Ye, Shuhan, et al.
Published: (2026)
iPay: Integrated Payment Action Recognition via Multimodal Networks and Adaptive Spatial Prior Learning
by: Huang, Kaicong, et al.
Published: (2026)
by: Huang, Kaicong, et al.
Published: (2026)
From MNIST to ImageNet: Understanding the Scalability Boundaries of Differentiable Logic Gate Networks
by: Brändle, Sven, et al.
Published: (2025)
by: Brändle, Sven, et al.
Published: (2025)
SUPClust: Active Learning at the Boundaries
by: Ono, Yuta, et al.
Published: (2024)
by: Ono, Yuta, et al.
Published: (2024)
Bridging Diversity and Uncertainty in Active learning with Self-Supervised Pre-Training
by: Doucet, Paul, et al.
Published: (2024)
by: Doucet, Paul, et al.
Published: (2024)
Enhanced Sparse Point Cloud Data Processing for Privacy-aware Human Action Recognition
by: Tunau, Maimunatu, et al.
Published: (2025)
by: Tunau, Maimunatu, et al.
Published: (2025)
Exploring Explainability in Video Action Recognition
by: Saha, Avinab, et al.
Published: (2024)
by: Saha, Avinab, et al.
Published: (2024)
Enhancing Action Recognition by Leveraging the Hierarchical Structure of Actions and Textual Context
by: Benavent-Lledo, Manuel, et al.
Published: (2024)
by: Benavent-Lledo, Manuel, et al.
Published: (2024)
EgoCross: Benchmarking Multimodal Large Language Models for Cross-Domain Egocentric Video Question Answering
by: Li, Yanjun, et al.
Published: (2025)
by: Li, Yanjun, et al.
Published: (2025)
COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition
by: Chen, Baiyu, et al.
Published: (2025)
by: Chen, Baiyu, et al.
Published: (2025)
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
by: Cai, Xiongyi, et al.
Published: (2025)
by: Cai, Xiongyi, et al.
Published: (2025)
OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation
by: Jawaid, Ahad, et al.
Published: (2025)
by: Jawaid, Ahad, et al.
Published: (2025)
Distilling Privileged Multimodal Information for Expression Recognition using Optimal Transport
by: Aslam, Muhammad Haseeb, et al.
Published: (2024)
by: Aslam, Muhammad Haseeb, et al.
Published: (2024)
Is 'Right' Right? Enhancing Object Orientation Understanding in Multimodal Large Language Models through Egocentric Instruction Tuning
by: Jung, Ji Hyeok, et al.
Published: (2024)
by: Jung, Ji Hyeok, et al.
Published: (2024)
Cognitively-Inspired Tokens Overcome Egocentric Bias in Multimodal Models
by: Leonard, Bridget, et al.
Published: (2026)
by: Leonard, Bridget, et al.
Published: (2026)
Egocentric Bias in Vision-Language Models
by: Wang, Maijunxian, et al.
Published: (2026)
by: Wang, Maijunxian, et al.
Published: (2026)
Exploring Ordinal Bias in Action Recognition for Instructional Videos
by: Kim, Joochan, et al.
Published: (2025)
by: Kim, Joochan, et al.
Published: (2025)
Variational Contrastive Learning for Skeleton-based Action Recognition
by: Nguyen, Dang Dinh, et al.
Published: (2026)
by: Nguyen, Dang Dinh, et al.
Published: (2026)
SkateboardAI: The Coolest Video Action Recognition for Skateboarding
by: Chen, Hanxiao
Published: (2023)
by: Chen, Hanxiao
Published: (2023)
TASAR: Transfer-based Attack on Skeletal Action Recognition
by: Diao, Yunfeng, et al.
Published: (2024)
by: Diao, Yunfeng, et al.
Published: (2024)
A Survey on Backbones for Deep Video Action Recognition
by: Tang, Zixuan, et al.
Published: (2024)
by: Tang, Zixuan, et al.
Published: (2024)
Flatten: Video Action Recognition is an Image Classification task
by: Chen, Junlin, et al.
Published: (2024)
by: Chen, Junlin, et al.
Published: (2024)
Benchmarking Egocentric Clinical Intent Understanding Capability for Medical Multimodal Large Language Models
by: Liu, Shaonan, et al.
Published: (2026)
by: Liu, Shaonan, et al.
Published: (2026)
Similar Items
-
AEye: A Visualization Tool for Image Datasets
by: Grötschla, Florian, et al.
Published: (2024) -
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
by: Mosconi, Matteo, et al.
Published: (2024) -
Continual Multimodal Egocentric Activity Recognition via Modality-Aware Novel Detection
by: Lim, Wonseon, et al.
Published: (2026) -
Object Aware Egocentric Online Action Detection
by: An, Joungbin, et al.
Published: (2024) -
The Impact of Scaling Training Data on Adversarial Robustness
by: Zimmerli, Marco, et al.
Published: (2025)