EPIR: An Efficient Patch Tokenization, Integration and Representation Framework for Micro-expression Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Junbo, Fu, Liangyu, Li, Yuke, Zhu, Yining, Wu, Xuecheng, Hu, Kun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning
by: Wang, Junbo, et al.
Published: (2026)
by: Wang, Junbo, et al.
Published: (2026)
FAMNet: Integrating 2D and 3D Features for Micro-expression Recognition via Multi-task Learning and Hierarchical Attention
by: Fu, Liangyu, et al.
Published: (2025)
by: Fu, Liangyu, et al.
Published: (2025)
HaltingVT: Adaptive Token Halting Transformer for Efficient Video Recognition
by: Wu, Qian, et al.
Published: (2024)
by: Wu, Qian, et al.
Published: (2024)
TACR-YOLO: A Real-time Detection Framework for Abnormal Human Behaviors Enhanced with Coordinate and Task-Aware Representations
by: Yin, Xinyi, et al.
Published: (2025)
by: Yin, Xinyi, et al.
Published: (2025)
3A-YOLO: New Real-Time Object Detectors with Triple Discriminative Awareness and Coordinated Representations
by: Wu, Xuecheng, et al.
Published: (2024)
by: Wu, Xuecheng, et al.
Published: (2024)
Benchmarking Micro-action Recognition: Dataset, Methods, and Applications
by: Guo, Dan, et al.
Published: (2024)
by: Guo, Dan, et al.
Published: (2024)
A Benchmark for Incremental Micro-expression Recognition
by: Lai, Zhengqin, et al.
Published: (2025)
by: Lai, Zhengqin, et al.
Published: (2025)
Context Patch Fusion With Class Token Enhancement for Weakly Supervised Semantic Segmentation
by: Fu, Yiyang, et al.
Published: (2026)
by: Fu, Yiyang, et al.
Published: (2026)
Scalable Audio-Visual Masked Autoencoders for Efficient Affective Video Facial Analysis
by: Wu, Xuecheng, et al.
Published: (2025)
by: Wu, Xuecheng, et al.
Published: (2025)
PESFormer: Boosting Macro- and Micro-expression Spotting with Direct Timestamp Encoding
by: Yu, Wang-Wang, et al.
Published: (2024)
by: Yu, Wang-Wang, et al.
Published: (2024)
FocusLLaVA: A Coarse-to-Fine Approach for Efficient and Effective Visual Token Compression
by: Zhu, Yuke, et al.
Published: (2024)
by: Zhu, Yuke, et al.
Published: (2024)
DFME: A New Benchmark for Dynamic Facial Micro-expression Recognition
by: Zhao, Sirui, et al.
Published: (2023)
by: Zhao, Sirui, et al.
Published: (2023)
A Trustworthy Method for Multimodal Emotion Recognition
by: Xue, Junxiao, et al.
Published: (2025)
by: Xue, Junxiao, et al.
Published: (2025)
ICANet: A Method of Short Video Emotion Recognition Driven by Multimodal Data
by: Wu, Xuecheng, et al.
Published: (2022)
by: Wu, Xuecheng, et al.
Published: (2022)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Weak Supervision with Arbitrary Single Frame for Micro- and Macro-expression Spotting
by: Yu, Wang-Wang, et al.
Published: (2024)
by: Yu, Wang-Wang, et al.
Published: (2024)
MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition
by: Hu, Chenxing, et al.
Published: (2026)
by: Hu, Chenxing, et al.
Published: (2026)
Micro-gesture Online Recognition using Learnable Query Points
by: Liu, Pengyu, et al.
Published: (2024)
by: Liu, Pengyu, et al.
Published: (2024)
Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action Recognition
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Micro-expression Recognition Based on Dual-branch Feature Extraction and Fusion
by: Zhang, Mingjie, et al.
Published: (2026)
by: Zhang, Mingjie, et al.
Published: (2026)
Is Micro-expression Ethnic Leaning?
by: Khor, Huai-Qian, et al.
Published: (2025)
by: Khor, Huai-Qian, et al.
Published: (2025)
Adaptive Temporal Motion Guided Graph Convolution Network for Micro-expression Recognition
by: Zhang, Fengyuan, et al.
Published: (2024)
by: Zhang, Fengyuan, et al.
Published: (2024)
Rethinking Key-frame-based Micro-expression Recognition: A Robust and Accurate Framework Against Key-frame Errors
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
Exploring Diverse Representations for Open Set Recognition
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
InfoSyncNet: Information Synchronization Temporal Convolutional Network for Visual Speech Recognition
by: Xue, Junxiao, et al.
Published: (2025)
by: Xue, Junxiao, et al.
Published: (2025)
AdaTok: Adaptive Token Compression with Object-Aware Representations for Efficient Multimodal LLMs
by: Zhang, Xinliang, et al.
Published: (2025)
by: Zhang, Xinliang, et al.
Published: (2025)
Prototypical Calibrating Ambiguous Samples for Micro-Action Recognition
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
by: Liu, Pengyu, et al.
Published: (2025)
by: Liu, Pengyu, et al.
Published: (2025)
Spiking Patches: Asynchronous, Sparse, and Efficient Tokens for Event Cameras
by: Øhrstrøm, Christoffer Koo, et al.
Published: (2025)
by: Øhrstrøm, Christoffer Koo, et al.
Published: (2025)
VisionTrim: Unified Vision Token Compression for Training-Free MLLM Acceleration
by: Yu, Hanxun, et al.
Published: (2026)
by: Yu, Hanxun, et al.
Published: (2026)
Patch-based Representation and Learning for Efficient Deformation Modeling
by: Chen, Ruochen, et al.
Published: (2026)
by: Chen, Ruochen, et al.
Published: (2026)
Patch as Node: Human-Centric Graph Representation Learning for Multimodal Action Recognition
by: Liang, Zeyu, et al.
Published: (2025)
by: Liang, Zeyu, et al.
Published: (2025)
PatchScaler: An Efficient Patch-Independent Diffusion Model for Image Super-Resolution
by: Liu, Yong, et al.
Published: (2024)
by: Liu, Yong, et al.
Published: (2024)
TokenFocus-VQA: Enhancing Text-to-Image Alignment with Position-Aware Focus and Multi-Perspective Aggregations on LVLMs
by: Zhang, Zijian, et al.
Published: (2025)
by: Zhang, Zijian, et al.
Published: (2025)
Temporal and Spatial Feature Fusion Framework for Dynamic Micro Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Adaptive Fusion Network with Temporal-Ranked and Motion-Intensity Dynamic Images for Micro-expression Recognition
by: Man, Thi Bich Phuong, et al.
Published: (2025)
by: Man, Thi Bich Phuong, et al.
Published: (2025)
SDD-YOLO: A Small-Target Detection Framework for Ground-to-Air Anti-UAV Surveillance with Edge-Efficient Deployment
by: Chen, Pengyu, et al.
Published: (2026)
by: Chen, Pengyu, et al.
Published: (2026)
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches
by: Wu, Cheng-En, et al.
Published: (2024)
by: Wu, Cheng-En, et al.
Published: (2024)
Structure-Preserving Patch Decoding for Efficient Neural Video Representation
by: Hayami, Taiga, et al.
Published: (2025)
by: Hayami, Taiga, et al.
Published: (2025)
[CLS] is Not Enough: Multi-Label Recognition via Patch-Level Inference and Adaptive Aggregation
by: Wang, Akang, et al.
Published: (2026)
by: Wang, Akang, et al.
Published: (2026)
Similar Items
-
DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning
by: Wang, Junbo, et al.
Published: (2026) -
FAMNet: Integrating 2D and 3D Features for Micro-expression Recognition via Multi-task Learning and Hierarchical Attention
by: Fu, Liangyu, et al.
Published: (2025) -
HaltingVT: Adaptive Token Halting Transformer for Efficient Video Recognition
by: Wu, Qian, et al.
Published: (2024) -
TACR-YOLO: A Real-time Detection Framework for Abnormal Human Behaviors Enhanced with Coordinate and Task-Aware Representations
by: Yin, Xinyi, et al.
Published: (2025) -
3A-YOLO: New Real-Time Object Detectors with Triple Discriminative Awareness and Coordinated Representations
by: Wu, Xuecheng, et al.
Published: (2024)