SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting
Fuente:
arXiv
Salvato in:
| Autori principali: | Deng, Yicheng, Hayashi, Hideaki, Nagahara, Hajime |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Scale Spatio-Temporal Graph Convolutional Network for Facial Expression Spotting
di: Deng, Yicheng, et al.
Pubblicazione: (2024)
di: Deng, Yicheng, et al.
Pubblicazione: (2024)
Point-Supervised Facial Expression Spotting with Gaussian-Based Instance-Adaptive Intensity Modeling
di: Deng, Yicheng, et al.
Pubblicazione: (2025)
di: Deng, Yicheng, et al.
Pubblicazione: (2025)
MIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025)
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025)
Enhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025)
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025)
CALICO: Confident Active Learning with Integrated Calibration
di: Querol, Lorenzo S., et al.
Pubblicazione: (2024)
di: Querol, Lorenzo S., et al.
Pubblicazione: (2024)
MSSTNet: A Multi-Scale Spatio-Temporal CNN-Transformer Network for Dynamic Facial Expression Recognition
di: Wang, Linhuang, et al.
Pubblicazione: (2024)
di: Wang, Linhuang, et al.
Pubblicazione: (2024)
Synergistic Spotting and Recognition of Micro-Expression via Temporal State Transition
di: Zou, Bochao, et al.
Pubblicazione: (2024)
di: Zou, Bochao, et al.
Pubblicazione: (2024)
Multi-Focus Temporal Shifting for Precise Event Spotting in Sports Videos
di: Xu, Hao, et al.
Pubblicazione: (2025)
di: Xu, Hao, et al.
Pubblicazione: (2025)
RareSpot: Spotting Small and Rare Wildlife in Aerial Imagery with Multi-Scale Consistency and Context-Aware Augmentation
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
Causal-Ex: Causal Graph-based Micro and Macro Expression Spotting
di: Tan, Pei-Sze, et al.
Pubblicazione: (2025)
di: Tan, Pei-Sze, et al.
Pubblicazione: (2025)
EdgeSpotter: Multi-Scale Dense Text Spotting for Industrial Panel Monitoring
di: Fu, Changhong, et al.
Pubblicazione: (2025)
di: Fu, Changhong, et al.
Pubblicazione: (2025)
AdaSpot: Spend Resolution Where It Matters for Precise Event Spotting
di: Xarles, Artur, et al.
Pubblicazione: (2026)
di: Xarles, Artur, et al.
Pubblicazione: (2026)
OSL-ActionSpotting: A Unified Library for Action Spotting in Sports Videos
di: Benzakour, Yassine, et al.
Pubblicazione: (2024)
di: Benzakour, Yassine, et al.
Pubblicazione: (2024)
Deep Polarization Cues for Single-shot Shape and Subsurface Scattering Estimation
di: Li, Chenhao, et al.
Pubblicazione: (2024)
di: Li, Chenhao, et al.
Pubblicazione: (2024)
Improving Pain Classification using Spatio-Temporal Deep Learning Approaches with Facial Expressions
di: Ridouan, Aafaf, et al.
Pubblicazione: (2025)
di: Ridouan, Aafaf, et al.
Pubblicazione: (2025)
Survey of Action Recognition, Spotting and Spatio-Temporal Localization in Soccer -- Current Trends and Research Perspectives
di: Seweryn, Karolina, et al.
Pubblicazione: (2023)
di: Seweryn, Karolina, et al.
Pubblicazione: (2023)
Hyper-Local Deformable Transformers for Text Spotting on Historical Maps
di: Lin, Yijun, et al.
Pubblicazione: (2025)
di: Lin, Yijun, et al.
Pubblicazione: (2025)
From Spots to Pixels: Dense Spatial Gene Expression Prediction from Histology Images
di: Zhang, Ruikun, et al.
Pubblicazione: (2025)
di: Zhang, Ruikun, et al.
Pubblicazione: (2025)
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
HyPCV-Former: Hyperbolic Spatio-Temporal Transformer for 3D Point Cloud Video Anomaly Detection
di: Cao, Jiaping, et al.
Pubblicazione: (2025)
di: Cao, Jiaping, et al.
Pubblicazione: (2025)
CADSpotting: Robust Panoptic Symbol Spotting on Large-Scale CAD Drawings
di: Yang, Fuyi, et al.
Pubblicazione: (2024)
di: Yang, Fuyi, et al.
Pubblicazione: (2024)
MELDAE: A Framework for Micro-Expression Spotting, Detection, and Automatic Evaluation in In-the-Wild Conversational Scenes
di: Feng, Yigui, et al.
Pubblicazione: (2025)
di: Feng, Yigui, et al.
Pubblicazione: (2025)
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
di: Das, Alloy, et al.
Pubblicazione: (2023)
di: Das, Alloy, et al.
Pubblicazione: (2023)
GloTSFormer: Global Video Text Spotting Transformer
di: Wang, Han, et al.
Pubblicazione: (2024)
di: Wang, Han, et al.
Pubblicazione: (2024)
SpotEdit: Selective Region Editing in Diffusion Transformers
di: Qin, Zhibin, et al.
Pubblicazione: (2025)
di: Qin, Zhibin, et al.
Pubblicazione: (2025)
MEGC2025: Micro-Expression Grand Challenge on Spot Then Recognize and Visual Question Answering
di: Fan, Xinqi, et al.
Pubblicazione: (2025)
di: Fan, Xinqi, et al.
Pubblicazione: (2025)
T-DEED: Temporal-Discriminability Enhancer Encoder-Decoder for Precise Event Spotting in Sports Videos
di: Xarles, Artur, et al.
Pubblicazione: (2024)
di: Xarles, Artur, et al.
Pubblicazione: (2024)
Block-level Text Spotting with LLMs
di: Bannur, Ganesh, et al.
Pubblicazione: (2024)
di: Bannur, Ganesh, et al.
Pubblicazione: (2024)
STGFormer: Spatio-Temporal GraphFormer for 3D Human Pose Estimation in Video
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Lips Are Lying: Spotting the Temporal Inconsistency between Audio and Visual in Lip-Syncing DeepFakes
di: Liu, Weifeng, et al.
Pubblicazione: (2024)
di: Liu, Weifeng, et al.
Pubblicazione: (2024)
Open-Set Facial Expression Recognition
di: Zhang, Yuhang, et al.
Pubblicazione: (2024)
di: Zhang, Yuhang, et al.
Pubblicazione: (2024)
A Consistency-Aware Spot-Guided Transformer for Versatile and Hierarchical Point Cloud Registration
di: Huang, Renlang, et al.
Pubblicazione: (2024)
di: Huang, Renlang, et al.
Pubblicazione: (2024)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
di: Das, Alloy, et al.
Pubblicazione: (2024)
di: Das, Alloy, et al.
Pubblicazione: (2024)
DeepSolo++: Let Transformer Decoder with Explicit Points Solo for Multilingual Text Spotting
di: Ye, Maoyuan, et al.
Pubblicazione: (2023)
di: Ye, Maoyuan, et al.
Pubblicazione: (2023)
Watermark Text Pattern Spotting in Document Images
di: Krubiński, Mateusz, et al.
Pubblicazione: (2024)
di: Krubiński, Mateusz, et al.
Pubblicazione: (2024)
Hear the Scene: Audio-Enhanced Text Spotting
di: Li, Jing, et al.
Pubblicazione: (2024)
di: Li, Jing, et al.
Pubblicazione: (2024)
ASTRA: An Action Spotting TRAnsformer for Soccer Videos
di: Xarles, Artur, et al.
Pubblicazione: (2024)
di: Xarles, Artur, et al.
Pubblicazione: (2024)
Exploring Architectures for CNN-Based Word Spotting
di: Rusakov, Eugen, et al.
Pubblicazione: (2018)
di: Rusakov, Eugen, et al.
Pubblicazione: (2018)
Explainable Image Recognition via Enhanced Slot-attention Based Classifier
di: Wang, Bowen, et al.
Pubblicazione: (2024)
di: Wang, Bowen, et al.
Pubblicazione: (2024)
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Multi-Scale Spatio-Temporal Graph Convolutional Network for Facial Expression Spotting
di: Deng, Yicheng, et al.
Pubblicazione: (2024) -
Point-Supervised Facial Expression Spotting with Gaussian-Based Instance-Adaptive Intensity Modeling
di: Deng, Yicheng, et al.
Pubblicazione: (2025) -
MIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025) -
Enhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
di: Kawamura, Ryosuke, et al.
Pubblicazione: (2025) -
CALICO: Confident Active Learning with Integrated Calibration
di: Querol, Lorenzo S., et al.
Pubblicazione: (2024)