Spatio-temporal Transformers for Action Unit Classification with Event Cameras
Fuente:
arXiv
Saved in:
| Main Authors: | Cultrera, Luca, Becattini, Federico, Berlincioni, Lorenzo, Ferrari, Claudio, Del Bimbo, Alberto |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neuromorphic Valence and Arousal Estimation
by: Berlincioni, Lorenzo, et al.
Published: (2024)
by: Berlincioni, Lorenzo, et al.
Published: (2024)
Neuromorphic Face Analysis: a Survey
by: Becattini, Federico, et al.
Published: (2024)
by: Becattini, Federico, et al.
Published: (2024)
Neuromorphic Facial Analysis with Cross-Modal Supervision
by: Becattini, Federico, et al.
Published: (2024)
by: Becattini, Federico, et al.
Published: (2024)
Spike-TBR: a Noise Resilient Neuromorphic Event Representation
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
Drone Detection with Event Cameras
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
Garment Attribute Manipulation with Multi-level Attention
by: Casula, Vittorio, et al.
Published: (2024)
by: Casula, Vittorio, et al.
Published: (2024)
Diffusion Based Augmentation for Captioning and Retrieval in Cultural Heritage
by: Cioni, Dario, et al.
Published: (2023)
by: Cioni, Dario, et al.
Published: (2023)
FRED: The Florence RGB-Event Drone Dataset
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
SMEMO: Social Memory for Trajectory Forecasting
by: Marchetti, Francesco, et al.
Published: (2022)
by: Marchetti, Francesco, et al.
Published: (2022)
Prompt and Prejudice
by: Berlincioni, Lorenzo, et al.
Published: (2024)
by: Berlincioni, Lorenzo, et al.
Published: (2024)
Deepfake detection by exploiting surface anomalies: the SurFake approach
by: Ciamarra, Andrea, et al.
Published: (2023)
by: Ciamarra, Andrea, et al.
Published: (2023)
Neuromorphic Drone Detection: an Event-RGB Multimodal Approach
by: Magrini, Gabriele, et al.
Published: (2024)
by: Magrini, Gabriele, et al.
Published: (2024)
Backward-Compatible Aligned Representations via an Orthogonal Transformation Layer
by: Ricci, Simone, et al.
Published: (2024)
by: Ricci, Simone, et al.
Published: (2024)
Interactive Garment Recommendation with User in the Loop
by: Becattini, Federico, et al.
Published: (2024)
by: Becattini, Federico, et al.
Published: (2024)
Stationary Representations: Optimally Approximating Compatibility and Implications for Improved Model Replacements
by: Biondi, Niccolò, et al.
Published: (2024)
by: Biondi, Niccolò, et al.
Published: (2024)
Immunizing Images from Text to Image Editing via Adversarial Cross-Attention
by: Trippodo, Matteo, et al.
Published: (2025)
by: Trippodo, Matteo, et al.
Published: (2025)
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval
by: Agnolucci, Lorenzo, et al.
Published: (2024)
by: Agnolucci, Lorenzo, et al.
Published: (2024)
PEPR: Privileged Event-based Predictive Regularization for Domain Generalization
by: Magrini, Gabriele, et al.
Published: (2026)
by: Magrini, Gabriele, et al.
Published: (2026)
Mitigating Negative Flips via Margin Preserving Training
by: Ricci, Simone, et al.
Published: (2025)
by: Ricci, Simone, et al.
Published: (2025)
EV-Flying: an Event-based Dataset for In-The-Wild Recognition of Flying Objects
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
Transformer-based Fusion of 2D-pose and Spatio-temporal Embeddings for Distracted Driver Action Recognition
by: Akdag, Erkut, et al.
Published: (2024)
by: Akdag, Erkut, et al.
Published: (2024)
Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
Learning advisor networks for noisy image classification
by: Ricci, Simone, et al.
Published: (2022)
by: Ricci, Simone, et al.
Published: (2022)
Spatio-temporal Decoupled Knowledge Compensator for Few-Shot Action Recognition
by: Qu, Hongyu, et al.
Published: (2026)
by: Qu, Hongyu, et al.
Published: (2026)
Blur-aware Spatio-temporal Sparse Transformer for Video Deblurring
by: Zhang, Huicong, et al.
Published: (2024)
by: Zhang, Huicong, et al.
Published: (2024)
The Spatio-Temporal Poisson Point Process: A Simple Model for the Alignment of Event Camera Data
by: Gu, Cheng, et al.
Published: (2021)
by: Gu, Cheng, et al.
Published: (2021)
FineParser: A Fine-grained Spatio-temporal Action Parser for Human-centric Action Quality Assessment
by: Xu, Jinglin, et al.
Published: (2024)
by: Xu, Jinglin, et al.
Published: (2024)
Parents and Children: Distinguishing Multimodal DeepFakes from Natural Images
by: Amoroso, Roberto, et al.
Published: (2023)
by: Amoroso, Roberto, et al.
Published: (2023)
EventSleep: Sleep Activity Recognition with Event Cameras
by: Plou, Carlos, et al.
Published: (2024)
by: Plou, Carlos, et al.
Published: (2024)
Rethinking Efficient and Effective Point-based Networks for Event Camera Classification and Regression: EventMamba
by: Ren, Hongwei, et al.
Published: (2024)
by: Ren, Hongwei, et al.
Published: (2024)
Raw2Event: Converting Raw Frame Camera into Event Camera
by: Ning, Zijie, et al.
Published: (2025)
by: Ning, Zijie, et al.
Published: (2025)
EmoVOCA: Speech-Driven Emotional 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
Efficient Spatio-Temporal Vegetation Pixel Classification with Vision Transformers
by: Gomes, Alan, et al.
Published: (2026)
by: Gomes, Alan, et al.
Published: (2026)
Generalized Event Cameras
by: Sundar, Varun, et al.
Published: (2024)
by: Sundar, Varun, et al.
Published: (2024)
Autobiasing Event Cameras
by: Dilmaghani, Mehdi Sefidgar, et al.
Published: (2024)
by: Dilmaghani, Mehdi Sefidgar, et al.
Published: (2024)
ASMa: Asymmetric Spatio-temporal Masking for Skeleton Action Representation Learning
by: Anand, Aman, et al.
Published: (2026)
by: Anand, Aman, et al.
Published: (2026)
PSTTS: A Plug-and-Play Token Selector for Efficient Event-based Spatio-temporal Representation Learning
by: Zhao, Xiangmo, et al.
Published: (2025)
by: Zhao, Xiangmo, et al.
Published: (2025)
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
by: Yamane, Taiga, et al.
Published: (2025)
by: Yamane, Taiga, et al.
Published: (2025)
SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding
by: Hannan, Tanveer, et al.
Published: (2025)
by: Hannan, Tanveer, et al.
Published: (2025)
SpikePoint: An Efficient Point-based Spiking Neural Network for Event Cameras Action Recognition
by: Ren, Hongwei, et al.
Published: (2023)
by: Ren, Hongwei, et al.
Published: (2023)
Similar Items
-
Neuromorphic Valence and Arousal Estimation
by: Berlincioni, Lorenzo, et al.
Published: (2024) -
Neuromorphic Face Analysis: a Survey
by: Becattini, Federico, et al.
Published: (2024) -
Neuromorphic Facial Analysis with Cross-Modal Supervision
by: Becattini, Federico, et al.
Published: (2024) -
Spike-TBR: a Noise Resilient Neuromorphic Event Representation
by: Magrini, Gabriele, et al.
Published: (2025) -
Drone Detection with Event Cameras
by: Magrini, Gabriele, et al.
Published: (2025)