Latent Uncertainty Representations for Video-based Driver Action and Intention Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Vellenga, Koen, Steinhauer, H. Joe, Andersson, Jonas, Sjögren, Anders |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Last Layer Hamiltonian Monte Carlo
di: Vellenga, Koen, et al.
Pubblicazione: (2025)
di: Vellenga, Koen, et al.
Pubblicazione: (2025)
Taylor Videos for Action Recognition
di: Wang, Lei, et al.
Pubblicazione: (2024)
di: Wang, Lei, et al.
Pubblicazione: (2024)
Video RWKV:Video Action Recognition Based RWKV
di: Yin, Zhuowen, et al.
Pubblicazione: (2024)
di: Yin, Zhuowen, et al.
Pubblicazione: (2024)
CM2-Net: Continual Cross-Modal Mapping Network for Driver Action Recognition
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
di: Wang, Ruoyu, et al.
Pubblicazione: (2024)
EZ-CLIP: Efficient Zeroshot Video Action Recognition
di: Ahmad, Shahzad, et al.
Pubblicazione: (2023)
di: Ahmad, Shahzad, et al.
Pubblicazione: (2023)
Uncertainties of Latent Representations in Computer Vision
di: Kirchhof, Michael
Pubblicazione: (2024)
di: Kirchhof, Michael
Pubblicazione: (2024)
Midway Network: Learning Representations for Recognition and Motion from Latent Dynamics
di: Hoang, Christopher, et al.
Pubblicazione: (2025)
di: Hoang, Christopher, et al.
Pubblicazione: (2025)
Latent Action Pretraining from Videos
di: Ye, Seonghyeon, et al.
Pubblicazione: (2024)
di: Ye, Seonghyeon, et al.
Pubblicazione: (2024)
VEDIT: Latent Prediction Architecture For Procedural Video Representation Learning
di: Lin, Han, et al.
Pubblicazione: (2024)
di: Lin, Han, et al.
Pubblicazione: (2024)
Action-slot: Visual Action-centric Representations for Multi-label Atomic Activity Recognition in Traffic Scenes
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2023)
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2023)
Advancing Compressed Video Action Recognition through Progressive Knowledge Distillation
di: Soufleri, Efstathia, et al.
Pubblicazione: (2024)
di: Soufleri, Efstathia, et al.
Pubblicazione: (2024)
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
di: Yadav, Tanush, et al.
Pubblicazione: (2026)
di: Yadav, Tanush, et al.
Pubblicazione: (2026)
Olaf-World: Orienting Latent Actions for Video World Modeling
di: Jiang, Yuxin, et al.
Pubblicazione: (2026)
di: Jiang, Yuxin, et al.
Pubblicazione: (2026)
Being-H0.7: A Latent World-Action Model from Egocentric Videos
di: Luo, Hao, et al.
Pubblicazione: (2026)
di: Luo, Hao, et al.
Pubblicazione: (2026)
Improving Out-of-distribution Human Activity Recognition via IMU-Video Cross-modal Representation Learning
di: Cheshmi, Seyyed Saeid, et al.
Pubblicazione: (2025)
di: Cheshmi, Seyyed Saeid, et al.
Pubblicazione: (2025)
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
di: Yin, Xiangchen, et al.
Pubblicazione: (2025)
di: Yin, Xiangchen, et al.
Pubblicazione: (2025)
When Spatial meets Temporal in Action Recognition
di: Chen, Huilin, et al.
Pubblicazione: (2024)
di: Chen, Huilin, et al.
Pubblicazione: (2024)
Designing deep neural networks for driver intention recognition
di: Vellenga, Koen, et al.
Pubblicazione: (2024)
di: Vellenga, Koen, et al.
Pubblicazione: (2024)
Canonical Latent Representations in Conditional Diffusion Models
di: Xu, Yitao, et al.
Pubblicazione: (2025)
di: Xu, Yitao, et al.
Pubblicazione: (2025)
Segment to Focus: Guiding Latent Action Models in the Presence of Distractors
di: Fechner, Marcus, et al.
Pubblicazione: (2026)
di: Fechner, Marcus, et al.
Pubblicazione: (2026)
LLVD: LSTM-based Explicit Motion Modeling in Latent Space for Blind Video Denoising
di: Rashid, Loay, et al.
Pubblicazione: (2025)
di: Rashid, Loay, et al.
Pubblicazione: (2025)
Box2Flow: Instance-based Action Flow Graphs from Videos
di: Li, Jiatong, et al.
Pubblicazione: (2024)
di: Li, Jiatong, et al.
Pubblicazione: (2024)
Adaptive Hyper-Graph Convolution Network for Skeleton-based Human Action Recognition with Virtual Connections
di: Zhou, Youwei, et al.
Pubblicazione: (2024)
di: Zhou, Youwei, et al.
Pubblicazione: (2024)
ReSpike: Residual Frames-based Hybrid Spiking Neural Networks for Efficient Action Recognition
di: Xiao, Shiting, et al.
Pubblicazione: (2024)
di: Xiao, Shiting, et al.
Pubblicazione: (2024)
CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition
di: Wen, Yuhang, et al.
Pubblicazione: (2024)
di: Wen, Yuhang, et al.
Pubblicazione: (2024)
GraphAU-Pain: Graph-based Action Unit Representation for Pain Intensity Estimation
di: Wang, Zhiyu, et al.
Pubblicazione: (2025)
di: Wang, Zhiyu, et al.
Pubblicazione: (2025)
Latent Equivariant Operators for Robust Object Recognition: Promises and Challenges
di: Dinh, Minh, et al.
Pubblicazione: (2026)
di: Dinh, Minh, et al.
Pubblicazione: (2026)
Prototypical Calibrating Ambiguous Samples for Micro-Action Recognition
di: Li, Kun, et al.
Pubblicazione: (2024)
di: Li, Kun, et al.
Pubblicazione: (2024)
On the Utility of 3D Hand Poses for Action Recognition
di: Shamil, Md Salman, et al.
Pubblicazione: (2024)
di: Shamil, Md Salman, et al.
Pubblicazione: (2024)
Enhancing Interpretability of Sparse Latent Representations with Class Information
di: Abiz, Farshad Sangari, et al.
Pubblicazione: (2025)
di: Abiz, Farshad Sangari, et al.
Pubblicazione: (2025)
Machine Learning-Based Vehicle Intention Trajectory Recognition and Prediction for Autonomous Driving
di: Yu, Hanyi, et al.
Pubblicazione: (2024)
di: Yu, Hanyi, et al.
Pubblicazione: (2024)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
di: Fan, Linjuan, et al.
Pubblicazione: (2025)
di: Fan, Linjuan, et al.
Pubblicazione: (2025)
TrajFusionNet: Pedestrian Crossing Intention Prediction via Fusion of Sequential and Visual Trajectory Representations
di: Landry, François G., et al.
Pubblicazione: (2025)
di: Landry, François G., et al.
Pubblicazione: (2025)
Including Semantic Information via Word Embeddings for Skeleton-based Action Recognition
di: Aganian, Dustin, et al.
Pubblicazione: (2025)
di: Aganian, Dustin, et al.
Pubblicazione: (2025)
Balancing the Scales: Enhancing Fairness in Facial Expression Recognition with Latent Alignment
di: Rizvi, Syed Sameen Ahmad, et al.
Pubblicazione: (2024)
di: Rizvi, Syed Sameen Ahmad, et al.
Pubblicazione: (2024)
Isometric Representation Learning for Disentangled Latent Space of Diffusion Models
di: Hahm, Jaehoon, et al.
Pubblicazione: (2024)
di: Hahm, Jaehoon, et al.
Pubblicazione: (2024)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
di: Maduabuchi, Chika, et al.
Pubblicazione: (2025)
di: Maduabuchi, Chika, et al.
Pubblicazione: (2025)
Motus: A Unified Latent Action World Model
di: Bi, Hongzhe, et al.
Pubblicazione: (2025)
di: Bi, Hongzhe, et al.
Pubblicazione: (2025)
Gaze-Guided Graph Neural Network for Action Anticipation Conditioned on Intention
di: Ozdel, Suleyman, et al.
Pubblicazione: (2024)
di: Ozdel, Suleyman, et al.
Pubblicazione: (2024)
Advancing Vision-based Human Action Recognition: Exploring Vision-Language CLIP Model for Generalisation in Domain-Independent Tasks
di: Shandilya, Utkarsh, et al.
Pubblicazione: (2025)
di: Shandilya, Utkarsh, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Last Layer Hamiltonian Monte Carlo
di: Vellenga, Koen, et al.
Pubblicazione: (2025) -
Taylor Videos for Action Recognition
di: Wang, Lei, et al.
Pubblicazione: (2024) -
Video RWKV:Video Action Recognition Based RWKV
di: Yin, Zhuowen, et al.
Pubblicazione: (2024) -
CM2-Net: Continual Cross-Modal Mapping Network for Driver Action Recognition
di: Wang, Ruoyu, et al.
Pubblicazione: (2024) -
EZ-CLIP: Efficient Zeroshot Video Action Recognition
di: Ahmad, Shahzad, et al.
Pubblicazione: (2023)