Predicting Penalty Kick Direction Using Multi-Modal Deep Learning with Pose-Guided Attention
Fuente:
arXiv
Guardado en:
| Autores principales: | Ranasinghe, Pasindu, Ranasinghe, Pamudu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Deep Learning Approach to Identify Rock Bolts in Complex 3D Point Clouds of Underground Mines Captured Using Mobile Laser Scanners
por: Patra, Dibyayan, et al.
Publicado: (2025)
por: Patra, Dibyayan, et al.
Publicado: (2025)
LiDAR Point Cloud Colourisation Using Multi-Camera Fusion and Low-Light Image Enhancement
por: Ranasinghe, Pasindu, et al.
Publicado: (2025)
por: Ranasinghe, Pasindu, et al.
Publicado: (2025)
Predicting Soccer Penalty Kick Direction Using Human Action Recognition
por: Freire-Obregón, David, et al.
Publicado: (2025)
por: Freire-Obregón, David, et al.
Publicado: (2025)
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
por: Patra, Dibyayan, et al.
Publicado: (2026)
por: Patra, Dibyayan, et al.
Publicado: (2026)
MambaKick: Early Penalty Direction Prediction from HAR Embeddings
por: Velesaca, Henry O., et al.
Publicado: (2026)
por: Velesaca, Henry O., et al.
Publicado: (2026)
Towards Integrated Rock Support Visualisation in 3D Point Cloud of Underground Mines
por: Patra, Dibyayan, et al.
Publicado: (2026)
por: Patra, Dibyayan, et al.
Publicado: (2026)
Team-Aware Football Player Tracking with SAM: An Appearance-Based Approach to Occlusion Recovery
por: Ranasinghe, Chamath, et al.
Publicado: (2025)
por: Ranasinghe, Chamath, et al.
Publicado: (2025)
Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection
por: Ranasinghe, Yasiru, et al.
Publicado: (2026)
por: Ranasinghe, Yasiru, et al.
Publicado: (2026)
Multi-scale Attention Guided Pose Transfer
por: Roy, Prasun, et al.
Publicado: (2022)
por: Roy, Prasun, et al.
Publicado: (2022)
Language Repository for Long Video Understanding
por: Kahatapitiya, Kumara, et al.
Publicado: (2024)
por: Kahatapitiya, Kumara, et al.
Publicado: (2024)
Understanding Long Videos with Multimodal Language Models
por: Ranasinghe, Kanchana, et al.
Publicado: (2024)
por: Ranasinghe, Kanchana, et al.
Publicado: (2024)
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2023)
por: Ranasinghe, Yasiru, et al.
Publicado: (2023)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
CoPT: Unsupervised Domain Adaptive Segmentation using Domain-Agnostic Text Embeddings
por: Mata, Cristina, et al.
Publicado: (2025)
por: Mata, Cristina, et al.
Publicado: (2025)
Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception
por: Thushara, Rusiru, et al.
Publicado: (2026)
por: Thushara, Rusiru, et al.
Publicado: (2026)
Mysteries of the Deep: Role of Intermediate Representations in Out of Distribution Detection
por: De la Jara, I. M., et al.
Publicado: (2025)
por: De la Jara, I. M., et al.
Publicado: (2025)
Learning to Localize Objects Improves Spatial Reasoning in Visual-LLMs
por: Ranasinghe, Kanchana, et al.
Publicado: (2024)
por: Ranasinghe, Kanchana, et al.
Publicado: (2024)
Multi-Modal Monocular Endoscopic Depth and Pose Estimation with Edge-Guided Self-Supervision
por: Ju, Xinwei, et al.
Publicado: (2026)
por: Ju, Xinwei, et al.
Publicado: (2026)
Hierarchical Text-to-Vision Self Supervised Alignment for Improved Histopathology Representation Learning
por: Watawana, Hasindri, et al.
Publicado: (2024)
por: Watawana, Hasindri, et al.
Publicado: (2024)
Clinical-Prior Guided Multi-Modal Learning with Latent Attention Pooling for Gait-Based Scoliosis Screening
por: Chen, Dong, et al.
Publicado: (2026)
por: Chen, Dong, et al.
Publicado: (2026)
BEVPose: Unveiling Scene Semantics through Pose-Guided Multi-Modal BEV Alignment
por: Hosseinzadeh, Mehdi, et al.
Publicado: (2024)
por: Hosseinzadeh, Mehdi, et al.
Publicado: (2024)
ACIT: Attention-Guided Cross-Modal Interaction Transformer for Pedestrian Crossing Intention Prediction
por: Li, Yuanzhe, et al.
Publicado: (2025)
por: Li, Yuanzhe, et al.
Publicado: (2025)
Too Many Frames, Not All Useful: Efficient Strategies for Long-Form Video QA
por: Park, Jongwoo, et al.
Publicado: (2024)
por: Park, Jongwoo, et al.
Publicado: (2024)
MM-GTUNets: Unified Multi-Modal Graph Deep Learning for Brain Disorders Prediction
por: Cai, Luhui, et al.
Publicado: (2024)
por: Cai, Luhui, et al.
Publicado: (2024)
Pixel Motion Diffusion is What We Need for Robot Control
por: Nguyen, E-Ro, et al.
Publicado: (2025)
por: Nguyen, E-Ro, et al.
Publicado: (2025)
Recognition of Daily Activities through Multi-Modal Deep Learning: A Video, Pose, and Object-Aware Approach for Ambient Assisted Living
por: Hashemifard, Kooshan, et al.
Publicado: (2026)
por: Hashemifard, Kooshan, et al.
Publicado: (2026)
OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation
por: Huang, Qidong, et al.
Publicado: (2023)
por: Huang, Qidong, et al.
Publicado: (2023)
MultiSense-Pneumo: A Multimodal Learning Framework for Pneumonia Screening in Resource-Constrained Settings
por: Jayakody, Dineth, et al.
Publicado: (2026)
por: Jayakody, Dineth, et al.
Publicado: (2026)
Activating Self-Attention for Multi-Scene Absolute Pose Regression
por: Lee, Miso, et al.
Publicado: (2024)
por: Lee, Miso, et al.
Publicado: (2024)
Cross-Modal Attention Guided Unlearning in Vision-Language Models
por: Bhaila, Karuna, et al.
Publicado: (2025)
por: Bhaila, Karuna, et al.
Publicado: (2025)
Future Optical Flow Prediction Improves Robot Control & Video Generation
por: Ranasinghe, Kanchana, et al.
Publicado: (2026)
por: Ranasinghe, Kanchana, et al.
Publicado: (2026)
GateAttentionPose: Enhancing Pose Estimation with Agent Attention and Improved Gated Convolutions
por: Feng, Liang, et al.
Publicado: (2024)
por: Feng, Liang, et al.
Publicado: (2024)
Deep Learning Framework for Early Detection of Pancreatic Cancer Using Multi-Modal Medical Imaging Analysis
por: Slobodzian, Dennis, et al.
Publicado: (2025)
por: Slobodzian, Dennis, et al.
Publicado: (2025)
Prediction of Distant Metastasis in Head and Neck Cancer Patients Using Tumor and Peritumoral Multi-Modal Deep Learning
por: Tong, Nuo, et al.
Publicado: (2025)
por: Tong, Nuo, et al.
Publicado: (2025)
Pose2Trajectory: Using Transformers on Body Pose to Predict Tennis Player's Trajectory
por: AlShami, Ali K., et al.
Publicado: (2024)
por: AlShami, Ali K., et al.
Publicado: (2024)
Pixel Motion as Universal Representation for Robot Control
por: Ranasinghe, Kanchana, et al.
Publicado: (2025)
por: Ranasinghe, Kanchana, et al.
Publicado: (2025)
Test-Time Optimization for Domain Adaptive Open Vocabulary Segmentation
por: De Silva, Ulindu, et al.
Publicado: (2025)
por: De Silva, Ulindu, et al.
Publicado: (2025)
COSMO-INR: Complex Sinusoidal Modulation for Implicit Neural Representations
por: Thennakoon, Pandula, et al.
Publicado: (2025)
por: Thennakoon, Pandula, et al.
Publicado: (2025)
MultiAnimate: Pose-Guided Image Animation Made Extensible
por: Hu, Yingcheng, et al.
Publicado: (2026)
por: Hu, Yingcheng, et al.
Publicado: (2026)
Towards Balanced Multi-Modal Learning in 3D Human Pose Estimation
por: Qi, Mengshi, et al.
Publicado: (2025)
por: Qi, Mengshi, et al.
Publicado: (2025)
Ejemplares similares
-
A Deep Learning Approach to Identify Rock Bolts in Complex 3D Point Clouds of Underground Mines Captured Using Mobile Laser Scanners
por: Patra, Dibyayan, et al.
Publicado: (2025) -
LiDAR Point Cloud Colourisation Using Multi-Camera Fusion and Low-Light Image Enhancement
por: Ranasinghe, Pasindu, et al.
Publicado: (2025) -
Predicting Soccer Penalty Kick Direction Using Human Action Recognition
por: Freire-Obregón, David, et al.
Publicado: (2025) -
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
por: Patra, Dibyayan, et al.
Publicado: (2026) -
MambaKick: Early Penalty Direction Prediction from HAR Embeddings
por: Velesaca, Henry O., et al.
Publicado: (2026)