MulCPred: Learning Multi-modal Concepts for Explainable Pedestrian Action Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Feng, Yan, Carballo, Alexander, Fujii, Keisuke, Karlsson, Robin, Ding, Ming, Takeda, Kazuya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sparse Prototype Network for Explainable Pedestrian Behavior Prediction
von: Feng, Yan, et al.
Veröffentlicht: (2024)
von: Feng, Yan, et al.
Veröffentlicht: (2024)
Runner re-identification from single-view running video in the open-world setting
von: Suzuki, Tomohiro, et al.
Veröffentlicht: (2023)
von: Suzuki, Tomohiro, et al.
Veröffentlicht: (2023)
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
von: Karlsson, Robin, et al.
Veröffentlicht: (2023)
von: Karlsson, Robin, et al.
Veröffentlicht: (2023)
VIFSS: View-Invariant and Figure Skating-Specific Pose Representation Learning for Temporal Action Segmentation
von: Tanaka, Ryota, et al.
Veröffentlicht: (2025)
von: Tanaka, Ryota, et al.
Veröffentlicht: (2025)
Shot2Tactic-Caption: Multi-Scale Captioning of Badminton Videos for Tactical Understanding
von: Ding, Ning, et al.
Veröffentlicht: (2025)
von: Ding, Ning, et al.
Veröffentlicht: (2025)
BFMD: A Full-Match Badminton Dense Dataset for Dense Shot Captioning
von: Ding, Ning, et al.
Veröffentlicht: (2026)
von: Ding, Ning, et al.
Veröffentlicht: (2026)
PCBEAR: Pose Concept Bottleneck for Explainable Action Recognition
von: Lee, Jongseo, et al.
Veröffentlicht: (2025)
von: Lee, Jongseo, et al.
Veröffentlicht: (2025)
3D Pose-Based Temporal Action Segmentation for Figure Skating: A Fine-Grained and Jump Procedure-Aware Annotation Approach
von: Tanaka, Ryota, et al.
Veröffentlicht: (2024)
von: Tanaka, Ryota, et al.
Veröffentlicht: (2024)
RealTraj: Towards Real-World Pedestrian Trajectory Forecasting
von: Fujii, Ryo, et al.
Veröffentlicht: (2024)
von: Fujii, Ryo, et al.
Veröffentlicht: (2024)
MSCoTDet: Language-driven Multi-modal Fusion for Improved Multispectral Pedestrian Detection
von: Kim, Taeheon, et al.
Veröffentlicht: (2024)
von: Kim, Taeheon, et al.
Veröffentlicht: (2024)
Context-aware Multi-task Learning for Pedestrian Intent and Trajectory Prediction
von: Munir, Farzeen, et al.
Veröffentlicht: (2024)
von: Munir, Farzeen, et al.
Veröffentlicht: (2024)
GDTS: Goal-Guided Diffusion Model with Tree Sampling for Multi-Modal Pedestrian Trajectory Prediction
von: Sun, Ge, et al.
Veröffentlicht: (2023)
von: Sun, Ge, et al.
Veröffentlicht: (2023)
Basketball-SORT: An Association Method for Complex Multi-object Occlusion Problems in Basketball Multi-object Tracking
von: Hu, Qingrui, et al.
Veröffentlicht: (2024)
von: Hu, Qingrui, et al.
Veröffentlicht: (2024)
Disentangled Concepts Speak Louder Than Words: Explainable Video Action Recognition
von: Lee, Jongseo, et al.
Veröffentlicht: (2025)
von: Lee, Jongseo, et al.
Veröffentlicht: (2025)
Fine-grained Action Analysis: A Multi-modality and Multi-task Dataset of Figure Skating
von: Liu, Sheng-Lan, et al.
Veröffentlicht: (2023)
von: Liu, Sheng-Lan, et al.
Veröffentlicht: (2023)
Multi-level and Multi-modal Action Anticipation
von: Kim, Seulgi, et al.
Veröffentlicht: (2025)
von: Kim, Seulgi, et al.
Veröffentlicht: (2025)
Multi-View Pedestrian Occupancy Prediction with a Novel Synthetic Dataset
von: Aung, Sithu, et al.
Veröffentlicht: (2024)
von: Aung, Sithu, et al.
Veröffentlicht: (2024)
A Multi-Stage Goal-Driven Network for Pedestrian Trajectory Prediction
von: Wu, Xiuen, et al.
Veröffentlicht: (2024)
von: Wu, Xiuen, et al.
Veröffentlicht: (2024)
Diving Deeper Into Pedestrian Behavior Understanding: Intention Estimation, Action Prediction, and Event Risk Assessment
von: Rasouli, Amir, et al.
Veröffentlicht: (2024)
von: Rasouli, Amir, et al.
Veröffentlicht: (2024)
MulCLIP: A Multi-level Alignment Framework for Enhancing Fine-grained Long-context CLIP
von: Truong, Chau, et al.
Veröffentlicht: (2025)
von: Truong, Chau, et al.
Veröffentlicht: (2025)
AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models
von: Chen, Haokun, et al.
Veröffentlicht: (2025)
von: Chen, Haokun, et al.
Veröffentlicht: (2025)
Learning the Pedestrian-Vehicle Interaction for Pedestrian Trajectory Prediction
von: Zhang, Chi, et al.
Veröffentlicht: (2022)
von: Zhang, Chi, et al.
Veröffentlicht: (2022)
Space evaluation based on pitch control using drone video in Ultimate
von: Iwashita, Shunsuke, et al.
Veröffentlicht: (2024)
von: Iwashita, Shunsuke, et al.
Veröffentlicht: (2024)
Pedestrian Attribute Recognition as Label-balanced Multi-label Learning
von: Zhou, Yibo, et al.
Veröffentlicht: (2024)
von: Zhou, Yibo, et al.
Veröffentlicht: (2024)
Resonance: Learning to Predict Social-Aware Pedestrian Trajectories as Co-Vibrations
von: Wong, Conghao, et al.
Veröffentlicht: (2024)
von: Wong, Conghao, et al.
Veröffentlicht: (2024)
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
von: Li, Zhe, et al.
Veröffentlicht: (2024)
von: Li, Zhe, et al.
Veröffentlicht: (2024)
SAM2Grasp: Resolve Multi-modal Grasping via Prompt-conditioned Temporal Action Prediction
von: Wu, Shengkai, et al.
Veröffentlicht: (2025)
von: Wu, Shengkai, et al.
Veröffentlicht: (2025)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
von: Arazi, Alan, et al.
Veröffentlicht: (2026)
von: Arazi, Alan, et al.
Veröffentlicht: (2026)
View-aware Cross-modal Distillation for Multi-view Action Recognition
von: Nguyen, Trung Thanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Trung Thanh, et al.
Veröffentlicht: (2025)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
von: Bai, Andrew, et al.
Veröffentlicht: (2025)
von: Bai, Andrew, et al.
Veröffentlicht: (2025)
SocialMOIF: Multi-Order Intention Fusion for Pedestrian Trajectory Prediction
von: Chen, Kai, et al.
Veröffentlicht: (2025)
von: Chen, Kai, et al.
Veröffentlicht: (2025)
SoccerTrack v2: A Full-Pitch Multi-View Soccer Dataset for Game State Reconstruction
von: Scott, Atom, et al.
Veröffentlicht: (2025)
von: Scott, Atom, et al.
Veröffentlicht: (2025)
Pedestrian Trajectory Prediction Based on Social Interactions Learning With Random Weights
von: Xie, Jiajia, et al.
Veröffentlicht: (2025)
von: Xie, Jiajia, et al.
Veröffentlicht: (2025)
Temporal-contextual Event Learning for Pedestrian Crossing Intent Prediction
von: Liang, Hongbin, et al.
Veröffentlicht: (2025)
von: Liang, Hongbin, et al.
Veröffentlicht: (2025)
SoccerSynth Field: enhancing field detection with synthetic data from virtual soccer simulator
von: Qin, HaoBin, et al.
Veröffentlicht: (2025)
von: Qin, HaoBin, et al.
Veröffentlicht: (2025)
Multimodal Cross-Domain Few-Shot Learning for Egocentric Action Recognition
von: Hatano, Masashi, et al.
Veröffentlicht: (2024)
von: Hatano, Masashi, et al.
Veröffentlicht: (2024)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
Attention-Aware Multi-View Pedestrian Tracking
von: Alturki, Reef, et al.
Veröffentlicht: (2025)
von: Alturki, Reef, et al.
Veröffentlicht: (2025)
Fish Tracking Challenge 2024: A Multi-Object Tracking Competition with Sweetfish Schooling Data
von: Itoh, Makoto M., et al.
Veröffentlicht: (2024)
von: Itoh, Makoto M., et al.
Veröffentlicht: (2024)
SCoCCA: Multi-modal Sparse Concept Decomposition via Canonical Correlation Analysis
von: Gordon, Ehud, et al.
Veröffentlicht: (2026)
von: Gordon, Ehud, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Sparse Prototype Network for Explainable Pedestrian Behavior Prediction
von: Feng, Yan, et al.
Veröffentlicht: (2024) -
Runner re-identification from single-view running video in the open-world setting
von: Suzuki, Tomohiro, et al.
Veröffentlicht: (2023) -
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
von: Karlsson, Robin, et al.
Veröffentlicht: (2023) -
VIFSS: View-Invariant and Figure Skating-Specific Pose Representation Learning for Temporal Action Segmentation
von: Tanaka, Ryota, et al.
Veröffentlicht: (2025) -
Shot2Tactic-Caption: Multi-Scale Captioning of Badminton Videos for Tactical Understanding
von: Ding, Ning, et al.
Veröffentlicht: (2025)