T-MASK: Temporal Masking for Probing Foundation Models across Camera Views in Driver Monitoring
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ponbagavathi, Thinesh Thiyakesan, Peng, Kunyu, Roitberg, Alina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Probing Fine-Grained Action Understanding and Cross-View Generalization of Foundation Models
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2024)
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2024)
Order Matters: On Parameter-Efficient Image-to-Video Probing for Recognizing Nearly Symmetric Actions
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025)
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025)
Structured Relational Reasoning for Group Activity Assessment
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025)
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025)
Frame2Freq: Spectral Adapters for Fine-Grained Video Understanding
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2026)
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2026)
Deep Learning for Metabolic Rate Estimation from Biosignals: A Comparative Study of Architectures and Signal Selection
von: Babakhani, Sarvenaz, et al.
Veröffentlicht: (2025)
von: Babakhani, Sarvenaz, et al.
Veröffentlicht: (2025)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
Advancing Open-Set Domain Generalization Using Evidential Bi-Level Hardest Domain Scheduler
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
Robust Multiview Multimodal Driver Monitoring System Using Masked Multi-Head Self-Attention
von: Ma, Yiming, et al.
Veröffentlicht: (2023)
von: Ma, Yiming, et al.
Veröffentlicht: (2023)
[MASK] is All You Need
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
Probing into Camera Control of Video Models
von: Hou, Chen, et al.
Veröffentlicht: (2026)
von: Hou, Chen, et al.
Veröffentlicht: (2026)
Camera Perspective Transformation to Bird's Eye View via Spatial Transformer Model for Road Intersection Monitoring
von: Prajapati, Rukesh, et al.
Veröffentlicht: (2024)
von: Prajapati, Rukesh, et al.
Veröffentlicht: (2024)
Occlusion-aware Driver Monitoring System using the Driver Monitoring Dataset
von: Cañas, Paola Natalia, et al.
Veröffentlicht: (2025)
von: Cañas, Paola Natalia, et al.
Veröffentlicht: (2025)
Towards Activated Muscle Group Estimation in the Wild
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
Towards Temporal Fusion Beyond the Field of View for Camera-based Semantic Scene Completion
von: Bae, Jongseong, et al.
Veröffentlicht: (2025)
von: Bae, Jongseong, et al.
Veröffentlicht: (2025)
EventMamba: Enhancing Spatio-Temporal Locality with State Space Models for Event-Based Video Reconstruction
von: Ge, Chengjie, et al.
Veröffentlicht: (2025)
von: Ge, Chengjie, et al.
Veröffentlicht: (2025)
MC-BEVRO: Multi-Camera Bird Eye View Road Occupancy Detection for Traffic Monitoring
von: Vaghela, Arpitsinh, et al.
Veröffentlicht: (2025)
von: Vaghela, Arpitsinh, et al.
Veröffentlicht: (2025)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
Vision-language Models for Driver Monitoring Systems: A Driver Activity Description Dataset
von: Lerch, David J., et al.
Veröffentlicht: (2026)
von: Lerch, David J., et al.
Veröffentlicht: (2026)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
Multi-View Foundation Models
von: Segre, Leo, et al.
Veröffentlicht: (2025)
von: Segre, Leo, et al.
Veröffentlicht: (2025)
Skeleton-Based Human Action Recognition with Noisy Labels
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Towards Synthetic Data Generation for Improved Pain Recognition in Videos under Patient Constraints
von: Nasimzada, Jonas, et al.
Veröffentlicht: (2024)
von: Nasimzada, Jonas, et al.
Veröffentlicht: (2024)
Cross-View Cross-Modal Unsupervised Domain Adaptation for Driver Monitoring System
von: Bhalla, Aditi, et al.
Veröffentlicht: (2025)
von: Bhalla, Aditi, et al.
Veröffentlicht: (2025)
Dynamic View Synthesis from Small Camera Motion Videos
von: Sun, Huiqiang, et al.
Veröffentlicht: (2025)
von: Sun, Huiqiang, et al.
Veröffentlicht: (2025)
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
von: Zhou, Jensen, et al.
Veröffentlicht: (2025)
von: Zhou, Jensen, et al.
Veröffentlicht: (2025)
Camera Splatting for Continuous View Optimization
von: Lee, Gahye, et al.
Veröffentlicht: (2025)
von: Lee, Gahye, et al.
Veröffentlicht: (2025)
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
MV2MAE: Multi-View Video Masked Autoencoders
von: Shah, Ketul, et al.
Veröffentlicht: (2024)
von: Shah, Ketul, et al.
Veröffentlicht: (2024)
Multiview Self-Representation Learning across Heterogeneous Views
von: Chen, Jie, et al.
Veröffentlicht: (2026)
von: Chen, Jie, et al.
Veröffentlicht: (2026)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
OnlineBEV: Recurrent Temporal Fusion in Bird's Eye View Representations for Multi-Camera 3D Perception
von: Koh, Junho, et al.
Veröffentlicht: (2025)
von: Koh, Junho, et al.
Veröffentlicht: (2025)
CalibAnyView: Beyond Single-View Camera Calibration in the Wild
von: Li, Boying, et al.
Veröffentlicht: (2026)
von: Li, Boying, et al.
Veröffentlicht: (2026)
Boosting Multi-View Stereo with Depth Foundation Model in the Absence of Real-World Labels
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
Asymmetric Masked Distillation for Pre-Training Small Foundation Models
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2023)
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2023)
Temporal-Mapping Photography for Event Cameras
von: Bao, Yuhan, et al.
Veröffentlicht: (2024)
von: Bao, Yuhan, et al.
Veröffentlicht: (2024)
MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
von: Zhu, Ruishu, et al.
Veröffentlicht: (2025)
von: Zhu, Ruishu, et al.
Veröffentlicht: (2025)
UMAMI: Unifying Masked Autoregressive Models and Deterministic Rendering for View Synthesis
von: Le, Thanh-Tung, et al.
Veröffentlicht: (2025)
von: Le, Thanh-Tung, et al.
Veröffentlicht: (2025)
Mitigating Label Noise using Prompt-Based Hyperbolic Meta-Learning in Open-Set Domain Generalization
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
AnyCalib: On-Manifold Learning for Model-Agnostic Single-View Camera Calibration
von: Tirado-Garín, Javier, et al.
Veröffentlicht: (2025)
von: Tirado-Garín, Javier, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Probing Fine-Grained Action Understanding and Cross-View Generalization of Foundation Models
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2024) -
Order Matters: On Parameter-Efficient Image-to-Video Probing for Recognizing Nearly Symmetric Actions
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025) -
Structured Relational Reasoning for Group Activity Assessment
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2025) -
Frame2Freq: Spectral Adapters for Fine-Grained Video Understanding
von: Ponbagavathi, Thinesh Thiyakesan, et al.
Veröffentlicht: (2026) -
Deep Learning for Metabolic Rate Estimation from Biosignals: A Comparative Study of Architectures and Signal Selection
von: Babakhani, Sarvenaz, et al.
Veröffentlicht: (2025)