Action-Agnostic Point-Level Supervision for Temporal Action Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Yoshida, Shuhei M., Shibata, Takashi, Terao, Makoto, Okatani, Takayuki, Sugiyama, Masashi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MS-DPPs: Multi-Source Determinantal Point Processes for Contextual Diversity Refinement of Composite Attributes in Text to Image Retrieval
by: Sogi, Naoya, et al.
Published: (2025)
by: Sogi, Naoya, et al.
Published: (2025)
Object-Aware Query Perturbation for Cross-Modal Image-Text Retrieval
by: Sogi, Naoya, et al.
Published: (2024)
by: Sogi, Naoya, et al.
Published: (2024)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
by: Xu, Jie, et al.
Published: (2025)
by: Xu, Jie, et al.
Published: (2025)
Latent Action Learning Requires Supervision in the Presence of Distractors
by: Nikulin, Alexander, et al.
Published: (2025)
by: Nikulin, Alexander, et al.
Published: (2025)
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
by: Sun, Li, et al.
Published: (2024)
by: Sun, Li, et al.
Published: (2024)
LAMS-Edit: Latent and Attention Mixing with Schedulers for Improved Content Preservation in Diffusion-Based Image and Style Editing
by: Fu, Wingwa, et al.
Published: (2026)
by: Fu, Wingwa, et al.
Published: (2026)
Towards Generalizing Temporal Action Segmentation to Unseen Views
by: Bahrami, Emad, et al.
Published: (2025)
by: Bahrami, Emad, et al.
Published: (2025)
Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning
by: Ganai, Milan, et al.
Published: (2026)
by: Ganai, Milan, et al.
Published: (2026)
One-Shot Machine Unlearning with Mnemonic Code
by: Yamashita, Tomoya, et al.
Published: (2023)
by: Yamashita, Tomoya, et al.
Published: (2023)
Towards Precise Action Spotting: Addressing Temporal Misalignment in Labels with Dynamic Label Assignment
by: Tamura, Masato
Published: (2025)
by: Tamura, Masato
Published: (2025)
Learning Contrastive Feature Representations for Facial Action Unit Detection
by: Shang, Ziqiao, et al.
Published: (2024)
by: Shang, Ziqiao, et al.
Published: (2024)
ActionParty: Multi-Subject Action Binding in Generative Video Games
by: Pondaven, Alexander, et al.
Published: (2026)
by: Pondaven, Alexander, et al.
Published: (2026)
Video Action Differencing
by: Burgess, James, et al.
Published: (2025)
by: Burgess, James, et al.
Published: (2025)
Impact of Noisy Supervision in Foundation Model Learning
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Survey of Action Recognition, Spotting and Spatio-Temporal Localization in Soccer -- Current Trends and Research Perspectives
by: Seweryn, Karolina, et al.
Published: (2023)
by: Seweryn, Karolina, et al.
Published: (2023)
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
by: Wang, Zhijie, et al.
Published: (2022)
by: Wang, Zhijie, et al.
Published: (2022)
2D-3D Interlaced Transformer for Point Cloud Segmentation with Scene-Level Supervision
by: Yang, Cheng-Kun, et al.
Published: (2023)
by: Yang, Cheng-Kun, et al.
Published: (2023)
Semantically Guided Action Anticipation
by: Diko, Anxhelo, et al.
Published: (2024)
by: Diko, Anxhelo, et al.
Published: (2024)
Action-Inspired Generative Models
by: A., Eshwar R., et al.
Published: (2026)
by: A., Eshwar R., et al.
Published: (2026)
Multimodal Attack Detection for Action Recognition Models
by: Mumcu, Furkan, et al.
Published: (2024)
by: Mumcu, Furkan, et al.
Published: (2024)
Real-Time Human Action Recognition on Embedded Platforms
by: Wang, Ruiqi, et al.
Published: (2024)
by: Wang, Ruiqi, et al.
Published: (2024)
Feature Hallucination for Self-supervised Action Recognition
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Evolving Skeletons: Motion Dynamics in Action Recognition
by: Qiu, Jushang, et al.
Published: (2025)
by: Qiu, Jushang, et al.
Published: (2025)
Zero-Shot Action Generalization with Limited Observations
by: Alchihabi, Abdullah, et al.
Published: (2025)
by: Alchihabi, Abdullah, et al.
Published: (2025)
Semantically Guided Representation Learning For Action Anticipation
by: Diko, Anxhelo, et al.
Published: (2024)
by: Diko, Anxhelo, et al.
Published: (2024)
From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
by: Zhang, Zhengshen, et al.
Published: (2025)
by: Zhang, Zhengshen, et al.
Published: (2025)
TTF-VLA: Temporal Token Fusion via Pixel-Attention Integration for Vision-Language-Action Models
by: Liu, Chenghao, et al.
Published: (2025)
by: Liu, Chenghao, et al.
Published: (2025)
VISAGE: Video Synthesis using Action Graphs for Surgery
by: Yeganeh, Yousef, et al.
Published: (2024)
by: Yeganeh, Yousef, et al.
Published: (2024)
About Time: Advances, Challenges, and Outlooks of Action Understanding
by: Stergiou, Alexandros, et al.
Published: (2024)
by: Stergiou, Alexandros, et al.
Published: (2024)
Direct Distillation between Different Domains
by: Tang, Jialiang, et al.
Published: (2024)
by: Tang, Jialiang, et al.
Published: (2024)
Explainability of Point Cloud Neural Networks Using SMILE: Statistical Model-Agnostic Interpretability with Local Explanations
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
Vision-Language Models Unlock Task-Centric Latent Actions
by: Nikulin, Alexander, et al.
Published: (2026)
by: Nikulin, Alexander, et al.
Published: (2026)
Olaf-World: Orienting Latent Actions for Video World Modeling
by: Jiang, Yuxin, et al.
Published: (2026)
by: Jiang, Yuxin, et al.
Published: (2026)
EgoBrain: Synergizing Minds and Eyes For Human Action Understanding
by: Lin, Nie, et al.
Published: (2025)
by: Lin, Nie, et al.
Published: (2025)
Do We Need Large VLMs for Spotting Soccer Actions?
by: Chakraborty, Ritabrata, et al.
Published: (2025)
by: Chakraborty, Ritabrata, et al.
Published: (2025)
Embracing Biased Transition Matrices for Complementary-Label Learning with Many Classes
by: Mai, Tan-Ha, et al.
Published: (2026)
by: Mai, Tan-Ha, et al.
Published: (2026)
Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks
by: Chen, Hao, et al.
Published: (2023)
by: Chen, Hao, et al.
Published: (2023)
Learning to Visually Connect Actions and their Effects
by: Parmar, Paritosh, et al.
Published: (2024)
by: Parmar, Paritosh, et al.
Published: (2024)
Boosting Facial Action Unit Detection Through Jointly Learning Facial Landmark Detection and Domain Separation and Reconstruction
by: Shang, Ziqiao, et al.
Published: (2023)
by: Shang, Ziqiao, et al.
Published: (2023)
Multi-Label Knowledge Distillation
by: Yang, Penghui, et al.
Published: (2023)
by: Yang, Penghui, et al.
Published: (2023)
Similar Items
-
MS-DPPs: Multi-Source Determinantal Point Processes for Contextual Diversity Refinement of Composite Attributes in Text to Image Retrieval
by: Sogi, Naoya, et al.
Published: (2025) -
Object-Aware Query Perturbation for Cross-Modal Image-Text Retrieval
by: Sogi, Naoya, et al.
Published: (2024) -
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
by: Xu, Jie, et al.
Published: (2025) -
Latent Action Learning Requires Supervision in the Presence of Distractors
by: Nikulin, Alexander, et al.
Published: (2025) -
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
by: Sun, Li, et al.
Published: (2024)