About Time: Advances, Challenges, and Outlooks of Action Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Stergiou, Alexandros, Poppe, Ronald |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Traffic Anomalies from Generative Models on Real-Time Observations
by: Giasemis, Fotis I., et al.
Published: (2025)
by: Giasemis, Fotis I., et al.
Published: (2025)
EgoBrain: Synergizing Minds and Eyes For Human Action Understanding
by: Lin, Nie, et al.
Published: (2025)
by: Lin, Nie, et al.
Published: (2025)
From Isolated Islands to Pangea: Unifying Semantic Space for Human Action Understanding
by: Li, Yong-Lu, et al.
Published: (2023)
by: Li, Yong-Lu, et al.
Published: (2023)
Real-Time Human Action Recognition on Embedded Platforms
by: Wang, Ruiqi, et al.
Published: (2024)
by: Wang, Ruiqi, et al.
Published: (2024)
The Promise of Analog Deep Learning: Recent Advances, Challenges and Opportunities
by: Datar, Aditya, et al.
Published: (2024)
by: Datar, Aditya, et al.
Published: (2024)
Do Language Models Understand Time?
by: Ding, Xi, et al.
Published: (2024)
by: Ding, Xi, et al.
Published: (2024)
EITNet: An IoT-Enhanced Framework for Real-Time Basketball Action Recognition
by: Liu, Jingyu, et al.
Published: (2024)
by: Liu, Jingyu, et al.
Published: (2024)
VideoRefer Suite: Advancing Spatial-Temporal Object Understanding with Video LLM
by: Yuan, Yuqian, et al.
Published: (2024)
by: Yuan, Yuqian, et al.
Published: (2024)
SAR-RARP50: Segmentation of surgical instrumentation and Action Recognition on Robot-Assisted Radical Prostatectomy Challenge
by: Psychogyios, Dimitrios, et al.
Published: (2023)
by: Psychogyios, Dimitrios, et al.
Published: (2023)
Towards Generalisable Time Series Understanding Across Domains
by: Turgut, Özgün, et al.
Published: (2024)
by: Turgut, Özgün, et al.
Published: (2024)
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)
by: Doumanoglou, Alexandros, et al.
Published: (2023)
Consistent Diffusion Meets Tweedie: Training Exact Ambient Diffusion Models with Noisy Data
by: Daras, Giannis, et al.
Published: (2024)
by: Daras, Giannis, et al.
Published: (2024)
PRISM: Distributed Inference for Foundation Models at Edge
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
OpenSUN3D: 1st Workshop Challenge on Open-Vocabulary 3D Scene Understanding
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
Action-Agnostic Point-Level Supervision for Temporal Action Detection
by: Yoshida, Shuhei M., et al.
Published: (2024)
by: Yoshida, Shuhei M., et al.
Published: (2024)
ActionParty: Multi-Subject Action Binding in Generative Video Games
by: Pondaven, Alexander, et al.
Published: (2026)
by: Pondaven, Alexander, et al.
Published: (2026)
OUI Need to Talk About Weight Decay: A New Perspective on Overfitting Detection
by: Fernández-Hernández, Alberto, et al.
Published: (2025)
by: Fernández-Hernández, Alberto, et al.
Published: (2025)
Video Action Differencing
by: Burgess, James, et al.
Published: (2025)
by: Burgess, James, et al.
Published: (2025)
Feature-Based Instance Neighbor Discovery: Advanced Stable Test-Time Adaptation in Dynamic World
by: Jiang, Qinting, et al.
Published: (2025)
by: Jiang, Qinting, et al.
Published: (2025)
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
by: Agnur, Bharath Kumar
Published: (2024)
by: Agnur, Bharath Kumar
Published: (2024)
Semantically Guided Action Anticipation
by: Diko, Anxhelo, et al.
Published: (2024)
by: Diko, Anxhelo, et al.
Published: (2024)
Action-Inspired Generative Models
by: A., Eshwar R., et al.
Published: (2026)
by: A., Eshwar R., et al.
Published: (2026)
Task adaptation of Vision-Language-Action model: 1st Place Solution for the 2025 BEHAVIOR Challenge
by: Larchenko, Ilia, et al.
Published: (2025)
by: Larchenko, Ilia, et al.
Published: (2025)
Semantically Guided Representation Learning For Action Anticipation
by: Diko, Anxhelo, et al.
Published: (2024)
by: Diko, Anxhelo, et al.
Published: (2024)
Feature Hallucination for Self-supervised Action Recognition
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Evolving Skeletons: Motion Dynamics in Action Recognition
by: Qiu, Jushang, et al.
Published: (2025)
by: Qiu, Jushang, et al.
Published: (2025)
Zero-Shot Action Generalization with Limited Observations
by: Alchihabi, Abdullah, et al.
Published: (2025)
by: Alchihabi, Abdullah, et al.
Published: (2025)
Advancing AI-Powered Medical Image Synthesis: Insights from MedVQA-GI Challenge Using CLIP, Fine-Tuned Stable Diffusion, and Dream-Booth + LoRA
by: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Published: (2025)
by: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Published: (2025)
FLIP Reasoning Challenge
by: Plesner, Andreas, et al.
Published: (2025)
by: Plesner, Andreas, et al.
Published: (2025)
VISAGE: Video Synthesis using Action Graphs for Surgery
by: Yeganeh, Yousef, et al.
Published: (2024)
by: Yeganeh, Yousef, et al.
Published: (2024)
Latent Action Learning Requires Supervision in the Presence of Distractors
by: Nikulin, Alexander, et al.
Published: (2025)
by: Nikulin, Alexander, et al.
Published: (2025)
Towards Generalizing Temporal Action Segmentation to Unseen Views
by: Bahrami, Emad, et al.
Published: (2025)
by: Bahrami, Emad, et al.
Published: (2025)
LAVIB: A Large-scale Video Interpolation Benchmark
by: Stergiou, Alexandros
Published: (2024)
by: Stergiou, Alexandros
Published: (2024)
TRANSPORTER: Transferring Visual Semantics from VLM Manifolds
by: Stergiou, Alexandros
Published: (2025)
by: Stergiou, Alexandros
Published: (2025)
Subspace-Boosted Model Merging
by: Skorobogat, Ronald, et al.
Published: (2025)
by: Skorobogat, Ronald, et al.
Published: (2025)
CLIP Can Understand Depth
by: Kim, Sohee, et al.
Published: (2024)
by: Kim, Sohee, et al.
Published: (2024)
Understanding Multi-View Transformers
by: Stary, Michal, et al.
Published: (2025)
by: Stary, Michal, et al.
Published: (2025)
Discover Your Neighbors: Advanced Stable Test-Time Adaptation in Dynamic World
by: Jiang, Qinting, et al.
Published: (2024)
by: Jiang, Qinting, et al.
Published: (2024)
Learning Contrastive Feature Representations for Facial Action Unit Detection
by: Shang, Ziqiao, et al.
Published: (2024)
by: Shang, Ziqiao, et al.
Published: (2024)
See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
Similar Items
-
Learning Traffic Anomalies from Generative Models on Real-Time Observations
by: Giasemis, Fotis I., et al.
Published: (2025) -
EgoBrain: Synergizing Minds and Eyes For Human Action Understanding
by: Lin, Nie, et al.
Published: (2025) -
From Isolated Islands to Pangea: Unifying Semantic Space for Human Action Understanding
by: Li, Yong-Lu, et al.
Published: (2023) -
Real-Time Human Action Recognition on Embedded Platforms
by: Wang, Ruiqi, et al.
Published: (2024) -
The Promise of Analog Deep Learning: Recent Advances, Challenges and Opportunities
by: Datar, Aditya, et al.
Published: (2024)