ACT360: An Efficient 360-Degree Action Detection and Summarization Framework for Mission-Critical Training and Debriefing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tiwari, Aditi, Nahrstedt, Klara |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
Fire360: A Benchmark for Robust Perception and Episodic Memory in Degraded 360-Degree Firefighting Videos
von: Tiwari, Aditi, et al.
Veröffentlicht: (2025)
von: Tiwari, Aditi, et al.
Veröffentlicht: (2025)
NeRV360: Neural Representation for 360-Degree Videos with a Viewport Decoder
von: Arai, Daichi, et al.
Veröffentlicht: (2025)
von: Arai, Daichi, et al.
Veröffentlicht: (2025)
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
OmniSense: Towards Edge-Assisted Online Analytics for 360-Degree Videos
von: Zhang, Miao, et al.
Veröffentlicht: (2025)
von: Zhang, Miao, et al.
Veröffentlicht: (2025)
Controllable Audio-Visual Viewpoint Generation from 360° Spatial Information
von: Marinoni, Christian, et al.
Veröffentlicht: (2025)
von: Marinoni, Christian, et al.
Veröffentlicht: (2025)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
Panonut360: A Head and Eye Tracking Dataset for Panoramic Video
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
Viewport-based Neural 360° Image Compression
von: Liao, Jingwei, et al.
Veröffentlicht: (2026)
von: Liao, Jingwei, et al.
Veröffentlicht: (2026)
Sphere-GAN: a GAN-based Approach for Saliency Estimation in 360° Videos
von: Wahba, Mahmoud Z. A., et al.
Veröffentlicht: (2025)
von: Wahba, Mahmoud Z. A., et al.
Veröffentlicht: (2025)
Efficiently Collecting Training Dataset for 2D Object Detection by Online Visual Feedback
von: Kiyokawa, Takuya, et al.
Veröffentlicht: (2023)
von: Kiyokawa, Takuya, et al.
Veröffentlicht: (2023)
Training-Free Adaptive 360-degree Video Streaming via Semantic Potential Fields
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
von: Aiersilan, Aizierjiang, et al.
Veröffentlicht: (2026)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
von: Xie, Liping, et al.
Veröffentlicht: (2025)
von: Xie, Liping, et al.
Veröffentlicht: (2025)
Unsupervised Transcript-assisted Video Summarization and Highlight Detection
von: Barbakos, Spyros, et al.
Veröffentlicht: (2025)
von: Barbakos, Spyros, et al.
Veröffentlicht: (2025)
Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
360+x: A Panoptic Multi-modal Scene Understanding Dataset
von: Chen, Hao, et al.
Veröffentlicht: (2024)
von: Chen, Hao, et al.
Veröffentlicht: (2024)
UAU-Net: Uncertainty-aware Representation Learning and Evidential Classification for Facial Action Unit Detection
von: Li, Yuze, et al.
Veröffentlicht: (2026)
von: Li, Yuze, et al.
Veröffentlicht: (2026)
UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos
von: Mei, Yuting, et al.
Veröffentlicht: (2024)
von: Mei, Yuting, et al.
Veröffentlicht: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
Every Painting Awakened: A Training-free Framework for Painting-to-Animation Generation
von: Liu, Lingyu, et al.
Veröffentlicht: (2025)
von: Liu, Lingyu, et al.
Veröffentlicht: (2025)
Lyra: An Efficient and Speech-Centric Framework for Omni-Cognition
von: Zhong, Zhisheng, et al.
Veröffentlicht: (2024)
von: Zhong, Zhisheng, et al.
Veröffentlicht: (2024)
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2025)
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
Human Action Recognition without Human
von: Kataoka, Hirokatsu, et al.
Veröffentlicht: (2016)
von: Kataoka, Hirokatsu, et al.
Veröffentlicht: (2016)
HMPE:HeatMap Embedding for Efficient Transformer-Based Small Object Detection
von: Zeng, YangChen
Veröffentlicht: (2025)
von: Zeng, YangChen
Veröffentlicht: (2025)
Robust Modality-incomplete Anomaly Detection: A Modality-instructive Framework with Benchmark
von: Miao, Bingchen, et al.
Veröffentlicht: (2024)
von: Miao, Bingchen, et al.
Veröffentlicht: (2024)
StableMoFusion: Towards Robust and Efficient Diffusion-based Motion Generation Framework
von: Huang, Yiheng, et al.
Veröffentlicht: (2024)
von: Huang, Yiheng, et al.
Veröffentlicht: (2024)
Noise-Tolerant Learning for Audio-Visual Action Recognition
von: Han, Haochen, et al.
Veröffentlicht: (2022)
von: Han, Haochen, et al.
Veröffentlicht: (2022)
SpecFLASH: A Latent-Guided Semi-autoregressive Speculative Decoding Framework for Efficient Multimodal Generation
von: Wang, Zihua, et al.
Veröffentlicht: (2025)
von: Wang, Zihua, et al.
Veröffentlicht: (2025)
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
von: Guo, Xiaopeng, et al.
Veröffentlicht: (2026)
von: Guo, Xiaopeng, et al.
Veröffentlicht: (2026)
360PanT: Training-Free Text-Driven 360-Degree Panorama-to-Panorama Translation
von: Wang, Hai, et al.
Veröffentlicht: (2024)
von: Wang, Hai, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Unsupervised Video Summarization with Reward Generator Training
von: Abbasi, Mehryar, et al.
Veröffentlicht: (2024)
von: Abbasi, Mehryar, et al.
Veröffentlicht: (2024)
HKD4VLM: A Progressive Hybrid Knowledge Distillation Framework for Robust Multimodal Hallucination and Factuality Detection in VLMs
von: Zhang, Zijian, et al.
Veröffentlicht: (2025)
von: Zhang, Zijian, et al.
Veröffentlicht: (2025)
FeatDistill: A Feature Distillation Enhanced Multi-Expert Ensemble Framework for Robust AI-generated Image Detection
von: Tu, Zhilin, et al.
Veröffentlicht: (2026)
von: Tu, Zhilin, et al.
Veröffentlicht: (2026)
CASR: Refining Action Segmentation via Marginalizing Frame-levle Causal Relationships
von: Du, Keqing, et al.
Veröffentlicht: (2023)
von: Du, Keqing, et al.
Veröffentlicht: (2023)
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
TSalV360: A Method and Dataset for Text-driven Saliency Detection in 360-Degrees Videos
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2025)
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2025)
Wavelet-Decoupling Contrastive Enhancement Network for Fine-Grained Skeleton-Based Action Recognition
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
von: Huang, Shunyu, et al.
Veröffentlicht: (2026)
von: Huang, Shunyu, et al.
Veröffentlicht: (2026)
Enhancing Self-Supervised Talking Head Forgery Detection via a Training-Free Dual-System Framework
von: Liu, Ke, et al.
Veröffentlicht: (2026)
von: Liu, Ke, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024) -
Fire360: A Benchmark for Robust Perception and Episodic Memory in Degraded 360-Degree Firefighting Videos
von: Tiwari, Aditi, et al.
Veröffentlicht: (2025) -
NeRV360: Neural Representation for 360-Degree Videos with a Viewport Decoder
von: Arai, Daichi, et al.
Veröffentlicht: (2025) -
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
von: Li, Wenrui, et al.
Veröffentlicht: (2024) -
OmniSense: Towards Edge-Assisted Online Analytics for 360-Degree Videos
von: Zhang, Miao, et al.
Veröffentlicht: (2025)