Visual Self-paced Iterative Learning for Unsupervised Temporal Action Localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Yupeng, Jiang, Han, Liu, Hao, Wang, Kun, Tang, Haoyu, Nie, Liqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Uncovering Hidden Connections: Iterative Search and Reasoning for Video-grounded Dialog
von: Zhang, Haoyu, et al.
Veröffentlicht: (2023)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2023)
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
Technical Report for Ego4D Long-Term Action Anticipation Challenge 2025
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
Open-Vocabulary Action Localization with Iterative Visual Prompting
von: Wake, Naoki, et al.
Veröffentlicht: (2024)
von: Wake, Naoki, et al.
Veröffentlicht: (2024)
Exploring Scalability of Self-Training for Open-Vocabulary Temporal Action Localization
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2024)
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2024)
FineCIR: Explicit Parsing of Fine-Grained Modification Semantics for Composed Image Retrieval
von: Li, Zixu, et al.
Veröffentlicht: (2025)
von: Li, Zixu, et al.
Veröffentlicht: (2025)
Object-Shot Enhanced Grounding Network for Egocentric Video
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
Advancing Brain Imaging Analysis Step-by-step via Progressive Self-paced Learning
von: Yang, Yanwu, et al.
Veröffentlicht: (2024)
von: Yang, Yanwu, et al.
Veröffentlicht: (2024)
Saliency-Guided Representation with Consistency Policy Learning for Visual Unsupervised Reinforcement Learning
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
VISTA: Technical Report for the Ego4D Short-Term Object Interaction Anticipation at EgoVis 2026
von: Chu, Qiaohui, et al.
Veröffentlicht: (2026)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2026)
The Solution for Temporal Action Localisation Task of Perception Test Challenge 2024
von: Han, Yinan, et al.
Veröffentlicht: (2024)
von: Han, Yinan, et al.
Veröffentlicht: (2024)
HCQA @ Ego4D EgoSchema Challenge 2024
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
ObjectNLQ @ Ego4D Episodic Memory Challenge 2024
von: Feng, Yisen, et al.
Veröffentlicht: (2024)
von: Feng, Yisen, et al.
Veröffentlicht: (2024)
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
von: Xu, Ziqiang, et al.
Veröffentlicht: (2025)
von: Xu, Ziqiang, et al.
Veröffentlicht: (2025)
Idempotent Unsupervised Representation Learning for Skeleton-Based Action Recognition
von: Lin, Lilang, et al.
Veröffentlicht: (2024)
von: Lin, Lilang, et al.
Veröffentlicht: (2024)
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
Spatial Understanding from Videos: Structured Prompts Meet Simulation Data
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
OSGNet @ Ego4D Episodic Memory Challenge 2025
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
HCQA-1.5 @ Ego4D EgoSchema Challenge 2025
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
CloDS: Visual-Only Unsupervised Cloth Dynamics Learning in Unknown Conditions
von: Zhan, Yuliang, et al.
Veröffentlicht: (2026)
von: Zhan, Yuliang, et al.
Veröffentlicht: (2026)
SHAPE : Self-Improved Visual Preference Alignment by Iteratively Generating Holistic Winner
von: Chen, Kejia, et al.
Veröffentlicht: (2025)
von: Chen, Kejia, et al.
Veröffentlicht: (2025)
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
von: Fu, Zhiheng, et al.
Veröffentlicht: (2026)
von: Fu, Zhiheng, et al.
Veröffentlicht: (2026)
Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
Open Multimodal Retrieval-Augmented Factual Image Generation
von: Tian, Yang, et al.
Veröffentlicht: (2025)
von: Tian, Yang, et al.
Veröffentlicht: (2025)
Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space
von: Chen, Chao, et al.
Veröffentlicht: (2025)
von: Chen, Chao, et al.
Veröffentlicht: (2025)
Cluster Contrast for Unsupervised Visual Representation Learning
von: Giakoumoglou, Nikolaos, et al.
Veröffentlicht: (2025)
von: Giakoumoglou, Nikolaos, et al.
Veröffentlicht: (2025)
Unsupervised Domain Adaptation for Action Recognition via Self-Ensembling and Conditional Embedding Alignment
von: Ghosh, Indrajeet, et al.
Veröffentlicht: (2024)
von: Ghosh, Indrajeet, et al.
Veröffentlicht: (2024)
MaeFuse: Transferring Omni Features with Pretrained Masked Autoencoders for Infrared and Visible Image Fusion via Guided Training
von: Li, Jiayang, et al.
Veröffentlicht: (2024)
von: Li, Jiayang, et al.
Veröffentlicht: (2024)
VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies
von: Gao, Mingjian, et al.
Veröffentlicht: (2026)
von: Gao, Mingjian, et al.
Veröffentlicht: (2026)
Iterative Feedback Network for Unsupervised Point Cloud Registration
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
Chain-of-Evidence Multimodal Reasoning for Few-shot Temporal Action Localization
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Enhancing Temporal Action Localization: Advanced S6 Modeling with Recurrent Mechanism
von: Lee, Sangyoun, et al.
Veröffentlicht: (2024)
von: Lee, Sangyoun, et al.
Veröffentlicht: (2024)
Towards Unified Semantic and Controllable Image Fusion: A Diffusion Transformer Approach
von: Li, Jiayang, et al.
Veröffentlicht: (2025)
von: Li, Jiayang, et al.
Veröffentlicht: (2025)
STAR: Semantic-Temporal Adaptive Representation Learning for Few-Shot Action Recognition
von: Liu, Hongli, et al.
Veröffentlicht: (2026)
von: Liu, Hongli, et al.
Veröffentlicht: (2026)
Online Iterative Self-Alignment for Radiology Report Generation
von: Xiao, Ting, et al.
Veröffentlicht: (2025)
von: Xiao, Ting, et al.
Veröffentlicht: (2025)
Unsupervised Synthetic Image Attribution: Alignment and Disentanglement
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
Think Twice to See More: Iterative Visual Reasoning in Medical VLMs
von: Chen, Kaitao, et al.
Veröffentlicht: (2025)
von: Chen, Kaitao, et al.
Veröffentlicht: (2025)
Leveraging Unsupervised Learning for Cost-Effective Visual Anomaly Detection
von: Long, Yunbo, et al.
Veröffentlicht: (2024)
von: Long, Yunbo, et al.
Veröffentlicht: (2024)
TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action
von: Cheng, Jen-Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Jen-Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Uncovering Hidden Connections: Iterative Search and Reasoning for Video-grounded Dialog
von: Zhang, Haoyu, et al.
Veröffentlicht: (2023) -
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025) -
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
von: Li, Zaijing, et al.
Veröffentlicht: (2026) -
Technical Report for Ego4D Long-Term Action Anticipation Challenge 2025
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025) -
Open-Vocabulary Action Localization with Iterative Visual Prompting
von: Wake, Naoki, et al.
Veröffentlicht: (2024)