Memory Storyboard: Leveraging Temporal Segmentation for Streaming Self-Supervised Learning from Egocentric Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yanlai, Ren, Mengye |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos
by: Wang, Ying, et al.
Published: (2023)
by: Wang, Ying, et al.
Published: (2023)
Seeking the Unfamiliar but Memorable: Conceptual Creativity as Meta-Learning
by: Ren, Mengye
Published: (2026)
by: Ren, Mengye
Published: (2026)
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
by: Yang, Yanlai, et al.
Published: (2025)
by: Yang, Yanlai, et al.
Published: (2025)
Midway Network: Learning Representations for Recognition and Motion from Latent Dynamics
by: Hoang, Christopher, et al.
Published: (2025)
by: Hoang, Christopher, et al.
Published: (2025)
MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents
by: Kim, Kangsan, et al.
Published: (2026)
by: Kim, Kangsan, et al.
Published: (2026)
Integrating Present and Past in Unsupervised Continual Learning
by: Zhang, Yipeng, et al.
Published: (2024)
by: Zhang, Yipeng, et al.
Published: (2024)
Learning from Memory: Non-Parametric Memory Augmented Self-Supervised Learning of Visual Features
by: Silva, Thalles, et al.
Published: (2024)
by: Silva, Thalles, et al.
Published: (2024)
Opinion: Learning Intuitive Physics May Require More than Visual Data
by: Su, Ellen, et al.
Published: (2025)
by: Su, Ellen, et al.
Published: (2025)
Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design
by: Engel, Dominik, et al.
Published: (2023)
by: Engel, Dominik, et al.
Published: (2023)
Temporal-consistent CAMs for Weakly Supervised Video Segmentation in Waste Sorting
by: Marelli, Andrea, et al.
Published: (2025)
by: Marelli, Andrea, et al.
Published: (2025)
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
by: Bhalgat, Yash, et al.
Published: (2024)
by: Bhalgat, Yash, et al.
Published: (2024)
Selective Masking based Self-Supervised Learning for Image Semantic Segmentation
by: Wang, Yuemin, et al.
Published: (2025)
by: Wang, Yuemin, et al.
Published: (2025)
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos
by: Murtaza, Shakeeb, et al.
Published: (2024)
by: Murtaza, Shakeeb, et al.
Published: (2024)
Self-Supervised Discriminative Feature Learning for Deep Multi-View Clustering
by: Xu, Jie, et al.
Published: (2021)
by: Xu, Jie, et al.
Published: (2021)
Story2Board: A Training-Free Approach for Expressive Storyboard Generation
by: Dinkevich, David, et al.
Published: (2025)
by: Dinkevich, David, et al.
Published: (2025)
ChronoSelect: Robust Learning with Noisy Labels via Dynamics Temporal Memory
by: Wang, Jianchao, et al.
Published: (2025)
by: Wang, Jianchao, et al.
Published: (2025)
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding
by: Nagrani, Arsha, et al.
Published: (2026)
by: Nagrani, Arsha, et al.
Published: (2026)
DINOv2 based Self Supervised Learning For Few Shot Medical Image Segmentation
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
by: Hoque, Ryan, et al.
Published: (2025)
by: Hoque, Ryan, et al.
Published: (2025)
TESPEC: Temporally-Enhanced Self-Supervised Pretraining for Event Cameras
by: Mohammadi, Mohammad, et al.
Published: (2025)
by: Mohammadi, Mohammad, et al.
Published: (2025)
PooDLe: Pooled and dense self-supervised learning from naturalistic videos
by: Wang, Alex N., et al.
Published: (2024)
by: Wang, Alex N., et al.
Published: (2024)
Segment Anything without Supervision
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
Ego-VPA: Egocentric Video Understanding with Parameter-efficient Adaptation
by: Wu, Tz-Ying, et al.
Published: (2024)
by: Wu, Tz-Ying, et al.
Published: (2024)
Variational Self-Supervised Learning
by: Yavuz, Mehmet Can, et al.
Published: (2025)
by: Yavuz, Mehmet Can, et al.
Published: (2025)
PSPU: Enhanced Positive and Unlabeled Learning by Leveraging Pseudo Supervision
by: Wang, Chengjie, et al.
Published: (2024)
by: Wang, Chengjie, et al.
Published: (2024)
RADLER: Radar Object Detection Leveraging Semantic 3D City Models and Self-Supervised Radar-Image Learning
by: Luo, Yuan, et al.
Published: (2025)
by: Luo, Yuan, et al.
Published: (2025)
Improving Video Instance Segmentation by Light-weight Temporal Uncertainty Estimates
by: Maag, Kira, et al.
Published: (2020)
by: Maag, Kira, et al.
Published: (2020)
Lifelong Learning of Video Diffusion Models From a Single Video Stream
by: Yoo, Jason, et al.
Published: (2024)
by: Yoo, Jason, et al.
Published: (2024)
PRCL: Probabilistic Representation Contrastive Learning for Semi-Supervised Semantic Segmentation
by: Xie, Haoyu, et al.
Published: (2024)
by: Xie, Haoyu, et al.
Published: (2024)
Self-Supervised Learning by Curvature Alignment
by: Ghojogh, Benyamin, et al.
Published: (2025)
by: Ghojogh, Benyamin, et al.
Published: (2025)
Leveraging Label Proportion Prior for Class-Imbalanced Semi-Supervised Learning
by: Akiba, Kohki, et al.
Published: (2026)
by: Akiba, Kohki, et al.
Published: (2026)
EgoSurgery-HTS: A Dataset for Egocentric Hand-Tool Segmentation in Open Surgery Videos
by: Darjana, Nathan, et al.
Published: (2025)
by: Darjana, Nathan, et al.
Published: (2025)
StreamDiffusionV2: A Streaming System for Dynamic and Interactive Video Generation
by: Feng, Tianrui, et al.
Published: (2025)
by: Feng, Tianrui, et al.
Published: (2025)
Self-Supervised Moving Object Segmentation of Sparse and Noisy Radar Point Clouds
by: Schwarzer, Leon, et al.
Published: (2025)
by: Schwarzer, Leon, et al.
Published: (2025)
The Impact of Semi-Supervised Learning on Line Segment Detection
by: Engman, Johanna, et al.
Published: (2024)
by: Engman, Johanna, et al.
Published: (2024)
Learning to Segment using Summary Statistics and Weak Supervision
by: Kulkarni, Omkar, et al.
Published: (2026)
by: Kulkarni, Omkar, et al.
Published: (2026)
Weakly-Supervised Anomaly Detection in Surveillance Videos Based on Two-Stream I3D Convolution Network
by: Nejad, Sareh Soltani, et al.
Published: (2024)
by: Nejad, Sareh Soltani, et al.
Published: (2024)
Self-Supervised Disentanglement by Leveraging Structure in Data Augmentations
by: Eastwood, Cian, et al.
Published: (2023)
by: Eastwood, Cian, et al.
Published: (2023)
Generalized Semi-Supervised Learning via Self-Supervised Feature Adaptation
by: Liang, Jiachen, et al.
Published: (2024)
by: Liang, Jiachen, et al.
Published: (2024)
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
by: Peirone, Simone Alberto, et al.
Published: (2024)
by: Peirone, Simone Alberto, et al.
Published: (2024)
Similar Items
-
LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos
by: Wang, Ying, et al.
Published: (2023) -
Seeking the Unfamiliar but Memorable: Conceptual Creativity as Meta-Learning
by: Ren, Mengye
Published: (2026) -
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
by: Yang, Yanlai, et al.
Published: (2025) -
Midway Network: Learning Representations for Recognition and Motion from Latent Dynamics
by: Hoang, Christopher, et al.
Published: (2025) -
MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents
by: Kim, Kangsan, et al.
Published: (2026)