VONet: Unsupervised Video Object Learning With Parallel U-Net Attention and Object-wise Sequential VAE
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Haonan, Xu, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Compact Clustering Attention (COCA) for Unsupervised Object-Centric Learning
by: Küçüksözen, Can, et al.
Published: (2025)
by: Küçüksözen, Can, et al.
Published: (2025)
RS-TinyNet: Stage-wise Feature Fusion Network for Detecting Tiny Objects in Remote Sensing Images
by: Jiang, Xiaozheng, et al.
Published: (2025)
by: Jiang, Xiaozheng, et al.
Published: (2025)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026)
by: Seong, Hyun Seok, et al.
Published: (2026)
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
by: Jeon, Inseok, et al.
Published: (2026)
by: Jeon, Inseok, et al.
Published: (2026)
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
by: Daniel, Tal, et al.
Published: (2023)
by: Daniel, Tal, et al.
Published: (2023)
Saliency-Motion Guided Trunk-Collateral Network for Unsupervised Video Object Segmentation
by: Zheng, Xiangyu, et al.
Published: (2025)
by: Zheng, Xiangyu, et al.
Published: (2025)
Unsupervised Object-Centric Learning from Multiple Unspecified Viewpoints
by: Yuan, Jinyang, et al.
Published: (2024)
by: Yuan, Jinyang, et al.
Published: (2024)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
by: Pramanik, Rishav, et al.
Published: (2024)
by: Pramanik, Rishav, et al.
Published: (2024)
Conformal Object Detection by Sequential Risk Control
by: andéol, Léo, et al.
Published: (2025)
by: andéol, Léo, et al.
Published: (2025)
Unsupervised Machine Learning for Detecting and Locating Human-Made Objects in 3D Point Cloud
by: Zhao, Hong, et al.
Published: (2024)
by: Zhao, Hong, et al.
Published: (2024)
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
by: Lao, Dong, et al.
Published: (2023)
by: Lao, Dong, et al.
Published: (2023)
Multiple Object Stitching for Unsupervised Representation Learning
by: Shen, Chengchao, et al.
Published: (2025)
by: Shen, Chengchao, et al.
Published: (2025)
DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric Voxelization
by: Zhao, Yanpeng, et al.
Published: (2023)
by: Zhao, Yanpeng, et al.
Published: (2023)
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
by: Moon, WonJun, et al.
Published: (2026)
by: Moon, WonJun, et al.
Published: (2026)
NOAH: Learning Pairwise Object Category Attentions for Image Classification
by: Li, Chao, et al.
Published: (2024)
by: Li, Chao, et al.
Published: (2024)
FORLA: Federated Object-centric Representation Learning with Slot Attention
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Spatiotemporal Attention Learning Framework for Event-Driven Object Recognition
by: Xie, Tiantian, et al.
Published: (2025)
by: Xie, Tiantian, et al.
Published: (2025)
Edge Attention Module for Object Classification
by: Roy, Santanu, et al.
Published: (2025)
by: Roy, Santanu, et al.
Published: (2025)
VideoOrion: Tokenizing Object Dynamics in Videos
by: Feng, Yicheng, et al.
Published: (2024)
by: Feng, Yicheng, et al.
Published: (2024)
Incremental Object Keypoint Learning
by: Liang, Mingfu, et al.
Published: (2025)
by: Liang, Mingfu, et al.
Published: (2025)
CSA-Net: Channel-wise Spatially Autocorrelated Attention Networks
by: Nikzad, Nick, et al.
Published: (2024)
by: Nikzad, Nick, et al.
Published: (2024)
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational Videos
by: Biswas, Dipayan, et al.
Published: (2025)
by: Biswas, Dipayan, et al.
Published: (2025)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
AnomalousNet: A Hybrid Approach with Attention U-Nets and Change Point Detection for Accurate Characterization of Anomalous Diffusion in Video Data
by: Ahsini, Yusef, et al.
Published: (2025)
by: Ahsini, Yusef, et al.
Published: (2025)
Decoupling Amplitude and Phase Attention in Frequency Domain for RGB-Event based Visual Object Tracking
by: Wang, Shiao, et al.
Published: (2026)
by: Wang, Shiao, et al.
Published: (2026)
Object-Centric Diffusion for Efficient Video Editing
by: Kahatapitiya, Kumara, et al.
Published: (2024)
by: Kahatapitiya, Kumara, et al.
Published: (2024)
Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation
by: Shi, Ge, et al.
Published: (2024)
by: Shi, Ge, et al.
Published: (2024)
Boosting Object Representation Learning via Motion and Object Continuity
by: Delfosse, Quentin, et al.
Published: (2022)
by: Delfosse, Quentin, et al.
Published: (2022)
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities
by: Zadaianchuk, Andrii, et al.
Published: (2023)
by: Zadaianchuk, Andrii, et al.
Published: (2023)
DoughNet: A Visual Predictive Model for Topological Manipulation of Deformable Objects
by: Bauer, Dominik, et al.
Published: (2024)
by: Bauer, Dominik, et al.
Published: (2024)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
by: Fan, Ke, et al.
Published: (2024)
by: Fan, Ke, et al.
Published: (2024)
Reasoning-Enhanced Object-Centric Learning for Videos
by: Li, Jian, et al.
Published: (2024)
by: Li, Jian, et al.
Published: (2024)
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps
by: Raj, Arjun, et al.
Published: (2024)
by: Raj, Arjun, et al.
Published: (2024)
Unsupervised Representation Learning by Balanced Self Attention Matching
by: Shalam, Daniel, et al.
Published: (2024)
by: Shalam, Daniel, et al.
Published: (2024)
VidTwin: Video VAE with Decoupled Structure and Dynamics
by: Wang, Yuchi, et al.
Published: (2024)
by: Wang, Yuchi, et al.
Published: (2024)
Dual Prototype Attention for Unsupervised Video Object Segmentation
by: Cho, Suhwan, et al.
Published: (2022)
by: Cho, Suhwan, et al.
Published: (2022)
Guided Slot Attention for Unsupervised Video Object Segmentation
by: Lee, Minhyeok, et al.
Published: (2023)
by: Lee, Minhyeok, et al.
Published: (2023)
VQPy: An Object-Oriented Approach to Modern Video Analytics
by: Yu, Shan, et al.
Published: (2023)
by: Yu, Shan, et al.
Published: (2023)
GTA: Guided Transfer of Spatial Attention from Object-Centric Representations
by: Seo, SeokHyun, et al.
Published: (2024)
by: Seo, SeokHyun, et al.
Published: (2024)
Similar Items
-
Hierarchical Compact Clustering Attention (COCA) for Unsupervised Object-Centric Learning
by: Küçüksözen, Can, et al.
Published: (2025) -
RS-TinyNet: Stage-wise Feature Fusion Network for Detecting Tiny Objects in Remote Sensing Images
by: Jiang, Xiaozheng, et al.
Published: (2025) -
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026) -
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
by: Jeon, Inseok, et al.
Published: (2026) -
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
by: Daniel, Tal, et al.
Published: (2023)