Predicting Video Slot Attention Queries from Random Slot-Feature Pairs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhao, Rongzhen, Li, Jian, Kannala, Juho, Pajarinen, Joni |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Smoothing Slot Attention Iterations and Recurrences
par: Zhao, Rongzhen, et autres
Publié: (2025)
par: Zhao, Rongzhen, et autres
Publié: (2025)
Slot Attention with Re-Initialization and Self-Distillation
par: Zhao, Rongzhen, et autres
Publié: (2025)
par: Zhao, Rongzhen, et autres
Publié: (2025)
Internalizing Temporal Consistency in Video Object-Centric Learning without Explicit Regularization
par: Zhao, Rongzhen, et autres
Publié: (2026)
par: Zhao, Rongzhen, et autres
Publié: (2026)
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
par: Liu, Hongjia, et autres
Publié: (2025)
par: Liu, Hongjia, et autres
Publié: (2025)
Cycle Consistency in Video Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2026)
par: Zhao, Rongzhen, et autres
Publié: (2026)
Vector-Quantized Vision Foundation Models for Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2025)
par: Zhao, Rongzhen, et autres
Publié: (2025)
Multi-Scale Fusion for Object Representation
par: Zhao, Rongzhen, et autres
Publié: (2024)
par: Zhao, Rongzhen, et autres
Publié: (2024)
Organized Grouped Discrete Representation for Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2024)
par: Zhao, Rongzhen, et autres
Publié: (2024)
Grouped Discrete Representation for Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2024)
par: Zhao, Rongzhen, et autres
Publié: (2024)
Grouped Discrete Representation Guides Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2024)
par: Zhao, Rongzhen, et autres
Publié: (2024)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
par: Wang, Yanbo, et autres
Publié: (2023)
par: Wang, Yanbo, et autres
Publié: (2023)
Rethinking Temporal Consistency in Video Object-Centric Learning: From Prediction to Correspondence
par: Li, Zhiyuan, et autres
Publié: (2026)
par: Li, Zhiyuan, et autres
Publié: (2026)
Slot-VLM: SlowFast Slots for Video-Language Modeling
par: Xu, Jiaqi, et autres
Publié: (2024)
par: Xu, Jiaqi, et autres
Publié: (2024)
Guided Slot Attention for Unsupervised Video Object Segmentation
par: Lee, Minhyeok, et autres
Publié: (2023)
par: Lee, Minhyeok, et autres
Publié: (2023)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
par: Fan, Ke, et autres
Publié: (2024)
par: Fan, Ke, et autres
Publié: (2024)
Slot Attention-based Feature Filtering for Few-Shot Learning
par: Rodenas, Javier, et autres
Publié: (2025)
par: Rodenas, Javier, et autres
Publié: (2025)
Object-Centric Vision Token Pruning for Vision Language Models
par: Li, Guangyuan, et autres
Publié: (2025)
par: Li, Guangyuan, et autres
Publié: (2025)
Attention Normalization Impacts Cardinality Generalization in Slot Attention
par: Krimmel, Markus, et autres
Publié: (2024)
par: Krimmel, Markus, et autres
Publié: (2024)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
par: Liao, Guiqiu, et autres
Publié: (2025)
par: Liao, Guiqiu, et autres
Publié: (2025)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
par: Pramanik, Rishav, et autres
Publié: (2024)
par: Pramanik, Rishav, et autres
Publié: (2024)
When Slots Compete: Slot Merging in Object-Centric Learning
par: Chatzisavvas, Christos, et autres
Publié: (2026)
par: Chatzisavvas, Christos, et autres
Publié: (2026)
PAWS: Perception of Articulation in the Wild at Scale from Egocentric Videos
par: Wang, Yihao, et autres
Publié: (2026)
par: Wang, Yihao, et autres
Publié: (2026)
MUFASA: A Multi-Layer Framework for Slot Attention
par: Bock, Sebastian, et autres
Publié: (2026)
par: Bock, Sebastian, et autres
Publié: (2026)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
par: Lai, Yixuan, et autres
Publié: (2026)
par: Lai, Yixuan, et autres
Publié: (2026)
Learning Global Object-Centric Representations via Disentangled Slot Attention
par: Chen, Tonglin, et autres
Publié: (2024)
par: Chen, Tonglin, et autres
Publié: (2024)
Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
par: Dedhia, Bhishma, et autres
Publié: (2024)
par: Dedhia, Bhishma, et autres
Publié: (2024)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
par: Liu, Yu, et autres
Publié: (2024)
par: Liu, Yu, et autres
Publié: (2024)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
par: Dou, Weijia, et autres
Publié: (2026)
par: Dou, Weijia, et autres
Publié: (2026)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
par: Han, Jiwook, et autres
Publié: (2026)
par: Han, Jiwook, et autres
Publié: (2026)
Improving Sound Source Localization with Joint Slot Attention on Image and Audio
par: Kim, Inho, et autres
Publié: (2025)
par: Kim, Inho, et autres
Publié: (2025)
FORLA: Federated Object-centric Representation Learning with Slot Attention
par: Liao, Guiqiu, et autres
Publié: (2025)
par: Liao, Guiqiu, et autres
Publié: (2025)
Bootstrapping Top-down Information for Self-modulating Slot Attention
par: Kim, Dongwon, et autres
Publié: (2024)
par: Kim, Dongwon, et autres
Publié: (2024)
PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning
par: Villar-Corrales, Angel, et autres
Publié: (2025)
par: Villar-Corrales, Angel, et autres
Publié: (2025)
Slot Structured World Models
par: Collu, Jonathan, et autres
Publié: (2024)
par: Collu, Jonathan, et autres
Publié: (2024)
Unsupervised Structural Scene Decomposition via Foreground-Aware Slot Attention with Pseudo-Mask Guidance
par: Sheng, Huankun, et autres
Publié: (2025)
par: Sheng, Huankun, et autres
Publié: (2025)
Slot-guided Volumetric Object Radiance Fields
par: Qi, Di, et autres
Publié: (2024)
par: Qi, Di, et autres
Publié: (2024)
QASA: Quality-Guided K-Adaptive Slot Attention for Unsupervised Object-Centric Learning
par: Ouyang, Tianran, et autres
Publié: (2026)
par: Ouyang, Tianran, et autres
Publié: (2026)
PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery
par: Park, Jicheol, et autres
Publié: (2024)
par: Park, Jicheol, et autres
Publié: (2024)
A2-GNN: Angle-Annular GNN for Visual Descriptor-free Camera Relocalization
par: Zhang, Yejun, et autres
Publié: (2025)
par: Zhang, Yejun, et autres
Publié: (2025)
DGC-GNN: Leveraging Geometry and Color Cues for Visual Descriptor-Free 2D-3D Matching
par: Wang, Shuzhe, et autres
Publié: (2023)
par: Wang, Shuzhe, et autres
Publié: (2023)
Documents similaires
-
Smoothing Slot Attention Iterations and Recurrences
par: Zhao, Rongzhen, et autres
Publié: (2025) -
Slot Attention with Re-Initialization and Self-Distillation
par: Zhao, Rongzhen, et autres
Publié: (2025) -
Internalizing Temporal Consistency in Video Object-Centric Learning without Explicit Regularization
par: Zhao, Rongzhen, et autres
Publié: (2026) -
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
par: Liu, Hongjia, et autres
Publié: (2025) -
Cycle Consistency in Video Object-Centric Learning
par: Zhao, Rongzhen, et autres
Publié: (2026)