PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Villar-Corrales, Angel, Behnke, Sven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TextOCVP: Object-Centric Video Prediction with Language Guidance
by: Villar-Corrales, Angel, et al.
Published: (2025)
by: Villar-Corrales, Angel, et al.
Published: (2025)
MCDS-VSS: Moving Camera Dynamic Scene Video Semantic Segmentation by Filtering with Self-Supervised Geometry and Motion
by: Villar-Corrales, Angel, et al.
Published: (2024)
by: Villar-Corrales, Angel, et al.
Published: (2024)
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
by: Spieler, Jonathan, et al.
Published: (2026)
by: Spieler, Jonathan, et al.
Published: (2026)
OC-SOP: Enhancing Vision-Based 3D Semantic Occupancy Prediction by Object-Centric Awareness
by: Cao, Helin, et al.
Published: (2025)
by: Cao, Helin, et al.
Published: (2025)
SOLD: Slot Object-Centric Latent Dynamics Models for Relational Manipulation Learning from Pixels
by: Mosbach, Malte, et al.
Published: (2024)
by: Mosbach, Malte, et al.
Published: (2024)
Temporally Consistent Object-Centric Learning by Contrasting Slots
by: Manasyan, Anna, et al.
Published: (2024)
by: Manasyan, Anna, et al.
Published: (2024)
VideoPCDNet: Video Parsing and Prediction with Phase Correlation Networks
by: Vicente, Noel José Rodrigues, et al.
Published: (2025)
by: Vicente, Noel José Rodrigues, et al.
Published: (2025)
Inverse++: Vision-Centric 3D Semantic Occupancy Prediction Assisted with 3D Object Detection
by: Ming, Zhenxing, et al.
Published: (2025)
by: Ming, Zhenxing, et al.
Published: (2025)
VILP: Imitation Learning with Latent Video Planning
by: Xu, Zhengtong, et al.
Published: (2025)
by: Xu, Zhengtong, et al.
Published: (2025)
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
by: Pätzold, Bastian, et al.
Published: (2025)
by: Pätzold, Bastian, et al.
Published: (2025)
SWA-SOP: Spatially-aware Window Attention for Semantic Occupancy Prediction in Autonomous Driving
by: Cao, Helin, et al.
Published: (2025)
by: Cao, Helin, et al.
Published: (2025)
DiffSSC: Semantic LiDAR Scan Completion using Denoising Diffusion Probabilistic Models
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
SLCF-Net: Sequential LiDAR-Camera Fusion for Semantic Scene Completion using a 3D Recurrent U-Net
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
Efficient Image Annotation via Semi-Supervised Object Segmentation with Label Propagation
by: Tutevych, Vitalii, et al.
Published: (2026)
by: Tutevych, Vitalii, et al.
Published: (2026)
Person Segmentation and Action Classification for Multi-Channel Hemisphere Field of View LiDAR Sensors
by: Seliunina, Svetlana, et al.
Published: (2024)
by: Seliunina, Svetlana, et al.
Published: (2024)
LiDAR-based Registration against Georeferenced Models for Globally Consistent Allocentric Maps
by: Quenzel, Jan, et al.
Published: (2024)
by: Quenzel, Jan, et al.
Published: (2024)
When Slots Compete: Slot Merging in Object-Centric Learning
by: Chatzisavvas, Christos, et al.
Published: (2026)
by: Chatzisavvas, Christos, et al.
Published: (2026)
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
by: Hanyu, Taisei, et al.
Published: (2025)
by: Hanyu, Taisei, et al.
Published: (2025)
GLASS: Guided Latent Slot Diffusion for Object-Centric Learning
by: Singh, Krishnakant, et al.
Published: (2024)
by: Singh, Krishnakant, et al.
Published: (2024)
LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps
by: Wang, Yihao, et al.
Published: (2025)
by: Wang, Yihao, et al.
Published: (2025)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Zero-shot Object-Centric Instruction Following: Integrating Foundation Models with Traditional Navigation
by: Raychaudhuri, Sonia, et al.
Published: (2024)
by: Raychaudhuri, Sonia, et al.
Published: (2024)
Slot-Level Robotic Placement via Visual Imitation from Single Human Video
by: Shan, Dandan, et al.
Published: (2025)
by: Shan, Dandan, et al.
Published: (2025)
FuncGrasp: Learning Object-Centric Neural Grasp Functions from Single Annotated Example Object
by: Chen, Hanzhi, et al.
Published: (2024)
by: Chen, Hanzhi, et al.
Published: (2024)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
by: Giannakakis, Nikos, et al.
Published: (2025)
by: Giannakakis, Nikos, et al.
Published: (2025)
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
by: Wang, Yanbo, et al.
Published: (2023)
by: Wang, Yanbo, et al.
Published: (2023)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
3D Feature Distillation with Object-Centric Priors
by: Tziafas, Georgios, et al.
Published: (2024)
by: Tziafas, Georgios, et al.
Published: (2024)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
by: Han, Jiwook, et al.
Published: (2026)
by: Han, Jiwook, et al.
Published: (2026)
LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation
by: Guan, Tianrui, et al.
Published: (2024)
by: Guan, Tianrui, et al.
Published: (2024)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
GenFlowRL: Shaping Rewards with Generative Object-Centric Flow in Visual Reinforcement Learning
by: Yu, Kelin, et al.
Published: (2025)
by: Yu, Kelin, et al.
Published: (2025)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
by: Wang, Kuanning, et al.
Published: (2026)
by: Wang, Kuanning, et al.
Published: (2026)
MaskPlanner: Learning-Based Object-Centric Motion Generation from 3D Point Clouds
by: Tiboni, Gabriele, et al.
Published: (2025)
by: Tiboni, Gabriele, et al.
Published: (2025)
Rethinking Progression of Memory State in Robotic Manipulation: An Object-Centric Perspective
by: Chung, Nhat, et al.
Published: (2025)
by: Chung, Nhat, et al.
Published: (2025)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
by: Dou, Weijia, et al.
Published: (2026)
by: Dou, Weijia, et al.
Published: (2026)
Learning Action-Conditional and Object-Centric Gaussian Splatting World Models for Rigid Objects
by: Kreber, Jens U., et al.
Published: (2026)
by: Kreber, Jens U., et al.
Published: (2026)
CarFormer: Self-Driving with Learned Object-Centric Representations
by: Hamdan, Shadi, et al.
Published: (2024)
by: Hamdan, Shadi, et al.
Published: (2024)
Entity-Centric Reinforcement Learning for Object Manipulation from Pixels
by: Haramati, Dan, et al.
Published: (2024)
by: Haramati, Dan, et al.
Published: (2024)
Similar Items
-
TextOCVP: Object-Centric Video Prediction with Language Guidance
by: Villar-Corrales, Angel, et al.
Published: (2025) -
MCDS-VSS: Moving Camera Dynamic Scene Video Semantic Segmentation by Filtering with Self-Supervised Geometry and Motion
by: Villar-Corrales, Angel, et al.
Published: (2024) -
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
by: Spieler, Jonathan, et al.
Published: (2026) -
OC-SOP: Enhancing Vision-Based 3D Semantic Occupancy Prediction by Object-Centric Awareness
by: Cao, Helin, et al.
Published: (2025) -
SOLD: Slot Object-Centric Latent Dynamics Models for Relational Manipulation Learning from Pixels
by: Mosbach, Malte, et al.
Published: (2024)