On the Benefits of Instance Decomposition in Video Prediction Models
Fuente:
arXiv
Saved in:
| Main Authors: | Suleyman, Eliyas, Henderson, Paul, Pugeault, Nicolas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flow and Depth Assisted Video Prediction with Latent Transformer
by: Suleyman, Eliyas, et al.
Published: (2025)
by: Suleyman, Eliyas, et al.
Published: (2025)
Beyond Reconstruction: A Physics Based Neural Deferred Shader for Photo-realistic Rendering
by: He, Zhuo, et al.
Published: (2025)
by: He, Zhuo, et al.
Published: (2025)
Generative Fields: Uncovering Hierarchical Feature Control for StyleGAN via Inverted Receptive Fields
by: He, Zhuo, et al.
Published: (2025)
by: He, Zhuo, et al.
Published: (2025)
A Convolutional Neural Deferred Shader for Physics Based Rendering
by: He, Zhuo, et al.
Published: (2025)
by: He, Zhuo, et al.
Published: (2025)
Splat-Portrait: Generalizing Talking Heads with Gaussian Splatting
by: Shi, Tong, et al.
Published: (2026)
by: Shi, Tong, et al.
Published: (2026)
Detail-Enhanced Intra- and Inter-modal Interaction for Audio-Visual Emotion Recognition
by: Shi, Tong, et al.
Published: (2024)
by: Shi, Tong, et al.
Published: (2024)
The Bad Batches: Enhancing Self-Supervised Learning in Image Classification Through Representative Batch Curation
by: Goksu, Ozgu, et al.
Published: (2024)
by: Goksu, Ozgu, et al.
Published: (2024)
FedQuad: Federated Stochastic Quadruplet Learning to Mitigate Data Heterogeneity
by: Goksu, Ozgu, et al.
Published: (2025)
by: Goksu, Ozgu, et al.
Published: (2025)
Enhancing Federated Quadruplet Learning: Stochastic Client Selection and Embedding Stability Analysis
by: Goksu, Ozgu, et al.
Published: (2026)
by: Goksu, Ozgu, et al.
Published: (2026)
Hybrid-Regularized Magnitude Pruning for Robust Federated Learning under Covariate Shift
by: Goksu, Ozgu, et al.
Published: (2024)
by: Goksu, Ozgu, et al.
Published: (2024)
Modeling 3D Pedestrian-Vehicle Interactions for Vehicle-Conditioned Pose Forecasting
by: Zhu, Guangxun, et al.
Published: (2026)
by: Zhu, Guangxun, et al.
Published: (2026)
InstanceV: Instance-Level Video Generation
by: Chen, Yuheng, et al.
Published: (2025)
by: Chen, Yuheng, et al.
Published: (2025)
InstanceAnimator: Multi-Instance Sketch Video Colorization
by: Zhang, Yinhan, et al.
Published: (2026)
by: Zhang, Yinhan, et al.
Published: (2026)
Virtually Unrolling the Herculaneum Papyri by Diffeomorphic Spiral Fitting
by: Henderson, Paul
Published: (2025)
by: Henderson, Paul
Published: (2025)
A Temporal Modeling Framework for Video Pre-Training on Video Instance Segmentation
by: Zhong, Qing, et al.
Published: (2025)
by: Zhong, Qing, et al.
Published: (2025)
State-space Decomposition Model for Video Prediction Considering Long-term Motion Trend
by: Cui, Fei, et al.
Published: (2024)
by: Cui, Fei, et al.
Published: (2024)
Instance Brownian Bridge as Texts for Open-vocabulary Video Instance Segmentation
by: Cheng, Zesen, et al.
Published: (2024)
by: Cheng, Zesen, et al.
Published: (2024)
Foundation Models for Amodal Video Instance Segmentation in Automated Driving
by: Breitenstein, Jasmin, et al.
Published: (2024)
by: Breitenstein, Jasmin, et al.
Published: (2024)
UVIS: Unsupervised Video Instance Segmentation
by: Huang, Shuaiyi, et al.
Published: (2024)
by: Huang, Shuaiyi, et al.
Published: (2024)
SDI-Paste: Synthetic Dynamic Instance Copy-Paste for Video Instance Segmentation
by: Shrestha, Sahir, et al.
Published: (2024)
by: Shrestha, Sahir, et al.
Published: (2024)
Does Synthetic Layered Design Data Benefit Layered Design Decomposition?
by: Wu, Kam Man, et al.
Published: (2026)
by: Wu, Kam Man, et al.
Published: (2026)
Physics-Aware Video Instance Removal Benchmark
by: Li, Zirui, et al.
Published: (2026)
by: Li, Zirui, et al.
Published: (2026)
SyncVIS: Synchronized Video Instance Segmentation
by: Zheng, Rongkun, et al.
Published: (2024)
by: Zheng, Rongkun, et al.
Published: (2024)
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
iMOVE: Instance-Motion-Aware Video Understanding
by: Li, Jiaze, et al.
Published: (2025)
by: Li, Jiaze, et al.
Published: (2025)
Video Instance Shadow Detection Under the Sun and Sky
by: Xing, Zhenghao, et al.
Published: (2022)
by: Xing, Zhenghao, et al.
Published: (2022)
OpenVIS: Open-vocabulary Video Instance Segmentation
by: Guo, Pinxue, et al.
Published: (2023)
by: Guo, Pinxue, et al.
Published: (2023)
VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement
by: Kim, Hanjung, et al.
Published: (2023)
by: Kim, Hanjung, et al.
Published: (2023)
Instance-Aligned Captions for Explainable Video Anomaly Detection
by: Song, Inpyo, et al.
Published: (2026)
by: Song, Inpyo, et al.
Published: (2026)
What is Point Supervision Worth in Video Instance Segmentation?
by: Huang, Shuaiyi, et al.
Published: (2024)
by: Huang, Shuaiyi, et al.
Published: (2024)
PM-VIS+: High-Performance Video Instance Segmentation without Video Annotation
by: Yang, Zhangjing, et al.
Published: (2024)
by: Yang, Zhangjing, et al.
Published: (2024)
UnIRe: Unsupervised Instance Decomposition for Dynamic Urban Scene Reconstruction
by: Mao, Yunxuan, et al.
Published: (2025)
by: Mao, Yunxuan, et al.
Published: (2025)
InstaDrive: Instance-Aware Driving World Models for Realistic and Consistent Video Generation
by: Yang, Zhuoran, et al.
Published: (2026)
by: Yang, Zhuoran, et al.
Published: (2026)
A Lightweight Video Anomaly Detection Model with Weak Supervision and Adaptive Instance Selection
by: Wang, Yang, et al.
Published: (2023)
by: Wang, Yang, et al.
Published: (2023)
ConsisDrive: Identity-Preserving Driving World Models for Video Generation by Instance Mask
by: Yang, Zhuoran, et al.
Published: (2026)
by: Yang, Zhuoran, et al.
Published: (2026)
InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes
by: Yang, Zesong, et al.
Published: (2025)
by: Yang, Zesong, et al.
Published: (2025)
Video Patch Pruning: Efficient Video Instance Segmentation via Early Token Reduction
by: Glandorf, Patrick, et al.
Published: (2026)
by: Glandorf, Patrick, et al.
Published: (2026)
Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation
by: Dong, Jiahua, et al.
Published: (2025)
by: Dong, Jiahua, et al.
Published: (2025)
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
by: Niu, Quanzhu, et al.
Published: (2025)
by: Niu, Quanzhu, et al.
Published: (2025)
Similar Items
-
Flow and Depth Assisted Video Prediction with Latent Transformer
by: Suleyman, Eliyas, et al.
Published: (2025) -
Beyond Reconstruction: A Physics Based Neural Deferred Shader for Photo-realistic Rendering
by: He, Zhuo, et al.
Published: (2025) -
Generative Fields: Uncovering Hierarchical Feature Control for StyleGAN via Inverted Receptive Fields
by: He, Zhuo, et al.
Published: (2025) -
A Convolutional Neural Deferred Shader for Physics Based Rendering
by: He, Zhuo, et al.
Published: (2025) -
Splat-Portrait: Generalizing Talking Heads with Gaussian Splatting
by: Shi, Tong, et al.
Published: (2026)