CanvasMAR: Improving Masked Autoregressive Video Prediction With Canvas
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zian, Zhang, Muhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring MLLM-Diffusion Information Transfer with MetaCanvas
von: Lin, Han, et al.
Veröffentlicht: (2025)
von: Lin, Han, et al.
Veröffentlicht: (2025)
Graph Canvas for Controllable 3D Scene Generation
von: Liu, Libin, et al.
Veröffentlicht: (2024)
von: Liu, Libin, et al.
Veröffentlicht: (2024)
Meta Pruning via Graph Metanetworks : A Universal Meta Learning Framework for Network Pruning
von: Liu, Yewei, et al.
Veröffentlicht: (2025)
von: Liu, Yewei, et al.
Veröffentlicht: (2025)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
von: Yariv, Guy, et al.
Veröffentlicht: (2025)
von: Yariv, Guy, et al.
Veröffentlicht: (2025)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
von: Huang, Xun, et al.
Veröffentlicht: (2025)
von: Huang, Xun, et al.
Veröffentlicht: (2025)
Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2026)
von: Kwon, Taesung, et al.
Veröffentlicht: (2026)
VideoMAR: Autoregressive Video Generatio with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
Look Through Masks: Towards Masked Face Recognition with De-Occlusion Distillation
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
Annealed Relaxation of Speculative Decoding for Faster Autoregressive Image Generation
von: Li, Xingyao, et al.
Veröffentlicht: (2026)
von: Li, Xingyao, et al.
Veröffentlicht: (2026)
SeqSAM: Autoregressive Multiple Hypothesis Prediction for Medical Image Segmentation using SAM
von: Towle, Benjamin, et al.
Veröffentlicht: (2025)
von: Towle, Benjamin, et al.
Veröffentlicht: (2025)
Masked Face Recognition with Generative-to-Discriminative Representations
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
von: Zhang, David Junhao, et al.
Veröffentlicht: (2024)
von: Zhang, David Junhao, et al.
Veröffentlicht: (2024)
Multi-modal Masked Siamese Network Improves Chest X-Ray Representation Learning
von: Shurrab, Saeed, et al.
Veröffentlicht: (2024)
von: Shurrab, Saeed, et al.
Veröffentlicht: (2024)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
Mamba-3D as Masked Autoencoders for Accurate and Data-Efficient Analysis of Medical Ultrasound Videos
von: Zhou, Jiaheng, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaheng, et al.
Veröffentlicht: (2025)
PointNSP: Autoregressive 3D Point Cloud Generation with Next-Scale Level-of-Detail Prediction
von: Meng, Ziqiao, et al.
Veröffentlicht: (2025)
von: Meng, Ziqiao, et al.
Veröffentlicht: (2025)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
von: Chin, Zhi-Yi, et al.
Veröffentlicht: (2023)
von: Chin, Zhi-Yi, et al.
Veröffentlicht: (2023)
Improving Diffusion-Based Image Synthesis with Context Prediction
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Improving Personalisation in Valence and Arousal Prediction using Data Augmentation
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2024)
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2024)
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
von: Tang, Haotian, et al.
Veröffentlicht: (2024)
von: Tang, Haotian, et al.
Veröffentlicht: (2024)
MixMask: Revisiting Masking Strategy for Siamese ConvNets
von: Vishniakov, Kirill, et al.
Veröffentlicht: (2022)
von: Vishniakov, Kirill, et al.
Veröffentlicht: (2022)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
PhiNet v2: A Mask-Free Brain-Inspired Vision Foundation Model from Video
von: Yamada, Makoto, et al.
Veröffentlicht: (2025)
von: Yamada, Makoto, et al.
Veröffentlicht: (2025)
Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation
von: Lokegaonkar, Vaibhavi, et al.
Veröffentlicht: (2026)
von: Lokegaonkar, Vaibhavi, et al.
Veröffentlicht: (2026)
MASC: Boosting Autoregressive Image Generation with a Manifold-Aligned Semantic Clustering
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
CBM: Curriculum by Masking
von: Jarca, Andrei, et al.
Veröffentlicht: (2024)
von: Jarca, Andrei, et al.
Veröffentlicht: (2024)
SpectralAR: Spectral Autoregressive Visual Generation
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
Control-Augmented Autoregressive Diffusion for Data Assimilation
von: Srivastava, Prakhar, et al.
Veröffentlicht: (2025)
von: Srivastava, Prakhar, et al.
Veröffentlicht: (2025)
MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2024)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2024)
i-MAE: Are Latent Representations in Masked Autoencoders Linearly Separable?
von: Zhang, Kevin, et al.
Veröffentlicht: (2022)
von: Zhang, Kevin, et al.
Veröffentlicht: (2022)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
von: Jeong, Hyeonho, et al.
Veröffentlicht: (2024)
von: Jeong, Hyeonho, et al.
Veröffentlicht: (2024)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
von: Lee, Dohun, et al.
Veröffentlicht: (2024)
von: Lee, Dohun, et al.
Veröffentlicht: (2024)
Taming the Entropy Cliff: Variable Codebook Size Quantization for Autoregressive Visual Generation
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
Guiding Video Prediction with Explicit Procedural Knowledge
von: Takenaka, Patrick, et al.
Veröffentlicht: (2024)
von: Takenaka, Patrick, et al.
Veröffentlicht: (2024)
Visual Autoregressive Transformers Must Use $Ω(n^2 d)$ Memory
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
von: Hu, Zhangyi, et al.
Veröffentlicht: (2025)
von: Hu, Zhangyi, et al.
Veröffentlicht: (2025)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
Provably Robust Conformal Prediction with Improved Efficiency
von: Yan, Ge, et al.
Veröffentlicht: (2024)
von: Yan, Ge, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring MLLM-Diffusion Information Transfer with MetaCanvas
von: Lin, Han, et al.
Veröffentlicht: (2025) -
Graph Canvas for Controllable 3D Scene Generation
von: Liu, Libin, et al.
Veröffentlicht: (2024) -
Meta Pruning via Graph Metanetworks : A Universal Meta Learning Framework for Network Pruning
von: Liu, Yewei, et al.
Veröffentlicht: (2025) -
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
von: Yariv, Guy, et al.
Veröffentlicht: (2025) -
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
von: Huang, Xun, et al.
Veröffentlicht: (2025)