Saved in:
| Main Authors: | Ming, Ruibo, Wu, Jingwei, Huang, Zhewei, Ju, Zhuoxuan, HU, Jianming, Peng, Lihui, Zhou, Shuchang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.03758 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey on Future Frame Synthesis: Bridging Deterministic and Generative Approaches
by: Ming, Ruibo, et al.
Published: (2024)
by: Ming, Ruibo, et al.
Published: (2024)
Advancing Video Self-Supervised Learning via Image Foundation Models
by: Wu, Jingwei, et al.
Published: (2025)
by: Wu, Jingwei, et al.
Published: (2025)
A Style-Based Profiling Framework for Quantifying the Synthetic-to-Real Gap in Autonomous Driving Datasets
by: Yao, Dingyi, et al.
Published: (2025)
by: Yao, Dingyi, et al.
Published: (2025)
A Watermark for Auto-Regressive Image Generation Models
by: Wu, Yihan, et al.
Published: (2025)
by: Wu, Yihan, et al.
Published: (2025)
Synthetic Datasets for Autonomous Driving: A Survey
by: Song, Zhihang, et al.
Published: (2023)
by: Song, Zhihang, et al.
Published: (2023)
FAR-Drive: Frame-AutoRegressive Video Generation in Closed-Loop Autonomous Driving
by: Li, Yaoru, et al.
Published: (2026)
by: Li, Yaoru, et al.
Published: (2026)
Synthetic Dataset Evaluation Based on Generalized Cross Validation
by: Song, Zhihang, et al.
Published: (2025)
by: Song, Zhihang, et al.
Published: (2025)
A re-calibration method for object detection with multi-modal alignment bias in autonomous driving
by: Song, Zhihang, et al.
Published: (2024)
by: Song, Zhihang, et al.
Published: (2024)
Timealign: A multi-modal object detection method for time misalignment fusing in autonomous driving
by: Song, Zhihang, et al.
Published: (2024)
by: Song, Zhihang, et al.
Published: (2024)
SBF: An Effective Representation to Augment Skeleton for Video-based Human Action Recognition
by: Peng, Zhuoxuan, et al.
Published: (2026)
by: Peng, Zhuoxuan, et al.
Published: (2026)
Single Domain Generalization for Crowd Counting
by: Peng, Zhuoxuan, et al.
Published: (2024)
by: Peng, Zhuoxuan, et al.
Published: (2024)
Auto-Regressive Surface Cutting
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
AR-Diffusion: Asynchronous Video Generation with Auto-Regressive Diffusion
by: Sun, Mingzhen, et al.
Published: (2025)
by: Sun, Mingzhen, et al.
Published: (2025)
Nested AutoRegressive Models
by: Wu, Hongyu, et al.
Published: (2025)
by: Wu, Hongyu, et al.
Published: (2025)
GAReT: Cross-view Video Geolocalization with Adapters and Auto-Regressive Transformers
by: Pillai, Manu S, et al.
Published: (2024)
by: Pillai, Manu S, et al.
Published: (2024)
InfinityStar: Unified Spacetime AutoRegressive Modeling for Visual Generation
by: Liu, Jinlai, et al.
Published: (2025)
by: Liu, Jinlai, et al.
Published: (2025)
DeeAD: Dynamic Early Exit of Vision-Language Action for Efficient Autonomous Driving
by: HU, Haibo, et al.
Published: (2025)
by: HU, Haibo, et al.
Published: (2025)
AvatarPointillist: AutoRegressive 4D Gaussian Avatarization
by: Liu, Hongyu, et al.
Published: (2026)
by: Liu, Hongyu, et al.
Published: (2026)
Quality-Aware Language-Conditioned Local Auto-Regressive Anomaly Synthesis and Detection
by: Qian, Long, et al.
Published: (2025)
by: Qian, Long, et al.
Published: (2025)
Auto-Regressively Generating Multi-View Consistent Images
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
AutoWeather4D: Autonomous Driving Video Weather Conversion via G-Buffer Dual-Pass Editing
by: Liu, Tianyu, et al.
Published: (2026)
by: Liu, Tianyu, et al.
Published: (2026)
FARMER: Flow AutoRegressive Transformer over Pixels
by: Zheng, Guangting, et al.
Published: (2025)
by: Zheng, Guangting, et al.
Published: (2025)
OneCAT: Decoder-Only Auto-Regressive Model for Unified Understanding and Generation
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice
by: Liu, Shuming, et al.
Published: (2026)
by: Liu, Shuming, et al.
Published: (2026)
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
ProGVC: Progressive-based Generative Video Compression via Auto-Regressive Context Modeling
by: Li, Daowen, et al.
Published: (2026)
by: Li, Daowen, et al.
Published: (2026)
FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge
by: Shen, Xuan, et al.
Published: (2025)
by: Shen, Xuan, et al.
Published: (2025)
See4D: Pose-Free 4D Generation via Auto-Regressive Video Inpainting
by: Lu, Dongyue, et al.
Published: (2025)
by: Lu, Dongyue, et al.
Published: (2025)
HiMix: Hierarchical Artifact-aware Mixup for Generalized Synthetic Image Detection
by: Zhou, Shuchang, et al.
Published: (2026)
by: Zhou, Shuchang, et al.
Published: (2026)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
by: Xin, Yi, et al.
Published: (2025)
by: Xin, Yi, et al.
Published: (2025)
Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling
by: Xin, Yi, et al.
Published: (2025)
by: Xin, Yi, et al.
Published: (2025)
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
by: Yang, Jian, et al.
Published: (2024)
by: Yang, Jian, et al.
Published: (2024)
Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization
by: Peng, Jingwei, et al.
Published: (2025)
by: Peng, Jingwei, et al.
Published: (2025)
STAR: Scale-wise Text-conditioned AutoRegressive image generation
by: Ma, Xiaoxiao, et al.
Published: (2024)
by: Ma, Xiaoxiao, et al.
Published: (2024)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
by: Han, Jian, et al.
Published: (2024)
by: Han, Jian, et al.
Published: (2024)
Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressive Diffusion
by: Shiu, Hau-Shiang, et al.
Published: (2025)
by: Shiu, Hau-Shiang, et al.
Published: (2025)
Expanding mmWave Datasets for Human Pose Estimation with Unlabeled Data and LiDAR Datasets
by: Peng, Zhuoxuan, et al.
Published: (2026)
by: Peng, Zhuoxuan, et al.
Published: (2026)
MASRA: MLLM-Assisted Semantic-Relational Consistent Alignment for Video Temporal Grounding
by: Ran, Ran, et al.
Published: (2026)
by: Ran, Ran, et al.
Published: (2026)
POBEVM: Real-time Video Matting via Progressively Optimize the Target Body and Edge
by: Xian, Jianming
Published: (2024)
by: Xian, Jianming
Published: (2024)
Similar Items
-
A Survey on Future Frame Synthesis: Bridging Deterministic and Generative Approaches
by: Ming, Ruibo, et al.
Published: (2024) -
Advancing Video Self-Supervised Learning via Image Foundation Models
by: Wu, Jingwei, et al.
Published: (2025) -
A Style-Based Profiling Framework for Quantifying the Synthetic-to-Real Gap in Autonomous Driving Datasets
by: Yao, Dingyi, et al.
Published: (2025) -
A Watermark for Auto-Regressive Image Generation Models
by: Wu, Yihan, et al.
Published: (2025) -
Synthetic Datasets for Autonomous Driving: A Survey
by: Song, Zhihang, et al.
Published: (2023)