ParallelVLM: Lossless Video-LLM Acceleration with Visual Alignment Aware Parallel Speculative Decoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kong, Quan, Shen, Yuhao, Ji, Yicheng, Li, Huan, Wang, Cong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SpecVLM: Enhancing Speculative Decoding of Video LLMs via Verifier-Guided Token Pruning
von: Ji, Yicheng, et al.
Veröffentlicht: (2025)
von: Ji, Yicheng, et al.
Veröffentlicht: (2025)
Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation
von: So, Junhyuk, et al.
Veröffentlicht: (2025)
von: So, Junhyuk, et al.
Veröffentlicht: (2025)
HIPPO: Accelerating Video Large Language Models Inference via Holistic-aware Parallel Speculative Decoding
von: Lv, Qitan, et al.
Veröffentlicht: (2026)
von: Lv, Qitan, et al.
Veröffentlicht: (2026)
Vision-TTT: Efficient and Expressive Visual Representation Learning with Test-Time Training
von: Kong, Quan, et al.
Veröffentlicht: (2026)
von: Kong, Quan, et al.
Veröffentlicht: (2026)
Autoregressive Image Generation with Randomized Parallel Decoding
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
Accelerating Parallel Diffusion Model Serving with Residual Compression
von: Luo, Jiajun, et al.
Veröffentlicht: (2025)
von: Luo, Jiajun, et al.
Veröffentlicht: (2025)
Rethinking Autoregressive Models for Lossless Image Compression via Hierarchical Parallelism and Progressive Adaptation
von: Li, Daxin, et al.
Veröffentlicht: (2025)
von: Li, Daxin, et al.
Veröffentlicht: (2025)
SpecVLM: Fast Speculative Decoding in Vision-Language Models
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Dynamic-VLM: Simple Dynamic Visual Token Compression for VideoLLM
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding
von: Jang, Doohyuk, et al.
Veröffentlicht: (2024)
von: Jang, Doohyuk, et al.
Veröffentlicht: (2024)
Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism
von: Shen, Yuhao, et al.
Veröffentlicht: (2026)
von: Shen, Yuhao, et al.
Veröffentlicht: (2026)
SpecBranch: Speculative Decoding via Hybrid Drafting and Rollback-Aware Branch Parallelism
von: Shen, Yuhao, et al.
Veröffentlicht: (2025)
von: Shen, Yuhao, et al.
Veröffentlicht: (2025)
ViSpec: Accelerating Vision-Language Models with Vision-Aware Speculative Decoding
von: Kang, Jialiang, et al.
Veröffentlicht: (2025)
von: Kang, Jialiang, et al.
Veröffentlicht: (2025)
VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping
von: Dong, Haotian, et al.
Veröffentlicht: (2025)
von: Dong, Haotian, et al.
Veröffentlicht: (2025)
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
Parallelized Autoregressive Visual Generation
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuqing, et al.
Veröffentlicht: (2024)
Minute-Long Videos with Dual Parallelisms
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
FREE: Uncertainty-Aware Autoregression for Parallel Diffusion Transformers
von: Wen, Xinwan, et al.
Veröffentlicht: (2025)
von: Wen, Xinwan, et al.
Veröffentlicht: (2025)
PD-APE: A Parallel Decoding Framework with Adaptive Position Encoding for 3D Visual Grounding
von: Hou, Chenshu, et al.
Veröffentlicht: (2024)
von: Hou, Chenshu, et al.
Veröffentlicht: (2024)
Speculative Decoding for Autoregressive Video Generation
von: Hu, Yuezhou, et al.
Veröffentlicht: (2026)
von: Hu, Yuezhou, et al.
Veröffentlicht: (2026)
HSD: Training-Free Acceleration for Document Parsing Vision-Language Model with Hierarchical Speculative Decoding
von: Liao, Wenhui, et al.
Veröffentlicht: (2026)
von: Liao, Wenhui, et al.
Veröffentlicht: (2026)
Accelerating Auto-regressive Text-to-Image Generation with Training-free Speculative Jacobi Decoding
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
CogVLM2: Visual Language Models for Image and Video Understanding
von: Hong, Wenyi, et al.
Veröffentlicht: (2024)
von: Hong, Wenyi, et al.
Veröffentlicht: (2024)
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Sparrow: Text-Anchored Window Attention with Visual-Semantic Glimpsing for Speculative Decoding in Video LLMs
von: Zhang, Libo, et al.
Veröffentlicht: (2026)
von: Zhang, Libo, et al.
Veröffentlicht: (2026)
Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding
von: Yin, Kun, et al.
Veröffentlicht: (2026)
von: Yin, Kun, et al.
Veröffentlicht: (2026)
SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation
von: Kang, Jialiang, et al.
Veröffentlicht: (2026)
von: Kang, Jialiang, et al.
Veröffentlicht: (2026)
Embodied Multi-Modal Agent trained by an LLM from a Parallel TextWorld
von: Yang, Yijun, et al.
Veröffentlicht: (2023)
von: Yang, Yijun, et al.
Veröffentlicht: (2023)
Continuous Speculative Decoding for Autoregressive Image Generation
von: Wang, Zili, et al.
Veröffentlicht: (2024)
von: Wang, Zili, et al.
Veröffentlicht: (2024)
REF-VLM: Triplet-Based Referring Paradigm for Unified Visual Decoding
von: Tai, Yan, et al.
Veröffentlicht: (2025)
von: Tai, Yan, et al.
Veröffentlicht: (2025)
End-to-End Dense Video Grounding via Parallel Regression
von: Shi, Fengyuan, et al.
Veröffentlicht: (2021)
von: Shi, Fengyuan, et al.
Veröffentlicht: (2021)
CASCADE: Context-Aware Relaxation for Speculative Image Decoding
von: Yildirim, Selin, et al.
Veröffentlicht: (2026)
von: Yildirim, Selin, et al.
Veröffentlicht: (2026)
SAGE: Accelerating Vision-Language Models via Entropy-Guided Adaptive Speculative Decoding
von: Tong, Yujia, et al.
Veröffentlicht: (2026)
von: Tong, Yujia, et al.
Veröffentlicht: (2026)
db-SP: Accelerating Sparse Attention for Visual Generative Models with Dual-Balanced Sequence Parallelism
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
Cross-Image Contrastive Decoding: Precise, Lossless Suppression of Language Priors in Large Vision-Language Models
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
von: Zhao, Jianfei, et al.
Veröffentlicht: (2025)
Generating, Fast and Slow: Scalable Parallel Video Generation with Video Interface Networks
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
SJD++: Improved Speculative Jacobi Decoding for Training-free Acceleration of Discrete Auto-regressive Text-to-Image Generation
von: Teng, Yao, et al.
Veröffentlicht: (2025)
von: Teng, Yao, et al.
Veröffentlicht: (2025)
MMSpec: Benchmarking Speculative Decoding for Vision-Language Models
von: Shen, Hui, et al.
Veröffentlicht: (2026)
von: Shen, Hui, et al.
Veröffentlicht: (2026)
VideoExpert: Augmented LLM for Temporal-Sensitive Video Understanding
von: Zhao, Henghao, et al.
Veröffentlicht: (2025)
von: Zhao, Henghao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SpecVLM: Enhancing Speculative Decoding of Video LLMs via Verifier-Guided Token Pruning
von: Ji, Yicheng, et al.
Veröffentlicht: (2025) -
Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation
von: So, Junhyuk, et al.
Veröffentlicht: (2025) -
HIPPO: Accelerating Video Large Language Models Inference via Holistic-aware Parallel Speculative Decoding
von: Lv, Qitan, et al.
Veröffentlicht: (2026) -
Vision-TTT: Efficient and Expressive Visual Representation Learning with Test-Time Training
von: Kong, Quan, et al.
Veröffentlicht: (2026) -
Autoregressive Image Generation with Randomized Parallel Decoding
von: Li, Haopeng, et al.
Veröffentlicht: (2025)