Gespeichert in:
| Hauptverfasser: | Zeng, Weili, Huang, Ziyuan, Ji, Kaixiang, Yan, Yichao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2503.21817 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GM-Skip: Metric-Guided Transformer Block Skipping for Efficient Vision-Language Models
von: Huang, Lianming, et al.
Veröffentlicht: (2025)
von: Huang, Lianming, et al.
Veröffentlicht: (2025)
Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves
von: Wu, Shihan, et al.
Veröffentlicht: (2024)
von: Wu, Shihan, et al.
Veröffentlicht: (2024)
SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping
von: Li, Jiajun, et al.
Veröffentlicht: (2025)
von: Li, Jiajun, et al.
Veröffentlicht: (2025)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
von: Hartman, Max, et al.
Veröffentlicht: (2025)
von: Hartman, Max, et al.
Veröffentlicht: (2025)
SkipViT: Speeding Up Vision Transformers with a Token-Level Skip Connection
von: Ataiefard, Foozhan, et al.
Veröffentlicht: (2024)
von: Ataiefard, Foozhan, et al.
Veröffentlicht: (2024)
SkipSR: Faster Super Resolution with Token Skipping
von: Choudhury, Rohan, et al.
Veröffentlicht: (2025)
von: Choudhury, Rohan, et al.
Veröffentlicht: (2025)
Skip and Skip: Segmenting Medical Images with Prompts
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
von: Chen, Jiawei, et al.
Veröffentlicht: (2024)
SkipGS: Post-Densification Backward Skipping for Efficient 3DGS Training
von: Li, Jingxing, et al.
Veröffentlicht: (2026)
von: Li, Jingxing, et al.
Veröffentlicht: (2026)
Energy-Efficient & Real-Time Computer Vision with Intelligent Skipping via Reconfigurable CMOS Image Sensors
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
von: Kaiser, Md Abdullah-Al, et al.
Veröffentlicht: (2024)
Always Skip Attention
von: Ji, Yiping, et al.
Veröffentlicht: (2025)
von: Ji, Yiping, et al.
Veröffentlicht: (2025)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
SkipcrossNets: Adaptive Skip-cross Fusion for Road Detection
von: Gong, Yan, et al.
Veröffentlicht: (2023)
von: Gong, Yan, et al.
Veröffentlicht: (2023)
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models
von: Liu, Jizhihui, et al.
Veröffentlicht: (2025)
von: Liu, Jizhihui, et al.
Veröffentlicht: (2025)
On the Vulnerability of Skip Connections to Model Inversion Attacks
von: Koh, Jun Hao, et al.
Veröffentlicht: (2024)
von: Koh, Jun Hao, et al.
Veröffentlicht: (2024)
SSI-DM: Singularity Skipping Inversion of Diffusion Models
von: Min, Chen, et al.
Veröffentlicht: (2026)
von: Min, Chen, et al.
Veröffentlicht: (2026)
Skipping Computations in Multimodal LLMs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
Efficient Vision-Language Reasoning via Adaptive Token Pruning
von: Li, Xue, et al.
Veröffentlicht: (2025)
von: Li, Xue, et al.
Veröffentlicht: (2025)
VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping
von: Dong, Haotian, et al.
Veröffentlicht: (2025)
von: Dong, Haotian, et al.
Veröffentlicht: (2025)
Predict to Skip: Linear Multistep Feature Forecasting for Efficient Diffusion Transformers
von: Cui, Hanshuai, et al.
Veröffentlicht: (2026)
von: Cui, Hanshuai, et al.
Veröffentlicht: (2026)
TRIO: Token Reduction via Inference-Objective Guidance for Efficient Vision-Language Models
von: Zhang, Haokui, et al.
Veröffentlicht: (2026)
von: Zhang, Haokui, et al.
Veröffentlicht: (2026)
Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer
von: Huang, Ziyuan, et al.
Veröffentlicht: (2025)
von: Huang, Ziyuan, et al.
Veröffentlicht: (2025)
Prompt-Aware Adaptive Elastic Weight Consolidation for Continual Learning in Medical Vision-Language Models
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
Vision-Centric Activation and Coordination for Multimodal Large Language Models
von: Wang, Yunnan, et al.
Veröffentlicht: (2025)
von: Wang, Yunnan, et al.
Veröffentlicht: (2025)
SkiP: When to Skip and When to Refine for Efficient Robot Manipulation
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
Accelerating Pre-training of Multimodal LLMs via Chain-of-Sight
von: Huang, Ziyuan, et al.
Veröffentlicht: (2024)
von: Huang, Ziyuan, et al.
Veröffentlicht: (2024)
Pretrained Diffusion Models Are Inherently Skipped-Step Samplers
von: Xu, Wenju
Veröffentlicht: (2025)
von: Xu, Wenju
Veröffentlicht: (2025)
Raw Instinct: Trust Your Classifiers and Skip the Conversion
von: Kantas, Christos, et al.
Veröffentlicht: (2024)
von: Kantas, Christos, et al.
Veröffentlicht: (2024)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
von: Hyung, Junha, et al.
Veröffentlicht: (2024)
von: Hyung, Junha, et al.
Veröffentlicht: (2024)
MAIL++: Multi-Modal Bi-directional Agent Layer for Vision-Language Models
von: Chen, Kaixiang, et al.
Veröffentlicht: (2026)
von: Chen, Kaixiang, et al.
Veröffentlicht: (2026)
VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
von: Xu, Siyu, et al.
Veröffentlicht: (2025)
von: Xu, Siyu, et al.
Veröffentlicht: (2025)
Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
Towards Stabilized and Efficient Diffusion Transformers through Long-Skip-Connections with Spectral Constraints
von: Chen, Guanjie, et al.
Veröffentlicht: (2024)
von: Chen, Guanjie, et al.
Veröffentlicht: (2024)
Energy-Driven Adaptive Visual Token Pruning for Efficient Vision-Language Models
von: He, Jialuo, et al.
Veröffentlicht: (2026)
von: He, Jialuo, et al.
Veröffentlicht: (2026)
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Beyond Skip Connection: Pooling and Unpooling Design for Elimination Singularities
von: Sun, Chengkun, et al.
Veröffentlicht: (2024)
von: Sun, Chengkun, et al.
Veröffentlicht: (2024)
Attention Debiasing for Token Pruning in Vision Language Models
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
PPT: Token Pruning and Pooling for Efficient Vision Transformers
von: Wu, Xinjian, et al.
Veröffentlicht: (2023)
von: Wu, Xinjian, et al.
Veröffentlicht: (2023)
S$^{2}$-DMs:Skip-Step Diffusion Models
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration
von: Endo, Mark, et al.
Veröffentlicht: (2024)
von: Endo, Mark, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GM-Skip: Metric-Guided Transformer Block Skipping for Efficient Vision-Language Models
von: Huang, Lianming, et al.
Veröffentlicht: (2025) -
Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves
von: Wu, Shihan, et al.
Veröffentlicht: (2024) -
SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping
von: Li, Jiajun, et al.
Veröffentlicht: (2025) -
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
von: Hartman, Max, et al.
Veröffentlicht: (2025) -
SkipViT: Speeding Up Vision Transformers with a Token-Level Skip Connection
von: Ataiefard, Foozhan, et al.
Veröffentlicht: (2024)