Rethinking Token Reduction for Diffusion Models via Output-Similarity-Awareness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Hangyeol, Lee, Hyojeong, Kim, Joo-Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
Target-Aware Video Diffusion Models
von: Kim, Taeksoo, et al.
Veröffentlicht: (2025)
von: Kim, Taeksoo, et al.
Veröffentlicht: (2025)
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
von: Kim, Minseo, et al.
Veröffentlicht: (2026)
von: Kim, Minseo, et al.
Veröffentlicht: (2026)
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
GranQ: Efficient Channel-wise Quantization via Vectorized Pre-Scaling for Zero-Shot QAT
von: Hong, Inpyo, et al.
Veröffentlicht: (2025)
von: Hong, Inpyo, et al.
Veröffentlicht: (2025)
DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models
von: Min, Jaewon, et al.
Veröffentlicht: (2026)
von: Min, Jaewon, et al.
Veröffentlicht: (2026)
FastSTAR: Spatiotemporal Token Pruning for Efficient Autoregressive Video Synthesis
von: Yune, Sungwoong, et al.
Veröffentlicht: (2026)
von: Yune, Sungwoong, et al.
Veröffentlicht: (2026)
RAD: Region-Aware Diffusion Models for Image Inpainting
von: Kim, Sora, et al.
Veröffentlicht: (2024)
von: Kim, Sora, et al.
Veröffentlicht: (2024)
Rethinking Token Reduction for Large Vision-Language Models
von: Wang, Yi, et al.
Veröffentlicht: (2026)
von: Wang, Yi, et al.
Veröffentlicht: (2026)
Adjusting Initial Noise to Mitigate Memorization in Text-to-Image Diffusion Models
von: Han, Hyeonggeun, et al.
Veröffentlicht: (2025)
von: Han, Hyeonggeun, et al.
Veröffentlicht: (2025)
ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning
von: Lee, Yuna, et al.
Veröffentlicht: (2026)
von: Lee, Yuna, et al.
Veröffentlicht: (2026)
Similarity-Aware Token Pruning: Your VLM but Faster
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2025)
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2025)
Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models
von: Ma, Kexin, et al.
Veröffentlicht: (2026)
von: Ma, Kexin, et al.
Veröffentlicht: (2026)
Sortblock: Similarity-Aware Feature Reuse for Diffusion Model
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
Ditto: Accelerating Diffusion Model via Temporal Value Similarity
von: Kim, Sungbin, et al.
Veröffentlicht: (2025)
von: Kim, Sungbin, et al.
Veröffentlicht: (2025)
Dexterous World Models
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
VScan: Rethinking Visual Token Reduction for Efficient Large Vision-Language Models
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
Neighbor-Aware Token Reduction via Hilbert Curve for Vision Transformers
von: Li, Yunge, et al.
Veröffentlicht: (2025)
von: Li, Yunge, et al.
Veröffentlicht: (2025)
Rethinking Visual Token Reduction in LVLMs Under Cross-Modal Misalignment
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
Cached Adaptive Token Merging: Dynamic Token Reduction and Redundant Computation Elimination in Diffusion Model
von: Saghatchian, Omid, et al.
Veröffentlicht: (2025)
von: Saghatchian, Omid, et al.
Veröffentlicht: (2025)
Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2025)
von: Hyun, Jeongseok, et al.
Veröffentlicht: (2025)
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
Let Triggers Control: Frequency-Aware Dropout for Effective Token Control
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
How Diffusion Models Memorize
von: Kim, Juyeop, et al.
Veröffentlicht: (2025)
von: Kim, Juyeop, et al.
Veröffentlicht: (2025)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
von: Lee, Inhee, et al.
Veröffentlicht: (2024)
von: Lee, Inhee, et al.
Veröffentlicht: (2024)
REAR: Rethinking Visual Autoregressive Models via Generator-Tokenizer Consistency Regularization
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
von: Lew, Jaihyun, et al.
Veröffentlicht: (2024)
Advanced Knowledge Transfer: Refined Feature Distillation for Zero-Shot Quantization in Edge Computing
von: Hong, Inpyo, et al.
Veröffentlicht: (2024)
von: Hong, Inpyo, et al.
Veröffentlicht: (2024)
SUPER Decoder Block for Reconstruction-Aware U-Net Variants
von: Joo, Siheon, et al.
Veröffentlicht: (2025)
von: Joo, Siheon, et al.
Veröffentlicht: (2025)
Treating Motion as Option with Output Selection for Unsupervised Video Object Segmentation
von: Cho, Suhwan, et al.
Veröffentlicht: (2023)
von: Cho, Suhwan, et al.
Veröffentlicht: (2023)
TARA: Token-Aware LoRA for Composable Personalization in Diffusion Models
von: Peng, Yuqi, et al.
Veröffentlicht: (2025)
von: Peng, Yuqi, et al.
Veröffentlicht: (2025)
GOATex: Geometry & Occlusion-Aware Texturing
von: Kim, Hyunjin, et al.
Veröffentlicht: (2025)
von: Kim, Hyunjin, et al.
Veröffentlicht: (2025)
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2025)
Geometry-Aware Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2026)
von: Lee, Junho, et al.
Veröffentlicht: (2026)
Rethinking Pose Refinement in 3D Gaussian Splatting under Pose Prior and Geometric Uncertainty
von: Kong, Mangyu, et al.
Veröffentlicht: (2026)
von: Kong, Mangyu, et al.
Veröffentlicht: (2026)
Semi-Supervised Domain Adaptation for Wildfire Detection
von: Jang, JooYoung, et al.
Veröffentlicht: (2024)
von: Jang, JooYoung, et al.
Veröffentlicht: (2024)
Landscape-Awareness for Geometric View Diffusion Model
von: Chen, Yan-Ting, et al.
Veröffentlicht: (2026)
von: Chen, Yan-Ting, et al.
Veröffentlicht: (2026)
Blockwise Flow Matching: Improving Flow Matching Models For Efficient High-Quality Generation
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
Lesion-Aware Post-Training of Latent Diffusion Models for Synthesizing Diffusion MRI from CT Perfusion
von: Lee, Junhyeok, et al.
Veröffentlicht: (2025)
von: Lee, Junhyeok, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026) -
Target-Aware Video Diffusion Models
von: Kim, Taeksoo, et al.
Veröffentlicht: (2025) -
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
von: Kim, Minseo, et al.
Veröffentlicht: (2026) -
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025) -
GranQ: Efficient Channel-wise Quantization via Vectorized Pre-Scaling for Zero-Shot QAT
von: Hong, Inpyo, et al.
Veröffentlicht: (2025)