SemanticDialect: Semantic-Aware Mixed-Format Quantization for Video Diffusion Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Jang, Wonsuk, Tambe, Thierry |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BlockDialect: Block-wise Fine-grained Mixed Format Quantization for Energy-Efficient LLM Inference
di: Jang, Wonsuk, et al.
Pubblicazione: (2025)
di: Jang, Wonsuk, et al.
Pubblicazione: (2025)
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
di: Zhang, Jiaji, et al.
Pubblicazione: (2025)
di: Zhang, Jiaji, et al.
Pubblicazione: (2025)
Semantic Alignment and Reinforcement for Data-Free Quantization of Vision Transformers
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
Vision Transformer-based Semantic Communications With Importance-Aware Quantization
di: Park, Joohyuk, et al.
Pubblicazione: (2024)
di: Park, Joohyuk, et al.
Pubblicazione: (2024)
6Bit-Diffusion: Inference-Time Mixed-Precision Quantization for Video Diffusion Models
di: Su, Rundong, et al.
Pubblicazione: (2026)
di: Su, Rundong, et al.
Pubblicazione: (2026)
SDiT: Semantic Region-Adaptive for Diffusion Transformers
di: Lin, Bowen, et al.
Pubblicazione: (2026)
di: Lin, Bowen, et al.
Pubblicazione: (2026)
Video Anomaly Detection with Semantics-Aware Information Bottleneck
di: Li, Juntong, et al.
Pubblicazione: (2025)
di: Li, Juntong, et al.
Pubblicazione: (2025)
DVD-Quant: Data-free Video Diffusion Transformers Quantization
di: Li, Zhiteng, et al.
Pubblicazione: (2025)
di: Li, Zhiteng, et al.
Pubblicazione: (2025)
Multimodal Semantic-Aware Automatic Colorization with Diffusion Prior
di: Wang, Han, et al.
Pubblicazione: (2024)
di: Wang, Han, et al.
Pubblicazione: (2024)
Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
di: Luo, Dezhao, et al.
Pubblicazione: (2024)
di: Luo, Dezhao, et al.
Pubblicazione: (2024)
Pixel-Perfect Depth with Semantics-Prompted Diffusion Transformers
di: Xu, Gangwei, et al.
Pubblicazione: (2025)
di: Xu, Gangwei, et al.
Pubblicazione: (2025)
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning
di: Liu, Caihua, et al.
Pubblicazione: (2025)
di: Liu, Caihua, et al.
Pubblicazione: (2025)
Semantic Aware Diffusion Inverse Tone Mapping
di: Goswami, Abhishek, et al.
Pubblicazione: (2024)
di: Goswami, Abhishek, et al.
Pubblicazione: (2024)
SemanticGen: Video Generation in Semantic Space
di: Bai, Jianhong, et al.
Pubblicazione: (2025)
di: Bai, Jianhong, et al.
Pubblicazione: (2025)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
di: Yu, Zhu, et al.
Pubblicazione: (2024)
di: Yu, Zhu, et al.
Pubblicazione: (2024)
SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models
di: Lian, Jiesong, et al.
Pubblicazione: (2026)
di: Lian, Jiesong, et al.
Pubblicazione: (2026)
A Hidden Semantic Bottleneck in Conditional Embeddings of Diffusion Transformers
di: Pham, Trung X., et al.
Pubblicazione: (2026)
di: Pham, Trung X., et al.
Pubblicazione: (2026)
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
An Analysis on Quantizing Diffusion Transformers
di: Yang, Yuewei, et al.
Pubblicazione: (2024)
di: Yang, Yuewei, et al.
Pubblicazione: (2024)
Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers
di: Feng, Weilun, et al.
Pubblicazione: (2025)
di: Feng, Weilun, et al.
Pubblicazione: (2025)
Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing
di: Liu, Lin, et al.
Pubblicazione: (2026)
di: Liu, Lin, et al.
Pubblicazione: (2026)
Dual Semantic-Aware Network for Noise Suppressed Ultrasound Video Segmentation
di: Zhou, Ling, et al.
Pubblicazione: (2025)
di: Zhou, Ling, et al.
Pubblicazione: (2025)
Multilevel Semantic-Aware Model for AI-Generated Video Quality Assessment
di: Li, Jiaze, et al.
Pubblicazione: (2025)
di: Li, Jiaze, et al.
Pubblicazione: (2025)
A Semantic and Motion-Aware Spatiotemporal Transformer Network for Action Detection
di: Korban, Matthew, et al.
Pubblicazione: (2024)
di: Korban, Matthew, et al.
Pubblicazione: (2024)
NEGATE: Constrained Semantic Guidance for Linguistic Negation in Text-to-Video Diffusion
di: Kang, Taewon, et al.
Pubblicazione: (2026)
di: Kang, Taewon, et al.
Pubblicazione: (2026)
Single-step Diffusion-based Video Coding with Semantic-Temporal Guidance
di: Xue, Naifu, et al.
Pubblicazione: (2025)
di: Xue, Naifu, et al.
Pubblicazione: (2025)
Diffusion-based RGB-D Semantic Segmentation with Deformable Attention Transformer
di: Bui, Minh, et al.
Pubblicazione: (2024)
di: Bui, Minh, et al.
Pubblicazione: (2024)
Sampling-Aware Quantization for Diffusion Models
di: Zeng, Qian, et al.
Pubblicazione: (2025)
di: Zeng, Qian, et al.
Pubblicazione: (2025)
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
di: Feng, Weilun, et al.
Pubblicazione: (2025)
di: Feng, Weilun, et al.
Pubblicazione: (2025)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
Balancing Saliency and Coverage: Semantic Prominence-Aware Budgeting for Visual Token Compression in VLMs
di: Lee, Jaehoon, et al.
Pubblicazione: (2026)
di: Lee, Jaehoon, et al.
Pubblicazione: (2026)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
di: Lew, Jaihyun, et al.
Pubblicazione: (2024)
di: Lew, Jaihyun, et al.
Pubblicazione: (2024)
TreeQ: Pushing the Quantization Boundary of Diffusion Transformer via Tree-Structured Mixed-Precision Search
di: Yang, Kaicheng, et al.
Pubblicazione: (2025)
di: Yang, Kaicheng, et al.
Pubblicazione: (2025)
AttAnchor: Guiding Cross-Modal Token Alignment in VLMs with Attention Anchors
di: Zhang, Junyang, et al.
Pubblicazione: (2025)
di: Zhang, Junyang, et al.
Pubblicazione: (2025)
Layered 3D Human Generation via Semantic-Aware Diffusion Model
di: Wang, Yi, et al.
Pubblicazione: (2023)
di: Wang, Yi, et al.
Pubblicazione: (2023)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
di: Zhu, Jingyuan, et al.
Pubblicazione: (2026)
di: Zhu, Jingyuan, et al.
Pubblicazione: (2026)
Training-Free Semantic Video Composition via Pre-trained Diffusion Model
di: Guo, Jiaqi, et al.
Pubblicazione: (2024)
di: Guo, Jiaqi, et al.
Pubblicazione: (2024)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
di: Wang, Qian, et al.
Pubblicazione: (2024)
di: Wang, Qian, et al.
Pubblicazione: (2024)
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers
di: Yi, Sanghyun, et al.
Pubblicazione: (2025)
di: Yi, Sanghyun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
BlockDialect: Block-wise Fine-grained Mixed Format Quantization for Energy-Efficient LLM Inference
di: Jang, Wonsuk, et al.
Pubblicazione: (2025) -
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
di: Zhang, Jiaji, et al.
Pubblicazione: (2025) -
Semantic Alignment and Reinforcement for Data-Free Quantization of Vision Transformers
di: Zhong, Yunshan, et al.
Pubblicazione: (2024) -
Vision Transformer-based Semantic Communications With Importance-Aware Quantization
di: Park, Joohyuk, et al.
Pubblicazione: (2024) -
6Bit-Diffusion: Inference-Time Mixed-Precision Quantization for Video Diffusion Models
di: Su, Rundong, et al.
Pubblicazione: (2026)