Saved in:
| Main Authors: | Moon, Jaehyeon, Kim, Dohyung, Cheon, Junyong, Ham, Bumsub |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2404.00928 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward INT4 Fixed-Point Training via Exploring Quantization Error for Gradients
by: Kim, Dohyung, et al.
Published: (2024)
by: Kim, Dohyung, et al.
Published: (2024)
Scheduling Weight Transitions for Quantization-Aware Training
by: Lee, Junghyup, et al.
Published: (2024)
by: Lee, Junghyup, et al.
Published: (2024)
Maximizing the Position Embedding for Vision Transformers with Global Average Pooling
by: Lee, Wonjun, et al.
Published: (2025)
by: Lee, Wonjun, et al.
Published: (2025)
AZ-NAS: Assembling Zero-Cost Proxies for Network Architecture Search
by: Lee, Junghyup, et al.
Published: (2024)
by: Lee, Junghyup, et al.
Published: (2024)
AccuQuant: Simulating Multiple Denoising Steps for Quantizing Diffusion Models
by: Lee, Seunghoon, et al.
Published: (2025)
by: Lee, Seunghoon, et al.
Published: (2025)
Relational Feature Caching for Accelerating Diffusion Transformers
by: Son, Byunggwan, et al.
Published: (2026)
by: Son, Byunggwan, et al.
Published: (2026)
Subnet-Aware Dynamic Supernet Training for Neural Architecture Search
by: Jeon, Jeimin, et al.
Published: (2025)
by: Jeon, Jeimin, et al.
Published: (2025)
GrowTAS: Progressive Expansion from Small to Large Subnets for Efficient ViT Architecture Search
by: Lee, Hyunju, et al.
Published: (2025)
by: Lee, Hyunju, et al.
Published: (2025)
TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts
by: Jeon, Jeimin, et al.
Published: (2026)
by: Jeon, Jeimin, et al.
Published: (2026)
ZIP: An Efficient Zeroth-order Prompt Tuning for Black-box Vision-Language Models
by: Park, Seonghwan, et al.
Published: (2025)
by: Park, Seonghwan, et al.
Published: (2025)
Demonstrating the Efficacy of Kolmogorov-Arnold Networks in Vision Tasks
by: Cheon, Minjong
Published: (2024)
by: Cheon, Minjong
Published: (2024)
Efficient Few-Shot Neural Architecture Search by Counting the Number of Nonlinear Functions
by: Oh, Youngmin, et al.
Published: (2024)
by: Oh, Youngmin, et al.
Published: (2024)
Activation Quantization of Vision Encoders Needs Prefixing Registers
by: Kim, Seunghyeon, et al.
Published: (2025)
by: Kim, Seunghyeon, et al.
Published: (2025)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
by: Li, Zhengang, et al.
Published: (2024)
by: Li, Zhengang, et al.
Published: (2024)
DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models
by: Ryu, Hyogon, et al.
Published: (2025)
by: Ryu, Hyogon, et al.
Published: (2025)
JLT: Clean-Latent Prediction in Latent Diffusion Transformers
by: Fu, Funing, et al.
Published: (2026)
by: Fu, Funing, et al.
Published: (2026)
Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference
by: Cho, Hyeonwoo, et al.
Published: (2026)
by: Cho, Hyeonwoo, et al.
Published: (2026)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
by: Sharify, Sayeh, et al.
Published: (2026)
by: Sharify, Sayeh, et al.
Published: (2026)
Set2Seq Transformer: Temporal and Position-Aware Set Representations for Sequential Multiple-Instance Learning
by: Efthymiou, Athanasios, et al.
Published: (2024)
by: Efthymiou, Athanasios, et al.
Published: (2024)
Cluster-Aware Similarity Diffusion for Instance Retrieval
by: Luo, Jifei, et al.
Published: (2024)
by: Luo, Jifei, et al.
Published: (2024)
Learning to Transform for Generalizable Instance-wise Invariance
by: Singhal, Utkarsh, et al.
Published: (2023)
by: Singhal, Utkarsh, et al.
Published: (2023)
FYI: Flip Your Images for Dataset Distillation
by: Son, Byunggwan, et al.
Published: (2024)
by: Son, Byunggwan, et al.
Published: (2024)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
by: Lee, Sanghoon, et al.
Published: (2026)
by: Lee, Sanghoon, et al.
Published: (2026)
3DPillars: Pillar-based two-stage 3D object detection
by: Noh, Jongyoun, et al.
Published: (2025)
by: Noh, Jongyoun, et al.
Published: (2025)
The Effects of Grouped Structural Global Pruning of Vision Transformers on Domain Generalisation
by: Riaz, Hamza, et al.
Published: (2025)
by: Riaz, Hamza, et al.
Published: (2025)
SIS-Challenge: Event-based Spatio-temporal Instance Segmentation Challenge at the CVPR 2025 Event-based Vision Workshop
by: Hamann, Friedhelm, et al.
Published: (2025)
by: Hamann, Friedhelm, et al.
Published: (2025)
GNN-ViTCap: GNN-Enhanced Multiple Instance Learning with Vision Transformers for Whole Slide Image Classification and Captioning
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
QuantAttack: Exploiting Dynamic Quantization to Attack Vision Transformers
by: Baras, Amit, et al.
Published: (2023)
by: Baras, Amit, et al.
Published: (2023)
QGen: On the Ability to Generalize in Quantization Aware Training
by: AskariHemmat, MohammadHossein, et al.
Published: (2024)
by: AskariHemmat, MohammadHossein, et al.
Published: (2024)
Data-Augmented Quantization-Aware Knowledge Distillation
by: Kur, Justin, et al.
Published: (2025)
by: Kur, Justin, et al.
Published: (2025)
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control
by: Park, Seongmin, et al.
Published: (2024)
by: Park, Seongmin, et al.
Published: (2024)
From Pixels to Perception: Interpretable Predictions via Instance-wise Grouped Feature Selection
by: Vandenhirtz, Moritz, et al.
Published: (2025)
by: Vandenhirtz, Moritz, et al.
Published: (2025)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
KAN-CL: Per-Knot Importance Regularization for Continual Learning with Kolmogorov-Arnold Networks
by: Cheon, Minjong
Published: (2026)
by: Cheon, Minjong
Published: (2026)
Sharpness-Aware Data Generation for Zero-shot Quantization
by: Hoang-Anh, Dung, et al.
Published: (2025)
by: Hoang-Anh, Dung, et al.
Published: (2025)
Visualizing the loss landscape of Self-supervised Vision Transformer
by: Lee, Youngwan, et al.
Published: (2024)
by: Lee, Youngwan, et al.
Published: (2024)
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
by: Lee, Wonjun, et al.
Published: (2025)
by: Lee, Wonjun, et al.
Published: (2025)
CAPA: Contribution-Aware Pruning and FFN Approximation for Efficient Large Vision-Language Models
by: Jha, Samyak, et al.
Published: (2026)
by: Jha, Samyak, et al.
Published: (2026)
Disentangled Representations for Short-Term and Long-Term Person Re-Identification
by: Eom, Chanho, et al.
Published: (2024)
by: Eom, Chanho, et al.
Published: (2024)
Survey of Quantization Techniques for On-Device Vision-based Crack Detection
by: Zhang, Yuxuan, et al.
Published: (2025)
by: Zhang, Yuxuan, et al.
Published: (2025)
Similar Items
-
Toward INT4 Fixed-Point Training via Exploring Quantization Error for Gradients
by: Kim, Dohyung, et al.
Published: (2024) -
Scheduling Weight Transitions for Quantization-Aware Training
by: Lee, Junghyup, et al.
Published: (2024) -
Maximizing the Position Embedding for Vision Transformers with Global Average Pooling
by: Lee, Wonjun, et al.
Published: (2025) -
AZ-NAS: Assembling Zero-Cost Proxies for Network Architecture Search
by: Lee, Junghyup, et al.
Published: (2024) -
AccuQuant: Simulating Multiple Denoising Steps for Quantizing Diffusion Models
by: Lee, Seunghoon, et al.
Published: (2025)