Inlier-Centric Post-Training Quantization for Object Detection Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Minsu, Lee, Dongyeun, Yu, Jaemyung, Hur, Jiwan, Kim, Giseop, Kim, Junmo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
SSG: Scaled Spatial Guidance for Multi-Scale Visual Autoregressive Generation
von: Shin, Youngwoo, et al.
Veröffentlicht: (2026)
von: Shin, Youngwoo, et al.
Veröffentlicht: (2026)
Addressing Diverging Training Costs using BEVRestore for High-resolution Bird's Eye View Map Construction
von: Kim, Minsu, et al.
Veröffentlicht: (2024)
von: Kim, Minsu, et al.
Veröffentlicht: (2024)
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
von: Kim, Minseo, et al.
Veröffentlicht: (2026)
von: Kim, Minseo, et al.
Veröffentlicht: (2026)
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
Comparison Reveals Commonality: Customized Image Generation through Contrastive Inversion
von: Kim, Minseo, et al.
Veröffentlicht: (2025)
von: Kim, Minseo, et al.
Veröffentlicht: (2025)
IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
von: Hur, Jiwan, et al.
Veröffentlicht: (2024)
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
von: Lee, Hansang, et al.
Veröffentlicht: (2017)
von: Lee, Hansang, et al.
Veröffentlicht: (2017)
Locality-Aware Zero-Shot Human-Object Interaction Detection
von: Kim, Sanghyun, et al.
Veröffentlicht: (2025)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2025)
Self-supervised Transformation Learning for Equivariant Representations
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
Post-Training Quantization via Residual Truncation and Zero Suppression for Diffusion Models
von: Kim, Donghoon, et al.
Veröffentlicht: (2025)
von: Kim, Donghoon, et al.
Veröffentlicht: (2025)
Harnessing the Power of Training-Free Techniques in Text-to-2D Generation for Text-to-3D Generation via Score Distillation Sampling
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
von: Lee, Junhong, et al.
Veröffentlicht: (2025)
TRAN-D: 2D Gaussian Splatting-based Sparse-view Transparent Object Depth Reconstruction via Physics Simulation for Scene Update
von: Kim, Jeongyun, et al.
Veröffentlicht: (2025)
von: Kim, Jeongyun, et al.
Veröffentlicht: (2025)
FreeAction: Training-Free Techniques for Enhanced Fidelity of Trajectory-to-Video Generation
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
von: Lee, Chanho, et al.
Veröffentlicht: (2023)
von: Lee, Chanho, et al.
Veröffentlicht: (2023)
VERIA: Verification-Centric Multimodal Instance Augmentation for Long-Tailed 3D Object Detection
von: Lee, Jumin, et al.
Veröffentlicht: (2026)
von: Lee, Jumin, et al.
Veröffentlicht: (2026)
PTQ4VM: Post-Training Quantization for Visual Mamba
von: Cho, Younghyun, et al.
Veröffentlicht: (2024)
von: Cho, Younghyun, et al.
Veröffentlicht: (2024)
Similarity-Aware Selective State-Space Modeling for Semantic Correspondence
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
von: Kim, Seungwook, et al.
Veröffentlicht: (2025)
Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding
von: Kim, Jiwan, et al.
Veröffentlicht: (2026)
von: Kim, Jiwan, et al.
Veröffentlicht: (2026)
Stream and Query-guided Feature Aggregation for Efficient and Effective 3D Occupancy Prediction
von: Moon, Seokha, et al.
Veröffentlicht: (2025)
von: Moon, Seokha, et al.
Veröffentlicht: (2025)
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
OSPO: Object-Centric Self-Improving Preference Optimization for Text-to-Image Generation
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
CR-QAT: Curriculum Relational Quantization-Aware Training for Open-Vocabulary Object Detection
von: Park, Jinyeong, et al.
Veröffentlicht: (2026)
von: Park, Jinyeong, et al.
Veröffentlicht: (2026)
The Effects of Mixed Sample Data Augmentation are Class Dependent
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
Inspecting Explainability of Transformer Models with Additional Statistical Information
von: Nguyen, Hoang C., et al.
Veröffentlicht: (2023)
von: Nguyen, Hoang C., et al.
Veröffentlicht: (2023)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
Part-Aware Bottom-Up Group Reasoning for Fine-Grained Social Interaction Detection
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
Scheduling Weight Transitions for Quantization-Aware Training
von: Lee, Junghyup, et al.
Veröffentlicht: (2024)
von: Lee, Junghyup, et al.
Veröffentlicht: (2024)
Modeling Stereo-Confidence Out of the End-to-End Stereo-Matching Network via Disparity Plane Sweep
von: Lee, Jae Young, et al.
Veröffentlicht: (2024)
von: Lee, Jae Young, et al.
Veröffentlicht: (2024)
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
Towards More Practical Group Activity Detection: A New Benchmark and Model
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025) -
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025) -
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025) -
SSG: Scaled Spatial Guidance for Multi-Scale Visual Autoregressive Generation
von: Shin, Youngwoo, et al.
Veröffentlicht: (2026) -
Addressing Diverging Training Costs using BEVRestore for High-resolution Bird's Eye View Map Construction
von: Kim, Minsu, et al.
Veröffentlicht: (2024)