Visual Detector Compression via Location-Aware Discriminant Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Lan, Qizhen, Choi, Jung Im, Tian, Qing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLoCKDistill: Consistent Location-and-Context-aware Knowledge Distillation for DETRs
by: Lan, Qizhen, et al.
Published: (2025)
by: Lan, Qizhen, et al.
Published: (2025)
ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation
by: Lan, Qizhen, et al.
Published: (2025)
by: Lan, Qizhen, et al.
Published: (2025)
Learnable Instance Attention Filtering for Adaptive Detector Distillation
by: Liu, Chen, et al.
Published: (2026)
by: Liu, Chen, et al.
Published: (2026)
DGL-GAN: Discriminator Guided Learning for GAN Compression
by: Tian, Yuesong, et al.
Published: (2021)
by: Tian, Yuesong, et al.
Published: (2021)
ReCo-KD: Region- and Context-Aware Knowledge Distillation for Efficient 3D Medical Image Segmentation
by: Lan, Qizhen, et al.
Published: (2026)
by: Lan, Qizhen, et al.
Published: (2026)
Balancing Saliency and Coverage: Semantic Prominence-Aware Budgeting for Visual Token Compression in VLMs
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
3A-YOLO: New Real-Time Object Detectors with Triple Discriminative Awareness and Coordinated Representations
by: Wu, Xuecheng, et al.
Published: (2024)
by: Wu, Xuecheng, et al.
Published: (2024)
FLoC: Facility Location-Based Efficient Visual Token Compression for Long Video Understanding
by: Cho, Janghoon, et al.
Published: (2025)
by: Cho, Janghoon, et al.
Published: (2025)
Multi-dimension Transformer with Attention-based Filtering for Medical Image Segmentation
by: Wang, Wentao, et al.
Published: (2024)
by: Wang, Wentao, et al.
Published: (2024)
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
Comparison Of Deep Object Detectors On A New Vulnerable Pedestrian Dataset
by: Sharma, Devansh, et al.
Published: (2022)
by: Sharma, Devansh, et al.
Published: (2022)
DMOFC: Discrimination Metric-Optimized Feature Compression
by: Gao, Changsheng, et al.
Published: (2024)
by: Gao, Changsheng, et al.
Published: (2024)
SARA: Scene-Aware Reconstruction Accelerator
by: Lee, Jee Won, et al.
Published: (2026)
by: Lee, Jee Won, et al.
Published: (2026)
From Performance to Practice: Knowledge-Distilled Segmentator for On-Premises Clinical Workflows
by: Lan, Qizhen, et al.
Published: (2026)
by: Lan, Qizhen, et al.
Published: (2026)
Understanding Deep Representation Learning via Layerwise Feature Compression and Discrimination
by: Wang, Peng, et al.
Published: (2023)
by: Wang, Peng, et al.
Published: (2023)
Location-Aware Pretraining for Medical Difference Visual Question Answering
by: Musinguzi, Denis, et al.
Published: (2026)
by: Musinguzi, Denis, et al.
Published: (2026)
Head-Aware KV Cache Compression for Efficient Visual Autoregressive Modeling
by: Qin, Ziran, et al.
Published: (2025)
by: Qin, Ziran, et al.
Published: (2025)
Can Visual Input Be Compressed? A Visual Token Compression Benchmark for Large Multimodal Models
by: Peng, Tianfan, et al.
Published: (2025)
by: Peng, Tianfan, et al.
Published: (2025)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
From Objects to Events: Unlocking Complex Visual Understanding in Object Detectors via LLM-guided Symbolic Reasoning
by: Zeng, Yuhui, et al.
Published: (2025)
by: Zeng, Yuhui, et al.
Published: (2025)
Interpretable Multimodal Framework for Human-Centered Street Assessment: Integrating Visual-Language Models for Perceptual Urban Diagnostics
by: Lan, HaoTian
Published: (2025)
by: Lan, HaoTian
Published: (2025)
On Learning Discriminative Features from Synthesized Data for Self-Supervised Fine-Grained Visual Recognition
by: Wang, Zihu, et al.
Published: (2024)
by: Wang, Zihu, et al.
Published: (2024)
Enhancing Visual Classification using Comparative Descriptors
by: Lee, Hankyeol, et al.
Published: (2024)
by: Lee, Hankyeol, et al.
Published: (2024)
Content-Aware Mamba for Learned Image Compression
by: Chen, Yunuo, et al.
Published: (2025)
by: Chen, Yunuo, et al.
Published: (2025)
Resolving Blind Inverse Problems under Dynamic Range Compression via Structured Forward Operator Modeling
by: Liu, Muyu, et al.
Published: (2026)
by: Liu, Muyu, et al.
Published: (2026)
Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression
by: Zhang, Dingyuan, et al.
Published: (2024)
by: Zhang, Dingyuan, et al.
Published: (2024)
Text Grouping Adapter: Adapting Pre-trained Text Detector for Layout Analysis
by: Bi, Tianci, et al.
Published: (2024)
by: Bi, Tianci, et al.
Published: (2024)
Adaptive-VoCo: Complexity-Aware Visual Token Compression for Vision-Language Models
by: Guo, Xiaoyang, et al.
Published: (2025)
by: Guo, Xiaoyang, et al.
Published: (2025)
PRISM: Color-Stratified Point Cloud Sampling
by: Lim, Hansol, et al.
Published: (2026)
by: Lim, Hansol, et al.
Published: (2026)
Region-based Cluster Discrimination for Visual Representation Learning
by: Xie, Yin, et al.
Published: (2025)
by: Xie, Yin, et al.
Published: (2025)
Multi-label Cluster Discrimination for Visual Representation Learning
by: An, Xiang, et al.
Published: (2024)
by: An, Xiang, et al.
Published: (2024)
SARD: Segmentation-Aware Anomaly Synthesis via Region-Constrained Diffusion with Discriminative Mask Guidance
by: Wang, Yanshu, et al.
Published: (2025)
by: Wang, Yanshu, et al.
Published: (2025)
Discriminative Perception via Anchored Description for Reasoning Segmentation
by: Yang, Tao, et al.
Published: (2026)
by: Yang, Tao, et al.
Published: (2026)
Visual-RRT: Finding Paths toward Visual-Goals via Differentiable Rendering
by: Lee, Sebin, et al.
Published: (2026)
by: Lee, Sebin, et al.
Published: (2026)
LocCa: Visual Pretraining with Location-aware Captioners
by: Wan, Bo, et al.
Published: (2024)
by: Wan, Bo, et al.
Published: (2024)
DocVXQA: Context-Aware Visual Explanations for Document Question Answering
by: Souibgui, Mohamed Ali, et al.
Published: (2025)
by: Souibgui, Mohamed Ali, et al.
Published: (2025)
Free-VSC: Free Semantics from Visual Foundation Models for Unsupervised Video Semantic Compression
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
Continual Multimodal Egocentric Activity Recognition via Modality-Aware Novel Detection
by: Lim, Wonseon, et al.
Published: (2026)
by: Lim, Wonseon, et al.
Published: (2026)
Modality-Decoupled RGB-Thermal Object Detector via Query Fusion
by: Tian, Chao, et al.
Published: (2026)
by: Tian, Chao, et al.
Published: (2026)
Lost in Distortion: Uncovering the Domain Gap Between Computer Vision and Brain Imaging -- A Study on Pretraining for Age Prediction
by: Zhang, Yanteng, et al.
Published: (2025)
by: Zhang, Yanteng, et al.
Published: (2025)
Similar Items
-
CLoCKDistill: Consistent Location-and-Context-aware Knowledge Distillation for DETRs
by: Lan, Qizhen, et al.
Published: (2025) -
ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation
by: Lan, Qizhen, et al.
Published: (2025) -
Learnable Instance Attention Filtering for Adaptive Detector Distillation
by: Liu, Chen, et al.
Published: (2026) -
DGL-GAN: Discriminator Guided Learning for GAN Compression
by: Tian, Yuesong, et al.
Published: (2021) -
ReCo-KD: Region- and Context-Aware Knowledge Distillation for Efficient 3D Medical Image Segmentation
by: Lan, Qizhen, et al.
Published: (2026)