Computer Vision Model Compression Techniques for Embedded Systems: A Survey
Fuente:
arXiv
Salvato in:
| Autori principali: | Lopes, Alexandre, Santos, Fernando Pereira dos, de Oliveira, Diulhio, Schiezaro, Mauricio, Pedrini, Helio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CCNeXt: An Effective Self-Supervised Stereo Depth Estimation Approach
di: Lopes, Alexandre, et al.
Pubblicazione: (2025)
di: Lopes, Alexandre, et al.
Pubblicazione: (2025)
Identification of Deforestation Areas in the Amazon Rainforest Using Change Detection Models
di: Konishi, Christian Massao, et al.
Pubblicazione: (2025)
di: Konishi, Christian Massao, et al.
Pubblicazione: (2025)
CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition
di: Estevam, Valter, et al.
Pubblicazione: (2026)
di: Estevam, Valter, et al.
Pubblicazione: (2026)
Dense Video Captioning Using Unsupervised Semantic Information
di: Estevam, Valter, et al.
Pubblicazione: (2021)
di: Estevam, Valter, et al.
Pubblicazione: (2021)
P-NOC: adversarial training of CAM generating networks for robust weakly supervised semantic segmentation priors
di: David, Lucas, et al.
Pubblicazione: (2023)
di: David, Lucas, et al.
Pubblicazione: (2023)
Weakly Supervised Attention-based Models Using Activation Maps for Citrus Mite and Insect Pest Classification
di: Bollis, Edson, et al.
Pubblicazione: (2021)
di: Bollis, Edson, et al.
Pubblicazione: (2021)
Self-Organizing Visual Prototypes for Non-Parametric Representation Learning
di: Silva, Thalles, et al.
Pubblicazione: (2025)
di: Silva, Thalles, et al.
Pubblicazione: (2025)
Learning from Memory: Non-Parametric Memory Augmented Self-Supervised Learning of Visual Features
di: Silva, Thalles, et al.
Pubblicazione: (2024)
di: Silva, Thalles, et al.
Pubblicazione: (2024)
Comprehensive Survey of Model Compression and Speed up for Vision Transformers
di: Chen, Feiyang, et al.
Pubblicazione: (2024)
di: Chen, Feiyang, et al.
Pubblicazione: (2024)
Model Compression Techniques in Biometrics Applications: A Survey
di: Caldeira, Eduarda, et al.
Pubblicazione: (2024)
di: Caldeira, Eduarda, et al.
Pubblicazione: (2024)
SMART-Vision: Survey of Modern Action Recognition Techniques in Vision
di: AlShami, Ali K., et al.
Pubblicazione: (2025)
di: AlShami, Ali K., et al.
Pubblicazione: (2025)
A Survey of Medical Vision-and-Language Applications and Their Techniques
di: Chen, Qi, et al.
Pubblicazione: (2024)
di: Chen, Qi, et al.
Pubblicazione: (2024)
Evaluating the Impact of Compression Techniques on the Robustness of CNNs under Natural Corruptions
di: Da Silva, Itallo Patrick Castro Alves, et al.
Pubblicazione: (2025)
di: Da Silva, Itallo Patrick Castro Alves, et al.
Pubblicazione: (2025)
A Framework for Generating Artificial Datasets to Validate Absolute and Relative Position Concepts
di: de Araújo, George Corrêa, et al.
Pubblicazione: (2025)
di: de Araújo, George Corrêa, et al.
Pubblicazione: (2025)
A Survey on the Robustness of Computer Vision Models against Common Corruptions
di: Wang, Shunxin, et al.
Pubblicazione: (2023)
di: Wang, Shunxin, et al.
Pubblicazione: (2023)
CompressNAS : A Fast and Efficient Technique for Model Compression using Decomposition
di: Sah, Sudhakar, et al.
Pubblicazione: (2025)
di: Sah, Sudhakar, et al.
Pubblicazione: (2025)
FairPIVARA: Reducing and Assessing Biases in CLIP-Based Multimodal Models
di: Moreira, Diego A. B., et al.
Pubblicazione: (2024)
di: Moreira, Diego A. B., et al.
Pubblicazione: (2024)
Fairness and Bias Mitigation in Computer Vision: A Survey
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
Vision Mamba in Remote Sensing: A Comprehensive Survey of Techniques, Applications and Outlook
di: Bao, Muyi, et al.
Pubblicazione: (2025)
di: Bao, Muyi, et al.
Pubblicazione: (2025)
Self-ReS: Self-Reflection in Large Vision-Language Models for Long Video Understanding
di: Pereira, Joao, et al.
Pubblicazione: (2025)
di: Pereira, Joao, et al.
Pubblicazione: (2025)
Vision Transformers on the Edge: A Comprehensive Survey of Model Compression and Acceleration Strategies
di: Saha, Shaibal, et al.
Pubblicazione: (2025)
di: Saha, Shaibal, et al.
Pubblicazione: (2025)
Vision-Language Models for Vision Tasks: A Survey
di: Zhang, Jingyi, et al.
Pubblicazione: (2023)
di: Zhang, Jingyi, et al.
Pubblicazione: (2023)
Token Compression Meets Compact Vision Transformers: A Survey and Comparative Evaluation for Edge AI
di: Nguyen, Phat, et al.
Pubblicazione: (2025)
di: Nguyen, Phat, et al.
Pubblicazione: (2025)
A Review of Transformer-Based Models for Computer Vision Tasks: Capturing Global Context and Spatial Relationships
di: Pereira, Gracile Astlin, et al.
Pubblicazione: (2024)
di: Pereira, Gracile Astlin, et al.
Pubblicazione: (2024)
Physical Adversarial Attack meets Computer Vision: A Decade Survey
di: Wei, Hui, et al.
Pubblicazione: (2022)
di: Wei, Hui, et al.
Pubblicazione: (2022)
Survey of Quantization Techniques for On-Device Vision-based Crack Detection
di: Zhang, Yuxuan, et al.
Pubblicazione: (2025)
di: Zhang, Yuxuan, et al.
Pubblicazione: (2025)
Survey of Multimodal Geospatial Foundation Models: Techniques, Applications, and Challenges
di: Yang, Liling, et al.
Pubblicazione: (2025)
di: Yang, Liling, et al.
Pubblicazione: (2025)
Extreme Model Compression for Edge Vision-Language Models: Sparse Temporal Token Fusion and Adaptive Neural Compression
di: Tanvir, Md Tasnin, et al.
Pubblicazione: (2025)
di: Tanvir, Md Tasnin, et al.
Pubblicazione: (2025)
PPE: Positional Preservation Embedding for Token Compression in Multimodal Large Language Models
di: Huang, Mouxiao, et al.
Pubblicazione: (2025)
di: Huang, Mouxiao, et al.
Pubblicazione: (2025)
A Survey on Efficient Vision-Language Models
di: Shinde, Gaurav, et al.
Pubblicazione: (2025)
di: Shinde, Gaurav, et al.
Pubblicazione: (2025)
Explainability for Vision Foundation Models: A Survey
di: Kazmierczak, Rémi, et al.
Pubblicazione: (2025)
di: Kazmierczak, Rémi, et al.
Pubblicazione: (2025)
Advancing Autonomous Driving Perception: Analysis of Sensor Fusion and Computer Vision Techniques
di: Bharti, Urvishkumar, et al.
Pubblicazione: (2024)
di: Bharti, Urvishkumar, et al.
Pubblicazione: (2024)
A Survey of Token Compression for Efficient Multimodal Large Language Models
di: Shao, Kele, et al.
Pubblicazione: (2025)
di: Shao, Kele, et al.
Pubblicazione: (2025)
Exploring Hardware Friendly Bottleneck Architecture in CNN for Embedded Computing Systems
di: Lei, Xing, et al.
Pubblicazione: (2024)
di: Lei, Xing, et al.
Pubblicazione: (2024)
Vision without Images: End-to-End Computer Vision from Single Compressive Measurements
di: Pan, Fengpu, et al.
Pubblicazione: (2025)
di: Pan, Fengpu, et al.
Pubblicazione: (2025)
Training-free Conditional Image Embedding Framework Leveraging Large Vision Language Models
di: Kawarada, Masayuki, et al.
Pubblicazione: (2025)
di: Kawarada, Masayuki, et al.
Pubblicazione: (2025)
Multimodal Model for Computational Pathology:Representation Learning and Image Compression
di: Wu, Peihang, et al.
Pubblicazione: (2026)
di: Wu, Peihang, et al.
Pubblicazione: (2026)
Contrast & Compress: Learning Lightweight Embeddings for Short Trajectories
di: Vivekanandan, Abhishek, et al.
Pubblicazione: (2025)
di: Vivekanandan, Abhishek, et al.
Pubblicazione: (2025)
Geometric Transformation-Embedded Mamba for Learned Video Compression
di: Wei, Hao, et al.
Pubblicazione: (2026)
di: Wei, Hao, et al.
Pubblicazione: (2026)
FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression
di: Tong, Bo, et al.
Pubblicazione: (2024)
di: Tong, Bo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CCNeXt: An Effective Self-Supervised Stereo Depth Estimation Approach
di: Lopes, Alexandre, et al.
Pubblicazione: (2025) -
Identification of Deforestation Areas in the Amazon Rainforest Using Change Detection Models
di: Konishi, Christian Massao, et al.
Pubblicazione: (2025) -
CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition
di: Estevam, Valter, et al.
Pubblicazione: (2026) -
Dense Video Captioning Using Unsupervised Semantic Information
di: Estevam, Valter, et al.
Pubblicazione: (2021) -
P-NOC: adversarial training of CAM generating networks for robust weakly supervised semantic segmentation priors
di: David, Lucas, et al.
Pubblicazione: (2023)