GhostNetV3: Exploring the Training Strategies for Compact Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zhenhua, Hao, Zhiwei, Han, Kai, Tang, Yehui, Wang, Yunhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GhostNetV3-Small: A Tailored Architecture and Comparative Study of Distillation Strategies for Tiny Images
von: Zager, Florian, et al.
Veröffentlicht: (2025)
von: Zager, Florian, et al.
Veröffentlicht: (2025)
ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
SAM-DiffSR: Structure-Modulated Diffusion Model for Image Super-Resolution
von: Wang, Chengcheng, et al.
Veröffentlicht: (2024)
von: Wang, Chengcheng, et al.
Veröffentlicht: (2024)
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
von: Rang, Miao, et al.
Veröffentlicht: (2025)
von: Rang, Miao, et al.
Veröffentlicht: (2025)
GPT4Image: Large Pre-trained Models Help Vision Models Learn Better on Perception Task
von: Ding, Ning, et al.
Veröffentlicht: (2023)
von: Ding, Ning, et al.
Veröffentlicht: (2023)
Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs
von: Han, Kai, et al.
Veröffentlicht: (2024)
von: Han, Kai, et al.
Veröffentlicht: (2024)
Data-efficient Large Vision Models through Sequential Autoregression
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
Token Compensator: Altering Inference Cost of Vision Transformer without Re-Tuning
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
SLAB: Efficient Transformers with Simplified Linear Attention and Progressive Re-parameterized Batch Normalization
von: Guo, Jialong, et al.
Veröffentlicht: (2024)
von: Guo, Jialong, et al.
Veröffentlicht: (2024)
ParameterNet: Parameters Are All You Need
von: Han, Kai, et al.
Veröffentlicht: (2023)
von: Han, Kai, et al.
Veröffentlicht: (2023)
No Time to Waste: Squeeze Time into Channel for Mobile Video Understanding
von: Zhai, Yingjie, et al.
Veröffentlicht: (2024)
von: Zhai, Yingjie, et al.
Veröffentlicht: (2024)
Context-Guided Spatial Feature Reconstruction for Efficient Semantic Segmentation
von: Ni, Zhenliang, et al.
Veröffentlicht: (2024)
von: Ni, Zhenliang, et al.
Veröffentlicht: (2024)
Post-Training Quantization for Diffusion Transformer via Hierarchical Timestep Grouping
von: Ding, Ning, et al.
Veröffentlicht: (2025)
von: Ding, Ning, et al.
Veröffentlicht: (2025)
Memory-Space Visual Prompting for Efficient Vision-Language Fine-Tuning
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
TinySAM: Pushing the Envelope for Efficient Segment Anything Model
von: Shu, Han, et al.
Veröffentlicht: (2023)
von: Shu, Han, et al.
Veröffentlicht: (2023)
A Survey on Transformer Compression
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
Ghost-Stereo: GhostNet-based Cost Volume Enhancement and Aggregation for Stereo Matching Networks
von: Jiang, Xingguang, et al.
Veröffentlicht: (2024)
von: Jiang, Xingguang, et al.
Veröffentlicht: (2024)
GenVidBench: A 6-Million Benchmark for AI-Generated Video Detection
von: Ni, Zhenliang, et al.
Veröffentlicht: (2025)
von: Ni, Zhenliang, et al.
Veröffentlicht: (2025)
Revealing the Power of Post-Training for Small Language Models via Knowledge Distillation
von: Rang, Miao, et al.
Veröffentlicht: (2025)
von: Rang, Miao, et al.
Veröffentlicht: (2025)
Ghost-dil-NetVLAD: A Lightweight Neural Network for Visual Place Recognition
von: Gong, Qingyuan, et al.
Veröffentlicht: (2021)
von: Gong, Qingyuan, et al.
Veröffentlicht: (2021)
An Empirical Study of Scaling Law for OCR
von: Rang, Miao, et al.
Veröffentlicht: (2023)
von: Rang, Miao, et al.
Veröffentlicht: (2023)
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
DECO: Unleashing the Potential of ConvNets for Query-based Detection and Segmentation
von: Chen, Xinghao, et al.
Veröffentlicht: (2023)
von: Chen, Xinghao, et al.
Veröffentlicht: (2023)
Training-Free Open-Ended Object Detection and Segmentation via Attention as Prompts
von: Lin, Zhiwei, et al.
Veröffentlicht: (2024)
von: Lin, Zhiwei, et al.
Veröffentlicht: (2024)
Diffusion Sampling Path Tells More: An Efficient Plug-and-Play Strategy for Sample Filtering
von: Wang, Sixian, et al.
Veröffentlicht: (2025)
von: Wang, Sixian, et al.
Veröffentlicht: (2025)
TTTFusion: A Test-Time Training-Based Strategy for Multimodal Medical Image Fusion in Surgical Robots
von: Xie, Qinhua, et al.
Veröffentlicht: (2025)
von: Xie, Qinhua, et al.
Veröffentlicht: (2025)
Compact Model Training by Low-Rank Projection with Energy Transfer
von: Guo, Kailing, et al.
Veröffentlicht: (2022)
von: Guo, Kailing, et al.
Veröffentlicht: (2022)
U-REPA: Aligning Diffusion U-Nets to ViTs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
Distilling Semantic Priors from SAM to Efficient Image Restoration Models
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
von: Zhu, Jinguo, et al.
Veröffentlicht: (2025)
von: Zhu, Jinguo, et al.
Veröffentlicht: (2025)
TSCM: A Teacher-Student Model for Vision Place Recognition Using Cross-Metric Knowledge Distillation
von: Shen, Yehui, et al.
Veröffentlicht: (2024)
von: Shen, Yehui, et al.
Veröffentlicht: (2024)
RAPID^3: Tri-Level Reinforced Acceleration Policies for Diffusion Transformer
von: Zhao, Wangbo, et al.
Veröffentlicht: (2025)
von: Zhao, Wangbo, et al.
Veröffentlicht: (2025)
Exploring Active Data Selection Strategies for Continuous Training in Deepfake Detection
von: Furuhashi, Yoshihiko, et al.
Veröffentlicht: (2025)
von: Furuhashi, Yoshihiko, et al.
Veröffentlicht: (2025)
Multi-Scale Correlation-Aware Transformer for Maritime Vessel Re-Identification
von: Liu, Yunhe
Veröffentlicht: (2025)
von: Liu, Yunhe
Veröffentlicht: (2025)
The Radar Ghost Dataset -- An Evaluation of Ghost Objects in Automotive Radar Data
von: Kraus, Florian, et al.
Veröffentlicht: (2024)
von: Kraus, Florian, et al.
Veröffentlicht: (2024)
An Empirical Study of World Model Quantization
von: Fu, Zhongqian, et al.
Veröffentlicht: (2026)
von: Fu, Zhongqian, et al.
Veröffentlicht: (2026)
RepGhost: A Hardware-Efficient Ghost Module via Re-parameterization
von: Chen, Chengpeng, et al.
Veröffentlicht: (2022)
von: Chen, Chengpeng, et al.
Veröffentlicht: (2022)
An Embeddable Implicit IUVD Representation for Part-based 3D Human Surface Reconstruction
von: Li, Baoxing, et al.
Veröffentlicht: (2024)
von: Li, Baoxing, et al.
Veröffentlicht: (2024)
DynGhost: Temporally-Modelled Transformer for Dynamic Ghost Imaging with Quantum Detectors
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
von: Lin, Zhiwei, et al.
Veröffentlicht: (2025)
von: Lin, Zhiwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GhostNetV3-Small: A Tailored Architecture and Comparative Study of Distillation Strategies for Tiny Images
von: Zager, Florian, et al.
Veröffentlicht: (2025) -
ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025) -
SAM-DiffSR: Structure-Modulated Diffusion Model for Image Super-Resolution
von: Wang, Chengcheng, et al.
Veröffentlicht: (2024) -
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
von: Rang, Miao, et al.
Veröffentlicht: (2025) -
GPT4Image: Large Pre-trained Models Help Vision Models Learn Better on Perception Task
von: Ding, Ning, et al.
Veröffentlicht: (2023)