When Training-Free NAS Meets Vision Transformer: A Neural Tangent Kernel Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Qiqi, Zhu, Yichen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025)
by: Kim, Beomsu, et al.
Published: (2025)
Neural Tangent Knowledge Distillation for Optical Convolutional Networks
by: Xiang, Jinlin, et al.
Published: (2025)
by: Xiang, Jinlin, et al.
Published: (2025)
A Survey on Graph Neural Networks and Graph Transformers in Computer Vision: A Task-Oriented Perspective
by: Chen, Chaoqi, et al.
Published: (2022)
by: Chen, Chaoqi, et al.
Published: (2022)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
by: Fontana, Maxime, et al.
Published: (2023)
by: Fontana, Maxime, et al.
Published: (2023)
Oscillation-Reduced MXFP4 Training for Vision Transformers
by: Chen, Yuxiang, et al.
Published: (2025)
by: Chen, Yuxiang, et al.
Published: (2025)
When Dynamic Data Selection Meets Data Augmentation
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models
by: Chu, Huanpeng, et al.
Published: (2025)
by: Chu, Huanpeng, et al.
Published: (2025)
InceptionNeXt: When Inception Meets ConvNeXt
by: Yu, Weihao, et al.
Published: (2023)
by: Yu, Weihao, et al.
Published: (2023)
ProKeR: A Kernel Perspective on Few-Shot Adaptation of Large Vision-Language Models
by: Bendou, Yassir, et al.
Published: (2025)
by: Bendou, Yassir, et al.
Published: (2025)
Training-Free Restoration of Pruned Neural Networks
by: Lee, Keonho, et al.
Published: (2025)
by: Lee, Keonho, et al.
Published: (2025)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
by: Bader, Jessica, et al.
Published: (2025)
by: Bader, Jessica, et al.
Published: (2025)
Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precision
by: Huang, Xijie, et al.
Published: (2023)
by: Huang, Xijie, et al.
Published: (2023)
MetaFormer Baselines for Vision
by: Yu, Weihao, et al.
Published: (2022)
by: Yu, Weihao, et al.
Published: (2022)
When Rubrics Fail: Error Enumeration as Reward in Reference-Free RL Post-Training for Virtual Try-On
by: Ikezogwo, Wisdom, et al.
Published: (2026)
by: Ikezogwo, Wisdom, et al.
Published: (2026)
Sparse Model Inversion: Efficient Inversion of Vision Transformers for Data-Free Applications
by: Hu, Zixuan, et al.
Published: (2025)
by: Hu, Zixuan, et al.
Published: (2025)
When Astronomy Meets AI: Manazel For Crescent Visibility Prediction in Morocco
by: Lairgi, Yassir
Published: (2025)
by: Lairgi, Yassir
Published: (2025)
When Does Supervised Training Pay Off? The Hidden Economics of Object Detection in the Era of Vision-Language Models
by: Al-Hamadani, Samer
Published: (2025)
by: Al-Hamadani, Samer
Published: (2025)
CLAReSNet: When Convolution Meets Latent Attention for Hyperspectral Image Classification
by: Bandyopadhyay, Asmit, et al.
Published: (2025)
by: Bandyopadhyay, Asmit, et al.
Published: (2025)
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
by: Rangwani, Harsh, et al.
Published: (2024)
by: Rangwani, Harsh, et al.
Published: (2024)
FIMA-Q: Post-Training Quantization for Vision Transformers by Fisher Information Matrix Approximation
by: Wu, Zhuguanyu, et al.
Published: (2025)
by: Wu, Zhuguanyu, et al.
Published: (2025)
Fairness-aware Vision Transformer via Debiased Self-Attention
by: Qiang, Yao, et al.
Published: (2023)
by: Qiang, Yao, et al.
Published: (2023)
Efficient Long-Horizon GUI Agents via Training-Free KV Cache Compression
by: Zhou, Bowen, et al.
Published: (2026)
by: Zhou, Bowen, et al.
Published: (2026)
RoPECraft: Training-Free Motion Transfer with Trajectory-Guided RoPE Optimization on Diffusion Transformers
by: Gokmen, Ahmet Berke, et al.
Published: (2025)
by: Gokmen, Ahmet Berke, et al.
Published: (2025)
Open-Vocabulary Panoptic Segmentation Using BERT Pre-Training of Vision-Language Multiway Transformer Model
by: Chen, Yi-Chia, et al.
Published: (2024)
by: Chen, Yi-Chia, et al.
Published: (2024)
SiamNAS: Siamese Surrogate Model for Dominance Relation Prediction in Multi-objective Neural Architecture Search
by: Zhou, Yuyang, et al.
Published: (2025)
by: Zhou, Yuyang, et al.
Published: (2025)
TopoPerception: A Shortcut-Free Evaluation of Global Visual Perception in Large Vision-Language Models
by: Zhou, Wenhao, et al.
Published: (2025)
by: Zhou, Wenhao, et al.
Published: (2025)
Adaptive Training Meets Progressive Scaling: Elevating Efficiency in Diffusion Models
by: Li, Wenhao, et al.
Published: (2023)
by: Li, Wenhao, et al.
Published: (2023)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
Stronger Normalization-Free Transformers
by: Chen, Mingzhi, et al.
Published: (2025)
by: Chen, Mingzhi, et al.
Published: (2025)
Proactive Gradient Conflict Mitigation in Multi-Task Learning: A Sparse Training Perspective
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Convolutional Neural Nets vs Vision Transformers: A SpaceNet Case Study with Balanced vs Imbalanced Regimes
by: Gothi, Akshar
Published: (2025)
by: Gothi, Akshar
Published: (2025)
CKGAN: Training Generative Adversarial Networks Using Characteristic Kernel Integral Probability Metrics
by: Zhang, Kuntian, et al.
Published: (2025)
by: Zhang, Kuntian, et al.
Published: (2025)
An Adaptive Tangent Feature Perspective of Neural Networks
by: LeJeune, Daniel, et al.
Published: (2023)
by: LeJeune, Daniel, et al.
Published: (2023)
Improving the Effectiveness and Efficiency of Stochastic Neighbour Embedding with Isolation Kernel
by: Zhu, Ye, et al.
Published: (2019)
by: Zhu, Ye, et al.
Published: (2019)
Intriguing Differences Between Zero-Shot and Systematic Evaluations of Vision-Language Transformer Models
by: Salman, Shaeke, et al.
Published: (2024)
by: Salman, Shaeke, et al.
Published: (2024)
Consistent Diffusion Meets Tweedie: Training Exact Ambient Diffusion Models with Noisy Data
by: Daras, Giannis, et al.
Published: (2024)
by: Daras, Giannis, et al.
Published: (2024)
Improving Interpretation Faithfulness for Vision Transformers
by: Hu, Lijie, et al.
Published: (2023)
by: Hu, Lijie, et al.
Published: (2023)
Block-Recurrent Dynamics in Vision Transformers
by: Jacobs, Mozes, et al.
Published: (2025)
by: Jacobs, Mozes, et al.
Published: (2025)
Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models
by: Tang, Kai, et al.
Published: (2025)
by: Tang, Kai, et al.
Published: (2025)
TRAVEL: Training-Free Retrieval and Alignment for Vision-and-Language Navigation
by: Rajabi, Navid, et al.
Published: (2025)
by: Rajabi, Navid, et al.
Published: (2025)
Similar Items
-
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025) -
Neural Tangent Knowledge Distillation for Optical Convolutional Networks
by: Xiang, Jinlin, et al.
Published: (2025) -
A Survey on Graph Neural Networks and Graph Transformers in Computer Vision: A Task-Oriented Perspective
by: Chen, Chaoqi, et al.
Published: (2022) -
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
by: Fontana, Maxime, et al.
Published: (2023) -
Oscillation-Reduced MXFP4 Training for Vision Transformers
by: Chen, Yuxiang, et al.
Published: (2025)