Weight-Entanglement Meets Gradient-Based Neural Architecture Search
Fuente:
arXiv
Saved in:
| Main Authors: | Sukthanker, Rhea Sanjay, Krishnakumar, Arjun, Safari, Mahmoud, Hutter, Frank |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Search for Customized Activation Functions with Gradient Descent
by: Strack, Lukas, et al.
Published: (2024)
by: Strack, Lukas, et al.
Published: (2024)
HW-GPT-Bench: Hardware-Aware Architecture Benchmark for Language Models
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
Multi-objective Differentiable Neural Architecture Search
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
Where to Begin: Efficient Pretraining via Subnetwork Selection and Distillation
by: Krishnakumar, Arjun, et al.
Published: (2025)
by: Krishnakumar, Arjun, et al.
Published: (2025)
confopt: A Library for Implementation and Evaluation of Gradient-based One-Shot NAS Methods
by: Jha, Abhash Kumar, et al.
Published: (2025)
by: Jha, Abhash Kumar, et al.
Published: (2025)
Gompertz Linear Units: Leveraging Asymmetry for Enhanced Learning Dynamics
by: Das, Indrashis, et al.
Published: (2025)
by: Das, Indrashis, et al.
Published: (2025)
Diffusion-Based Neural Network Weights Generation
by: Soro, Bedionita, et al.
Published: (2024)
by: Soro, Bedionita, et al.
Published: (2024)
Neural Architecture Search: Two Constant Shared Weights Initialisations
by: Gracheva, Ekaterina
Published: (2023)
by: Gracheva, Ekaterina
Published: (2023)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
Towards Efficient Few-shot Graph Neural Architecture Search via Partitioning Gradient Contribution
by: Song, Wenhao, et al.
Published: (2025)
by: Song, Wenhao, et al.
Published: (2025)
A Survey on Neural Architecture Search Based on Reinforcement Learning
by: Shao, Wenzhu
Published: (2024)
by: Shao, Wenzhu
Published: (2024)
Zero-Shot Neural Architecture Search with Weighted Response Correlation
by: Jing, Kun, et al.
Published: (2025)
by: Jing, Kun, et al.
Published: (2025)
Graph Neural Architecture Search with GPT-4
by: Wang, Haishuai, et al.
Published: (2023)
by: Wang, Haishuai, et al.
Published: (2023)
Multi-Objective Neural Architecture Search by Learning Search Space Partitions
by: Zhao, Yiyang, et al.
Published: (2024)
by: Zhao, Yiyang, et al.
Published: (2024)
Compressing Large Language Models with Automated Sub-Network Search
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024)
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
Gradient Flow Convergence Guarantee for General Neural Network Architectures
by: Jakhmola, Yash
Published: (2025)
by: Jakhmola, Yash
Published: (2025)
Transferrable Surrogates in Expressive Neural Architecture Search Spaces
by: Qin, Shiwen, et al.
Published: (2025)
by: Qin, Shiwen, et al.
Published: (2025)
XGrad: Boosting Gradient-Based Optimizers With Weight Prediction
by: Guan, Lei, et al.
Published: (2023)
by: Guan, Lei, et al.
Published: (2023)
Levin Tree Search with Context Models
by: Orseau, Laurent, et al.
Published: (2023)
by: Orseau, Laurent, et al.
Published: (2023)
SeqNAS: Neural Architecture Search for Event Sequence Classification
by: Udovichenko, Igor, et al.
Published: (2024)
by: Udovichenko, Igor, et al.
Published: (2024)
MicroNAS: Zero-Shot Neural Architecture Search for MCUs
by: Qiao, Ye, et al.
Published: (2024)
by: Qiao, Ye, et al.
Published: (2024)
Reinforced Compressive Neural Architecture Search for Versatile Adversarial Robustness
by: Wang, Dingrong, et al.
Published: (2024)
by: Wang, Dingrong, et al.
Published: (2024)
Unsupervised Graph Neural Architecture Search with Disentangled Self-supervision
by: Zhang, Zeyang, et al.
Published: (2024)
by: Zhang, Zeyang, et al.
Published: (2024)
Improving LLM-based Global Optimization with Search Space Partitioning
by: Schwanke, Andrej, et al.
Published: (2025)
by: Schwanke, Andrej, et al.
Published: (2025)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
by: Veness, Joel, et al.
Published: (2025)
by: Veness, Joel, et al.
Published: (2025)
LLM4GNAS: A Large Language Model Based Toolkit for Graph Neural Architecture Search
by: Gao, Yang, et al.
Published: (2025)
by: Gao, Yang, et al.
Published: (2025)
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
by: Wei, Zhao, et al.
Published: (2025)
by: Wei, Zhao, et al.
Published: (2025)
Finite-Particle Convergence Rates for Conservative and Non-Conservative Drifting Models
by: Balasubramanian, Krishnakumar
Published: (2026)
by: Balasubramanian, Krishnakumar
Published: (2026)
Large-Step Training Dynamics of a Two-Factor Linear Transformer Model
by: Balasubramanian, Krishnakumar
Published: (2026)
by: Balasubramanian, Krishnakumar
Published: (2026)
Teasing Apart Architecture and Initial Weights as Sources of Inductive Bias in Neural Networks
by: Bencomo, Gianluca, et al.
Published: (2025)
by: Bencomo, Gianluca, et al.
Published: (2025)
Speeding Up Multi-Objective Hyperparameter Optimization by Task Similarity-Based Meta-Learning for the Tree-Structured Parzen Estimator
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Models
by: Armstrong, Marcus, et al.
Published: (2026)
by: Armstrong, Marcus, et al.
Published: (2026)
Kernel-Level Energy-Efficient Neural Architecture Search for Tabular Dataset
by: La, Hoang-Loc, et al.
Published: (2025)
by: La, Hoang-Loc, et al.
Published: (2025)
Causal-aware Graph Neural Architecture Search under Distribution Shifts
by: Li, Peiwen, et al.
Published: (2024)
by: Li, Peiwen, et al.
Published: (2024)
NASH: Neural Architecture and Accelerator Search for Multiplication-Reduced Hybrid Models
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture Search
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
The Devil Is in Gradient Entanglement: Energy-Aware Gradient Coordinator for Robust Generalized Category Discovery
by: Zheng, Haiyang, et al.
Published: (2026)
by: Zheng, Haiyang, et al.
Published: (2026)
EquiTabPFN: A Target-Permutation Equivariant Prior Fitted Networks
by: Arbel, Michael, et al.
Published: (2025)
by: Arbel, Michael, et al.
Published: (2025)
Agentic NL2SQL to Reduce Computational Costs
by: Jehle, Dominik, et al.
Published: (2025)
by: Jehle, Dominik, et al.
Published: (2025)
Similar Items
-
Efficient Search for Customized Activation Functions with Gradient Descent
by: Strack, Lukas, et al.
Published: (2024) -
HW-GPT-Bench: Hardware-Aware Architecture Benchmark for Language Models
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024) -
Multi-objective Differentiable Neural Architecture Search
by: Sukthanker, Rhea Sanjay, et al.
Published: (2024) -
Where to Begin: Efficient Pretraining via Subnetwork Selection and Distillation
by: Krishnakumar, Arjun, et al.
Published: (2025) -
confopt: A Library for Implementation and Evaluation of Gradient-based One-Shot NAS Methods
by: Jha, Abhash Kumar, et al.
Published: (2025)