Saved in:
| Main Authors: | Chaudhry, Zan, Mizuno, Naoko |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.16975 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Implicit Statistical Inference in Transformers: Approximating Likelihood-Ratio Tests In-Context
by: Chaudhry, Faris, et al.
Published: (2026)
by: Chaudhry, Faris, et al.
Published: (2026)
Non-Interfering Weight Fields: Treating Model Parameters as a Continuously Extensible Function
by: Chaudhry, Sarim
Published: (2026)
by: Chaudhry, Sarim
Published: (2026)
Heterogeneous Learning Rate Scheduling for Neural Architecture Search on Long-Tailed Datasets
by: Tang, Chenxia
Published: (2024)
by: Tang, Chenxia
Published: (2024)
Test-Time Training with KV Binding Is Secretly Linear Attention
by: Liu, Junchen, et al.
Published: (2026)
by: Liu, Junchen, et al.
Published: (2026)
Recursive Concept Evolution for Compositional Reasoning in Large Language Models
by: Chaudhry, Sarim
Published: (2026)
by: Chaudhry, Sarim
Published: (2026)
Automated Design of Linear Bounding Functions for Sigmoidal Nonlinearities in Neural Networks
by: König, Matthias, et al.
Published: (2024)
by: König, Matthias, et al.
Published: (2024)
Edge Prompt Tuning for Graph Neural Networks
by: Fu, Xingbo, et al.
Published: (2025)
by: Fu, Xingbo, et al.
Published: (2025)
Universal Prompt Tuning for Graph Neural Networks
by: Fang, Taoran, et al.
Published: (2022)
by: Fang, Taoran, et al.
Published: (2022)
Combinatorial Optimization with Automated Graph Neural Networks
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Linear Mode Connectivity in Sparse Neural Networks
by: McDermott, Luke, et al.
Published: (2023)
by: McDermott, Luke, et al.
Published: (2023)
Optimal Linear Decay Learning Rate Schedules and Further Refinements
by: Defazio, Aaron, et al.
Published: (2023)
by: Defazio, Aaron, et al.
Published: (2023)
On the Impact of Class Imbalance on the Learning Dynamics of Deep Neural Networks:An Intuitive Insight
by: Mustapha, Ismail B., et al.
Published: (2026)
by: Mustapha, Ismail B., et al.
Published: (2026)
Lecture Notes on Linear Neural Networks: A Tale of Optimization and Generalization in Deep Learning
by: Cohen, Nadav, et al.
Published: (2024)
by: Cohen, Nadav, et al.
Published: (2024)
Full Bayesian Significance Testing for Neural Networks
by: Liu, Zehua, et al.
Published: (2024)
by: Liu, Zehua, et al.
Published: (2024)
Testing Individual Fairness in Graph Neural Networks
by: Nasiri, Roya
Published: (2025)
by: Nasiri, Roya
Published: (2025)
Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
Multi-Objective Neural Architecture Search by Learning Search Space Partitions
by: Zhao, Yiyang, et al.
Published: (2024)
by: Zhao, Yiyang, et al.
Published: (2024)
Decoupling Search and Learning in Neural Net Training
by: Vegesna, Akshay, et al.
Published: (2025)
by: Vegesna, Akshay, et al.
Published: (2025)
AutoSTF: Decoupled Neural Architecture Search for Cost-Effective Automated Spatio-Temporal Forecasting
by: Lyu, Tengfei, et al.
Published: (2024)
by: Lyu, Tengfei, et al.
Published: (2024)
Neural Parameter Search for Slimmer Fine-Tuned Models and Better Transfer
by: Du, Guodong, et al.
Published: (2025)
by: Du, Guodong, et al.
Published: (2025)
Generating In-store Customer Journeys from Scratch with GPT Architectures
by: Horikomi, Taizo, et al.
Published: (2024)
by: Horikomi, Taizo, et al.
Published: (2024)
TopoTune : A Framework for Generalized Combinatorial Complex Neural Networks
by: Papillon, Mathilde, et al.
Published: (2024)
by: Papillon, Mathilde, et al.
Published: (2024)
Tuning for Trustworthiness -- Balancing Performance and Explanation Consistency in Neural Network Optimization
by: Hinterleitner, Alexander, et al.
Published: (2025)
by: Hinterleitner, Alexander, et al.
Published: (2025)
Learning to Solve Combinatorial Optimization under Positive Linear Constraints via Non-Autoregressive Neural Networks
by: Wang, Runzhong, et al.
Published: (2024)
by: Wang, Runzhong, et al.
Published: (2024)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
by: Onorato, Gabriele
Published: (2024)
by: Onorato, Gabriele
Published: (2024)
Distilling Linearized Behavior into Non-Linear Fine-Tuning for Effective Task Arithmetic
by: Sommariva, Thomas, et al.
Published: (2026)
by: Sommariva, Thomas, et al.
Published: (2026)
SEANN: A Domain-Informed Neural Network for Epidemiological Insights
by: Guimbaud, Jean-Baptiste, et al.
Published: (2025)
by: Guimbaud, Jean-Baptiste, et al.
Published: (2025)
A Simple Sparse Matrix Vector Multiplication Approach to Padded Convolution
by: Chaudhry, Zan
Published: (2024)
by: Chaudhry, Zan
Published: (2024)
Improving Line Search Methods for Large Scale Neural Network Training
by: Kenneweg, Philip, et al.
Published: (2024)
by: Kenneweg, Philip, et al.
Published: (2024)
Testing Components of the Attention Schema Theory in Artificial Neural Networks
by: Farrell, Kathryn T., et al.
Published: (2024)
by: Farrell, Kathryn T., et al.
Published: (2024)
On the Rate of Convergence of GD in Non-linear Neural Networks: An Adversarial Robustness Perspective
by: Smorodinsky, Guy, et al.
Published: (2026)
by: Smorodinsky, Guy, et al.
Published: (2026)
Prediction of Bank Credit Ratings using Heterogeneous Topological Graph Neural Networks
by: Liu, Junyi, et al.
Published: (2025)
by: Liu, Junyi, et al.
Published: (2025)
On the Sensitivity of Firing Rate-Based Federated Spiking Neural Networks to Differential Privacy
by: Pereira, Luiz, et al.
Published: (2026)
by: Pereira, Luiz, et al.
Published: (2026)
Self-Pro: A Self-Prompt and Tuning Framework for Graph Neural Networks
by: Gong, Chenghua, et al.
Published: (2023)
by: Gong, Chenghua, et al.
Published: (2023)
Graph Neural Networks for Learning Equivariant Representations of Neural Networks
by: Kofinas, Miltiadis, et al.
Published: (2024)
by: Kofinas, Miltiadis, et al.
Published: (2024)
A Survey on Neural Architecture Search Based on Reinforcement Learning
by: Shao, Wenzhu
Published: (2024)
by: Shao, Wenzhu
Published: (2024)
Linearization Explains Fine-Tuning in Large Language Models
by: Afzal, Zahra Rahimi, et al.
Published: (2026)
by: Afzal, Zahra Rahimi, et al.
Published: (2026)
Linear $Q$-Learning Does Not Diverge in $L^2$: Convergence Rates to a Bounded Set
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Explainable Heterogeneous Anomaly Detection in Financial Networks via Adaptive Expert Routing
by: Li, Zan, et al.
Published: (2025)
by: Li, Zan, et al.
Published: (2025)
NdLinear: Preserving Multi-Dimensional Structure for Parameter-Efficient Neural Networks
by: Reneau, Alex, et al.
Published: (2025)
by: Reneau, Alex, et al.
Published: (2025)
Similar Items
-
Implicit Statistical Inference in Transformers: Approximating Likelihood-Ratio Tests In-Context
by: Chaudhry, Faris, et al.
Published: (2026) -
Non-Interfering Weight Fields: Treating Model Parameters as a Continuously Extensible Function
by: Chaudhry, Sarim
Published: (2026) -
Heterogeneous Learning Rate Scheduling for Neural Architecture Search on Long-Tailed Datasets
by: Tang, Chenxia
Published: (2024) -
Test-Time Training with KV Binding Is Secretly Linear Attention
by: Liu, Junchen, et al.
Published: (2026) -
Recursive Concept Evolution for Compositional Reasoning in Large Language Models
by: Chaudhry, Sarim
Published: (2026)