Neural Organ Transplantation (NOT): Checkpoint-Based Modular Adaptation for Transformer Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Al-Zuraiqi, Ahmad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FERI: A Multitask-based Fairness Achieving Algorithm with Applications to Fair Organ Transplantation
von: Li, Can, et al.
Veröffentlicht: (2023)
von: Li, Can, et al.
Veröffentlicht: (2023)
Neural ODE and SDE Models for Adaptation and Planning in Model-Based Reinforcement Learning
von: Han, Chao, et al.
Veröffentlicht: (2026)
von: Han, Chao, et al.
Veröffentlicht: (2026)
Transformers as Neural Operators for Solutions of Differential Equations with Finite Regularity
von: Shih, Benjamin, et al.
Veröffentlicht: (2024)
von: Shih, Benjamin, et al.
Veröffentlicht: (2024)
Training Neural Networks for Modularity aids Interpretability
von: Golechha, Satvik, et al.
Veröffentlicht: (2024)
von: Golechha, Satvik, et al.
Veröffentlicht: (2024)
Studying Cross-cluster Modularity in Neural Networks
von: Golechha, Satvik, et al.
Veröffentlicht: (2025)
von: Golechha, Satvik, et al.
Veröffentlicht: (2025)
GoldenTransformer: A Modular Fault Injection Framework for Transformer Robustness Research
von: Howard, Luke
Veröffentlicht: (2025)
von: Howard, Luke
Veröffentlicht: (2025)
Engineering Verifiable Modularity in Transformers via Per-Layer Supervision
von: Kerce, J. Clayton
Veröffentlicht: (2026)
von: Kerce, J. Clayton
Veröffentlicht: (2026)
READ: Recurrent Adaptation of Large Transformers
von: Nguyen, John, et al.
Veröffentlicht: (2023)
von: Nguyen, John, et al.
Veröffentlicht: (2023)
On the Structural Limitations of Weight-Based Neural Adaptation and the Role of Reversible Behavioral Learning
von: Konduru, Pardhu Sri Rushi Varma
Veröffentlicht: (2026)
von: Konduru, Pardhu Sri Rushi Varma
Veröffentlicht: (2026)
Learning Modular Exponentiation with Transformers
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
Mashup Learning: Faster Finetuning by Remixing Past Checkpoints
von: Vaina, Sofia Maria Lo Cicero, et al.
Veröffentlicht: (2026)
von: Vaina, Sofia Maria Lo Cicero, et al.
Veröffentlicht: (2026)
AttentionSmithy: A Modular Framework for Rapid Transformer Development and Customization
von: Cranney, Caleb, et al.
Veröffentlicht: (2025)
von: Cranney, Caleb, et al.
Veröffentlicht: (2025)
Uncovering a Universal Abstract Algorithm for Modular Addition in Neural Networks
von: McCracken, Gavin, et al.
Veröffentlicht: (2025)
von: McCracken, Gavin, et al.
Veröffentlicht: (2025)
DHA: Learning Decoupled-Head Attention from Transformer Checkpoints via Adaptive Heads Fusion
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
Towards Empirical Interpretation of Internal Circuits and Properties in Grokked Transformers on Modular Polynomials
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
Model Fusion via Neuron Transplantation
von: Öz, Muhammed, et al.
Veröffentlicht: (2025)
von: Öz, Muhammed, et al.
Veröffentlicht: (2025)
AdaDim: Dimensionality Adaptation for SSL Representational Dynamics
von: Kokilepersaud, Kiran, et al.
Veröffentlicht: (2025)
von: Kokilepersaud, Kiran, et al.
Veröffentlicht: (2025)
FAA-CLIP: Federated Adversarial Adaptation of CLIP
von: Wu, Yihang, et al.
Veröffentlicht: (2025)
von: Wu, Yihang, et al.
Veröffentlicht: (2025)
Plug-and-Play Transformer Modules for Test-Time Adaptation
von: Chang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Chang, Xiangyu, et al.
Veröffentlicht: (2024)
Modular Boundaries in Recurrent Neural Networks
von: Tanner, Jacob, et al.
Veröffentlicht: (2023)
von: Tanner, Jacob, et al.
Veröffentlicht: (2023)
Dispatch-Aware Ragged Attention for Pruned Vision Transformers
von: Abdellatif, Seifeldin, et al.
Veröffentlicht: (2026)
von: Abdellatif, Seifeldin, et al.
Veröffentlicht: (2026)
Boosting Cross-problem Generalization in Diffusion-Based Neural Combinatorial Solver via Inference Time Adaptation
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
SURGE: Surrogate Gradient Adaptation in Binary Neural Networks
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
Unlocking Emergent Modularity in Large Language Models
von: Qiu, Zihan, et al.
Veröffentlicht: (2023)
von: Qiu, Zihan, et al.
Veröffentlicht: (2023)
CoSA: Compressed Sensing-Based Adaptation of Large Language Models
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
Simulations of Common Unsupervised Domain Adaptation Algorithms for Image Classification
von: Chaddad, Ahmad, et al.
Veröffentlicht: (2025)
von: Chaddad, Ahmad, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning for Phishing Detection with Transformer-Based Semantic Features
von: Faisal, Aseer Al
Veröffentlicht: (2025)
von: Faisal, Aseer Al
Veröffentlicht: (2025)
FocusLearn: Fully-Interpretable, High-Performance Modular Neural Networks for Time Series
von: Su, Qiqi, et al.
Veröffentlicht: (2023)
von: Su, Qiqi, et al.
Veröffentlicht: (2023)
NASP-T: A Fuzzy Neuro-Symbolic Transformer for Logic-Constrained Aviation Safety Report Classification
von: Machot, Fadi Al, et al.
Veröffentlicht: (2025)
von: Machot, Fadi Al, et al.
Veröffentlicht: (2025)
Lightweight Geometric Adaptation for Training Physics-Informed Neural Networks
von: An, Kang, et al.
Veröffentlicht: (2026)
von: An, Kang, et al.
Veröffentlicht: (2026)
Spectral Transformer Neural Processes
von: Chen, Xianhe, et al.
Veröffentlicht: (2026)
von: Chen, Xianhe, et al.
Veröffentlicht: (2026)
Weight Space Representation Learning via Neural Field Adaptation
von: Yang, Zhuoqian, et al.
Veröffentlicht: (2025)
von: Yang, Zhuoqian, et al.
Veröffentlicht: (2025)
A Lightweight DL Model for Smart Grid Power Forecasting with Feature and Resolution Mismatch
von: Al-Shareeda, Sarah, et al.
Veröffentlicht: (2025)
von: Al-Shareeda, Sarah, et al.
Veröffentlicht: (2025)
Neural Decompiling of Tracr Transformers
von: Thurnherr, Hannes, et al.
Veröffentlicht: (2024)
von: Thurnherr, Hannes, et al.
Veröffentlicht: (2024)
Transformers are Graph Neural Networks
von: Joshi, Chaitanya K.
Veröffentlicht: (2025)
von: Joshi, Chaitanya K.
Veröffentlicht: (2025)
NeuralGrok: Accelerate Grokking by Neural Gradient Transformation
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
METHOD: Modular Efficient Transformer for Health Outcome Discovery
von: Qian, Linglong, et al.
Veröffentlicht: (2025)
von: Qian, Linglong, et al.
Veröffentlicht: (2025)
Cliqueformer: Model-Based Optimization with Structured Transformers
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
SeqBattNet: A Discrete-State Physics-Informed Neural Network with Aging Adaptation for Battery Modeling
von: Tran, Khoa, et al.
Veröffentlicht: (2025)
von: Tran, Khoa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FERI: A Multitask-based Fairness Achieving Algorithm with Applications to Fair Organ Transplantation
von: Li, Can, et al.
Veröffentlicht: (2023) -
Neural ODE and SDE Models for Adaptation and Planning in Model-Based Reinforcement Learning
von: Han, Chao, et al.
Veröffentlicht: (2026) -
Transformers as Neural Operators for Solutions of Differential Equations with Finite Regularity
von: Shih, Benjamin, et al.
Veröffentlicht: (2024) -
Training Neural Networks for Modularity aids Interpretability
von: Golechha, Satvik, et al.
Veröffentlicht: (2024) -
Studying Cross-cluster Modularity in Neural Networks
von: Golechha, Satvik, et al.
Veröffentlicht: (2025)