NoProp: Training Neural Networks without Full Back-propagation or Full Forward-propagation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Qinyu, Teh, Yee Whye, Pascanu, Razvan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
von: Galashov, Alexandre, et al.
Veröffentlicht: (2024)
von: Galashov, Alexandre, et al.
Veröffentlicht: (2024)
Kalman Filter for Online Classification of Non-Stationary Data
von: Titsias, Michalis K., et al.
Veröffentlicht: (2023)
von: Titsias, Michalis K., et al.
Veröffentlicht: (2023)
Incorporating Unlabelled Data into Bayesian Neural Networks
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
Deep Grokking: Would Deep Neural Networks Generalize Better?
von: Fan, Simin, et al.
Veröffentlicht: (2024)
von: Fan, Simin, et al.
Veröffentlicht: (2024)
Revisiting Dynamic Evaluation: Online Adaptation for Large Language Models
von: Rannen-Triki, Amal, et al.
Veröffentlicht: (2024)
von: Rannen-Triki, Amal, et al.
Veröffentlicht: (2024)
SymDiff: Equivariant Diffusion via Stochastic Symmetrisation
von: Zhang, Leo, et al.
Veröffentlicht: (2024)
von: Zhang, Leo, et al.
Veröffentlicht: (2024)
Latent Space Representations of Neural Algorithmic Reasoners
von: Mirjanić, Vladimir V., et al.
Veröffentlicht: (2023)
von: Mirjanić, Vladimir V., et al.
Veröffentlicht: (2023)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024)
von: Sims, Anya, et al.
Veröffentlicht: (2024)
ssProp: Energy-Efficient Training for Convolutional Neural Networks with Scheduled Sparse Back Propagation
von: Zhong, Lujia, et al.
Veröffentlicht: (2024)
von: Zhong, Lujia, et al.
Veröffentlicht: (2024)
L3Ms -- Lagrange Large Language Models
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
Rao-Blackwellised Reparameterisation Gradients
von: Lam, Kevin H., et al.
Veröffentlicht: (2025)
von: Lam, Kevin H., et al.
Veröffentlicht: (2025)
Manifold Aware Denoising Score Matching (MAD)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
Disentangling the Causes of Plasticity Loss in Neural Networks
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
von: Lyle, Clare, et al.
Veröffentlicht: (2024)
Selective Safety Steering via Value-Filtered Decoding
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
Lattice: Learning to Efficiently Compress the Memory
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
Meta Flow Maps enable scalable reward alignment
von: Potaptchik, Peter, et al.
Veröffentlicht: (2026)
von: Potaptchik, Peter, et al.
Veröffentlicht: (2026)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
Meta-Learning Objectives for Preference Optimization
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
Full-Spectrum Graph Neural Networks: Expressive and Scalable
von: Wang, Xiaohan, et al.
Veröffentlicht: (2026)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2026)
EvIL: Evolution Strategies for Generalisable Imitation Learning
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
Full Bayesian Significance Testing for Neural Networks
von: Liu, Zehua, et al.
Veröffentlicht: (2024)
von: Liu, Zehua, et al.
Veröffentlicht: (2024)
What Can Grokking Teach Us About Learning Under Nonstationarity?
von: Lyle, Clare, et al.
Veröffentlicht: (2025)
von: Lyle, Clare, et al.
Veröffentlicht: (2025)
Meta-learning how to Share Credit among Macro-Actions
von: Hosu, Ionel-Alexandru, et al.
Veröffentlicht: (2025)
von: Hosu, Ionel-Alexandru, et al.
Veröffentlicht: (2025)
Revisiting Adam for Streaming Reinforcement Learning
von: Gogianu, Florin, et al.
Veröffentlicht: (2026)
von: Gogianu, Florin, et al.
Veröffentlicht: (2026)
SigmaDock: Untwisting Molecular Docking With Fragment-Based SE(3) Diffusion
von: Prat, Alvaro, et al.
Veröffentlicht: (2025)
von: Prat, Alvaro, et al.
Veröffentlicht: (2025)
ECO: Quantized Training without Full-Precision Master Weights
von: Nikdan, Mahdi, et al.
Veröffentlicht: (2026)
von: Nikdan, Mahdi, et al.
Veröffentlicht: (2026)
Metropolis-Adjusted Diffusion Models
von: Lam, Kevin H., et al.
Veröffentlicht: (2026)
von: Lam, Kevin H., et al.
Veröffentlicht: (2026)
Derivation of Back-propagation for Graph Convolutional Networks using Matrix Calculus and its Application to Explainable Artificial Intelligence
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
Amortized Probabilistic Detection of Communities in Graphs
von: Wang, Yueqi, et al.
Veröffentlicht: (2020)
von: Wang, Yueqi, et al.
Veröffentlicht: (2020)
Online Adaptation of Language Models with a Memory of Amortized Contexts
von: Tack, Jihoon, et al.
Veröffentlicht: (2024)
von: Tack, Jihoon, et al.
Veröffentlicht: (2024)
Prompting Strategies for Enabling Large Language Models to Infer Causation from Correlation
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
Context-Guided Diffusion for Out-of-Distribution Molecular and Protein Design
von: Klarner, Leo, et al.
Veröffentlicht: (2024)
von: Klarner, Leo, et al.
Veröffentlicht: (2024)
True 4-Bit Quantized Convolutional Neural Network Training on CPU: Achieving Full-Precision Parity
von: Tathe, Shivnath
Veröffentlicht: (2026)
von: Tathe, Shivnath
Veröffentlicht: (2026)
FullCert: Deterministic End-to-End Certification for Training and Inference of Neural Networks
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
Lightweight Dataset Pruning without Full Training via Example Difficulty and Prediction Uncertainty
von: Cho, Yeseul, et al.
Veröffentlicht: (2025)
von: Cho, Yeseul, et al.
Veröffentlicht: (2025)
MS-SSM: A Multi-Scale State Space Model for Efficient Sequence Modeling
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
von: Karami, Mahdi, et al.
Veröffentlicht: (2025)
No Representation, No Trust: Connecting Representation, Collapse, and Trust Issues in PPO
von: Moalla, Skander, et al.
Veröffentlicht: (2024)
von: Moalla, Skander, et al.
Veröffentlicht: (2024)
Attention as a Hypernetwork
von: Schug, Simon, et al.
Veröffentlicht: (2024)
von: Schug, Simon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
von: Galashov, Alexandre, et al.
Veröffentlicht: (2024) -
Kalman Filter for Online Classification of Non-Stationary Data
von: Titsias, Michalis K., et al.
Veröffentlicht: (2023) -
Incorporating Unlabelled Data into Bayesian Neural Networks
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023) -
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
von: Lai, Yuhang, et al.
Veröffentlicht: (2026) -
Deep Grokking: Would Deep Neural Networks Generalize Better?
von: Fan, Simin, et al.
Veröffentlicht: (2024)