Saved in:
| Main Authors: | Liu, Taowen, Andronic, Marta, Gündüz, Deniz, Constantinides, George A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.00874 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Stochastic Rounding with Few Random Bits
by: Fitzgibbon, Andrew, et al.
Published: (2025)
by: Fitzgibbon, Andrew, et al.
Published: (2025)
Direction-Preserving Number Representations
by: Zadeh, Bardia, et al.
Published: (2026)
by: Zadeh, Bardia, et al.
Published: (2026)
Stochastic Rounding Increases Small Singular Values
by: Ma, Linkai, et al.
Published: (2026)
by: Ma, Linkai, et al.
Published: (2026)
NeuraLUT-Assemble: Hardware-aware Assembling of Sub-Neural Networks for Efficient LUT Inference
by: Andronic, Marta, et al.
Published: (2025)
by: Andronic, Marta, et al.
Published: (2025)
NeuraLUT: Hiding Neural Network Density in Boolean Synthesizable Functions
by: Andronic, Marta, et al.
Published: (2024)
by: Andronic, Marta, et al.
Published: (2024)
PolyLUT: Learning Piecewise Polynomials for Ultra-Low Latency FPGA LUT-based Inference
by: Andronic, Marta, et al.
Published: (2023)
by: Andronic, Marta, et al.
Published: (2023)
A PDE-based Explanation of Extreme Numerical Sensitivities and Edge of Stability in Training Neural Networks
by: Sun, Yuxin, et al.
Published: (2022)
by: Sun, Yuxin, et al.
Published: (2022)
Training-Free Looped Transformers
by: Chen, Lizhang, et al.
Published: (2026)
by: Chen, Lizhang, et al.
Published: (2026)
On the Convergence of the Gradient Descent Method with Stochastic Fixed-point Rounding Errors under the Polyak-Lojasiewicz Inequality
by: Xia, Lu, et al.
Published: (2023)
by: Xia, Lu, et al.
Published: (2023)
Training Hamiltonian neural networks without backpropagation
by: Rahma, Atamert, et al.
Published: (2024)
by: Rahma, Atamert, et al.
Published: (2024)
Multi-Level Monte Carlo Training of Neural Operators
by: Rowbottom, James, et al.
Published: (2025)
by: Rowbottom, James, et al.
Published: (2025)
Generative Feature Training of Thin 2-Layer Networks
by: Hertrich, Johannes, et al.
Published: (2024)
by: Hertrich, Johannes, et al.
Published: (2024)
Multi-Preconditioned LBFGS for Training Finite-Basis PINNs
by: Salvadó-Benasco, Marc, et al.
Published: (2026)
by: Salvadó-Benasco, Marc, et al.
Published: (2026)
DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-Training
by: Hao, Zhongkai, et al.
Published: (2024)
by: Hao, Zhongkai, et al.
Published: (2024)
Data-Parallel Neural Network Training via Nonlinearly Preconditioned Trust-Region Method
by: Alegría, Samuel A. Cruz, et al.
Published: (2025)
by: Alegría, Samuel A. Cruz, et al.
Published: (2025)
Time Extrapolation with Graph Convolutional Autoencoder and Tensor Train Decomposition
by: Chen, Yuanhong, et al.
Published: (2025)
by: Chen, Yuanhong, et al.
Published: (2025)
PolyLUT: Ultra-low Latency Polynomial Inference with Hardware-Aware Structured Pruning
by: Andronic, Marta, et al.
Published: (2025)
by: Andronic, Marta, et al.
Published: (2025)
Mixture-of-Experts Operator Transformer for Large-Scale PDE Pre-Training
by: Wang, Hong, et al.
Published: (2025)
by: Wang, Hong, et al.
Published: (2025)
Stabilizing Physics-Informed Consistency Models via Structure-Preserving Training
by: Chang, Che-Chia, et al.
Published: (2026)
by: Chang, Che-Chia, et al.
Published: (2026)
A New Tensor Network: Tubal Tensor Train and Its Applications
by: Ahmadi-Asl, Salman, et al.
Published: (2026)
by: Ahmadi-Asl, Salman, et al.
Published: (2026)
Automatic Differentiation is Essential in Training Neural Networks for Solving Differential Equations
by: Chen, Chuqi, et al.
Published: (2024)
by: Chen, Chuqi, et al.
Published: (2024)
Why Does Stochastic Gradient Descent Slow Down in Low-Precision Training?
by: Yun, Vincent-Daniel
Published: (2025)
by: Yun, Vincent-Daniel
Published: (2025)
An Augmented Backward-Corrected Projector Splitting Integrator for Dynamical Low-Rank Training
by: Kusch, Jonas, et al.
Published: (2025)
by: Kusch, Jonas, et al.
Published: (2025)
RL-PINNs: Reinforcement Learning-Driven Adaptive Sampling for Efficient Training of PINNs
by: Song, Zhenao
Published: (2025)
by: Song, Zhenao
Published: (2025)
Quantifying Training Difficulty and Accelerating Convergence in Neural Network-Based PDE Solvers
by: Chen, Chuqi, et al.
Published: (2024)
by: Chen, Chuqi, et al.
Published: (2024)
DeepONet for Solving Nonlinear Partial Differential Equations with Physics-Informed Training
by: Yang, Yahong
Published: (2024)
by: Yang, Yahong
Published: (2024)
$ε$-rank and the Staircase Phenomenon: New Insights into Neural Network Training Dynamics
by: Yang, Jiang, et al.
Published: (2024)
by: Yang, Jiang, et al.
Published: (2024)
Training-free score-based diffusion for parameter-dependent stochastic dynamical systems
by: Yang, Minglei, et al.
Published: (2026)
by: Yang, Minglei, et al.
Published: (2026)
Enhanced BPINN Training Convergence in Solving General and Multi-scale Elliptic PDEs with Noise
by: Hou, Yilong, et al.
Published: (2024)
by: Hou, Yilong, et al.
Published: (2024)
Are Deep Learning Based Hybrid PDE Solvers Reliable? Why Training Paradigms and Update Strategies Matter
by: Wu, Yuhan, et al.
Published: (2026)
by: Wu, Yuhan, et al.
Published: (2026)
Market-Driven Subset Selection for Budgeted Training
by: Jha, Ashish, et al.
Published: (2025)
by: Jha, Ashish, et al.
Published: (2025)
BitMoD: Bit-serial Mixture-of-Datatype LLM Acceleration
by: Chen, Yuzong, et al.
Published: (2024)
by: Chen, Yuzong, et al.
Published: (2024)
Unveiling the Power of Multiple Gossip Steps: A Stability-Based Generalization Analysis in Decentralized Training
by: Li, Qinglun, et al.
Published: (2025)
by: Li, Qinglun, et al.
Published: (2025)
WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points
by: Li, Dongyue, et al.
Published: (2026)
by: Li, Dongyue, et al.
Published: (2026)
Dual Cone Gradient Descent for Training Physics-Informed Neural Networks
by: Hwang, Youngsik, et al.
Published: (2024)
by: Hwang, Youngsik, et al.
Published: (2024)
Streaming Krylov-Accelerated Stochastic Gradient Descent
by: Thomas, Stephen
Published: (2025)
by: Thomas, Stephen
Published: (2025)
Learning Stochastic Dynamical Systems with Structured Noise
by: Guo, Ziheng, et al.
Published: (2025)
by: Guo, Ziheng, et al.
Published: (2025)
Stochastic diagonal estimation with adaptive parameter selection
by: Han, Zongyuan, et al.
Published: (2024)
by: Han, Zongyuan, et al.
Published: (2024)
Spectral Audit of In-Context Operator Networks
by: Gao, Zhiwei, et al.
Published: (2026)
by: Gao, Zhiwei, et al.
Published: (2026)
A Natural Primal-Dual Hybrid Gradient Method for Adversarial Neural Network Training on Solving Partial Differential Equations
by: Liu, Shu, et al.
Published: (2024)
by: Liu, Shu, et al.
Published: (2024)
Similar Items
-
On Stochastic Rounding with Few Random Bits
by: Fitzgibbon, Andrew, et al.
Published: (2025) -
Direction-Preserving Number Representations
by: Zadeh, Bardia, et al.
Published: (2026) -
Stochastic Rounding Increases Small Singular Values
by: Ma, Linkai, et al.
Published: (2026) -
NeuraLUT-Assemble: Hardware-aware Assembling of Sub-Neural Networks for Efficient LUT Inference
by: Andronic, Marta, et al.
Published: (2025) -
NeuraLUT: Hiding Neural Network Density in Boolean Synthesizable Functions
by: Andronic, Marta, et al.
Published: (2024)