Mixture-of-Experts Operator Transformer for Large-Scale PDE Pre-Training
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hong, Xin, Haiyang, Wang, Jie, Yang, Xuanze, Zha, Fei, Dong, Huanshuo, Jiang, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STNet: Spectral Transformation Network for Solving Operator Eigenvalue Problem
by: Wang, Hong, et al.
Published: (2025)
by: Wang, Hong, et al.
Published: (2025)
DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-Training
by: Hao, Zhongkai, et al.
Published: (2024)
by: Hao, Zhongkai, et al.
Published: (2024)
Latent Neural Operator for Solving Forward and Inverse PDE Problems
by: Wang, Tian, et al.
Published: (2024)
by: Wang, Tian, et al.
Published: (2024)
Mixture of Experts Softens the Curse of Dimensionality in Operator Learning
by: Kratsios, Anastasis, et al.
Published: (2024)
by: Kratsios, Anastasis, et al.
Published: (2024)
Deep Learning-Enhanced Preconditioning for Efficient Conjugate Gradient Solvers in Large-Scale PDE Systems
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE Solvers
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
CATO: Charted Attention for Neural PDE Operators
by: Cheng, Chun-Wun, et al.
Published: (2026)
by: Cheng, Chun-Wun, et al.
Published: (2026)
Quantifying Training Difficulty and Accelerating Convergence in Neural Network-Based PDE Solvers
by: Chen, Chuqi, et al.
Published: (2024)
by: Chen, Chuqi, et al.
Published: (2024)
A PDE-based Explanation of Extreme Numerical Sensitivities and Edge of Stability in Training Neural Networks
by: Sun, Yuxin, et al.
Published: (2022)
by: Sun, Yuxin, et al.
Published: (2022)
Blending Neural Operators and Relaxation Methods in PDE Numerical Solvers
by: Zhang, Enrui, et al.
Published: (2022)
by: Zhang, Enrui, et al.
Published: (2022)
Neural Parameter Regression for Explicit Representations of PDE Solution Operators
by: Mundinger, Konrad, et al.
Published: (2024)
by: Mundinger, Konrad, et al.
Published: (2024)
DeltaPhi: Physical States Residual Learning for Neural Operators in Data-Limited PDE Solving
by: Yue, Xihang, et al.
Published: (2024)
by: Yue, Xihang, et al.
Published: (2024)
DeepONet Augmented by Randomized Neural Networks for Efficient Operator Learning in PDEs
by: Jiang, Zhaoxi, et al.
Published: (2025)
by: Jiang, Zhaoxi, et al.
Published: (2025)
Scaling up Probabilistic PDE Simulators with Structured Volumetric Information
by: Weiland, Tim, et al.
Published: (2024)
by: Weiland, Tim, et al.
Published: (2024)
A persistent-homology-based Bayesian prior for potential coefficient reconstruction in an elliptic PDE
by: Deng, Zhiliang, et al.
Published: (2025)
by: Deng, Zhiliang, et al.
Published: (2025)
U-HNO: A U-shaped Hybrid Neural Operator with Sparse-Point Adaptive Routing for Non-stationary PDE Dynamics
by: Ma, Yingzhe, et al.
Published: (2026)
by: Ma, Yingzhe, et al.
Published: (2026)
BCAT: A Block Causal Transformer for PDE Foundation Models for Fluid Dynamics
by: Liu, Yuxuan, et al.
Published: (2025)
by: Liu, Yuxuan, et al.
Published: (2025)
Physics-Informed Inference Time Scaling for Solving High-Dimensional PDE via Defect Correction
by: Fan, Zexi, et al.
Published: (2025)
by: Fan, Zexi, et al.
Published: (2025)
Universal Approximation of Operators with Transformers and Neural Integral Operators
by: Zappala, Emanuele, et al.
Published: (2024)
by: Zappala, Emanuele, et al.
Published: (2024)
Are Deep Learning Based Hybrid PDE Solvers Reliable? Why Training Paradigms and Update Strategies Matter
by: Wu, Yuhan, et al.
Published: (2026)
by: Wu, Yuhan, et al.
Published: (2026)
PDE Generalization of In-Context Operator Networks: A Study on 1D Scalar Nonlinear Conservation Laws
by: Yang, Liu, et al.
Published: (2024)
by: Yang, Liu, et al.
Published: (2024)
From Simple to Complex: Curriculum-Guided Physics-Informed Neural Networks via Gaussian Mixture Models
by: Yang, Jianan, et al.
Published: (2026)
by: Yang, Jianan, et al.
Published: (2026)
PDE Solvers Should Be Local: Fast, Stable Rollouts with Learned Local Stencils
by: Cheng, Chun-Wun, et al.
Published: (2025)
by: Cheng, Chun-Wun, et al.
Published: (2025)
HAMLET: Graph Transformer Neural Operator for Partial Differential Equations
by: Bryutkin, Andrey, et al.
Published: (2024)
by: Bryutkin, Andrey, et al.
Published: (2024)
Accelerating Data Generation for Neural Operators via Krylov Subspace Recycling
by: Wang, Hong, et al.
Published: (2024)
by: Wang, Hong, et al.
Published: (2024)
Adaptive-Distribution Randomized Neural Networks for PDEs: A Low-Dimensional Distribution-Learning Framework
by: Yang, You, et al.
Published: (2026)
by: Yang, You, et al.
Published: (2026)
When Attention Beats Fourier: Multi-Scale Transformers for PDE Solving on Irregular Domains
by: Yee, Brandon, et al.
Published: (2026)
by: Yee, Brandon, et al.
Published: (2026)
Monte Carlo Neural PDE Solver for Learning PDEs via Probabilistic Representation
by: Zhang, Rui, et al.
Published: (2023)
by: Zhang, Rui, et al.
Published: (2023)
Kernel Learning of PDE Solution Operators
by: Hu, Jianyu, et al.
Published: (2026)
by: Hu, Jianyu, et al.
Published: (2026)
A Multimodal PDE Foundation Model for Prediction and Scientific Text Descriptions
by: Negrini, Elisa, et al.
Published: (2025)
by: Negrini, Elisa, et al.
Published: (2025)
Training-Free Looped Transformers
by: Chen, Lizhang, et al.
Published: (2026)
by: Chen, Lizhang, et al.
Published: (2026)
Beyond Loss Guidance: Using PDE Residuals as Spectral Attention in Diffusion Neural Operators
by: Sawhney, Medha, et al.
Published: (2025)
by: Sawhney, Medha, et al.
Published: (2025)
CFO: Learning Continuous-Time PDE Dynamics via Flow-Matched Neural Operators
by: Hou, Xianglong, et al.
Published: (2025)
by: Hou, Xianglong, et al.
Published: (2025)
P$^2$C$^2$Net: PDE-Preserved Coarse Correction Network for efficient prediction of spatiotemporal dynamics
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
Graph Neural Regularizers for PDE Inverse Problems
by: Lauga, William, et al.
Published: (2025)
by: Lauga, William, et al.
Published: (2025)
Multi-Level Monte Carlo Training of Neural Operators
by: Rowbottom, James, et al.
Published: (2025)
by: Rowbottom, James, et al.
Published: (2025)
Scaling Laws and Pathologies of Single-Layer PINNs: Network Width and PDE Nonlinearity
by: Chaudhry, Faris
Published: (2026)
by: Chaudhry, Faris
Published: (2026)
Mathematics of Digital Twins and Transfer Learning for PDE Models
by: Zong, Yifei, et al.
Published: (2025)
by: Zong, Yifei, et al.
Published: (2025)
Mamba Neural Operator: Who Wins? Transformers vs. State-Space Models for PDEs
by: Cheng, Chun-Wun, et al.
Published: (2024)
by: Cheng, Chun-Wun, et al.
Published: (2024)
Latent Neural Operator Pretraining for Solving Time-Dependent PDEs
by: Wang, Tian, et al.
Published: (2024)
by: Wang, Tian, et al.
Published: (2024)
Similar Items
-
STNet: Spectral Transformation Network for Solving Operator Eigenvalue Problem
by: Wang, Hong, et al.
Published: (2025) -
DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-Training
by: Hao, Zhongkai, et al.
Published: (2024) -
Latent Neural Operator for Solving Forward and Inverse PDE Problems
by: Wang, Tian, et al.
Published: (2024) -
Mixture of Experts Softens the Curse of Dimensionality in Operator Learning
by: Kratsios, Anastasis, et al.
Published: (2024) -
Deep Learning-Enhanced Preconditioning for Efficient Conjugate Gradient Solvers in Large-Scale PDE Systems
by: Li, Rui, et al.
Published: (2024)