Where to Add PDE Diffusion in Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yukun, Zhou, Xueqing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers
by: Zhang, Yukun, et al.
Published: (2025)
by: Zhang, Yukun, et al.
Published: (2025)
Understanding Transformer Architecture through Continuous Dynamics: A Partial Differential Equation Perspective
by: Zhang, Yukun, et al.
Published: (2024)
by: Zhang, Yukun, et al.
Published: (2024)
ShiftAddViT: Mixture of Multiplication Primitives Towards Efficient Vision Transformer
by: You, Haoran, et al.
Published: (2023)
by: You, Haoran, et al.
Published: (2023)
Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE Solvers
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
SwiftDiffusion: Efficient Diffusion Model Serving with Add-on Modules
by: Li, Suyi, et al.
Published: (2024)
by: Li, Suyi, et al.
Published: (2024)
A PDE-Informed Latent Diffusion Model for 2-m Temperature Downscaling
by: Rosu, Paul, et al.
Published: (2025)
by: Rosu, Paul, et al.
Published: (2025)
DiffusionRollout: Uncertainty-Aware Rollout Planning in Long-Horizon PDE Solving
by: Yoo, Seungwoo, et al.
Published: (2026)
by: Yoo, Seungwoo, et al.
Published: (2026)
PDE-regularized Dynamics-informed Diffusion with Uncertainty-aware Filtering for Long-Horizon Dynamics
by: Baeg, Min Young, et al.
Published: (2026)
by: Baeg, Min Young, et al.
Published: (2026)
ShiftAddNAS: Hardware-Inspired Search for More Accurate and Efficient Neural Networks
by: You, Haoran, et al.
Published: (2022)
by: You, Haoran, et al.
Published: (2022)
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
by: Armandpour, Mohammadreza, et al.
Published: (2026)
by: Armandpour, Mohammadreza, et al.
Published: (2026)
Generative Latent Neural PDE Solver using Flow Matching
by: Li, Zijie, et al.
Published: (2025)
by: Li, Zijie, et al.
Published: (2025)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
LMILAtt: A Deep Learning Model for Depression Detection from Social Media Users Enhanced by Multi-Instance Learning Based on Attention Mechanism
by: Yang, Yukun
Published: (2025)
by: Yang, Yukun
Published: (2025)
A&B BNN: Add&Bit-Operation-Only Hardware-Friendly Binary Neural Network
by: Ma, Ruichen, et al.
Published: (2024)
by: Ma, Ruichen, et al.
Published: (2024)
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
by: You, Haoran, et al.
Published: (2024)
by: You, Haoran, et al.
Published: (2024)
Traj-Transformer: Diffusion Models with Transformer for GPS Trajectory Generation
by: Zhang, Zhiyang, et al.
Published: (2025)
by: Zhang, Zhiyang, et al.
Published: (2025)
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
by: Zhang, Yukun
Published: (2024)
by: Zhang, Yukun
Published: (2024)
How Few-Shot Examples Add Up: A Causal Decomposition of Function Vectors in In-Context Learning
by: Wang, Entang, et al.
Published: (2026)
by: Wang, Entang, et al.
Published: (2026)
Operator Learning with Domain Decomposition for Geometry Generalization in PDE Solving
by: Huang, Jianing, et al.
Published: (2025)
by: Huang, Jianing, et al.
Published: (2025)
VideoPDE: Unified Generative PDE Solving via Video Inpainting Diffusion Models
by: Li, Edward, et al.
Published: (2025)
by: Li, Edward, et al.
Published: (2025)
Planner-Admissible Graph-PDE Value Extensions for Sparse Goal-Conditioned Planning
by: Zhang, Shiheng
Published: (2026)
by: Zhang, Shiheng
Published: (2026)
On conditional diffusion models for PDE simulations
by: Shysheya, Aliaksandra, et al.
Published: (2024)
by: Shysheya, Aliaksandra, et al.
Published: (2024)
PolyLUT-Add: FPGA-based LUT Inference with Wide Inputs
by: Lou, Binglei, et al.
Published: (2024)
by: Lou, Binglei, et al.
Published: (2024)
LVM-GP: Uncertainty-Aware PDE Solver via coupling latent variable model and Gaussian process
by: Feng, Xiaodong, et al.
Published: (2025)
by: Feng, Xiaodong, et al.
Published: (2025)
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
by: Shen, Xuan, et al.
Published: (2024)
by: Shen, Xuan, et al.
Published: (2024)
Gradient-Informed Temporal Sampling Improves Rollout Accuracy in PDE Surrogate Training
by: Wang, Wenshuo, et al.
Published: (2026)
by: Wang, Wenshuo, et al.
Published: (2026)
DiffusionPDE: Generative PDE-Solving Under Partial Observation
by: Huang, Jiahe, et al.
Published: (2024)
by: Huang, Jiahe, et al.
Published: (2024)
Noisy PDE Training Requires Bigger PINNs
by: Andre-Sloan, Sebastien, et al.
Published: (2025)
by: Andre-Sloan, Sebastien, et al.
Published: (2025)
LE-PDE++: Mamba for accelerating PDEs Simulations
by: Liang, Aoming, et al.
Published: (2024)
by: Liang, Aoming, et al.
Published: (2024)
Diffusion Transformers as Open-World Spatiotemporal Foundation Models
by: Yuan, Yuan, et al.
Published: (2024)
by: Yuan, Yuan, et al.
Published: (2024)
JTreeformer: Graph-Transformer via Latent-Diffusion Model for Molecular Generation
by: Shi, Ji, et al.
Published: (2025)
by: Shi, Ji, et al.
Published: (2025)
Flow marching for a generative PDE foundation model
by: Chen, Zituo, et al.
Published: (2025)
by: Chen, Zituo, et al.
Published: (2025)
SFO: Learning PDE Operators via Spectral Filtering
by: Koren, Noam, et al.
Published: (2026)
by: Koren, Noam, et al.
Published: (2026)
Beyond Loss Guidance: Using PDE Residuals as Spectral Attention in Diffusion Neural Operators
by: Sawhney, Medha, et al.
Published: (2025)
by: Sawhney, Medha, et al.
Published: (2025)
Where Pretraining writes and Alignment reads: the asymmetry of Transformer weight space
by: Ruscio, Valeria, et al.
Published: (2026)
by: Ruscio, Valeria, et al.
Published: (2026)
Supercharging Graph Transformers with Advective Diffusion
by: Wu, Qitian, et al.
Published: (2023)
by: Wu, Qitian, et al.
Published: (2023)
Di-BiLPS: Denoising induced Bidirectional Latent-PDE-Solver under Sparse Observations
by: Li, Zhonghao, et al.
Published: (2026)
by: Li, Zhonghao, et al.
Published: (2026)
Posterior-First Neural PDE Simulation: Inferring Hidden Problem State from a Single Field
by: Wang, Wenshuo, et al.
Published: (2026)
by: Wang, Wenshuo, et al.
Published: (2026)
Revealing the Attention Floating Mechanism in Masked Diffusion Models
by: Dai, Xin, et al.
Published: (2026)
by: Dai, Xin, et al.
Published: (2026)
TF-CoDiT: Conditional Time Series Synthesis with Diffusion Transformers for Treasury Futures
by: Zhang, Yingxiao, et al.
Published: (2026)
by: Zhang, Yingxiao, et al.
Published: (2026)
Similar Items
-
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers
by: Zhang, Yukun, et al.
Published: (2025) -
Understanding Transformer Architecture through Continuous Dynamics: A Partial Differential Equation Perspective
by: Zhang, Yukun, et al.
Published: (2024) -
ShiftAddViT: Mixture of Multiplication Primitives Towards Efficient Vision Transformer
by: You, Haoran, et al.
Published: (2023) -
Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE Solvers
by: Zhou, Hang, et al.
Published: (2024) -
SwiftDiffusion: Efficient Diffusion Model Serving with Add-on Modules
by: Li, Suyi, et al.
Published: (2024)