Towards General Neural Surrogate Solvers with Specialized Neural Accelerators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mao, Chenkai, Lupoiu, Robert, Dai, Tianxiang, Chen, Mingkun, Fan, Jonathan A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unlocking Real-Time Fluorescence Lifetime Imaging: Multi-Pixel Parallelism for FPGA-Accelerated Processing
von: Erbas, Ismail, et al.
Veröffentlicht: (2024)
von: Erbas, Ismail, et al.
Veröffentlicht: (2024)
FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs
von: Dege, Pengcuo, et al.
Veröffentlicht: (2025)
von: Dege, Pengcuo, et al.
Veröffentlicht: (2025)
Privacy in Federated Learning with Spiking Neural Networks
von: Aksu, Dogukan, et al.
Veröffentlicht: (2025)
von: Aksu, Dogukan, et al.
Veröffentlicht: (2025)
Federated Neural Architecture Search with Model-Agnostic Meta Learning
von: Huang, Xinyuan, et al.
Veröffentlicht: (2025)
von: Huang, Xinyuan, et al.
Veröffentlicht: (2025)
A Parallel Alternative for Energy-Efficient Neural Network Training and Inferencing
von: Seal, Sudip K., et al.
Veröffentlicht: (2025)
von: Seal, Sudip K., et al.
Veröffentlicht: (2025)
Root Cause Analysis In Microservice Using Neural Granger Causal Discovery
von: Lin, Cheng-Ming, et al.
Veröffentlicht: (2024)
von: Lin, Cheng-Ming, et al.
Veröffentlicht: (2024)
TAPAS: Fast and Automatic Derivation of Tensor Parallel Strategies for Large Neural Networks
von: Shi, Ziji, et al.
Veröffentlicht: (2023)
von: Shi, Ziji, et al.
Veröffentlicht: (2023)
Semi-decentralized Training of Spatio-Temporal Graph Neural Networks for Traffic Prediction
von: Kralj, Ivan, et al.
Veröffentlicht: (2024)
von: Kralj, Ivan, et al.
Veröffentlicht: (2024)
D3-GNN: Dynamic Distributed Dataflow for Streaming Graph Neural Networks
von: Guliyev, Rustam, et al.
Veröffentlicht: (2024)
von: Guliyev, Rustam, et al.
Veröffentlicht: (2024)
Characterizing Mobile SoC for Accelerating Heterogeneous LLM Inference
von: Chen, Le, et al.
Veröffentlicht: (2025)
von: Chen, Le, et al.
Veröffentlicht: (2025)
Distributed Graph Neural Network Inference With Just-In-Time Compilation For Industry-Scale Graphs
von: Wu, Xiabao, et al.
Veröffentlicht: (2025)
von: Wu, Xiabao, et al.
Veröffentlicht: (2025)
EPSILON: Adaptive Fault Mitigation in Approximate Deep Neural Network using Statistical Signatures
von: Khalil, Khurram, et al.
Veröffentlicht: (2025)
von: Khalil, Khurram, et al.
Veröffentlicht: (2025)
DistrEE: Distributed Early Exit of Deep Neural Network Inference on Edge Devices
von: Peng, Xian, et al.
Veröffentlicht: (2025)
von: Peng, Xian, et al.
Veröffentlicht: (2025)
BitPipe: Bidirectional Interleaved Pipeline Parallelism for Accelerating Large Models Training
von: Wu, Houming, et al.
Veröffentlicht: (2024)
von: Wu, Houming, et al.
Veröffentlicht: (2024)
TawPipe: Topology-Aware Weight Pipeline Parallelism for Accelerating Long-Context Large Models Training
von: Wu, Houming, et al.
Veröffentlicht: (2025)
von: Wu, Houming, et al.
Veröffentlicht: (2025)
Mobility Accelerates Learning: Convergence Analysis on Hierarchical Federated Learning in Vehicular Networks
von: Chen, Tan, et al.
Veröffentlicht: (2024)
von: Chen, Tan, et al.
Veröffentlicht: (2024)
Accelerating Privacy-Preserving Federated Learning in Large-Scale LEO Satellite Systems
von: Guo, Binquan, et al.
Veröffentlicht: (2025)
von: Guo, Binquan, et al.
Veröffentlicht: (2025)
Syno: Structured Synthesis for Neural Operators
von: Zhuo, Yongqi, et al.
Veröffentlicht: (2024)
von: Zhuo, Yongqi, et al.
Veröffentlicht: (2024)
Accelerating MoE Model Inference with Expert Sharding
von: Balmau, Oana, et al.
Veröffentlicht: (2025)
von: Balmau, Oana, et al.
Veröffentlicht: (2025)
Mind the Gap: Revealing Inconsistencies Across Heterogeneous AI Accelerators
von: Wen, Elliott, et al.
Veröffentlicht: (2025)
von: Wen, Elliott, et al.
Veröffentlicht: (2025)
Calibre: Towards Fair and Accurate Personalized Federated Learning with Self-Supervised Learning
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Towards Straggler-Resilient Split Federated Learning: An Unbalanced Update Approach
von: Liang, Dandan, et al.
Veröffentlicht: (2025)
von: Liang, Dandan, et al.
Veröffentlicht: (2025)
Conflict-Free Replicated Data Types for Neural Network Model Merging: A Two-Layer Architecture Enabling CRDT-Compliant Model Merging Across 26 Strategies
von: Gillespie, Ryan
Veröffentlicht: (2026)
von: Gillespie, Ryan
Veröffentlicht: (2026)
Exploiting Inter-Layer Expert Affinity for Accelerating Mixture-of-Experts Model Inference
von: Yao, Jinghan, et al.
Veröffentlicht: (2024)
von: Yao, Jinghan, et al.
Veröffentlicht: (2024)
Acceleration for Deep Reinforcement Learning using Parallel and Distributed Computing: A Survey
von: Liu, Zhihong, et al.
Veröffentlicht: (2024)
von: Liu, Zhihong, et al.
Veröffentlicht: (2024)
Towards Quantum-Ready Blockchain Fraud Detection via Ensemble Graph Neural Networks
von: Haider, M. Z., et al.
Veröffentlicht: (2025)
von: Haider, M. Z., et al.
Veröffentlicht: (2025)
A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approximation
von: Li, Xiaocan, et al.
Veröffentlicht: (2025)
von: Li, Xiaocan, et al.
Veröffentlicht: (2025)
Tiny Deep Ensemble: Uncertainty Estimation in Edge AI Accelerators via Ensembling Normalization Layers with Shared Weights
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
FT-Transformer: Resilient and Reliable Transformer with End-to-End Fault Tolerant Attention
von: Dai, Huangliang, et al.
Veröffentlicht: (2025)
von: Dai, Huangliang, et al.
Veröffentlicht: (2025)
Frontier: Towards Comprehensive and Accurate LLM Inference Simulation
von: Feng, Yicheng, et al.
Veröffentlicht: (2026)
von: Feng, Yicheng, et al.
Veröffentlicht: (2026)
CommunityAI: Towards Community-based Federated Learning
von: Murturi, Ilir, et al.
Veröffentlicht: (2023)
von: Murturi, Ilir, et al.
Veröffentlicht: (2023)
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers
von: Daghero, Francesco, et al.
Veröffentlicht: (2025)
von: Daghero, Francesco, et al.
Veröffentlicht: (2025)
Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems
von: Miao, Xupeng, et al.
Veröffentlicht: (2023)
von: Miao, Xupeng, et al.
Veröffentlicht: (2023)
AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
Towards Robust and Efficient Federated Low-Rank Adaptation with Heterogeneous Clients
von: Koo, Jabin, et al.
Veröffentlicht: (2024)
von: Koo, Jabin, et al.
Veröffentlicht: (2024)
Towards One-shot Federated Learning: Advances, Challenges, and Future Directions
von: Amato, Flora, et al.
Veröffentlicht: (2025)
von: Amato, Flora, et al.
Veröffentlicht: (2025)
Measuring Heterogeneity in Machine Learning with Distributed Energy Distance
von: Fan, Mengchen, et al.
Veröffentlicht: (2025)
von: Fan, Mengchen, et al.
Veröffentlicht: (2025)
Cyclical Weight Consolidation: Towards Solving Catastrophic Forgetting in Serial Federated Learning
von: Song, Haoyue, et al.
Veröffentlicht: (2024)
von: Song, Haoyue, et al.
Veröffentlicht: (2024)
ECHO: Elastic Speculative Decoding with Sparse Gating for High-Concurrency Scenarios
von: Hu, Xinyi, et al.
Veröffentlicht: (2026)
von: Hu, Xinyi, et al.
Veröffentlicht: (2026)
Role-Based Fault Tolerance System for LLM RL Post-Training
von: Chen, Zhenqian, et al.
Veröffentlicht: (2025)
von: Chen, Zhenqian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unlocking Real-Time Fluorescence Lifetime Imaging: Multi-Pixel Parallelism for FPGA-Accelerated Processing
von: Erbas, Ismail, et al.
Veröffentlicht: (2024) -
FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs
von: Dege, Pengcuo, et al.
Veröffentlicht: (2025) -
Privacy in Federated Learning with Spiking Neural Networks
von: Aksu, Dogukan, et al.
Veröffentlicht: (2025) -
Federated Neural Architecture Search with Model-Agnostic Meta Learning
von: Huang, Xinyuan, et al.
Veröffentlicht: (2025) -
A Parallel Alternative for Energy-Efficient Neural Network Training and Inferencing
von: Seal, Sudip K., et al.
Veröffentlicht: (2025)