Communication Offloading on SmartNIC DPUs: A Quantitative Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wahlgren, Jacob, Hu, Andong, Pearce, Roger, Gokhale, Maya, Peng, Ivy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disaggregated Memory with SmartNIC Offloading: a Case Study on Graph Processing
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2024)
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2024)
Inter-APU Communication on AMD MI300A Systems via Infinity Fabric: a Deep Dive
von: Schieffer, Gabin, et al.
Veröffentlicht: (2025)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2025)
Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2025)
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2025)
Plug & Offload: Transparently Offloading TCP Stack onto Off-path SmartNIC with PnO-TCP
von: Nan, Hailong, et al.
Veröffentlicht: (2025)
von: Nan, Hailong, et al.
Veröffentlicht: (2025)
Multi-level Memory-Centric Profiling on ARM Processors with ARM SPE
von: Miksits, Samuel, et al.
Veröffentlicht: (2024)
von: Miksits, Samuel, et al.
Veröffentlicht: (2024)
Meili: Enabling SmartNIC as a Service in the Cloud
von: Su, Qiang, et al.
Veröffentlicht: (2023)
von: Su, Qiang, et al.
Veröffentlicht: (2023)
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
Understanding Layered Portability from HPC to Cloud in Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
Kub: Enabling Elastic HPC Workloads on Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2024)
The Forward-In-Time-Only Assumption in SmartNIC Resource Management: A Critique of Wave and the Case for Bilateral Interaction
von: Borrill, Paul
Veröffentlicht: (2026)
von: Borrill, Paul
Veröffentlicht: (2026)
SCENIC: Stream Computation-Enhanced SmartNIC
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2026)
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2026)
Harnessing Integrated CPU-GPU System Memory for HPC: a first look into Grace Hopper
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
Reliable Replication Protocols on SmartNICs
von: Katebzadeh, M. R. Siavash, et al.
Veröffentlicht: (2025)
von: Katebzadeh, M. R. Siavash, et al.
Veröffentlicht: (2025)
ARC-V: Vertical Resource Adaptivity for HPC Workloads in Containerized Environments
von: Medeiros, Daniel, et al.
Veröffentlicht: (2025)
von: Medeiros, Daniel, et al.
Veröffentlicht: (2025)
Closer in the Gap: Towards Portable Performance on RISC-V Vector Processors
von: Shi, Ruimin, et al.
Veröffentlicht: (2026)
von: Shi, Ruimin, et al.
Veröffentlicht: (2026)
High-performance Vector-length Agnostic Quantum Circuit Simulations on ARM Processors
von: Shi, Ruimin, et al.
Veröffentlicht: (2026)
von: Shi, Ruimin, et al.
Veröffentlicht: (2026)
ARM SVE Unleashed: Performance and Insights Across HPC Applications on Nvidia Grace
von: Shi, Ruimin, et al.
Veröffentlicht: (2025)
von: Shi, Ruimin, et al.
Veröffentlicht: (2025)
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
von: Schieffer, Gabin, et al.
Veröffentlicht: (2026)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2026)
A Chronological Analysis of the Evolution of SmartNICs
von: Ajayi, Olasupo, et al.
Veröffentlicht: (2025)
von: Ajayi, Olasupo, et al.
Veröffentlicht: (2025)
Employ SmartNICs' Data Path Accelerators for Ordered Key-Value Stores
von: Schimmelpfennig, Frederic, et al.
Veröffentlicht: (2026)
von: Schimmelpfennig, Frederic, et al.
Veröffentlicht: (2026)
Blink: CPU-Free LLM Inference by Delegating the Serving Stack to GPU and SmartNIC
von: Siavashi, Mohammad, et al.
Veröffentlicht: (2026)
von: Siavashi, Mohammad, et al.
Veröffentlicht: (2026)
OpenCUBE: Building an Open Source Cloud Blueprint with EPI Systems
von: Peng, Ivy, et al.
Veröffentlicht: (2024)
von: Peng, Ivy, et al.
Veröffentlicht: (2024)
Enabling AI Deep Potentials for Ab Initio-quality Molecular Dynamics Simulations in GROMACS
von: Hu, Andong, et al.
Veröffentlicht: (2026)
von: Hu, Andong, et al.
Veröffentlicht: (2026)
DPDPU: Data Processing with DPUs
von: Hu, Jiasheng, et al.
Veröffentlicht: (2024)
von: Hu, Jiasheng, et al.
Veröffentlicht: (2024)
FRESCO: Fast and Reliable Edge Offloading with Reputation-based Hybrid Smart Contracts
von: Zilic, Josip, et al.
Veröffentlicht: (2024)
von: Zilic, Josip, et al.
Veröffentlicht: (2024)
Accelerating Drug Discovery in AutoDock-GPU with Tensor Cores
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS
von: Pennati, Luca, et al.
Veröffentlicht: (2026)
von: Pennati, Luca, et al.
Veröffentlicht: (2026)
Boosting Performance of Iterative Applications on GPUs: Kernel Batching with CUDA Graphs
von: Ekelund, Jonah, et al.
Veröffentlicht: (2025)
von: Ekelund, Jonah, et al.
Veröffentlicht: (2025)
Proposal of Automatic Offloading Method in Mixed Offloading Destination Environment
von: Yamato, Yoji
Veröffentlicht: (2020)
von: Yamato, Yoji
Veröffentlicht: (2020)
Egret: Reinforcement Mechanism for Sequential Computation Offloading in Edge Computing
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
von: Peng, Haosong, et al.
Veröffentlicht: (2024)
Edge Offloading in Smart Grid
von: Arcas, Gabriel Ioan, et al.
Veröffentlicht: (2024)
von: Arcas, Gabriel Ioan, et al.
Veröffentlicht: (2024)
To Offload or Not To Offload: Model-driven Comparison of Edge-native and On-device Processing In the Era of Accelerators
von: Ng, Nathan, et al.
Veröffentlicht: (2025)
von: Ng, Nathan, et al.
Veröffentlicht: (2025)
FlexKV: Flexible Index Offloading for Memory-Disaggregated Key-Value Store
von: Hu, Zhisheng, et al.
Veröffentlicht: (2025)
von: Hu, Zhisheng, et al.
Veröffentlicht: (2025)
A Survey of Computation Offloading with Task Types
von: Zhang, Siqi, et al.
Veröffentlicht: (2023)
von: Zhang, Siqi, et al.
Veröffentlicht: (2023)
Collaborative Satellite Computing through Adaptive DNN Task Splitting and Offloading
von: Peng, Shifeng, et al.
Veröffentlicht: (2024)
von: Peng, Shifeng, et al.
Veröffentlicht: (2024)
CkIO: Parallel File Input for Over-Decomposed Task-Based Systems
von: Jacob, Mathew, et al.
Veröffentlicht: (2024)
von: Jacob, Mathew, et al.
Veröffentlicht: (2024)
Persistent and Partitioned MPI for Stencil Communication
von: Collom, Gerald, et al.
Veröffentlicht: (2025)
von: Collom, Gerald, et al.
Veröffentlicht: (2025)
dpBento: Benchmarking DPUs for Data Processing
von: Hu, Jiasheng, et al.
Veröffentlicht: (2025)
von: Hu, Jiasheng, et al.
Veröffentlicht: (2025)
Holmes: Towards Distributed Training Across Clusters with Heterogeneous NIC Environment
von: Yang, Fei, et al.
Veröffentlicht: (2023)
von: Yang, Fei, et al.
Veröffentlicht: (2023)
Accelerating Mixture-of-Experts Inference by Hiding Offloading Latency with Speculative Decoding
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Disaggregated Memory with SmartNIC Offloading: a Case Study on Graph Processing
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2024) -
Inter-APU Communication on AMD MI300A Systems via Infinity Fabric: a Deep Dive
von: Schieffer, Gabin, et al.
Veröffentlicht: (2025) -
Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs
von: Wahlgren, Jacob, et al.
Veröffentlicht: (2025) -
Plug & Offload: Transparently Offloading TCP Stack onto Off-path SmartNIC with PnO-TCP
von: Nan, Hailong, et al.
Veröffentlicht: (2025) -
Multi-level Memory-Centric Profiling on ARM Processors with ARM SPE
von: Miksits, Samuel, et al.
Veröffentlicht: (2024)