Towards CXL Resilience to CPU Failures
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Psistakis, Antonis, Ocalan, Burak, Alverti, Chloe, Chaix, Fabien, Alagappan, Ramnatthan, Torrellas, Josep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HEAL: Online Incremental Recovery for Leaderless Distributed Systems Across Persistency Models
von: Psistakis, Antonis, et al.
Veröffentlicht: (2026)
von: Psistakis, Antonis, et al.
Veröffentlicht: (2026)
IOMMU Support for Virtual-Address Remote DMA in an ARMv8 environment
von: Psistakis, Antonis
Veröffentlicht: (2025)
von: Psistakis, Antonis
Veröffentlicht: (2025)
Handling of Memory Page Faults during Virtual-Address RDMA
von: Psistakis, Antonis
Veröffentlicht: (2025)
von: Psistakis, Antonis
Veröffentlicht: (2025)
AgileLog: A Forkable Shared Log for Agents on Data Streams
von: Bhat, Shreesha G., et al.
Veröffentlicht: (2026)
von: Bhat, Shreesha G., et al.
Veröffentlicht: (2026)
Chameleon: Adaptive Caching and Scheduling for Many-Adapter LLM Inference Environments
von: Iliakopoulou, Nikoleta, et al.
Veröffentlicht: (2024)
von: Iliakopoulou, Nikoleta, et al.
Veröffentlicht: (2024)
Distributed-Memory Parallel Algorithms for Sparse Matrix and Sparse Tall-and-Skinny Matrix Multiplication
von: Ranawaka, Isuru, et al.
Veröffentlicht: (2024)
von: Ranawaka, Isuru, et al.
Veröffentlicht: (2024)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
von: Qiao, Tong, et al.
Veröffentlicht: (2025)
von: Qiao, Tong, et al.
Veröffentlicht: (2025)
CXL Shared Memory Programming: Barely Distributed and Almost Persistent
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Modeling the Potential of Message-Free Communication via CXL.mem
von: Vanecek, Stepan, et al.
Veröffentlicht: (2025)
von: Vanecek, Stepan, et al.
Veröffentlicht: (2025)
emucxl: an emulation framework for CXL-based disaggregated memory applications
von: Gond, Raja, et al.
Veröffentlicht: (2024)
von: Gond, Raja, et al.
Veröffentlicht: (2024)
MPI-over-CXL: Enhancing Communication Efficiency in Distributed HPC Systems
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
Serving Heterogeneous LoRA Adapters in Distributed LLM Inference Systems
von: Jaiswal, Shashwat, et al.
Veröffentlicht: (2025)
von: Jaiswal, Shashwat, et al.
Veröffentlicht: (2025)
Resilience Evaluation of Kubernetes in Cloud-Edge Environments via Failure Injection
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
Analysis and Optimized CXL-Attached Memory Allocation for Long-Context LLM Fine-Tuning
von: Liaw, Yong-Cheng, et al.
Veröffentlicht: (2025)
von: Liaw, Yong-Cheng, et al.
Veröffentlicht: (2025)
WindVE: Collaborative CPU-NPU Vector Embedding
von: Huang, Jinqi, et al.
Veröffentlicht: (2025)
von: Huang, Jinqi, et al.
Veröffentlicht: (2025)
A Unified CPU-GPU Protocol for GNN Training
von: Lin, Yi-Chien, et al.
Veröffentlicht: (2024)
von: Lin, Yi-Chien, et al.
Veröffentlicht: (2024)
Combining GPU and CPU for accelerating evolutionary computing workloads
von: Eynaliyev, Rustam, et al.
Veröffentlicht: (2025)
von: Eynaliyev, Rustam, et al.
Veröffentlicht: (2025)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
von: Ponce, Francisco, et al.
Veröffentlicht: (2026)
von: Ponce, Francisco, et al.
Veröffentlicht: (2026)
FailLite: Failure-Resilient Model Serving for Resource-Constrained Edge Environments
von: Wu, Li, et al.
Veröffentlicht: (2025)
von: Wu, Li, et al.
Veröffentlicht: (2025)
TraCT: Disaggregated LLM Serving with CXL Shared Memory KV Cache at Rack-Scale
von: Yoon, Dongha, et al.
Veröffentlicht: (2025)
von: Yoon, Dongha, et al.
Veröffentlicht: (2025)
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
von: Ding, Zhimin, et al.
Veröffentlicht: (2024)
ScalePool: Hybrid XLink-CXL Fabric for Composable Resource Disaggregation in Unified Scale-up Domains
von: Woo, Hyein, et al.
Veröffentlicht: (2025)
von: Woo, Hyein, et al.
Veröffentlicht: (2025)
Efficient Column-Wise N:M Pruning on RISC-V CPU
von: Chu, Chi-Wei, et al.
Veröffentlicht: (2025)
von: Chu, Chi-Wei, et al.
Veröffentlicht: (2025)
A Unified Programming Model for Heterogeneous Computing with CPU and Accelerator Technologies
von: Xiong, Yuqing
Veröffentlicht: (2022)
von: Xiong, Yuqing
Veröffentlicht: (2022)
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
Justin: Hybrid CPU/Memory Elastic Scaling for Distributed Stream Processing
von: Schmitz, Donatien, et al.
Veröffentlicht: (2025)
von: Schmitz, Donatien, et al.
Veröffentlicht: (2025)
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
von: Schieffer, Gabin, et al.
Veröffentlicht: (2026)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2026)
Performance characterisation of the 64-core SG2042 RISC-V CPU for HPC
von: Brown, Nick, et al.
Veröffentlicht: (2024)
von: Brown, Nick, et al.
Veröffentlicht: (2024)
MMStencil: Optimizing High-order Stencils on Multicore CPU using Matrix Unit
von: Wang, Yinuo, et al.
Veröffentlicht: (2025)
von: Wang, Yinuo, et al.
Veröffentlicht: (2025)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
von: Yi, Xinyao
Veröffentlicht: (2024)
von: Yi, Xinyao
Veröffentlicht: (2024)
Dual-pronged deep learning preprocessing on heterogeneous platforms with CPU, Accelerator and CSD
von: Wei, Jia, et al.
Veröffentlicht: (2024)
von: Wei, Jia, et al.
Veröffentlicht: (2024)
Co-Design and Evaluation of a CPU-Free MPI GPU Communication Abstraction and Implementation
von: Bridges, Patrick G., et al.
Veröffentlicht: (2026)
von: Bridges, Patrick G., et al.
Veröffentlicht: (2026)
Serving Hybrid LLM Loads with SLO Guarantees Using CPU-GPU Attention Piggybacking
von: Mo, Zizhao, et al.
Veröffentlicht: (2026)
von: Mo, Zizhao, et al.
Veröffentlicht: (2026)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
Aging-aware CPU Core Management for Embodied Carbon Amortization in Cloud LLM Inference
von: Hewage, Tharindu B., et al.
Veröffentlicht: (2025)
von: Hewage, Tharindu B., et al.
Veröffentlicht: (2025)
AcOrch: Accelerating Sampling-based GNN Training under CPU-NPU Heterogeneous Environments
von: Chen, Kefu, et al.
Veröffentlicht: (2026)
von: Chen, Kefu, et al.
Veröffentlicht: (2026)
CaraServe: CPU-Assisted and Rank-Aware LoRA Serving for Generative LLM Inference
von: Li, Suyi, et al.
Veröffentlicht: (2024)
von: Li, Suyi, et al.
Veröffentlicht: (2024)
Efficient CPU-GPU Collaborative Inference for MoE-based LLMs on Memory-Limited Systems
von: Huang, En-Ming, et al.
Veröffentlicht: (2025)
von: Huang, En-Ming, et al.
Veröffentlicht: (2025)
SiPipe: Bridging the CPU-GPU Utilization Gap for Efficient Pipeline-Parallel LLM Inference
von: He, Yongchao, et al.
Veröffentlicht: (2025)
von: He, Yongchao, et al.
Veröffentlicht: (2025)
Harnessing Integrated CPU-GPU System Memory for HPC: a first look into Grace Hopper
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
von: Schieffer, Gabin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
HEAL: Online Incremental Recovery for Leaderless Distributed Systems Across Persistency Models
von: Psistakis, Antonis, et al.
Veröffentlicht: (2026) -
IOMMU Support for Virtual-Address Remote DMA in an ARMv8 environment
von: Psistakis, Antonis
Veröffentlicht: (2025) -
Handling of Memory Page Faults during Virtual-Address RDMA
von: Psistakis, Antonis
Veröffentlicht: (2025) -
AgileLog: A Forkable Shared Log for Agents on Data Streams
von: Bhat, Shreesha G., et al.
Veröffentlicht: (2026) -
Chameleon: Adaptive Caching and Scheduling for Many-Adapter LLM Inference Environments
von: Iliakopoulou, Nikoleta, et al.
Veröffentlicht: (2024)