Saved in:
| Main Authors: | Khalil, Alex, Heilles, Guillaume, Parraga, Maria, Heilles, Simon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.23029 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Experimental Analysis of Server-Side Caching for Web Performance
by: Umar, Mohammad, et al.
Published: (2026)
by: Umar, Mohammad, et al.
Published: (2026)
SCARIF: Towards Carbon Modeling of Cloud Servers with Accelerators
by: Ji, Shixin, et al.
Published: (2024)
by: Ji, Shixin, et al.
Published: (2024)
PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers
by: Zhang, Hongbin, et al.
Published: (2026)
by: Zhang, Hongbin, et al.
Published: (2026)
Experiences with Model Context Protocol Servers for Science and High Performance Computing
by: Pan, Haochen, et al.
Published: (2025)
by: Pan, Haochen, et al.
Published: (2025)
TailBench++: Flexible Multi-Client, Multi-Server Benchmarking for Latency-Critical Workloads
by: Li, Zhilin, et al.
Published: (2025)
by: Li, Zhilin, et al.
Published: (2025)
Analysis of Server Throughput For Managed Big Data Analytics Frameworks
by: Anagnostakis, Emmanouil, et al.
Published: (2025)
by: Anagnostakis, Emmanouil, et al.
Published: (2025)
From Servers to Sites: Compositional Power Trace Generation of LLM Inference for Infrastructure Planning
by: Wilkins, Grant, et al.
Published: (2026)
by: Wilkins, Grant, et al.
Published: (2026)
Are Bus-Mounted Edge Servers Feasible?
by: Li, Xuezhi, et al.
Published: (2025)
by: Li, Xuezhi, et al.
Published: (2025)
Revisiting Parameter Server in LLM Post-Training
by: Wan, Xinyi, et al.
Published: (2026)
by: Wan, Xinyi, et al.
Published: (2026)
AgentServe: Algorithm-System Co-Design for Efficient Agentic AI Serving on a Consumer-Grade GPU
by: Zhang, Yuning, et al.
Published: (2026)
by: Zhang, Yuning, et al.
Published: (2026)
Benchmarking Compound AI Applications for Hardware-Software Co-Design
by: Samuthrsindh, Paramuth, et al.
Published: (2026)
by: Samuthrsindh, Paramuth, et al.
Published: (2026)
Gateways for Institutional-Grade Commerce and Interoperability of Digital Assets
by: Belchior, Rafael, et al.
Published: (2025)
by: Belchior, Rafael, et al.
Published: (2025)
Practical Federated Learning without a Server
by: Dhasade, Akash, et al.
Published: (2025)
by: Dhasade, Akash, et al.
Published: (2025)
Exploring the Viability of Unikernels for ARM-powered Edge Computing
by: Kaiser, Shahidullah, et al.
Published: (2024)
by: Kaiser, Shahidullah, et al.
Published: (2024)
A Multi-Server Information-Sharing Environment for Cross-Party Collaboration on A Private Cloud
by: Zhang, Jianping, et al.
Published: (2024)
by: Zhang, Jianping, et al.
Published: (2024)
Adversarial Analysis of the Differentially-Private Federated Learning in Cyber-Physical Critical Infrastructures
by: Hossain, Md Tamjid, et al.
Published: (2022)
by: Hossain, Md Tamjid, et al.
Published: (2022)
Performance Characterization of Distributed Deep Learning Strategies: A Quantitative Evaluation of DDP, FSDP, and Parameter Server Architectures on GPU Clusters
by: Ovi, Md Sultanul Islam
Published: (2025)
by: Ovi, Md Sultanul Islam
Published: (2025)
Leveraging Hardware Performance Counters for Predicting Workload Interference in Vector Supercomputers
by: Shubham, et al.
Published: (2024)
by: Shubham, et al.
Published: (2024)
LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU
by: Liao, Changyue, et al.
Published: (2024)
by: Liao, Changyue, et al.
Published: (2024)
Predictive Analysis of CFPB Consumer Complaints Using Machine Learning
by: Vaishnav, Dhwani, et al.
Published: (2024)
by: Vaishnav, Dhwani, et al.
Published: (2024)
Beyond End-to-End: Dynamic Chain Optimization for Private LLM Adaptation on the Edge
by: Wu, Yebo, et al.
Published: (2026)
by: Wu, Yebo, et al.
Published: (2026)
6G EdgeAI: Performance Evaluation and Analysis
by: Yang, Chien-Sheng, et al.
Published: (2025)
by: Yang, Chien-Sheng, et al.
Published: (2025)
DiSCo: Device-Server Collaborative LLM-Based Text Streaming Services
by: Sun, Ting, et al.
Published: (2025)
by: Sun, Ting, et al.
Published: (2025)
ReviveMoE: Fast Recovery for Hardware Failures in Large-Scale MoE LLM Inference Deployments
by: Li, Haley, et al.
Published: (2026)
by: Li, Haley, et al.
Published: (2026)
More is Different: Prototyping and Analyzing a New Form of Edge Server with Massive Mobile SoCs
by: Zhang, Li, et al.
Published: (2022)
by: Zhang, Li, et al.
Published: (2022)
BSODiag: A Global Diagnosis Framework for Batch Servers Outage in Large-scale Cloud Infrastructure Systems
by: Duan, Tao, et al.
Published: (2025)
by: Duan, Tao, et al.
Published: (2025)
Gaia: Hybrid Hardware Acceleration for Serverless AI in the 3D Compute Continuum
by: Reisecker, Maximilian, et al.
Published: (2025)
by: Reisecker, Maximilian, et al.
Published: (2025)
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks
by: Nader, Noujoud, et al.
Published: (2025)
by: Nader, Noujoud, et al.
Published: (2025)
Oases: Efficient Large-Scale Model Training on Commodity Servers via Overlapped and Automated Tensor Model Parallelism
by: Li, Shengwei, et al.
Published: (2023)
by: Li, Shengwei, et al.
Published: (2023)
Performance Evaluation of Hashing Algorithms on Commodity Hardware
by: Pandya, Marut
Published: (2024)
by: Pandya, Marut
Published: (2024)
Comparing the Performance of Heterogeneous Conjugate Gradient and Cholesky Solvers on Various Hardware Using SYCL
by: Thüring, Tim, et al.
Published: (2026)
by: Thüring, Tim, et al.
Published: (2026)
Understanding the Performance and Power of LLM Inferencing on Edge Accelerators
by: Arya, Mayank, et al.
Published: (2025)
by: Arya, Mayank, et al.
Published: (2025)
Data-Driven Analysis to Understand GPU Hardware Resource Usage of Optimizations
by: Islam, Tanzima Z., et al.
Published: (2024)
by: Islam, Tanzima Z., et al.
Published: (2024)
Benchmarking the Performance of Large Language Models on the Cerebras Wafer Scale Engine
by: Zhang, Zuoning, et al.
Published: (2024)
by: Zhang, Zuoning, et al.
Published: (2024)
Towards Energy-Efficient Serverless Computing with Hardware Isolation
by: Carl, Natalie, et al.
Published: (2025)
by: Carl, Natalie, et al.
Published: (2025)
IM-PIR: In-Memory Private Information Retrieval
by: Mwaisela, Mpoki, et al.
Published: (2025)
by: Mwaisela, Mpoki, et al.
Published: (2025)
Availability Modeling for Blockchain Provisioning in Private Clouds
by: Dantas, J, et al.
Published: (2025)
by: Dantas, J, et al.
Published: (2025)
SCOOT: SLO-Oriented Performance Tuning for LLM Inference Engines
by: Cheng, Ke, et al.
Published: (2024)
by: Cheng, Ke, et al.
Published: (2024)
City-Scale Visibility Graph Analysis via GPU-Accelerated HyperBall
by: Hodge, Alex, et al.
Published: (2026)
by: Hodge, Alex, et al.
Published: (2026)
Benchmarking Message Brokers for IoT Edge Computing: A Comprehensive Performance Study
by: Paul, Tapajit Chandra, et al.
Published: (2026)
by: Paul, Tapajit Chandra, et al.
Published: (2026)
Similar Items
-
Experimental Analysis of Server-Side Caching for Web Performance
by: Umar, Mohammad, et al.
Published: (2026) -
SCARIF: Towards Carbon Modeling of Cloud Servers with Accelerators
by: Ji, Shixin, et al.
Published: (2024) -
PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers
by: Zhang, Hongbin, et al.
Published: (2026) -
Experiences with Model Context Protocol Servers for Science and High Performance Computing
by: Pan, Haochen, et al.
Published: (2025) -
TailBench++: Flexible Multi-Client, Multi-Server Benchmarking for Latency-Critical Workloads
by: Li, Zhilin, et al.
Published: (2025)