Saved in:
| Main Author: | Telatin, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.04558 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The High Cost of Keeping Warm: Characterizing Overhead in Serverless Autoscaling Policies
by: Kondrashov, Leonid, et al.
Published: (2025)
by: Kondrashov, Leonid, et al.
Published: (2025)
Sky$^ε$-Tree: Embracing the Batch Updates of B$^ε$-trees through Access Port Parallelism on Skyrmion Racetrack Memory
by: Tsai, Yu-Shiang, et al.
Published: (2024)
by: Tsai, Yu-Shiang, et al.
Published: (2024)
Trident: Adaptive Scheduling for Heterogeneous Multimodal Data Pipelines
by: Pan, Ding, et al.
Published: (2026)
by: Pan, Ding, et al.
Published: (2026)
Serving LLMs in HPC Clusters: A Comparative Study of Qualcomm Cloud AI 100 Ultra and NVIDIA Data Center GPUs
by: Sada, Mohammad Firas, et al.
Published: (2025)
by: Sada, Mohammad Firas, et al.
Published: (2025)
Next-Generation Event-Driven Architectures: Performance, Scalability, and Intelligent Orchestration Across Messaging Frameworks
by: Arafat, Jahidul, et al.
Published: (2025)
by: Arafat, Jahidul, et al.
Published: (2025)
Optimizing Multi-DNN Inference on Mobile Devices through Heterogeneous Processor Co-Execution
by: Gao, Yunquan, et al.
Published: (2025)
by: Gao, Yunquan, et al.
Published: (2025)
Big Data Workload Profiling for Energy-Aware Cloud Resource Management
by: Parikh, Milan, et al.
Published: (2026)
by: Parikh, Milan, et al.
Published: (2026)
Attack-Centric by Design: A Program-Structure Taxonomy of Smart Contract Vulnerabilities
by: Hedayatnia, Parsa, et al.
Published: (2025)
by: Hedayatnia, Parsa, et al.
Published: (2025)
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
by: Lysenstøen, Christian
Published: (2026)
by: Lysenstøen, Christian
Published: (2026)
A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices
by: Jallouli, Chaimae, et al.
Published: (2026)
by: Jallouli, Chaimae, et al.
Published: (2026)
Work-Efficient Parallel Non-Maximum Suppression Kernels
by: Oro, David, et al.
Published: (2025)
by: Oro, David, et al.
Published: (2025)
Learning Interpretable Scheduling Algorithms for Data Processing Clusters
by: Hu, Zhibo, et al.
Published: (2024)
by: Hu, Zhibo, et al.
Published: (2024)
ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving
by: Li, Xiangchen, et al.
Published: (2026)
by: Li, Xiangchen, et al.
Published: (2026)
WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
by: Li, Xiangchen, et al.
Published: (2026)
by: Li, Xiangchen, et al.
Published: (2026)
Melding the Serverless Control Plane with the Conventional Cluster Manager for Speed and Resource Efficiency
by: Kondrashov, Leonid, et al.
Published: (2025)
by: Kondrashov, Leonid, et al.
Published: (2025)
Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents
by: Rao, Swanand, et al.
Published: (2026)
by: Rao, Swanand, et al.
Published: (2026)
Optimal Graph Stretching for Distributed Averaging
by: Dekker, Florine W., et al.
Published: (2025)
by: Dekker, Florine W., et al.
Published: (2025)
Addressing tokens dynamic generation, propagation, storage and renewal to secure the GlideinWMS pilot based jobs and system
by: Coimbra, Bruno Moreira, et al.
Published: (2025)
by: Coimbra, Bruno Moreira, et al.
Published: (2025)
Building the Palmetto API: Adding granular permissions and caching to the Slurm REST API without sacrificing compatibility
by: Godfrey, Ben, et al.
Published: (2026)
by: Godfrey, Ben, et al.
Published: (2026)
Design, Configuration, Implementation, and Performance of a Simple 32 Core Raspberry Pi Cluster
by: Cicirello, Vincent A.
Published: (2017)
by: Cicirello, Vincent A.
Published: (2017)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
by: Putra, Jody Almaida
Published: (2026)
by: Putra, Jody Almaida
Published: (2026)
A sandbox study proposal for private and distributed health data analysis
by: Brännvall, Rickard, et al.
Published: (2025)
by: Brännvall, Rickard, et al.
Published: (2025)
INAR-VL: Input-Aware Routing for Edge-Cloud Vision-Language Inference
by: Šabanović, Ahmed, et al.
Published: (2026)
by: Šabanović, Ahmed, et al.
Published: (2026)
On Reduction and Synthesis of Petri's Cycloids
by: Valk, Rüdiger, et al.
Published: (2025)
by: Valk, Rüdiger, et al.
Published: (2025)
Modelling cooperating failure-resilient Processes
by: Valk, Rüdiger
Published: (2024)
by: Valk, Rüdiger
Published: (2024)
Shattering the Ephemeral Storage Cost Barrier for Data-Intensive Serverless Workflows
by: Ustiugov, Dmitrii, et al.
Published: (2023)
by: Ustiugov, Dmitrii, et al.
Published: (2023)
Operational Memory Architecture for Kubernetes:Preserving Causal Context Across the Evidence Horizon
by: Khan, Shamsher
Published: (2026)
by: Khan, Shamsher
Published: (2026)
PIM-STM: Software Transactional Memory for Processing-In-Memory Systems
by: Lopes, André, et al.
Published: (2024)
by: Lopes, André, et al.
Published: (2024)
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
Solving Large Rank-Deficient Linear Least-Squares Problems on Shared-Memory CPU Architectures and GPU Architectures
by: Chillarón, Mónica, et al.
Published: (2024)
by: Chillarón, Mónica, et al.
Published: (2024)
Analysis of Design Patterns and Benchmark Practices in Apache Kafka Event-Streaming Systems
by: Mohammad, Muzeeb
Published: (2025)
by: Mohammad, Muzeeb
Published: (2025)
A C++17 Thread Pool for High-Performance Scientific Computing
by: Shoshany, Barak
Published: (2021)
by: Shoshany, Barak
Published: (2021)
LAMMPS-KOKKOS: Performance Portable Molecular Dynamics Across Exascale Architectures
by: Johansson, Anders, et al.
Published: (2025)
by: Johansson, Anders, et al.
Published: (2025)
Chain-Oriented Objective Logic with Neural Network Feedback Control and Cascade Filtering for Dynamic Multi-DSL Regulation
by: Han, Jipeng
Published: (2024)
by: Han, Jipeng
Published: (2024)
CAWAL: A novel unified analytics framework for enterprise web applications and multi-server environments
by: Canay, Özkan, et al.
Published: (2025)
by: Canay, Özkan, et al.
Published: (2025)
Model Gateway: Model Management Platform for Model-Driven Drug Discovery
by: Wu, Yan-Shiun, et al.
Published: (2025)
by: Wu, Yan-Shiun, et al.
Published: (2025)
Automated Dynamic AI Inference Scaling on HPC-Infrastructure: Integrating Kubernetes, Slurm and vLLM
by: Trappen, Tim, et al.
Published: (2025)
by: Trappen, Tim, et al.
Published: (2025)
Rust vs. C for Python Libraries: Evaluating Rust-Compatible Bindings Toolchains
by: Amaral, Isabella Basso do, et al.
Published: (2025)
by: Amaral, Isabella Basso do, et al.
Published: (2025)
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
by: Melnychuk, Anton, et al.
Published: (2025)
by: Melnychuk, Anton, et al.
Published: (2025)
Chat AI: A Seamless Slurm-Native Solution for HPC-Based Services
by: Doosthosseini, Ali, et al.
Published: (2024)
by: Doosthosseini, Ali, et al.
Published: (2024)
Similar Items
-
The High Cost of Keeping Warm: Characterizing Overhead in Serverless Autoscaling Policies
by: Kondrashov, Leonid, et al.
Published: (2025) -
Sky$^ε$-Tree: Embracing the Batch Updates of B$^ε$-trees through Access Port Parallelism on Skyrmion Racetrack Memory
by: Tsai, Yu-Shiang, et al.
Published: (2024) -
Trident: Adaptive Scheduling for Heterogeneous Multimodal Data Pipelines
by: Pan, Ding, et al.
Published: (2026) -
Serving LLMs in HPC Clusters: A Comparative Study of Qualcomm Cloud AI 100 Ultra and NVIDIA Data Center GPUs
by: Sada, Mohammad Firas, et al.
Published: (2025) -
Next-Generation Event-Driven Architectures: Performance, Scalability, and Intelligent Orchestration Across Messaging Frameworks
by: Arafat, Jahidul, et al.
Published: (2025)