How long can you sleep? Idle Time System Inefficiencies and Opportunities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Antoniou, Georgia, Volos, Haris, Yahya, Jawad Haj, Sazeides, Yiannakis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AgilePkgC: An Agile System Idle State Architecture for Energy Proportional Datacenter Servers
von: Antoniou, Georgia, et al.
Veröffentlicht: (2022)
von: Antoniou, Georgia, et al.
Veröffentlicht: (2022)
Taming Performance Variability caused by Client-Side Hardware Configuration
von: Antoniou, Georgia, et al.
Veröffentlicht: (2024)
von: Antoniou, Georgia, et al.
Veröffentlicht: (2024)
Analyzing a Two-Tier Disaggregated Memory Protection Scheme Based on Memory Replication
von: Volos, Haris, et al.
Veröffentlicht: (2025)
von: Volos, Haris, et al.
Veröffentlicht: (2025)
AgileWatts: An Energy-Efficient CPU Core Idle-State Architecture for Latency-Sensitive Server Applications
von: Yahya, Jawad Haj, et al.
Veröffentlicht: (2022)
von: Yahya, Jawad Haj, et al.
Veröffentlicht: (2022)
An opportunity to improve Data Center Efficiency: Optimizing the Server's Upgrade Cycle
von: Nikolaou, Panagiota, et al.
Veröffentlicht: (2025)
von: Nikolaou, Panagiota, et al.
Veröffentlicht: (2025)
The Non-Predictability of Mispredicted Branches using Timing Information
von: Constantinou, Ioannis, et al.
Veröffentlicht: (2026)
von: Constantinou, Ioannis, et al.
Veröffentlicht: (2026)
A Comparative Analysis of ARM and x86-64 Laptop-Class Processors: Architecture, Assembly-Level Performance, and Energy Efficiency
von: Özyılmaz, Mustafa Mert
Veröffentlicht: (2026)
von: Özyılmaz, Mustafa Mert
Veröffentlicht: (2026)
Tekum: Balanced Ternary Tapered Precision Real Arithmetic
von: Hunhold, Laslo
Veröffentlicht: (2025)
von: Hunhold, Laslo
Veröffentlicht: (2025)
The Case for Replication-Aware Memory-Error Protection in Disaggregated Memory
von: Volos, Haris
Veröffentlicht: (2023)
von: Volos, Haris
Veröffentlicht: (2023)
A Flexible Instruction Set Architecture for Efficient GEMMs
von: Santana, Alexandre de Limas, et al.
Veröffentlicht: (2025)
von: Santana, Alexandre de Limas, et al.
Veröffentlicht: (2025)
basic_RV32s: An Open-Source Microarchitectural Roadmap for RISC-V RV32I
von: Kang, Hyun Woo, et al.
Veröffentlicht: (2025)
von: Kang, Hyun Woo, et al.
Veröffentlicht: (2025)
RV-IM100: Quantifying ISA Extension, Datapath Width, and Pipeline Depth Trade-offs in RISC-V Microarchitectures
von: Kang, Hyunwoo
Veröffentlicht: (2026)
von: Kang, Hyunwoo
Veröffentlicht: (2026)
LUT Tensor Core: A Software-Hardware Co-Design for LUT-Based Low-Bit LLM Inference
von: Mo, Zhiwen, et al.
Veröffentlicht: (2024)
von: Mo, Zhiwen, et al.
Veröffentlicht: (2024)
RayFlex: An Open-Source RTL Implementation of the Hardware Ray Tracer Datapath
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
Data Gravity and the Energy Limits of Computation
von: Lee, Wonsuk, et al.
Veröffentlicht: (2026)
von: Lee, Wonsuk, et al.
Veröffentlicht: (2026)
SambaNova SN40L: Scaling the AI Memory Wall with Dataflow and Composition of Experts
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
Streamlining SIMD ISA Extensions with Takum Arithmetic: A Case Study on Intel AVX10.2
von: Hunhold, Laslo
Veröffentlicht: (2025)
von: Hunhold, Laslo
Veröffentlicht: (2025)
Application-Driven Exascale: The JUPITER Benchmark Suite
von: Herten, Andreas, et al.
Veröffentlicht: (2024)
von: Herten, Andreas, et al.
Veröffentlicht: (2024)
FREESS: A Web-Based Educational Simulator for a RISC-V-Inspired Superscalar Processor with Tomasulo-Style Dynamic Scheduling
von: Giorgi, Roberto, et al.
Veröffentlicht: (2026)
von: Giorgi, Roberto, et al.
Veröffentlicht: (2026)
Mestra: Exploring Migration on Virtualized CGRAs
von: Kyriazis, Agamemnon, et al.
Veröffentlicht: (2026)
von: Kyriazis, Agamemnon, et al.
Veröffentlicht: (2026)
Victima: Drastically Increasing Address Translation Reach by Leveraging Underutilized Cache Resources
von: Kanellopoulos, Konstantinos, et al.
Veröffentlicht: (2023)
von: Kanellopoulos, Konstantinos, et al.
Veröffentlicht: (2023)
PG-MDP: Profile-Guided Memory Dependence Prediction for Area-Constrained Cores
von: Panayi, Luke, et al.
Veröffentlicht: (2026)
von: Panayi, Luke, et al.
Veröffentlicht: (2026)
Multi-diseases detection with memristive system on chip
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Improving Memory Dependence Prediction with Static Analysis
von: Panayi, Luke, et al.
Veröffentlicht: (2024)
von: Panayi, Luke, et al.
Veröffentlicht: (2024)
Silent Data Corruption by 10x Test Escapes Threatens Reliable Computing
von: Mitra, Subhasish, et al.
Veröffentlicht: (2025)
von: Mitra, Subhasish, et al.
Veröffentlicht: (2025)
An SMT Formalization of Mixed-Precision Matrix Multiplication: Modeling Three Generations of Tensor Cores
von: Valpey, Benjamin, et al.
Veröffentlicht: (2025)
von: Valpey, Benjamin, et al.
Veröffentlicht: (2025)
AMC: Access to Miss Correlation Prefetcher for Evolving Graph Analytics
von: Singh, Abhishek, et al.
Veröffentlicht: (2024)
von: Singh, Abhishek, et al.
Veröffentlicht: (2024)
Fast NF4 Dequantization Kernels for Large Language Model Inference
von: Qi, Xiangbo, et al.
Veröffentlicht: (2026)
von: Qi, Xiangbo, et al.
Veröffentlicht: (2026)
Taming Wild Branches: Overcoming Hard-to-Predict Branches using the Bullseye Predictor
von: Behrendt, Emet, et al.
Veröffentlicht: (2025)
von: Behrendt, Emet, et al.
Veröffentlicht: (2025)
Glass-Box Analysis for Computer Systems: Transparency Index, Shapley Attribution, and Markov Models of Branch Prediction
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
von: Alpay, Faruk, et al.
Veröffentlicht: (2025)
parti-gem5: gem5's Timing Mode Parallelised
von: Cubero-Cascante, José, et al.
Veröffentlicht: (2023)
von: Cubero-Cascante, José, et al.
Veröffentlicht: (2023)
SoK: Where's the "up"?! A Comprehensive (bottom-up) Study on the Security of Arm Cortex-M Systems
von: Tan, Xi, et al.
Veröffentlicht: (2024)
von: Tan, Xi, et al.
Veröffentlicht: (2024)
Pushing the Memory Bandwidth Wall with CXL-enabled Idle I/O Bandwidth Harvesting
von: Kadiyala, Divya Kiran, et al.
Veröffentlicht: (2025)
von: Kadiyala, Divya Kiran, et al.
Veröffentlicht: (2025)
MEDEA: A Design-Time Multi-Objective Manager for Energy-Efficient DNN Inference on Heterogeneous Ultra-Low Power Platforms
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
Ten-Four: An Open-Source Fused Dot Product Unit for Mixed-Precision GPGPU Tensor Cores
von: Rout, Nikhil, et al.
Veröffentlicht: (2025)
von: Rout, Nikhil, et al.
Veröffentlicht: (2025)
Fast and Practical Strassen's Matrix Multiplication using FPGAs
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
SparseZipper: Enhancing Matrix Extensions to Accelerate SpGEMM on CPUs
von: Ta, Tuan, et al.
Veröffentlicht: (2025)
von: Ta, Tuan, et al.
Veröffentlicht: (2025)
Factor Machine: Mixed-signal Architecture for Fine-Grained Graph-Based Computing
von: Dudek, Piotr
Veröffentlicht: (2024)
von: Dudek, Piotr
Veröffentlicht: (2024)
Make LLM Inference Affordable to Everyone: Augmenting GPU Memory with NDP-DIMM
von: Liu, Lian, et al.
Veröffentlicht: (2025)
von: Liu, Lian, et al.
Veröffentlicht: (2025)
DARE: An Irregularity-Tolerant Matrix Processing Unit with a Densifying ISA and Filtered Runahead Execution
von: Yang, Xin, et al.
Veröffentlicht: (2025)
von: Yang, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AgilePkgC: An Agile System Idle State Architecture for Energy Proportional Datacenter Servers
von: Antoniou, Georgia, et al.
Veröffentlicht: (2022) -
Taming Performance Variability caused by Client-Side Hardware Configuration
von: Antoniou, Georgia, et al.
Veröffentlicht: (2024) -
Analyzing a Two-Tier Disaggregated Memory Protection Scheme Based on Memory Replication
von: Volos, Haris, et al.
Veröffentlicht: (2025) -
AgileWatts: An Energy-Efficient CPU Core Idle-State Architecture for Latency-Sensitive Server Applications
von: Yahya, Jawad Haj, et al.
Veröffentlicht: (2022) -
An opportunity to improve Data Center Efficiency: Optimizing the Server's Upgrade Cycle
von: Nikolaou, Panagiota, et al.
Veröffentlicht: (2025)