Beyond Static Policies: Exploring Dynamic Policy Selection for Single-Thread Performance Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yanxin, McDougall, Ian, Li, Junnan, Wadle, Shayne, Singh, Vikas, Sankaralingam, Karthikeyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SAHM: State-Aware Heterogeneous Multicore for Single-Thread Performance
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
NeuroScalar: A Deep Learning Framework for Fast, Accurate, and In-the-Wild Cycle-Level Performance Prediction
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
IPU: Flexible Hardware Introspection Units
von: McDougall, Ian, et al.
Veröffentlicht: (2023)
von: McDougall, Ian, et al.
Veröffentlicht: (2023)
Pedagogically Motivated and Composable Open-Source RISC-V Processors for Computer Science Education
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
Privacy-Preserving Performance Profiling of In-The-Wild GPUs
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
von: McDougall, Ian, et al.
Veröffentlicht: (2025)
LIMINAL: Exploring The Frontiers of LLM Decode Performance
von: Davies, Michael, et al.
Veröffentlicht: (2025)
von: Davies, Michael, et al.
Veröffentlicht: (2025)
Computer Architecture's AlphaZero Moment: Automated Discovery in an Encircled World
von: Sankaralingam, Karthikeyan
Veröffentlicht: (2026)
von: Sankaralingam, Karthikeyan
Veröffentlicht: (2026)
The Impact Market to Save Conference Peer Review: Decoupling Dissemination and Credentialing
von: Sankaralingam, Karthikeyan
Veröffentlicht: (2025)
von: Sankaralingam, Karthikeyan
Veröffentlicht: (2025)
Kitsune: Enabling Dataflow Execution on GPUs
von: Davies, Michael, et al.
Veröffentlicht: (2025)
von: Davies, Michael, et al.
Veröffentlicht: (2025)
Revet: A Language and Compiler for Dataflow Threads
von: Rucker, Alexander, et al.
Veröffentlicht: (2023)
von: Rucker, Alexander, et al.
Veröffentlicht: (2023)
A New Family of Thread to Core Allocation Policies for an SMT ARM Processor
von: Navarro, Marta, et al.
Veröffentlicht: (2025)
von: Navarro, Marta, et al.
Veröffentlicht: (2025)
A Statically and Dynamically Scalable Soft GPGPU
von: Langhammer, Martin, et al.
Veröffentlicht: (2024)
von: Langhammer, Martin, et al.
Veröffentlicht: (2024)
Optimizing and Exploring System Performance in Compact Processing-in-Memory-based Chips
von: Chen, Peilin, et al.
Veröffentlicht: (2025)
von: Chen, Peilin, et al.
Veröffentlicht: (2025)
Agile TLB Prefetching and Prediction Replacement Policy
von: Mersha, Melkamu, et al.
Veröffentlicht: (2024)
von: Mersha, Melkamu, et al.
Veröffentlicht: (2024)
SDT: Cutting Datacenter Tax Through Simultaneous Data-Delivery Threads
von: Mamandipoor, Amin, et al.
Veröffentlicht: (2025)
von: Mamandipoor, Amin, et al.
Veröffentlicht: (2025)
EMiX: Emulating Beyond Single-FPGA Limits
von: Kropotov, Alexander, et al.
Veröffentlicht: (2026)
von: Kropotov, Alexander, et al.
Veröffentlicht: (2026)
Switchable Single/Dual Edge Registers for Pipeline Architecture
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2024)
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2024)
Workload Characterization for Branch Predictability
von: Vikas, FNU, et al.
Veröffentlicht: (2025)
von: Vikas, FNU, et al.
Veröffentlicht: (2025)
Integrating Prefetcher Selection with Dynamic Request Allocation Improves Prefetching Efficiency
von: Li, Mengming, et al.
Veröffentlicht: (2025)
von: Li, Mengming, et al.
Veröffentlicht: (2025)
LFOC+: A Fair OS-level Cache-Clustering Policy for Commodity Multicore Systems
von: Saez, Juan Carlos, et al.
Veröffentlicht: (2024)
von: Saez, Juan Carlos, et al.
Veröffentlicht: (2024)
UniCAIM: A Unified CAM/CIM Architecture with Static-Dynamic KV Cache Pruning for Efficient Long-Context LLM Inference
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
Static Hardware Partitioning on RISC-V -- Shortcomings, Limitations, and Prospects
von: Ramsauer, Ralf, et al.
Veröffentlicht: (2022)
von: Ramsauer, Ralf, et al.
Veröffentlicht: (2022)
Modeling and Optimizing Performance Bottlenecks for Neuromorphic Accelerators
von: Yik, Jason, et al.
Veröffentlicht: (2025)
von: Yik, Jason, et al.
Veröffentlicht: (2025)
Manticore: Hardware-Accelerated RTL Simulation with Static Bulk-Synchronous Parallelism
von: Emami, Mahyar, et al.
Veröffentlicht: (2023)
von: Emami, Mahyar, et al.
Veröffentlicht: (2023)
Optimizing GEMM for Energy and Performance on Versal ACAP Architectures
von: Papalamprou, Ilias, et al.
Veröffentlicht: (2025)
von: Papalamprou, Ilias, et al.
Veröffentlicht: (2025)
3DGauCIM: Accelerating Static/Dynamic 3D Gaussian Splatting via Digital CIM for High Frame Rate Real-Time Edge Rendering
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
Exploiting Control-flow Enforcement Technology for Sound and Precise Static Binary Disassembly
von: Zhao, Brian, et al.
Veröffentlicht: (2025)
von: Zhao, Brian, et al.
Veröffentlicht: (2025)
Error Detection and Correction Codes for Safe In-Memory Computations
von: Parrini, Luca, et al.
Veröffentlicht: (2024)
von: Parrini, Luca, et al.
Veröffentlicht: (2024)
Global and Local Attention-based Inception U-Net for Static IR Drop Prediction
von: Chen, Yilu, et al.
Veröffentlicht: (2024)
von: Chen, Yilu, et al.
Veröffentlicht: (2024)
BreakHammer: Enhancing RowHammer Mitigations by Carefully Throttling Suspect Threads
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2024)
Ara2: Exploring Single- and Multi-Core Vector Processing with an Efficient RVV 1.0 Compliant Open-Source Processor
von: Perotti, Matteo, et al.
Veröffentlicht: (2023)
von: Perotti, Matteo, et al.
Veröffentlicht: (2023)
EN-T: Optimizing Tensor Computing Engines Performance via Encoder-Based Methodology
von: Wu, Qizhe, et al.
Veröffentlicht: (2024)
von: Wu, Qizhe, et al.
Veröffentlicht: (2024)
Striking the Balance: GEMM Performance Optimization Across Generations of Ryzen AI NPUs
von: Taka, Endri, et al.
Veröffentlicht: (2025)
von: Taka, Endri, et al.
Veröffentlicht: (2025)
ITHICA: Intra-Thread Instruction Checking Approach for Defect-Induced Silent Data Corruptions
von: Vavelidou, Ioanna, et al.
Veröffentlicht: (2026)
von: Vavelidou, Ioanna, et al.
Veröffentlicht: (2026)
AssertMiner: Module-Level Spec Generation and Assertion Mining using Static Analysis Guided LLMs
von: Lyu, Hongqin, et al.
Veröffentlicht: (2025)
von: Lyu, Hongqin, et al.
Veröffentlicht: (2025)
DICE: Enabling Efficient General-Purpose SIMT Execution with Statically Scheduled Coarse-Grained Reconfigurable Arrays
von: Wang, Jiayi, et al.
Veröffentlicht: (2026)
von: Wang, Jiayi, et al.
Veröffentlicht: (2026)
Real-Time Adaptive Neural Network on FPGA: Enhancing Adaptability through Dynamic Classifier Selection
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
UFO-MAC: A Unified Framework for Optimization of High-Performance Multipliers and Multiply-Accumulators
von: Zuo, Dongsheng, et al.
Veröffentlicht: (2024)
von: Zuo, Dongsheng, et al.
Veröffentlicht: (2024)
GAMA: High-Performance GEMM Acceleration on AMD Versal ML-Optimized AI Engines
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2025)
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2025)
N-TORC: Native Tensor Optimizer for Real-time Constraints
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2025)
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SAHM: State-Aware Heterogeneous Multicore for Single-Thread Performance
von: Wadle, Shayne, et al.
Veröffentlicht: (2025) -
NeuroScalar: A Deep Learning Framework for Fast, Accurate, and In-the-Wild Cycle-Level Performance Prediction
von: Wadle, Shayne, et al.
Veröffentlicht: (2025) -
IPU: Flexible Hardware Introspection Units
von: McDougall, Ian, et al.
Veröffentlicht: (2023) -
Pedagogically Motivated and Composable Open-Source RISC-V Processors for Computer Science Education
von: McDougall, Ian, et al.
Veröffentlicht: (2025) -
Privacy-Preserving Performance Profiling of In-The-Wild GPUs
von: McDougall, Ian, et al.
Veröffentlicht: (2025)