Saved in:
| Main Authors: | Inoie, Atsushi, Inoue, Yoshiaki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.12885 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ppOpen-AT: A Directive-base Auto-tuning Language
by: Katagiri, Takahiro
Published: (2024)
by: Katagiri, Takahiro
Published: (2024)
Age of Information with Age-Dependent Server Selection
by: Akar, Nail, et al.
Published: (2025)
by: Akar, Nail, et al.
Published: (2025)
Spatially Correlated multi-RIS Communication: The Effect of Inter-Operator Interference
by: Miridakis, Nikolaos I., et al.
Published: (2025)
by: Miridakis, Nikolaos I., et al.
Published: (2025)
2D-AoI: Age-of-Information of Distributed Sensors for Spatio-Temporal Processes
by: Fidler, Markus, et al.
Published: (2024)
by: Fidler, Markus, et al.
Published: (2024)
Performance Characterization of AutoNUMA Memory Tiering on Graph Analytics
by: Moura, Diego, et al.
Published: (2022)
by: Moura, Diego, et al.
Published: (2022)
MLKAPS: Machine Learning and Adaptive Sampling for HPC Kernel Auto-tuning
by: Jam, Mathys, et al.
Published: (2025)
by: Jam, Mathys, et al.
Published: (2025)
RAO-SS: A Prototype of Run-time Auto-tuning Facility for Sparse Direct Solvers
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
Stability and Heavy-traffic Delay Optimality of General Load Balancing Policies in Heterogeneous Service Systems
by: Luo, Yishun, et al.
Published: (2025)
by: Luo, Yishun, et al.
Published: (2025)
MambaCPU: Enhanced Correlation Mining with State Space Models for CPU Performance Prediction
by: Liu, Xiaoman
Published: (2024)
by: Liu, Xiaoman
Published: (2024)
AutoKernel: Autonomous GPU Kernel Optimization via Iterative Agent-Driven Search
by: Jaber, Jaber, et al.
Published: (2026)
by: Jaber, Jaber, et al.
Published: (2026)
WaveTune: Wave-aware Bilinear Modeling for Efficient GPU Kernel Auto-tuning
by: Zhang, Kaixuan, et al.
Published: (2026)
by: Zhang, Kaixuan, et al.
Published: (2026)
An Analysis of Performance Bottlenecks in MRI Pre-Processing
by: Dugré, Mathieu, et al.
Published: (2024)
by: Dugré, Mathieu, et al.
Published: (2024)
Information Retrieval in the Age of Generative AI: The RGB Model
by: Garetto, Michele, et al.
Published: (2025)
by: Garetto, Michele, et al.
Published: (2025)
AutoSAGE: Input-Aware CUDA Scheduling for Sparse GNN Aggregation (SpMM/SDDMM) and CSR Attention
by: Stankovic, Aleksandar
Published: (2025)
by: Stankovic, Aleksandar
Published: (2025)
Towards Multi-dimensional Elasticity for Pervasive Stream Processing Services
by: Sedlak, Boris, et al.
Published: (2025)
by: Sedlak, Boris, et al.
Published: (2025)
Opal: A Modular Framework for Optimizing Performance using Analytics and LLMs
by: Zaeed, Mohammad, et al.
Published: (2025)
by: Zaeed, Mohammad, et al.
Published: (2025)
A Microbenchmark Framework for Performance Evaluation of OpenMP Target Offloading
by: Atif, Mohammad, et al.
Published: (2025)
by: Atif, Mohammad, et al.
Published: (2025)
TINA: Acceleration of Non-NN Signal Processing Algorithms Using NN Accelerators
by: Boerkamp, Christiaan, et al.
Published: (2024)
by: Boerkamp, Christiaan, et al.
Published: (2024)
Xabclib:A Fully Auto-tuned Sparse Iterative Solver
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
End-to-End Throughput Benchmarking of Portable Deterministic CNN-Based Signal Processing Pipelines
by: Boerkamp, Christiaan, et al.
Published: (2026)
by: Boerkamp, Christiaan, et al.
Published: (2026)
AutoLALA: Automatic Loop Algebraic Locality Analysis for AI and HPC Kernels
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
Characterizing Machine Learning Force Fields as Emerging Molecular Dynamics Workloads on Graphics Processing Units
by: De Alwis, Udari, et al.
Published: (2026)
by: De Alwis, Udari, et al.
Published: (2026)
A Modular Graph-Native Query Optimization Framework
by: Lyu, Bingqing, et al.
Published: (2024)
by: Lyu, Bingqing, et al.
Published: (2024)
DSO: A GPU Energy Efficiency Optimizer by Fusing Dynamic and Static Information
by: Wang, Qiang, et al.
Published: (2024)
by: Wang, Qiang, et al.
Published: (2024)
Attributing the System's Overall Effect to its Components
by: Wang, Chenxi, et al.
Published: (2026)
by: Wang, Chenxi, et al.
Published: (2026)
Bringing Auto-tuning to HIP: Analysis of Tuning Impact and Difficulty on AMD and Nvidia GPUs
by: Lurati, Milo, et al.
Published: (2024)
by: Lurati, Milo, et al.
Published: (2024)
An Auto-tuning Method for Run-time Data Transformation for Sparse Matrix-Vector Multiplication
by: Katagiri, Takahiro, et al.
Published: (2024)
by: Katagiri, Takahiro, et al.
Published: (2024)
HD-MoE: Hybrid and Dynamic Parallelism for Mixture-of-Expert LLMs with 3D Near-Memory Processing
by: Huang, Haochen, et al.
Published: (2025)
by: Huang, Haochen, et al.
Published: (2025)
Age of Information Analysis for Multi-Priority Queue and NOMA Enabled C-V2X in IoV
by: Zhang, Zheng, et al.
Published: (2024)
by: Zhang, Zheng, et al.
Published: (2024)
Accelerating Machine Learning Queries with Linear Algebra Query Processing
by: Sun, Wenbo, et al.
Published: (2023)
by: Sun, Wenbo, et al.
Published: (2023)
Systematic Performance Evaluation Framework for LEO Mega-Constellation Satellite Networks
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Tuning the Tuner: Introducing Hyperparameter Optimization for Auto-Tuning
by: Willemsen, Floris-Jan, et al.
Published: (2025)
by: Willemsen, Floris-Jan, et al.
Published: (2025)
A$^3$PIM: An Automated, Analytic and Accurate Processing-in-Memory Offloader
by: Jiang, Qingcai, et al.
Published: (2024)
by: Jiang, Qingcai, et al.
Published: (2024)
Atys: An Efficient Profiling Framework for Identifying Hotspot Functions in Large-scale Cloud Microservices
by: Sun, Jiaqi, et al.
Published: (2025)
by: Sun, Jiaqi, et al.
Published: (2025)
A relação entre a «performance» social e a «performance» económico-financeira
by: Daniel Taborda
Published: (2007)
by: Daniel Taborda
Published: (2007)
A Framework for Distributed Resource Allocation in Quantum Networks
by: Panigrahy, Nitish K., et al.
Published: (2025)
by: Panigrahy, Nitish K., et al.
Published: (2025)
DIAL: Decentralized I/O AutoTuning via Learned Client-side Local Metrics for Parallel File System
by: Rashid, Md Hasanur, et al.
Published: (2026)
by: Rashid, Md Hasanur, et al.
Published: (2026)
Age of Information in a Single-Source Generate-at-Will Dual-Server Status Update System
by: Akar, Nail, et al.
Published: (2024)
by: Akar, Nail, et al.
Published: (2024)
Scalable Cyclic Schedulers for Age of Information Optimization in Large-Scale Status Update Systems
by: Akar, Nail, et al.
Published: (2024)
by: Akar, Nail, et al.
Published: (2024)
Hardware Acceleration for Knowledge Graph Processing: Challenges & Recent Developments
by: Besta, Maciej, et al.
Published: (2024)
by: Besta, Maciej, et al.
Published: (2024)
Similar Items
-
ppOpen-AT: A Directive-base Auto-tuning Language
by: Katagiri, Takahiro
Published: (2024) -
Age of Information with Age-Dependent Server Selection
by: Akar, Nail, et al.
Published: (2025) -
Spatially Correlated multi-RIS Communication: The Effect of Inter-Operator Interference
by: Miridakis, Nikolaos I., et al.
Published: (2025) -
2D-AoI: Age-of-Information of Distributed Sensors for Spatio-Temporal Processes
by: Fidler, Markus, et al.
Published: (2024) -
Performance Characterization of AutoNUMA Memory Tiering on Graph Analytics
by: Moura, Diego, et al.
Published: (2022)