Performance Characterization and Optimizations of Traditional ML Applications
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kumar, Harsh, Govindarajan, R. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Profiling Apple Silicon Performance for ML Training
von: Feng, Dahua, et al.
Veröffentlicht: (2025)
von: Feng, Dahua, et al.
Veröffentlicht: (2025)
Performance Characterization of Containers in Edge Computing
von: Gupta, Ragini, et al.
Veröffentlicht: (2025)
von: Gupta, Ragini, et al.
Veröffentlicht: (2025)
Characterize LSM-tree Compaction Performance via On-Device LLM Inference
von: Ding, Jiabiao, et al.
Veröffentlicht: (2026)
von: Ding, Jiabiao, et al.
Veröffentlicht: (2026)
The Next 700 ML-Enabled Compiler Optimizations
von: VenkataKeerthy, S., et al.
Veröffentlicht: (2023)
von: VenkataKeerthy, S., et al.
Veröffentlicht: (2023)
Characterization of Photovoltaic Performance through Current - Voltage Analysis
von: Priyalatha Alexander, et al.
Veröffentlicht: (2026)
von: Priyalatha Alexander, et al.
Veröffentlicht: (2026)
Two Criteria for Performance Analysis of Optimization Algorithms
von: Jing, Yunpeng, et al.
Veröffentlicht: (2024)
von: Jing, Yunpeng, et al.
Veröffentlicht: (2024)
Resource Allocation Influence on Application Performance in Sliced Testbeds
von: Moreira, Rodrigo, et al.
Veröffentlicht: (2024)
von: Moreira, Rodrigo, et al.
Veröffentlicht: (2024)
A Continuous Benchmarking Infrastructure for High-Performance Computing Applications
von: Alt, Christoph, et al.
Veröffentlicht: (2024)
von: Alt, Christoph, et al.
Veröffentlicht: (2024)
Scalable Packed Layouts for Vector-Length-Agnostic ML Code Generation
von: Beysel, Ege, et al.
Veröffentlicht: (2026)
von: Beysel, Ege, et al.
Veröffentlicht: (2026)
Accurate Performance Modeling And Uncertainty Analysis of Lossy Compression in Scientific Applications
von: Liu, Youyuan, et al.
Veröffentlicht: (2024)
von: Liu, Youyuan, et al.
Veröffentlicht: (2024)
Opal: A Modular Framework for Optimizing Performance using Analytics and LLMs
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2025)
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2025)
AI Application Benchmarking: Power-Aware Performance Analysis for Vision and Language Models
von: Mayr, Martin, et al.
Veröffentlicht: (2026)
von: Mayr, Martin, et al.
Veröffentlicht: (2026)
Understanding the Performance Horizon of the Latest ML Workloads with NonGEMM Workloads
von: Karami, Rachid, et al.
Veröffentlicht: (2024)
von: Karami, Rachid, et al.
Veröffentlicht: (2024)
Ecoscape: Fault Tolerance Benchmark for Adaptive Remediation Strategies in Real-Time Edge ML
von: Reiter, Hendrik, et al.
Veröffentlicht: (2025)
von: Reiter, Hendrik, et al.
Veröffentlicht: (2025)
Performance Optimization of 3D Stencil Computation on ARM Scalable Vector Extension
von: Chen, Hongguang
Veröffentlicht: (2025)
von: Chen, Hongguang
Veröffentlicht: (2025)
Performance Characterization of AutoNUMA Memory Tiering on Graph Analytics
von: Moura, Diego, et al.
Veröffentlicht: (2022)
von: Moura, Diego, et al.
Veröffentlicht: (2022)
OSCAR-P and aMLLibrary: Profiling and Predicting the Performance of FaaS-based Applications in Computing Continua
von: Sala, Roberto, et al.
Veröffentlicht: (2024)
von: Sala, Roberto, et al.
Veröffentlicht: (2024)
Concorde: Fast and Accurate CPU Performance Modeling with Compositional Analytical-ML Fusion
von: Nasr-Esfahany, Arash, et al.
Veröffentlicht: (2025)
von: Nasr-Esfahany, Arash, et al.
Veröffentlicht: (2025)
A Data-driven ML Approach for Maximizing Performance in LLM-Adapter Serving
von: Agullo, Ferran, et al.
Veröffentlicht: (2025)
von: Agullo, Ferran, et al.
Veröffentlicht: (2025)
Rethinking Temporal Models for TinyML: LSTM versus 1D-CNN in Resource-Constrained Devices
von: Saha, Bidyut, et al.
Veröffentlicht: (2026)
von: Saha, Bidyut, et al.
Veröffentlicht: (2026)
Characterizing and Optimizing Realistic Workloads on a Commercial Compute-in-SRAM Device
von: Zhang, Niansong, et al.
Veröffentlicht: (2025)
von: Zhang, Niansong, et al.
Veröffentlicht: (2025)
Tracing Optimization for Performance Modeling and Regression Detection
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
Overview of Web Application Performance Optimization Techniques
von: Vepsäläinen, Juho, et al.
Veröffentlicht: (2024)
von: Vepsäläinen, Juho, et al.
Veröffentlicht: (2024)
Performance of Confidential Computing GPUs
von: Ibarra, Antonio Martínez, et al.
Veröffentlicht: (2025)
von: Ibarra, Antonio Martínez, et al.
Veröffentlicht: (2025)
Scheduling in Quantum Satellite Networks: Fairness and Performance Optimization
von: Dikshit, Ashutosh Jayant, et al.
Veröffentlicht: (2025)
von: Dikshit, Ashutosh Jayant, et al.
Veröffentlicht: (2025)
Evaluating Compiler Optimization Impacts on zkVM Performance
von: Gassmann, Thomas, et al.
Veröffentlicht: (2025)
von: Gassmann, Thomas, et al.
Veröffentlicht: (2025)
PerfDojo: Automated ML Library Generation for Heterogeneous Architectures
von: Ivanov, Andrei, et al.
Veröffentlicht: (2025)
von: Ivanov, Andrei, et al.
Veröffentlicht: (2025)
Accelerating Sparse Ternary GEMM for Quantized ML on Apple Silicon
von: Lipshitz, Baraq, et al.
Veröffentlicht: (2025)
von: Lipshitz, Baraq, et al.
Veröffentlicht: (2025)
Application Research On Real-Time Perception Of Device Performance Status
von: Wang, Zhe, et al.
Veröffentlicht: (2024)
von: Wang, Zhe, et al.
Veröffentlicht: (2024)
ETM2: Empowering Traditional Memory Bandwidth Regulation using ETM
von: Zuepke, Alexander, et al.
Veröffentlicht: (2026)
von: Zuepke, Alexander, et al.
Veröffentlicht: (2026)
gigiProfiler: Diagnosing Performance Issues by Uncovering Application Resource Bottlenecks
von: Hu, Yigong, et al.
Veröffentlicht: (2025)
von: Hu, Yigong, et al.
Veröffentlicht: (2025)
Performance Characterization of Expert Router for Scalable LLM Inference
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
LoPace: A Lossless Optimized Prompt Accurate Compression Engine for Large Language Model Applications
von: Ulla, Aman
Veröffentlicht: (2026)
von: Ulla, Aman
Veröffentlicht: (2026)
Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders
von: Iglovikov, Vladimir, et al.
Veröffentlicht: (2026)
von: Iglovikov, Vladimir, et al.
Veröffentlicht: (2026)
Hardware-efficient tractable probabilistic inference for TinyML Neurosymbolic AI applications
von: Leslin, Jelin, et al.
Veröffentlicht: (2025)
von: Leslin, Jelin, et al.
Veröffentlicht: (2025)
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance Optimization
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
From Profiling to Optimization: Unveiling the Profile Guided Optimization
von: Liu, Bingxin, et al.
Veröffentlicht: (2025)
von: Liu, Bingxin, et al.
Veröffentlicht: (2025)
ADS Performance Revisited
von: Weber, Alexander, et al.
Veröffentlicht: (2024)
von: Weber, Alexander, et al.
Veröffentlicht: (2024)
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
von: Werner, Elias, et al.
Veröffentlicht: (2023)
von: Werner, Elias, et al.
Veröffentlicht: (2023)
U-TOE: Universal TinyML On-board Evaluation Toolkit for Low-Power IoT
von: Huang, Zhaolan, et al.
Veröffentlicht: (2023)
von: Huang, Zhaolan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Profiling Apple Silicon Performance for ML Training
von: Feng, Dahua, et al.
Veröffentlicht: (2025) -
Performance Characterization of Containers in Edge Computing
von: Gupta, Ragini, et al.
Veröffentlicht: (2025) -
Characterize LSM-tree Compaction Performance via On-Device LLM Inference
von: Ding, Jiabiao, et al.
Veröffentlicht: (2026) -
The Next 700 ML-Enabled Compiler Optimizations
von: VenkataKeerthy, S., et al.
Veröffentlicht: (2023) -
Characterization of Photovoltaic Performance through Current - Voltage Analysis
von: Priyalatha Alexander, et al.
Veröffentlicht: (2026)