Towards an Integrated Performance Framework for Fire Science and Management Workflows
Fuente:
arXiv
Salvato in:
| Autori principali: | Ahmed, H., Shende, R., Perez, I., Crawl, D., Purawat, S., Altintas, I. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
di: Werner, Elias, et al.
Pubblicazione: (2023)
di: Werner, Elias, et al.
Pubblicazione: (2023)
Prefetching Cache Optimization Using Graph Neural Networks: A Modular Framework and Conceptual Analysis
di: Qowy, F. I.
Pubblicazione: (2025)
di: Qowy, F. I.
Pubblicazione: (2025)
VDTuner: Automated Performance Tuning for Vector Data Management Systems
di: Yang, Tiannuo, et al.
Pubblicazione: (2024)
di: Yang, Tiannuo, et al.
Pubblicazione: (2024)
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
di: Duan, Shukai, et al.
Pubblicazione: (2024)
di: Duan, Shukai, et al.
Pubblicazione: (2024)
Towards Run Time Estimation of the Gaussian Chemistry Code for SEAGrid Science Gateway
di: Beltre, Angel, et al.
Pubblicazione: (2019)
di: Beltre, Angel, et al.
Pubblicazione: (2019)
An Autotuning-based Optimization Framework for Mixed-kernel SVM Classifications in Smart Pixel Datasets and Heterojunction Transistors
di: Wu, Xingfu, et al.
Pubblicazione: (2024)
di: Wu, Xingfu, et al.
Pubblicazione: (2024)
Forecasting GPU Performance for Deep Learning Training and Inference
di: Lee, Seonho, et al.
Pubblicazione: (2024)
di: Lee, Seonho, et al.
Pubblicazione: (2024)
Application Research On Real-Time Perception Of Device Performance Status
di: Wang, Zhe, et al.
Pubblicazione: (2024)
di: Wang, Zhe, et al.
Pubblicazione: (2024)
Plug-and-Play Performance Estimation for LLM Services without Relying on Labeled Data
di: Wang, Can, et al.
Pubblicazione: (2024)
di: Wang, Can, et al.
Pubblicazione: (2024)
A Kernel-Based Approach for Accurate Steady-State Detection in Performance Time Series
di: Beseda, Martin, et al.
Pubblicazione: (2025)
di: Beseda, Martin, et al.
Pubblicazione: (2025)
MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models
di: Chitty-Venkata, Krishna Teja, et al.
Pubblicazione: (2025)
di: Chitty-Venkata, Krishna Teja, et al.
Pubblicazione: (2025)
oneDNN Graph Compiler: A Hybrid Approach for High-Performance Deep Learning Compilation
di: Li, Jianhui, et al.
Pubblicazione: (2023)
di: Li, Jianhui, et al.
Pubblicazione: (2023)
DF-GNN: Dynamic Fusion Framework for Attention Graph Neural Networks on GPUs
di: Liu, Jiahui, et al.
Pubblicazione: (2024)
di: Liu, Jiahui, et al.
Pubblicazione: (2024)
SynthEval: A Framework for Detailed Utility and Privacy Evaluation of Tabular Synthetic Data
di: Lautrup, Anton Danholt, et al.
Pubblicazione: (2024)
di: Lautrup, Anton Danholt, et al.
Pubblicazione: (2024)
Towards Universal Performance Modeling for Machine Learning Training on Multi-GPU Platforms
di: Lin, Zhongyi, et al.
Pubblicazione: (2024)
di: Lin, Zhongyi, et al.
Pubblicazione: (2024)
Accelerating AI Performance using Anderson Extrapolation on GPUs
di: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Pubblicazione: (2024)
di: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Pubblicazione: (2024)
CPINN-ABPI: Physics-Informed Neural Networks for Accurate Power Estimation in MPSoCs
di: Elshamy, Mohamed R., et al.
Pubblicazione: (2025)
di: Elshamy, Mohamed R., et al.
Pubblicazione: (2025)
Towards A Flexible Accuracy-Oriented Deep Learning Module Inference Latency Prediction Framework for Adaptive Optimization Algorithms
di: Shen, Jingran, et al.
Pubblicazione: (2023)
di: Shen, Jingran, et al.
Pubblicazione: (2023)
Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask
di: Abraham, Ashley N., et al.
Pubblicazione: (2026)
di: Abraham, Ashley N., et al.
Pubblicazione: (2026)
Risk-Aware Batch Testing for Performance Regression Detection
di: Sayedsalehi, Ali, et al.
Pubblicazione: (2026)
di: Sayedsalehi, Ali, et al.
Pubblicazione: (2026)
APOLLO: SGD-like Memory, AdamW-level Performance
di: Zhu, Hanqing, et al.
Pubblicazione: (2024)
di: Zhu, Hanqing, et al.
Pubblicazione: (2024)
Performance Modeling of Data Storage Systems using Generative Models
di: Al-Maeeni, Abdalaziz Rashid, et al.
Pubblicazione: (2023)
di: Al-Maeeni, Abdalaziz Rashid, et al.
Pubblicazione: (2023)
Toward A Formalized Approach for Spike Sorting Algorithms and Hardware Evaluation
di: Zhang, Tim, et al.
Pubblicazione: (2022)
di: Zhang, Tim, et al.
Pubblicazione: (2022)
Understanding the Performance Horizon of the Latest ML Workloads with NonGEMM Workloads
di: Karami, Rachid, et al.
Pubblicazione: (2024)
di: Karami, Rachid, et al.
Pubblicazione: (2024)
Efficient Chromosome Parallelization for Precision Medicine Genomic Workflows
di: Montserrat, Daniel Mas, et al.
Pubblicazione: (2025)
di: Montserrat, Daniel Mas, et al.
Pubblicazione: (2025)
ReLATE: Learning Efficient Sparse Encoding for High-Performance Tensor Decomposition
di: Helal, Ahmed E., et al.
Pubblicazione: (2025)
di: Helal, Ahmed E., et al.
Pubblicazione: (2025)
Concorde: Fast and Accurate CPU Performance Modeling with Compositional Analytical-ML Fusion
di: Nasr-Esfahany, Arash, et al.
Pubblicazione: (2025)
di: Nasr-Esfahany, Arash, et al.
Pubblicazione: (2025)
Towards Generalized Parameter Tuning in Coherent Ising Machines: A Portfolio-Based Approach
di: Hanyu, Tatsuro, et al.
Pubblicazione: (2025)
di: Hanyu, Tatsuro, et al.
Pubblicazione: (2025)
Integration of a systolic array based hardware accelerator into a DNN operator auto-tuning framework
di: Peccia, F. N., et al.
Pubblicazione: (2022)
di: Peccia, F. N., et al.
Pubblicazione: (2022)
Improving the Serving Performance of Multi-LoRA Large Language Models via Efficient LoRA and KV Cache Management
di: Zhang, Hang, et al.
Pubblicazione: (2025)
di: Zhang, Hang, et al.
Pubblicazione: (2025)
USEFUSE: Uniform Stride for Enhanced Performance in Fused Layer Architecture of Deep Neural Networks
di: Ibrahim, Muhammad Sohail, et al.
Pubblicazione: (2024)
di: Ibrahim, Muhammad Sohail, et al.
Pubblicazione: (2024)
Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU
di: Jiang, Jevin, et al.
Pubblicazione: (2026)
di: Jiang, Jevin, et al.
Pubblicazione: (2026)
Rapid Augmentations for Time Series (RATS): A High-Performance Library for Time Series Augmentation
di: Skaf, Wadie, et al.
Pubblicazione: (2026)
di: Skaf, Wadie, et al.
Pubblicazione: (2026)
DaCe AD: Unifying High-Performance Automatic Differentiation for Machine Learning and Scientific Computing
di: Boudaoud, Afif, et al.
Pubblicazione: (2025)
di: Boudaoud, Afif, et al.
Pubblicazione: (2025)
LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
BitLogic: Training Framework for Gradient-Based FPGA-Native Neural Networks
di: Bührer, Simon, et al.
Pubblicazione: (2026)
di: Bührer, Simon, et al.
Pubblicazione: (2026)
Systematic Evaluation of Optimization Techniques for Long-Context Language Models
di: Ahmed, Ammar, et al.
Pubblicazione: (2025)
di: Ahmed, Ammar, et al.
Pubblicazione: (2025)
Quantum Neural Networks for Wind Energy Forecasting: A Comparative Study of Performance and Scalability with Classical Models
di: Hangun, Batuhan, et al.
Pubblicazione: (2025)
di: Hangun, Batuhan, et al.
Pubblicazione: (2025)
A Scalable k-Medoids Clustering via Whale Optimization Algorithm
di: Chenan, Huang, et al.
Pubblicazione: (2024)
di: Chenan, Huang, et al.
Pubblicazione: (2024)
Parallel Implementations Assessment of a Spatial-Spectral Classifier for Hyperspectral Clinical Applications
di: Lazcano, Raquel, et al.
Pubblicazione: (2024)
di: Lazcano, Raquel, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
di: Werner, Elias, et al.
Pubblicazione: (2023) -
Prefetching Cache Optimization Using Graph Neural Networks: A Modular Framework and Conceptual Analysis
di: Qowy, F. I.
Pubblicazione: (2025) -
VDTuner: Automated Performance Tuning for Vector Data Management Systems
di: Yang, Tiannuo, et al.
Pubblicazione: (2024) -
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
di: Duan, Shukai, et al.
Pubblicazione: (2024) -
Towards Run Time Estimation of the Gaussian Chemistry Code for SEAGrid Science Gateway
di: Beltre, Angel, et al.
Pubblicazione: (2019)