WANDER: An Explainable Decision-Support Framework for HPC
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lahiry, Ankur, Banday, Banooqa, Bhattarai, Yugesh, Islam, Tanzima Z. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
COMPASS: A Unified Decision-Intelligence System for Navigating Performance Trade-off in HPC
von: Lahiry, Ankur, et al.
Veröffentlicht: (2026)
von: Lahiry, Ankur, et al.
Veröffentlicht: (2026)
Scalable GPU Performance Variability Analysis framework
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025)
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025)
Attention-Informed Surrogates for Navigating Power-Performance Trade-offs in HPC
von: Ahmed, Ashna Nawar, et al.
Veröffentlicht: (2026)
von: Ahmed, Ashna Nawar, et al.
Veröffentlicht: (2026)
Opal: A Modular Framework for Optimizing Performance using Analytics and LLMs
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2025)
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2025)
A Distributed Framework for Causal Modeling of Performance Variability in GPU Traces
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025)
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025)
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance Optimization
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2026)
Energy Concerns with HPC Systems and Applications
von: Nana, Roblex, et al.
Veröffentlicht: (2023)
von: Nana, Roblex, et al.
Veröffentlicht: (2023)
AutoLALA: Automatic Loop Algebraic Locality Analysis for AI and HPC Kernels
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
On The Role of Prompt Construction In Enhancing Efficacy and Efficiency of LLM-Based Tabular Data Generation
von: Banday, Banooqa, et al.
Veröffentlicht: (2024)
von: Banday, Banooqa, et al.
Veröffentlicht: (2024)
Novel Representation Learning Technique using Graphs for Performance Analytics
von: Ramadan, Tarek, et al.
Veröffentlicht: (2024)
von: Ramadan, Tarek, et al.
Veröffentlicht: (2024)
FlexQuant: Elastic Quantization Framework for Locally Hosted LLM on Edge Devices
von: Chai, Yuji, et al.
Veröffentlicht: (2025)
von: Chai, Yuji, et al.
Veröffentlicht: (2025)
Data-Driven Analysis to Understand GPU Hardware Resource Usage of Optimizations
von: Islam, Tanzima Z., et al.
Veröffentlicht: (2024)
von: Islam, Tanzima Z., et al.
Veröffentlicht: (2024)
Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions
von: Teranishi, Keita, et al.
Veröffentlicht: (2025)
von: Teranishi, Keita, et al.
Veröffentlicht: (2025)
Binary Bleed: Fast Distributed and Parallel Method for Automatic Model Selection
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
von: Barron, Ryan, et al.
Veröffentlicht: (2024)
Impact of Data-Oriented and Object-Oriented Design on Performance and Cache Utilization with Artificial Intelligence Algorithms in Multi-Threaded CPUs
von: Arantes, Gabriel M., et al.
Veröffentlicht: (2025)
von: Arantes, Gabriel M., et al.
Veröffentlicht: (2025)
PerfDojo: Automated ML Library Generation for Heterogeneous Architectures
von: Ivanov, Andrei, et al.
Veröffentlicht: (2025)
von: Ivanov, Andrei, et al.
Veröffentlicht: (2025)
Edge Deployment of Small Language Models, a comprehensive comparison of CPU, GPU and NPU backends
von: Prieto, Pablo, et al.
Veröffentlicht: (2025)
von: Prieto, Pablo, et al.
Veröffentlicht: (2025)
PixLift: Accelerating Web Browsing via AI Upscaling
von: Atinafu, Yonas, et al.
Veröffentlicht: (2025)
von: Atinafu, Yonas, et al.
Veröffentlicht: (2025)
Tiny-QMoE
von: Cashman, Jack, et al.
Veröffentlicht: (2025)
von: Cashman, Jack, et al.
Veröffentlicht: (2025)
Understanding and Benchmarking Artificial Intelligence: OpenAI's o3 Is Not AGI
von: Pfister, Rolf, et al.
Veröffentlicht: (2025)
von: Pfister, Rolf, et al.
Veröffentlicht: (2025)
XTC, A Research Platform for Optimizing AI Workload Operators
von: Hugo, Pompougnac, et al.
Veröffentlicht: (2025)
von: Hugo, Pompougnac, et al.
Veröffentlicht: (2025)
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs
von: Cai, Yanan, et al.
Veröffentlicht: (2025)
von: Cai, Yanan, et al.
Veröffentlicht: (2025)
Are We Scaling the Right Thing? A System Perspective on Test-Time Scaling
von: Zhao, Youpeng, et al.
Veröffentlicht: (2025)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2025)
When Quantization Is Free: An int4 KV Cache That Outruns fp16 on Apple Silicon
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
Training Transformers in Cosine Coefficient Space
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
von: Bergach, Mohamed Amine
Veröffentlicht: (2026)
Reliability by design: quantifying and eliminating fabrication risk in LLMs. From generative to consultative AI: a comparative analysis in the legal domain and lessons for high-stakes knowledge bases
von: Dantart, Alex
Veröffentlicht: (2026)
von: Dantart, Alex
Veröffentlicht: (2026)
SweetSpot: An Analytical Model for Predicting Energy Efficiency of LLM Inference
von: Cavagna, Hiari Pizzini, et al.
Veröffentlicht: (2026)
von: Cavagna, Hiari Pizzini, et al.
Veröffentlicht: (2026)
Learning, Potential, and Retention: An Approach for Evaluating Adaptive AI-Enabled Medical Devices
von: Burgon, Alexis, et al.
Veröffentlicht: (2026)
von: Burgon, Alexis, et al.
Veröffentlicht: (2026)
Faster LLM Inference using DBMS-Inspired Preemption and Cache Replacement Policies
von: Kim, Kyoungmin, et al.
Veröffentlicht: (2024)
von: Kim, Kyoungmin, et al.
Veröffentlicht: (2024)
Improving LLM Performance Through Black-Box Online Tuning: A Case for Adding System Specs to Factsheets for Trusted AI
von: Atinafu, Yonas, et al.
Veröffentlicht: (2026)
von: Atinafu, Yonas, et al.
Veröffentlicht: (2026)
TurboSpec: Closed-loop Speculation Control System for Optimizing LLM Serving Goodput
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2024)
Personalized Model-Based Design of Human Centric AI enabled CPS for Long term usage
von: Ngabonziza, Bernard, et al.
Veröffentlicht: (2026)
von: Ngabonziza, Bernard, et al.
Veröffentlicht: (2026)
ALISE: Accelerating Large Language Model Serving with Speculative Scheduling
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
DeepContext: A Context-aware, Cross-platform, and Cross-framework Tool for Performance Profiling and Analysis of Deep Learning Workloads
von: Zhao, Qidong, et al.
Veröffentlicht: (2024)
von: Zhao, Qidong, et al.
Veröffentlicht: (2024)
Time is Not Compute: Scaling Laws for Wall-Clock Constrained Training on Consumer GPUs
von: Liu, Yi
Veröffentlicht: (2026)
von: Liu, Yi
Veröffentlicht: (2026)
Ensuring Reliability of Curated EHR-Derived Data: The Validation of Accuracy for LLM/ML-Extracted Information and Data (VALID) Framework
von: Estevez, Melissa, et al.
Veröffentlicht: (2025)
von: Estevez, Melissa, et al.
Veröffentlicht: (2025)
APOLLO: SGD-like Memory, AdamW-level Performance
von: Zhu, Hanqing, et al.
Veröffentlicht: (2024)
von: Zhu, Hanqing, et al.
Veröffentlicht: (2024)
Evaluating the Efficacy of Foundational Models: Advancing Benchmarking Practices to Enhance Fine-Tuning Decision-Making
von: Amujo, Oluyemi Enoch, et al.
Veröffentlicht: (2024)
von: Amujo, Oluyemi Enoch, et al.
Veröffentlicht: (2024)
Performance Evaluation of Bitstring Representations in a Linear Genetic Programming Framework
von: Meli, Clyde, et al.
Veröffentlicht: (2025)
von: Meli, Clyde, et al.
Veröffentlicht: (2025)
HRA: A Multi-Criteria Framework for Ranking Metaheuristic Optimization Algorithms
von: Goula, Evgenia-Maria K., et al.
Veröffentlicht: (2024)
von: Goula, Evgenia-Maria K., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
COMPASS: A Unified Decision-Intelligence System for Navigating Performance Trade-off in HPC
von: Lahiry, Ankur, et al.
Veröffentlicht: (2026) -
Scalable GPU Performance Variability Analysis framework
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025) -
Attention-Informed Surrogates for Navigating Power-Performance Trade-offs in HPC
von: Ahmed, Ashna Nawar, et al.
Veröffentlicht: (2026) -
Opal: A Modular Framework for Optimizing Performance using Analytics and LLMs
von: Zaeed, Mohammad, et al.
Veröffentlicht: (2025) -
A Distributed Framework for Causal Modeling of Performance Variability in GPU Traces
von: Lahiry, Ankur, et al.
Veröffentlicht: (2025)