Accelerating AI Performance using Anderson Extrapolation on GPUs
Fuente:
arXiv
Saved in:
| Main Authors: | Dajani, Saleem Abdul Fattah Ahmed Al, Keyes, David E. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
sTiles: An Accelerated Computational Framework for Sparse Factorizations of Structured Matrices
by: Fattah, Esmail Abdul, et al.
Published: (2025)
by: Fattah, Esmail Abdul, et al.
Published: (2025)
Constructing artificial life and materials scientists with accelerated AI using Deep AndersoNN
by: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Published: (2024)
by: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Published: (2024)
Online Pseudo-average Shifting Attention(PASA) for Robust Low-precision LLM Inference: Algorithms and Numerical Analysis
by: Cheng, Long, et al.
Published: (2025)
by: Cheng, Long, et al.
Published: (2025)
Pushing the Envelope of LLM Inference on AI-PC and Intel GPUs
by: Georganas, Evangelos, et al.
Published: (2025)
by: Georganas, Evangelos, et al.
Published: (2025)
Identifying Best Practice Melting Patterns in Induction Furnaces: A Data-Driven Approach Using Time Series KMeans Clustering and Multi-Criteria Decision Making
by: Howard, Daniel Anthony, et al.
Published: (2024)
by: Howard, Daniel Anthony, et al.
Published: (2024)
Physics-Informed Neural Networks for High-Frequency and Multi-Scale Problems using Transfer Learning
by: Mustajab, Abdul Hannan, et al.
Published: (2024)
by: Mustajab, Abdul Hannan, et al.
Published: (2024)
It's all about PR -- Smart Benchmarking AI Accelerators using Performance Representatives
by: Jung, Alexander Louis-Ferdinand, et al.
Published: (2024)
by: Jung, Alexander Louis-Ferdinand, et al.
Published: (2024)
Estudio de la eficiencia en la escalabilidad de GPUs para el entrenamiento de Inteligencia Artificial
by: Cortes, David, et al.
Published: (2025)
by: Cortes, David, et al.
Published: (2025)
Performance Modeling of Data Storage Systems using Generative Models
by: Al-Maeeni, Abdalaziz Rashid, et al.
Published: (2023)
by: Al-Maeeni, Abdalaziz Rashid, et al.
Published: (2023)
Accelerating Eigenvalue Dataset Generation via Chebyshev Subspace Filter
by: Wang, Hong, et al.
Published: (2025)
by: Wang, Hong, et al.
Published: (2025)
Accelerating Matrix Diagonalization through Decision Transformers with Epsilon-Greedy Optimization
by: Bhatta, Kshitij, et al.
Published: (2024)
by: Bhatta, Kshitij, et al.
Published: (2024)
Accelerating Data Generation for Neural Operators via Krylov Subspace Recycling
by: Wang, Hong, et al.
Published: (2024)
by: Wang, Hong, et al.
Published: (2024)
Private LLM Inference on Consumer Blackwell GPUs: A Practical Guide for Cost-Effective Local Deployment in SMEs
by: Knoop, Jonathan, et al.
Published: (2026)
by: Knoop, Jonathan, et al.
Published: (2026)
Profiling LoRA/QLoRA Fine-Tuning Efficiency on Consumer GPUs: An RTX 4060 Case Study
by: Avinash, MSR
Published: (2025)
by: Avinash, MSR
Published: (2025)
Characterizing and Understanding HGNN Training on GPUs
by: Han, Dengke, et al.
Published: (2024)
by: Han, Dengke, et al.
Published: (2024)
DeepContour: A Hybrid Deep Learning Framework for Accelerating Generalized Eigenvalue Problem Solving via Efficient Contour Design
by: Chen, Yeqiu, et al.
Published: (2025)
by: Chen, Yeqiu, et al.
Published: (2025)
Accelerated Gradient-based Design Optimization Via Differentiable Physics-Informed Neural Operator: A Composites Autoclave Processing Case Study
by: Patel, Janak M., et al.
Published: (2025)
by: Patel, Janak M., et al.
Published: (2025)
EXAQ: Exponent Aware Quantization For LLMs Acceleration
by: Shkolnik, Moran, et al.
Published: (2024)
by: Shkolnik, Moran, et al.
Published: (2024)
Leveraging Speculative Sampling and KV-Cache Optimizations Together for Generative AI using OpenVINO
by: Barad, Haim, et al.
Published: (2023)
by: Barad, Haim, et al.
Published: (2023)
Performance evaluation of accelerated complex multiple-precision LU decomposition
by: Kouya, Tomonori
Published: (2024)
by: Kouya, Tomonori
Published: (2024)
Operator learning without the adjoint
by: Boullé, Nicolas, et al.
Published: (2024)
by: Boullé, Nicolas, et al.
Published: (2024)
Performance evaluation of accelerated real and complex multiple-precision sparse matrix-vector multiplication
by: Kouya, Tomonori
Published: (2024)
by: Kouya, Tomonori
Published: (2024)
APOLLO: SGD-like Memory, AdamW-level Performance
by: Zhu, Hanqing, et al.
Published: (2024)
by: Zhu, Hanqing, et al.
Published: (2024)
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
by: Zhao, Youpeng, et al.
Published: (2024)
by: Zhao, Youpeng, et al.
Published: (2024)
Automatic Generation of Fast and Accurate Performance Models for Deep Neural Network Accelerators
by: Lübeck, Konstantin, et al.
Published: (2024)
by: Lübeck, Konstantin, et al.
Published: (2024)
GreedySnake: Accelerating SSD-Offloaded LLM Training with Efficient Scheduling and Optimizer Step Overlapping
by: Yin, Yishu, et al.
Published: (2025)
by: Yin, Yishu, et al.
Published: (2025)
PRISM: Distribution-free Adaptive Computation of Matrix Functions for Accelerating Neural Network Training
by: Yang, Shenghao, et al.
Published: (2026)
by: Yang, Shenghao, et al.
Published: (2026)
On the Sustainability of AI Inferences in the Edge
by: Sobhani, Ghazal, et al.
Published: (2025)
by: Sobhani, Ghazal, et al.
Published: (2025)
GPU-Accelerated Parallel Selected Inversion for Structured Matrices Using sTiles
by: Fattah, Esmail Abdul, et al.
Published: (2025)
by: Fattah, Esmail Abdul, et al.
Published: (2025)
The Race to Efficiency: A New Perspective on AI Scaling Laws
by: Lu, Chien-Ping
Published: (2025)
by: Lu, Chien-Ping
Published: (2025)
A Practical Approach to Causal Inference over Time
by: Cinquini, Martina, et al.
Published: (2024)
by: Cinquini, Martina, et al.
Published: (2024)
Sparse $L^1$-Autoencoders for Scientific Data Compression
by: Chung, Matthias, et al.
Published: (2024)
by: Chung, Matthias, et al.
Published: (2024)
Graph Neural Networks for Emulation of Finite-Element Ice Dynamics in Greenland and Antarctic Ice Sheets
by: Koo, Younghyun, et al.
Published: (2024)
by: Koo, Younghyun, et al.
Published: (2024)
Neural Operators with Localized Integral and Differential Kernels
by: Liu-Schiaffini, Miguel, et al.
Published: (2024)
by: Liu-Schiaffini, Miguel, et al.
Published: (2024)
Paired Autoencoders for Likelihood-free Estimation in Inverse Problems
by: Chung, Matthias, et al.
Published: (2024)
by: Chung, Matthias, et al.
Published: (2024)
Learning Semilinear Neural Operators : A Unified Recursive Framework For Prediction And Data Assimilation
by: Singh, Ashutosh, et al.
Published: (2024)
by: Singh, Ashutosh, et al.
Published: (2024)
Observation-specific explanations through scattered data approximation
by: Ghidini, Valentina, et al.
Published: (2024)
by: Ghidini, Valentina, et al.
Published: (2024)
Mixture of Experts Softens the Curse of Dimensionality in Operator Learning
by: Kratsios, Anastasis, et al.
Published: (2024)
by: Kratsios, Anastasis, et al.
Published: (2024)
Advancing the Understanding of Fixed Point Iterations in Deep Neural Networks: A Detailed Analytical Study
by: Ke, Yekun, et al.
Published: (2024)
by: Ke, Yekun, et al.
Published: (2024)
Low-Rank Adversarial PGD Attack
by: Savostianova, Dayana, et al.
Published: (2024)
by: Savostianova, Dayana, et al.
Published: (2024)
Similar Items
-
sTiles: An Accelerated Computational Framework for Sparse Factorizations of Structured Matrices
by: Fattah, Esmail Abdul, et al.
Published: (2025) -
Constructing artificial life and materials scientists with accelerated AI using Deep AndersoNN
by: Dajani, Saleem Abdul Fattah Ahmed Al, et al.
Published: (2024) -
Online Pseudo-average Shifting Attention(PASA) for Robust Low-precision LLM Inference: Algorithms and Numerical Analysis
by: Cheng, Long, et al.
Published: (2025) -
Pushing the Envelope of LLM Inference on AI-PC and Intel GPUs
by: Georganas, Evangelos, et al.
Published: (2025) -
Identifying Best Practice Melting Patterns in Induction Furnaces: A Data-Driven Approach Using Time Series KMeans Clustering and Multi-Criteria Decision Making
by: Howard, Daniel Anthony, et al.
Published: (2024)