Tri-Accel: Curvature-Aware Precision-Adaptive and Memory-Elastic Optimization for Efficient GPU Usage
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sheibanian, Mohsen, Shaeri, Pouya, Beigi, Alimohammad, Woo, Ryan T., Keluskar, Aryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Semi-supervised Fake News Detection using Sentiment Encoding and LSTM with Self-Attention
von: Shaeri, Pouya, et al.
Veröffentlicht: (2024)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2024)
Evaluating Adaptive Personalization of Educational Readings with Simulated Learners
von: Woo, Ryan T., et al.
Veröffentlicht: (2026)
von: Woo, Ryan T., et al.
Veröffentlicht: (2026)
MID-L: Matrix-Interpolated Dropout Layer with Layer-wise Neuron Selection
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
Explainable Human-in-the-Loop Segmentation via Critic Feedback Signals
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
Sentiment and Social Signals in the Climate Crisis: A Survey on Analyzing Social Media Responses to Extreme Weather Events
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
LOOKAT: Lookup-Optimized Key-Attention for Memory-Efficient Transformers
von: Karmore, Aryan
Veröffentlicht: (2026)
von: Karmore, Aryan
Veröffentlicht: (2026)
MNIST-Gen: A Modular MNIST-Style Dataset Generation Using Hierarchical Semantics, Reinforcement Learning, and Category Theory
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
LLMem: Estimating GPU Memory Usage for Fine-Tuning Pre-Trained LLMs
von: Kim, Taeho, et al.
Veröffentlicht: (2024)
von: Kim, Taeho, et al.
Veröffentlicht: (2024)
CurvZO: Adaptive Curvature-Guided Sparse Zeroth-Order Optimization for Efficient LLM Fine-Tuning
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
Llamas on the Web: Memory-Efficient, Performance-Portable, and Multi-Precision LLM Inference with WebGPU
von: Levine, Reese, et al.
Veröffentlicht: (2026)
von: Levine, Reese, et al.
Veröffentlicht: (2026)
Memory-Efficient Sequential Pattern Mining with Hybrid Tries
von: Hosseininasab, Amin, et al.
Veröffentlicht: (2022)
von: Hosseininasab, Amin, et al.
Veröffentlicht: (2022)
Sample-Efficient Bayesian Optimization with Transfer Learning for Heterogeneous Search Spaces
von: Deshwal, Aryan, et al.
Veröffentlicht: (2024)
von: Deshwal, Aryan, et al.
Veröffentlicht: (2024)
CoMERA: Computing- and Memory-Efficient Training via Rank-Adaptive Tensor Optimization
von: Yang, Zi, et al.
Veröffentlicht: (2024)
von: Yang, Zi, et al.
Veröffentlicht: (2024)
Efficient Arbitrary Precision Acceleration for Large Language Models on GPU Tensor Cores
von: Ma, Shaobo, et al.
Veröffentlicht: (2024)
von: Ma, Shaobo, et al.
Veröffentlicht: (2024)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
COM-BOM: Bayesian Exemplar Search for Efficiently Exploring the Accuracy-Calibration Pareto Frontier
von: Luo, Gaoxiang, et al.
Veröffentlicht: (2025)
von: Luo, Gaoxiang, et al.
Veröffentlicht: (2025)
CurvaDion: Curvature-Adaptive Distributed Orthonormalization
von: Kumar, Bhavesh, et al.
Veröffentlicht: (2025)
von: Kumar, Bhavesh, et al.
Veröffentlicht: (2025)
CUROCKET: Optimizing ROCKET for GPU
von: Stüven, Ole, et al.
Veröffentlicht: (2026)
von: Stüven, Ole, et al.
Veröffentlicht: (2026)
GPU Memory Requirement Prediction for Deep Learning Task Based on Bidirectional Gated Recurrent Unit Optimization Transformer
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
FlashOptim: Optimizers for Memory-Efficient Training
von: Ortiz, Jose Javier Gonzalez, et al.
Veröffentlicht: (2026)
von: Ortiz, Jose Javier Gonzalez, et al.
Veröffentlicht: (2026)
Dynamic Memory Based Adaptive Optimization
von: Szegedy, Balázs, et al.
Veröffentlicht: (2024)
von: Szegedy, Balázs, et al.
Veröffentlicht: (2024)
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
RAMP: Reinforcement Adaptive Mixed Precision Quantization for Efficient On Device LLM Inference
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
von: Younesian, Sharareh, et al.
Veröffentlicht: (2026)
von: Younesian, Sharareh, et al.
Veröffentlicht: (2026)
Towards Optimizing the Costs of LLM Usage
von: Shekhar, Shivanshu, et al.
Veröffentlicht: (2024)
von: Shekhar, Shivanshu, et al.
Veröffentlicht: (2024)
Bayesian Optimization for Function-Valued Responses under Min-Max Criteria
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
Geometry-Aware Backdoor Attacks: Leveraging Curvature in Hyperbolic Embeddings
von: Baheri, Ali
Veröffentlicht: (2025)
von: Baheri, Ali
Veröffentlicht: (2025)
ButterflyMoE: Sub-Linear Ternary Experts via Structured Butterfly Orbits
von: Karmore, Aryan
Veröffentlicht: (2026)
von: Karmore, Aryan
Veröffentlicht: (2026)
Tri-MTL: A Triple Multitask Learning Approach for Respiratory Disease Diagnosis
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
SMMF: Square-Matricized Momentum Factorization for Memory-Efficient Optimization
von: Park, Kwangryeol, et al.
Veröffentlicht: (2024)
von: Park, Kwangryeol, et al.
Veröffentlicht: (2024)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
von: Mahdavinia, Pouria, et al.
Veröffentlicht: (2025)
von: Mahdavinia, Pouria, et al.
Veröffentlicht: (2025)
Curvature-Aware Optimization for High-Accuracy Physics-Informed Neural Networks
von: Jnini, Anas, et al.
Veröffentlicht: (2026)
von: Jnini, Anas, et al.
Veröffentlicht: (2026)
CA-HFP: Curvature-Aware Heterogeneous Federated Pruning with Model Reconstruction
von: Hu, Gang, et al.
Veröffentlicht: (2026)
von: Hu, Gang, et al.
Veröffentlicht: (2026)
Carnatic Raga Identification System using Rigorous Time-Delay Neural Network
von: Natesan, Sanjay, et al.
Veröffentlicht: (2024)
von: Natesan, Sanjay, et al.
Veröffentlicht: (2024)
Efficient Skill Discovery via Regret-Aware Optimization
von: Zhang, He, et al.
Veröffentlicht: (2025)
von: Zhang, He, et al.
Veröffentlicht: (2025)
Online GPU Energy Optimization with Switching-Aware Bandits
von: Xu, Xiongxiao, et al.
Veröffentlicht: (2024)
von: Xu, Xiongxiao, et al.
Veröffentlicht: (2024)
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
von: Beigi, Alimohammad, et al.
Veröffentlicht: (2024)
von: Beigi, Alimohammad, et al.
Veröffentlicht: (2024)
MISA: Memory-Efficient LLMs Optimization with Module-wise Importance Sampling
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
Memory-Efficient Gradient Unrolling for Large-Scale Bi-level Optimization
von: Shen, Qianli, et al.
Veröffentlicht: (2024)
von: Shen, Qianli, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Semi-supervised Fake News Detection using Sentiment Encoding and LSTM with Self-Attention
von: Shaeri, Pouya, et al.
Veröffentlicht: (2024) -
Evaluating Adaptive Personalization of Educational Readings with Simulated Learners
von: Woo, Ryan T., et al.
Veröffentlicht: (2026) -
MID-L: Matrix-Interpolated Dropout Layer with Layer-wise Neuron Selection
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025) -
Explainable Human-in-the-Loop Segmentation via Critic Feedback Signals
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025) -
Sentiment and Social Signals in the Climate Crisis: A Survey on Analyzing Social Media Responses to Extreme Weather Events
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)