Guardado en:
| Autores principales: | Iqbal, Zain, Valerio, Lorenzo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2601.05205 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GreenServ: Energy-Efficient Context-Aware Dynamic Routing for Multi-Model LLM Inference
por: Ziller, Thomas, et al.
Publicado: (2026)
por: Ziller, Thomas, et al.
Publicado: (2026)
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
por: Xue, Leyang, et al.
Publicado: (2024)
por: Xue, Leyang, et al.
Publicado: (2024)
Energy-Aware LLMs: A step towards sustainable AI for downstream applications
por: Tran, Nguyen Phuc, et al.
Publicado: (2025)
por: Tran, Nguyen Phuc, et al.
Publicado: (2025)
Cloud Computing Energy Consumption Prediction Based on Kernel Extreme Learning Machine Algorithm Improved by Vector Weighted Average Algorithm
por: Wang, Yuqing, et al.
Publicado: (2025)
por: Wang, Yuqing, et al.
Publicado: (2025)
Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems
por: Panigrahy, Deepak, et al.
Publicado: (2026)
por: Panigrahy, Deepak, et al.
Publicado: (2026)
KForge: Program Synthesis for Diverse AI Hardware Accelerators
por: Sereda, Taras, et al.
Publicado: (2025)
por: Sereda, Taras, et al.
Publicado: (2025)
Energy-Efficient Transformer Inference: Optimization Strategies for Time Series Classification
por: Kermani, Arshia, et al.
Publicado: (2025)
por: Kermani, Arshia, et al.
Publicado: (2025)
ALERT: Accurate Learning for Energy and Timeliness
por: Wan, Chengcheng, et al.
Publicado: (2019)
por: Wan, Chengcheng, et al.
Publicado: (2019)
Enhancing Energy-Awareness in Deep Learning through Fine-Grained Energy Measurement
por: Rajput, Saurabhsingh, et al.
Publicado: (2023)
por: Rajput, Saurabhsingh, et al.
Publicado: (2023)
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
por: Duan, Shukai, et al.
Publicado: (2024)
por: Duan, Shukai, et al.
Publicado: (2024)
Hardware optimization on Android for inference of AI models
por: Gherasim, Iulius, et al.
Publicado: (2025)
por: Gherasim, Iulius, et al.
Publicado: (2025)
PrETi: Predicting Execution Time in Early Stage with LLVM and Machine Learning
por: Xu, Risheng, et al.
Publicado: (2025)
por: Xu, Risheng, et al.
Publicado: (2025)
One Size Does Not Fit All: Architecture-Aware Adaptive Batch Scheduling with DEBA
por: Belias, François, et al.
Publicado: (2025)
por: Belias, François, et al.
Publicado: (2025)
Knowledge Grafting: A Mechanism for Optimizing AI Model Deployment in Resource-Constrained Environments
por: Almurshed, Osama, et al.
Publicado: (2025)
por: Almurshed, Osama, et al.
Publicado: (2025)
Greener Deep Reinforcement Learning: Analysis of Energy and Carbon Efficiency Across Atari Benchmarks
por: Gardner, Jason, et al.
Publicado: (2025)
por: Gardner, Jason, et al.
Publicado: (2025)
WCDT: Systematic WCET Optimization for Decision Tree Implementations
por: Hölscher, Nils, et al.
Publicado: (2025)
por: Hölscher, Nils, et al.
Publicado: (2025)
Leveraging Speculative Sampling and KV-Cache Optimizations Together for Generative AI using OpenVINO
por: Barad, Haim, et al.
Publicado: (2023)
por: Barad, Haim, et al.
Publicado: (2023)
Automating Energy-Efficient GPU Kernel Generation: A Fast Search-Based Compilation Approach
por: Zhang, Yijia, et al.
Publicado: (2024)
por: Zhang, Yijia, et al.
Publicado: (2024)
LiveTune: Dynamic Parameter Tuning for Feedback-Driven Optimization
por: Shabgahi, Soheil Zibakhsh, et al.
Publicado: (2023)
por: Shabgahi, Soheil Zibakhsh, et al.
Publicado: (2023)
AutoSAGE: Input-Aware CUDA Scheduling for Sparse GNN Aggregation (SpMM/SDDMM) and CSR Attention
por: Stankovic, Aleksandar
Publicado: (2025)
por: Stankovic, Aleksandar
Publicado: (2025)
A Scalable k-Medoids Clustering via Whale Optimization Algorithm
por: Chenan, Huang, et al.
Publicado: (2024)
por: Chenan, Huang, et al.
Publicado: (2024)
Hardware-efficient tractable probabilistic inference for TinyML Neurosymbolic AI applications
por: Leslin, Jelin, et al.
Publicado: (2025)
por: Leslin, Jelin, et al.
Publicado: (2025)
A Kernel-Based Approach for Accurate Steady-State Detection in Performance Time Series
por: Beseda, Martin, et al.
Publicado: (2025)
por: Beseda, Martin, et al.
Publicado: (2025)
Feature Optimization for Time Series Forecasting via Novel Randomized Uphill Climbing
por: Van Thanh, Nguyen
Publicado: (2025)
por: Van Thanh, Nguyen
Publicado: (2025)
AutoKernel: Autonomous GPU Kernel Optimization via Iterative Agent-Driven Search
por: Jaber, Jaber, et al.
Publicado: (2026)
por: Jaber, Jaber, et al.
Publicado: (2026)
Accuracy and Consumption analysis from a compressed model by CompactifAI from Multiverse Computing
por: Fovet, Damien, et al.
Publicado: (2025)
por: Fovet, Damien, et al.
Publicado: (2025)
cedar: Optimized and Unified Machine Learning Input Data Pipelines
por: Zhao, Mark, et al.
Publicado: (2024)
por: Zhao, Mark, et al.
Publicado: (2024)
Throughput Optimization as a Strategic Lever in Large-Scale AI Systems: Evidence from Dataloader and Memory Profiling Innovations
por: Jha, Mayank
Publicado: (2026)
por: Jha, Mayank
Publicado: (2026)
An Autotuning-based Optimization Framework for Mixed-kernel SVM Classifications in Smart Pixel Datasets and Heterojunction Transistors
por: Wu, Xingfu, et al.
Publicado: (2024)
por: Wu, Xingfu, et al.
Publicado: (2024)
Machine Learning Models for Reinforced Concrete Pipes Condition Prediction: The State-of-the-Art Using Artificial Neural Networks and Multiple Linear Regression in a Wisconsin Case Study
por: Mohammadagha, Mohsen, et al.
Publicado: (2025)
por: Mohammadagha, Mohsen, et al.
Publicado: (2025)
Who Wins the Race? (R Vs Python) - An Exploratory Study on Energy Consumption of Machine Learning Algorithms
por: Chattaraj, Rajrupa, et al.
Publicado: (2025)
por: Chattaraj, Rajrupa, et al.
Publicado: (2025)
EXAQ: Exponent Aware Quantization For LLMs Acceleration
por: Shkolnik, Moran, et al.
Publicado: (2024)
por: Shkolnik, Moran, et al.
Publicado: (2024)
Risk-Aware Batch Testing for Performance Regression Detection
por: Sayedsalehi, Ali, et al.
Publicado: (2026)
por: Sayedsalehi, Ali, et al.
Publicado: (2026)
On the Sustainability of AI Inferences in the Edge
por: Sobhani, Ghazal, et al.
Publicado: (2025)
por: Sobhani, Ghazal, et al.
Publicado: (2025)
A2Q+: Improving Accumulator-Aware Weight Quantization
por: Colbert, Ian, et al.
Publicado: (2024)
por: Colbert, Ian, et al.
Publicado: (2024)
ASPO: Constraint-Aware Bayesian Optimization for FPGA-based Soft Processors
por: Wu, Haoran, et al.
Publicado: (2025)
por: Wu, Haoran, et al.
Publicado: (2025)
Machine Learning Methods for Evaluating Public Crisis: Meta-Analysis
por: Okpala, Izunna, et al.
Publicado: (2023)
por: Okpala, Izunna, et al.
Publicado: (2023)
Adaptive Workload Distribution for Accuracy-aware DNN Inference on Collaborative Edge Platforms
por: Taufique, Zain, et al.
Publicado: (2023)
por: Taufique, Zain, et al.
Publicado: (2023)
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
por: Zhao, Youpeng, et al.
Publicado: (2024)
por: Zhao, Youpeng, et al.
Publicado: (2024)
Offline Reinforcement-Learning-Based Power Control for Application-Agnostic Energy Efficiency
por: Raj, Akhilesh, et al.
Publicado: (2026)
por: Raj, Akhilesh, et al.
Publicado: (2026)
Ejemplares similares
-
GreenServ: Energy-Efficient Context-Aware Dynamic Routing for Multi-Model LLM Inference
por: Ziller, Thomas, et al.
Publicado: (2026) -
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
por: Xue, Leyang, et al.
Publicado: (2024) -
Energy-Aware LLMs: A step towards sustainable AI for downstream applications
por: Tran, Nguyen Phuc, et al.
Publicado: (2025) -
Cloud Computing Energy Consumption Prediction Based on Kernel Extreme Learning Machine Algorithm Improved by Vector Weighted Average Algorithm
por: Wang, Yuqing, et al.
Publicado: (2025) -
Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems
por: Panigrahy, Deepak, et al.
Publicado: (2026)