Salvato in:
| Autori principali: | Desai, Omkar, Jiao, Ziyang, Pei, Shuyi, Bhimani, Janki, Kim, Bryan S. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2511.13724 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting
di: Sage, Manuel, et al.
Pubblicazione: (2024)
di: Sage, Manuel, et al.
Pubblicazione: (2024)
Sawtooth Wavefront Reordering: Enhanced CuTile FlashAttention on NVIDIA GB10
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
di: Abhyankar, Reyna, et al.
Pubblicazione: (2025)
di: Abhyankar, Reyna, et al.
Pubblicazione: (2025)
An Integrated Artificial Intelligence Operating System for Advanced Low-Altitude Aviation Applications
di: Tan, Minzhe, et al.
Pubblicazione: (2024)
di: Tan, Minzhe, et al.
Pubblicazione: (2024)
AdaptCache: KV Cache Native Storage Hierarchy for Low-Delay and High-Quality Language Model Serving
di: Feng, Shaoting, et al.
Pubblicazione: (2025)
di: Feng, Shaoting, et al.
Pubblicazione: (2025)
Semantic Scheduling for LLM Inference
di: Hua, Wenyue, et al.
Pubblicazione: (2025)
di: Hua, Wenyue, et al.
Pubblicazione: (2025)
EVICPRESS: Joint KV-Cache Compression and Eviction for Efficient LLM Serving
di: Feng, Shaoting, et al.
Pubblicazione: (2025)
di: Feng, Shaoting, et al.
Pubblicazione: (2025)
From Imperative to Declarative: Towards LLM-friendly OS Interfaces for Boosted Computer-Use Agents
di: Wang, Yuan, et al.
Pubblicazione: (2025)
di: Wang, Yuan, et al.
Pubblicazione: (2025)
Neuralink: Fast LLM Inference on Smartphones with Neuron Co-Activation Linking
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
di: Wang, Tuowei, et al.
Pubblicazione: (2024)
Hardware-Assisted Virtualization of Neural Processing Units for Cloud Platforms
di: Xue, Yuqi, et al.
Pubblicazione: (2024)
di: Xue, Yuqi, et al.
Pubblicazione: (2024)
Leveraging Machine Learning for Accurate IoT Device Identification in Dynamic Wireless Contexts
di: Tushir, Bhagyashri, et al.
Pubblicazione: (2024)
di: Tushir, Bhagyashri, et al.
Pubblicazione: (2024)
Machine Learning (ML) library in Linux kernel
di: Dubeyko, Viacheslav
Pubblicazione: (2026)
di: Dubeyko, Viacheslav
Pubblicazione: (2026)
Samoyeds: Accelerating MoE Models with Structured Sparsity Leveraging Sparse Tensor Cores
di: Wu, Chenpeng, et al.
Pubblicazione: (2025)
di: Wu, Chenpeng, et al.
Pubblicazione: (2025)
ConsumerBench: Benchmarking Generative AI Applications on End-User Devices
di: Gu, Yile, et al.
Pubblicazione: (2025)
di: Gu, Yile, et al.
Pubblicazione: (2025)
Fiddler: CPU-GPU Orchestration for Fast Inference of Mixture-of-Experts Models
di: Kamahori, Keisuke, et al.
Pubblicazione: (2024)
di: Kamahori, Keisuke, et al.
Pubblicazione: (2024)
Idiosyncrasies of Programmable Caching Engines
di: Peixoto, José, et al.
Pubblicazione: (2026)
di: Peixoto, José, et al.
Pubblicazione: (2026)
Accelerated Training on Low-Power Edge Devices
di: Ahmed, Mohamed Aboelenien, et al.
Pubblicazione: (2025)
di: Ahmed, Mohamed Aboelenien, et al.
Pubblicazione: (2025)
Crash-Consistent Checkpointing for AI Training on macOS/APFS
di: Jeon, Juha
Pubblicazione: (2025)
di: Jeon, Juha
Pubblicazione: (2025)
Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented Generation
di: Agarwal, Shubham, et al.
Pubblicazione: (2025)
di: Agarwal, Shubham, et al.
Pubblicazione: (2025)
Everything You Always Wanted to Know About Storage Compressibility of Pre-Trained ML Models but Were Afraid to Ask
di: Su, Zhaoyuan, et al.
Pubblicazione: (2024)
di: Su, Zhaoyuan, et al.
Pubblicazione: (2024)
When eBPF Meets Machine Learning: On-the-fly OS Kernel Compartmentalization
di: Wang, Zicheng, et al.
Pubblicazione: (2024)
di: Wang, Zicheng, et al.
Pubblicazione: (2024)
Puzzle: Scheduling Multiple Deep Learning Models on Mobile Device with Heterogeneous Processors
di: Kang, Duseok, et al.
Pubblicazione: (2025)
di: Kang, Duseok, et al.
Pubblicazione: (2025)
A unified framework for detecting point and collective anomalies in operating system logs via collaborative transformers
di: Nasirzadeh, Mohammad, et al.
Pubblicazione: (2025)
di: Nasirzadeh, Mohammad, et al.
Pubblicazione: (2025)
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents
di: Huang, Yutong, et al.
Pubblicazione: (2026)
di: Huang, Yutong, et al.
Pubblicazione: (2026)
DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback
di: Dong, Yunpeng, et al.
Pubblicazione: (2026)
di: Dong, Yunpeng, et al.
Pubblicazione: (2026)
Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes
di: Wu, Tianyuan, et al.
Pubblicazione: (2026)
di: Wu, Tianyuan, et al.
Pubblicazione: (2026)
Composable OS Kernel Architectures for Autonomous Intelligence
di: Singh, Rajpreet, et al.
Pubblicazione: (2025)
di: Singh, Rajpreet, et al.
Pubblicazione: (2025)
FlexInfer: Breaking Memory Constraint via Flexible and Efficient Offloading for On-Device LLM Inference
di: Du, Hongchao, et al.
Pubblicazione: (2025)
di: Du, Hongchao, et al.
Pubblicazione: (2025)
NaSh: Guardrails for an LLM-Powered Natural Language Shell
di: Gyawali, Bimal Raj, et al.
Pubblicazione: (2025)
di: Gyawali, Bimal Raj, et al.
Pubblicazione: (2025)
Integrating Artificial Intelligence into Operating Systems: A Survey on Techniques, Applications, and Future Directions
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
ProphetKV: User-Query-Driven Selective Recomputation for Efficient KV Cache Reuse in Retrieval-Augmented Generation
di: Wang, Shihao, et al.
Pubblicazione: (2026)
di: Wang, Shihao, et al.
Pubblicazione: (2026)
AgentCgroup: Understanding and Controlling OS Resources of AI Agents
di: Zheng, Yusheng, et al.
Pubblicazione: (2026)
di: Zheng, Yusheng, et al.
Pubblicazione: (2026)
Skim: Speculative Execution for Fast and Efficient Web Agents
di: Wong, Mike, et al.
Pubblicazione: (2026)
di: Wong, Mike, et al.
Pubblicazione: (2026)
An AI Agent Execution Environment to Safeguard User Data
di: Stanley, Robert, et al.
Pubblicazione: (2026)
di: Stanley, Robert, et al.
Pubblicazione: (2026)
CaMDN: Enhancing Cache Efficiency for Multi-tenant DNNs on Integrated NPUs
di: Cai, Tianhao, et al.
Pubblicazione: (2025)
di: Cai, Tianhao, et al.
Pubblicazione: (2025)
A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers
di: Saleh, Mohammad AlShaikh, et al.
Pubblicazione: (2026)
di: Saleh, Mohammad AlShaikh, et al.
Pubblicazione: (2026)
Towards Agentic OS: An LLM Agent Framework for Linux Schedulers
di: Zheng, Yusheng, et al.
Pubblicazione: (2025)
di: Zheng, Yusheng, et al.
Pubblicazione: (2025)
Dynamic Adaptation in Data Storage: Real-Time Machine Learning for Enhanced Prefetching
di: Cheng, Chiyu, et al.
Pubblicazione: (2024)
di: Cheng, Chiyu, et al.
Pubblicazione: (2024)
MaLV-OS: Rethinking the Operating System Architecture for Machine Learning in Virtualized Clouds
di: Bitchebe, Stella, et al.
Pubblicazione: (2025)
di: Bitchebe, Stella, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting
di: Sage, Manuel, et al.
Pubblicazione: (2024) -
Sawtooth Wavefront Reordering: Enhanced CuTile FlashAttention on NVIDIA GB10
di: Zhu, Yifan, et al.
Pubblicazione: (2026) -
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
di: Abhyankar, Reyna, et al.
Pubblicazione: (2025) -
An Integrated Artificial Intelligence Operating System for Advanced Low-Altitude Aviation Applications
di: Tan, Minzhe, et al.
Pubblicazione: (2024) -
AdaptCache: KV Cache Native Storage Hierarchy for Low-Delay and High-Quality Language Model Serving
di: Feng, Shaoting, et al.
Pubblicazione: (2025)