Preparation Meets Opportunity: Enhancing Data Preprocessing for ML Training With Seneca
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Desai, Omkar, Jiao, Ziyang, Pei, Shuyi, Bhimani, Janki, Kim, Bryan S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting
von: Sage, Manuel, et al.
Veröffentlicht: (2024)
von: Sage, Manuel, et al.
Veröffentlicht: (2024)
Sawtooth Wavefront Reordering: Enhanced CuTile FlashAttention on NVIDIA GB10
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
von: Abhyankar, Reyna, et al.
Veröffentlicht: (2025)
von: Abhyankar, Reyna, et al.
Veröffentlicht: (2025)
An Integrated Artificial Intelligence Operating System for Advanced Low-Altitude Aviation Applications
von: Tan, Minzhe, et al.
Veröffentlicht: (2024)
von: Tan, Minzhe, et al.
Veröffentlicht: (2024)
AdaptCache: KV Cache Native Storage Hierarchy for Low-Delay and High-Quality Language Model Serving
von: Feng, Shaoting, et al.
Veröffentlicht: (2025)
von: Feng, Shaoting, et al.
Veröffentlicht: (2025)
Semantic Scheduling for LLM Inference
von: Hua, Wenyue, et al.
Veröffentlicht: (2025)
von: Hua, Wenyue, et al.
Veröffentlicht: (2025)
EVICPRESS: Joint KV-Cache Compression and Eviction for Efficient LLM Serving
von: Feng, Shaoting, et al.
Veröffentlicht: (2025)
von: Feng, Shaoting, et al.
Veröffentlicht: (2025)
From Imperative to Declarative: Towards LLM-friendly OS Interfaces for Boosted Computer-Use Agents
von: Wang, Yuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuan, et al.
Veröffentlicht: (2025)
Neuralink: Fast LLM Inference on Smartphones with Neuron Co-Activation Linking
von: Wang, Tuowei, et al.
Veröffentlicht: (2024)
von: Wang, Tuowei, et al.
Veröffentlicht: (2024)
Hardware-Assisted Virtualization of Neural Processing Units for Cloud Platforms
von: Xue, Yuqi, et al.
Veröffentlicht: (2024)
von: Xue, Yuqi, et al.
Veröffentlicht: (2024)
Machine Learning (ML) library in Linux kernel
von: Dubeyko, Viacheslav
Veröffentlicht: (2026)
von: Dubeyko, Viacheslav
Veröffentlicht: (2026)
Leveraging Machine Learning for Accurate IoT Device Identification in Dynamic Wireless Contexts
von: Tushir, Bhagyashri, et al.
Veröffentlicht: (2024)
von: Tushir, Bhagyashri, et al.
Veröffentlicht: (2024)
Accelerated Training on Low-Power Edge Devices
von: Ahmed, Mohamed Aboelenien, et al.
Veröffentlicht: (2025)
von: Ahmed, Mohamed Aboelenien, et al.
Veröffentlicht: (2025)
Samoyeds: Accelerating MoE Models with Structured Sparsity Leveraging Sparse Tensor Cores
von: Wu, Chenpeng, et al.
Veröffentlicht: (2025)
von: Wu, Chenpeng, et al.
Veröffentlicht: (2025)
ConsumerBench: Benchmarking Generative AI Applications on End-User Devices
von: Gu, Yile, et al.
Veröffentlicht: (2025)
von: Gu, Yile, et al.
Veröffentlicht: (2025)
Fiddler: CPU-GPU Orchestration for Fast Inference of Mixture-of-Experts Models
von: Kamahori, Keisuke, et al.
Veröffentlicht: (2024)
von: Kamahori, Keisuke, et al.
Veröffentlicht: (2024)
Crash-Consistent Checkpointing for AI Training on macOS/APFS
von: Jeon, Juha
Veröffentlicht: (2025)
von: Jeon, Juha
Veröffentlicht: (2025)
Puzzle: Scheduling Multiple Deep Learning Models on Mobile Device with Heterogeneous Processors
von: Kang, Duseok, et al.
Veröffentlicht: (2025)
von: Kang, Duseok, et al.
Veröffentlicht: (2025)
When eBPF Meets Machine Learning: On-the-fly OS Kernel Compartmentalization
von: Wang, Zicheng, et al.
Veröffentlicht: (2024)
von: Wang, Zicheng, et al.
Veröffentlicht: (2024)
Idiosyncrasies of Programmable Caching Engines
von: Peixoto, José, et al.
Veröffentlicht: (2026)
von: Peixoto, José, et al.
Veröffentlicht: (2026)
Everything You Always Wanted to Know About Storage Compressibility of Pre-Trained ML Models but Were Afraid to Ask
von: Su, Zhaoyuan, et al.
Veröffentlicht: (2024)
von: Su, Zhaoyuan, et al.
Veröffentlicht: (2024)
Diagnosing and Resolving Cloud Platform Instability with Multi-modal RAG LLMs
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents
von: Huang, Yutong, et al.
Veröffentlicht: (2026)
von: Huang, Yutong, et al.
Veröffentlicht: (2026)
DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback
von: Dong, Yunpeng, et al.
Veröffentlicht: (2026)
von: Dong, Yunpeng, et al.
Veröffentlicht: (2026)
Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes
von: Wu, Tianyuan, et al.
Veröffentlicht: (2026)
von: Wu, Tianyuan, et al.
Veröffentlicht: (2026)
Composable OS Kernel Architectures for Autonomous Intelligence
von: Singh, Rajpreet, et al.
Veröffentlicht: (2025)
von: Singh, Rajpreet, et al.
Veröffentlicht: (2025)
FlexInfer: Breaking Memory Constraint via Flexible and Efficient Offloading for On-Device LLM Inference
von: Du, Hongchao, et al.
Veröffentlicht: (2025)
von: Du, Hongchao, et al.
Veröffentlicht: (2025)
NaSh: Guardrails for an LLM-Powered Natural Language Shell
von: Gyawali, Bimal Raj, et al.
Veröffentlicht: (2025)
von: Gyawali, Bimal Raj, et al.
Veröffentlicht: (2025)
Integrating Artificial Intelligence into Operating Systems: A Survey on Techniques, Applications, and Future Directions
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
ProphetKV: User-Query-Driven Selective Recomputation for Efficient KV Cache Reuse in Retrieval-Augmented Generation
von: Wang, Shihao, et al.
Veröffentlicht: (2026)
von: Wang, Shihao, et al.
Veröffentlicht: (2026)
AgentCgroup: Understanding and Controlling OS Resources of AI Agents
von: Zheng, Yusheng, et al.
Veröffentlicht: (2026)
von: Zheng, Yusheng, et al.
Veröffentlicht: (2026)
Skim: Speculative Execution for Fast and Efficient Web Agents
von: Wong, Mike, et al.
Veröffentlicht: (2026)
von: Wong, Mike, et al.
Veröffentlicht: (2026)
Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented Generation
von: Agarwal, Shubham, et al.
Veröffentlicht: (2025)
von: Agarwal, Shubham, et al.
Veröffentlicht: (2025)
An AI Agent Execution Environment to Safeguard User Data
von: Stanley, Robert, et al.
Veröffentlicht: (2026)
von: Stanley, Robert, et al.
Veröffentlicht: (2026)
CaMDN: Enhancing Cache Efficiency for Multi-tenant DNNs on Integrated NPUs
von: Cai, Tianhao, et al.
Veröffentlicht: (2025)
von: Cai, Tianhao, et al.
Veröffentlicht: (2025)
MaLV-OS: Rethinking the Operating System Architecture for Machine Learning in Virtualized Clouds
von: Bitchebe, Stella, et al.
Veröffentlicht: (2025)
von: Bitchebe, Stella, et al.
Veröffentlicht: (2025)
Herding LLaMaS: Using LLMs as an OS Module
von: Kamath, Aditya K, et al.
Veröffentlicht: (2024)
von: Kamath, Aditya K, et al.
Veröffentlicht: (2024)
PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU
von: Song, Yixin, et al.
Veröffentlicht: (2023)
von: Song, Yixin, et al.
Veröffentlicht: (2023)
Reinforcement Learning for Dynamic Memory Allocation
von: Lim, Arisrei, et al.
Veröffentlicht: (2024)
von: Lim, Arisrei, et al.
Veröffentlicht: (2024)
Energy-Efficient Computation with DVFS using Deep Reinforcement Learning for Multi-Task Systems in Edge Computing
von: Li, Xinyi, et al.
Veröffentlicht: (2024)
von: Li, Xinyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting
von: Sage, Manuel, et al.
Veröffentlicht: (2024) -
Sawtooth Wavefront Reordering: Enhanced CuTile FlashAttention on NVIDIA GB10
von: Zhu, Yifan, et al.
Veröffentlicht: (2026) -
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
von: Abhyankar, Reyna, et al.
Veröffentlicht: (2025) -
An Integrated Artificial Intelligence Operating System for Advanced Low-Altitude Aviation Applications
von: Tan, Minzhe, et al.
Veröffentlicht: (2024) -
AdaptCache: KV Cache Native Storage Hierarchy for Low-Delay and High-Quality Language Model Serving
von: Feng, Shaoting, et al.
Veröffentlicht: (2025)