Salient Store: Enabling Smart Storage for Continuous Learning Edge Servers
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mishra, Cyan Subhra, Chaudhary, Deeksha, Sampson, Jack, Knademir, Mahmut Taylan, Das, Chita |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Synergistic and Efficient Edge-Host Communication for Energy Harvesting Wireless Sensor Networks
par: Mishra, Cyan Subhra, et autres
Publié: (2024)
par: Mishra, Cyan Subhra, et autres
Publié: (2024)
Revisiting DNN Training for Intermittently-Powered Energy-Harvesting Micro-Computers
par: Mishra, Cyan Subhra, et autres
Publié: (2024)
par: Mishra, Cyan Subhra, et autres
Publié: (2024)
Diagnosing FP4 inference: a layer-wise and block-wise sensitivity analysis of NVFP4 and MXFP4
par: Cim, Musa, et autres
Publié: (2026)
par: Cim, Musa, et autres
Publié: (2026)
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
par: Kwon, Miryeong, et autres
Publié: (2025)
par: Kwon, Miryeong, et autres
Publié: (2025)
An RDMA-First Object Storage System with SmartNIC Offload
par: Zhu, Yu, et autres
Publié: (2025)
par: Zhu, Yu, et autres
Publié: (2025)
Reconfigurable Digital RRAM Logic Enables In-Situ Pruning and Learning for Edge AI
par: Wang, Songqi, et autres
Publié: (2025)
par: Wang, Songqi, et autres
Publié: (2025)
A Compact, Low Power Transprecision ALU for Smart Edge Devices
par: Dube, Ayushi, et autres
Publié: (2025)
par: Dube, Ayushi, et autres
Publié: (2025)
From PyTorch to Calyx: An Open-Source Compiler Toolchain for ML Accelerators
par: Xie, Jiahan, et autres
Publié: (2025)
par: Xie, Jiahan, et autres
Publié: (2025)
AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN Performance
par: Kang, Seungkwan, et autres
Publié: (2026)
par: Kang, Seungkwan, et autres
Publié: (2026)
Rainbow: A Composable Coherence Protocol for Multi-Chip Servers
par: Menezo, Lucia G., et autres
Publié: (2020)
par: Menezo, Lucia G., et autres
Publié: (2020)
Evaluating Large Language Models for Automatic Register Transfer Logic Generation via High-Level Synthesis
par: Swaroopa, Sneha, et autres
Publié: (2024)
par: Swaroopa, Sneha, et autres
Publié: (2024)
An opportunity to improve Data Center Efficiency: Optimizing the Server's Upgrade Cycle
par: Nikolaou, Panagiota, et autres
Publié: (2025)
par: Nikolaou, Panagiota, et autres
Publié: (2025)
PUFBind: PUF-Enabled Lightweight Program Binary Authentication for FPGA-based Embedded Systems
par: Swaroopa, Sneha, et autres
Publié: (2025)
par: Swaroopa, Sneha, et autres
Publié: (2025)
HillInfer: Efficient Long-Context LLM Inference on the Edge with Hierarchical KV Eviction using SmartSSD
par: Sun, He, et autres
Publié: (2026)
par: Sun, He, et autres
Publié: (2026)
Enabling Efficient Hardware Acceleration of Hybrid Vision Transformer (ViT) Networks at the Edge
par: Dumoulin, Joren, et autres
Publié: (2025)
par: Dumoulin, Joren, et autres
Publié: (2025)
A Host-SSD Collaborative Write Accelerator for LSM-Tree-Based Key-Value Stores
par: Kim, KiHwan, et autres
Publié: (2024)
par: Kim, KiHwan, et autres
Publié: (2024)
Pushing the Performance Envelope of DNN-based Recommendation Systems Inference on GPUs
par: Jain, Rishabh, et autres
Publié: (2024)
par: Jain, Rishabh, et autres
Publié: (2024)
AgilePkgC: An Agile System Idle State Architecture for Energy Proportional Datacenter Servers
par: Antoniou, Georgia, et autres
Publié: (2022)
par: Antoniou, Georgia, et autres
Publié: (2022)
A4: Microarchitecture-Aware LLC Management for Datacenter Servers with Emerging I/O Devices
par: Park, Haneul, et autres
Publié: (2025)
par: Park, Haneul, et autres
Publié: (2025)
NVLLM: A 3D NAND-Centric Architecture Enabling Edge on-Device LLM Inference
par: Hao, Mingbo, et autres
Publié: (2026)
par: Hao, Mingbo, et autres
Publié: (2026)
AgileWatts: An Energy-Efficient CPU Core Idle-State Architecture for Latency-Sensitive Server Applications
par: Yahya, Jawad Haj, et autres
Publié: (2022)
par: Yahya, Jawad Haj, et autres
Publié: (2022)
Garibaldi: A Pairwise Instruction-Data Management for Enhancing Shared Last-Level Cache Performance in Server Workloads
par: Kwon, Jaewon, et autres
Publié: (2025)
par: Kwon, Jaewon, et autres
Publié: (2025)
Performance Analysis of Matrix Multiplication for Deep Learning on the Edge
par: Ramírez, Cristian, et autres
Publié: (2024)
par: Ramírez, Cristian, et autres
Publié: (2024)
Design Environment of Quantization-Aware Edge AI Hardware for Few-Shot Learning
par: Kanda, R., et autres
Publié: (2026)
par: Kanda, R., et autres
Publié: (2026)
SmartQuant: CXL-based AI Model Store in Support of Runtime Configurable Weight Quantization
par: Xie, Rui, et autres
Publié: (2024)
par: Xie, Rui, et autres
Publié: (2024)
Bit-Width-Aware Design Environment for Few-Shot Learning on Edge AI Hardware
par: Kanda, R., et autres
Publié: (2026)
par: Kanda, R., et autres
Publié: (2026)
Formal Verification of Secure Encrypted Virtualization
par: Weerasena, Hansika, et autres
Publié: (2026)
par: Weerasena, Hansika, et autres
Publié: (2026)
PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference
par: Gu, Yufeng, et autres
Publié: (2025)
par: Gu, Yufeng, et autres
Publié: (2025)
EdgeLLM: A Highly Efficient CPU-FPGA Heterogeneous Edge Accelerator for Large Language Models
par: Huang, Mingqiang, et autres
Publié: (2024)
par: Huang, Mingqiang, et autres
Publié: (2024)
In-Storage Domain-Specific Acceleration for Serverless Computing
par: Mahapatra, Rohan, et autres
Publié: (2023)
par: Mahapatra, Rohan, et autres
Publié: (2023)
Neuro-Photonix: Enabling Near-Sensor Neuro-Symbolic AI Computing on Silicon Photonics Substrate
par: Najafi, Deniz, et autres
Publié: (2024)
par: Najafi, Deniz, et autres
Publié: (2024)
Flexible In-NAND Cryptographic Processing for Secure Flash Storage
par: Noh, Seock-Hwan, et autres
Publié: (2025)
par: Noh, Seock-Hwan, et autres
Publié: (2025)
EdgeMM: Multi-Core CPU with Heterogeneous AI-Extension and Activation-aware Weight Pruning for Multimodal LLMs at Edge
par: Bai, Kangbo, et autres
Publié: (2025)
par: Bai, Kangbo, et autres
Publié: (2025)
Trimma: Trimming Metadata Storage and Latency for Hybrid Memory Systems
par: Li, Yiwei, et autres
Publié: (2024)
par: Li, Yiwei, et autres
Publié: (2024)
Bandwidth-Effective DRAM Cache for GPUs with Storage-Class Memory
par: Hong, Jeongmin, et autres
Publié: (2024)
par: Hong, Jeongmin, et autres
Publié: (2024)
SecScale: A Scalable and Secure Trusted Execution Environment for Servers
par: Sunny, Ani, et autres
Publié: (2024)
par: Sunny, Ani, et autres
Publié: (2024)
Smart-Infinity: Fast Large Language Model Training using Near-Storage Processing on a Real System
par: Jang, Hongsun, et autres
Publié: (2024)
par: Jang, Hongsun, et autres
Publié: (2024)
Switchable Single/Dual Edge Registers for Pipeline Architecture
par: Singh, Suyash Vardhan, et autres
Publié: (2024)
par: Singh, Suyash Vardhan, et autres
Publié: (2024)
MING: An Automated CNN-to-Edge MLIR HLS framework
par: Bi, Jiahong, et autres
Publié: (2026)
par: Bi, Jiahong, et autres
Publié: (2026)
Hardware-Aware DNN Compression for Homogeneous Edge Devices
par: Zhang, Kunlong, et autres
Publié: (2025)
par: Zhang, Kunlong, et autres
Publié: (2025)
Documents similaires
-
Synergistic and Efficient Edge-Host Communication for Energy Harvesting Wireless Sensor Networks
par: Mishra, Cyan Subhra, et autres
Publié: (2024) -
Revisiting DNN Training for Intermittently-Powered Energy-Harvesting Micro-Computers
par: Mishra, Cyan Subhra, et autres
Publié: (2024) -
Diagnosing FP4 inference: a layer-wise and block-wise sensitivity analysis of NVFP4 and MXFP4
par: Cim, Musa, et autres
Publié: (2026) -
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
par: Kwon, Miryeong, et autres
Publié: (2025) -
An RDMA-First Object Storage System with SmartNIC Offload
par: Zhu, Yu, et autres
Publié: (2025)