PUSHtap: PIM-based In-Memory HTAP with Unified Data Storage Format
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Yilong, Gao, Mingyu, Zhang, Huanchen, Liu, Fangxin, Chen, Gongye, Xian, He, Guan, Haibing, Jiang, Li |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PIM-SHERPA: Software Method for On-device LLM Inference by Resolving PIM Memory Attribute and Layout Inconsistencies
by: Lee, Sunjung, et al.
Published: (2026)
by: Lee, Sunjung, et al.
Published: (2026)
UMDAM: A Unified Data Layout and DRAM Address Mapping for Heterogenous NPU-PIM
by: Huang, Hai
Published: (2025)
by: Huang, Hai
Published: (2025)
HyperOffload: Graph-Driven Hierarchical Memory Management for Large Language Models on SuperNode Architectures
by: Liu, Fangxin, et al.
Published: (2026)
by: Liu, Fangxin, et al.
Published: (2026)
Performance Models for a Two-tiered Storage System
by: Sasidharan, Aparna, et al.
Published: (2025)
by: Sasidharan, Aparna, et al.
Published: (2025)
Do We Need Tensor Cores for Stencil Computations?
by: Gu, Qiqi, et al.
Published: (2026)
by: Gu, Qiqi, et al.
Published: (2026)
SiDP: Memory-Efficient Data Parallelism for Offline LLM Inference
by: Zhao, Alan, et al.
Published: (2026)
by: Zhao, Alan, et al.
Published: (2026)
Argus: Token Aware Distributed LLM Inference Optimization
by: Wu, Panlong, et al.
Published: (2025)
by: Wu, Panlong, et al.
Published: (2025)
FWeb3: A Practical Incentive-Aware Federated Learning Framework
by: Yan, Peishen, et al.
Published: (2026)
by: Yan, Peishen, et al.
Published: (2026)
MILLION: Mastering Long-Context LLM Inference Via Outlier-Immunized KV Product Quantization
by: Wang, Zongwu, et al.
Published: (2025)
by: Wang, Zongwu, et al.
Published: (2025)
FedCod: An Efficient Communication Protocol for Cross-Silo Federated Learning with Coding
by: Yan, Peishen, et al.
Published: (2024)
by: Yan, Peishen, et al.
Published: (2024)
Flash-KMeans: Fast and Memory-Efficient Exact K-Means
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
DOLMA: A Data Object Level Memory Disaggregation Framework for HPC Applications
by: Zheng, Haoyu, et al.
Published: (2025)
by: Zheng, Haoyu, et al.
Published: (2025)
Demystifying Object-based Big Data Storage Systems
by: Mondal, Anindita Sarkar, et al.
Published: (2024)
by: Mondal, Anindita Sarkar, et al.
Published: (2024)
HMTRace: Hardware-Assisted Memory-Tagging based Dynamic Data Race Detection
by: Shastri, Jaidev, et al.
Published: (2024)
by: Shastri, Jaidev, et al.
Published: (2024)
LIME:Accelerating Collaborative Lossless LLM Inference on Memory-Constrained Edge Devices
by: Sun, Mingyu, et al.
Published: (2025)
by: Sun, Mingyu, et al.
Published: (2025)
GriNNder: Breaking the Memory Capacity Wall in Full-Graph GNN Training with Storage Offloading
by: Song, Jaeyong, et al.
Published: (2026)
by: Song, Jaeyong, et al.
Published: (2026)
ALPHA-PIM: Analysis of Linear Algebraic Processing for High-Performance Graph Applications on a Real Processing-In-Memory System
by: Barkhordar, Marzieh, et al.
Published: (2026)
by: Barkhordar, Marzieh, et al.
Published: (2026)
FlowWalker: A Memory-efficient and High-performance GPU-based Dynamic Graph Random Walk Framework
by: Mei, Junyi, et al.
Published: (2024)
by: Mei, Junyi, et al.
Published: (2024)
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
by: Fan, Yuankai, et al.
Published: (2025)
by: Fan, Yuankai, et al.
Published: (2025)
Parallel Seismic Data Processing Performance with Cloud-based Storage
by: Mohapatra, Sasmita, et al.
Published: (2025)
by: Mohapatra, Sasmita, et al.
Published: (2025)
Nimbus: A Unified Embodied Synthetic Data Generation Framework
by: He, Zeyu, et al.
Published: (2026)
by: He, Zeyu, et al.
Published: (2026)
Multi-Path Bound for DAG Tasks
by: He, Qingqiang, et al.
Published: (2023)
by: He, Qingqiang, et al.
Published: (2023)
Automatic BLAS Offloading on Unified Memory Architecture: A Study on NVIDIA Grace-Hopper
by: Li, Junjie, et al.
Published: (2024)
by: Li, Junjie, et al.
Published: (2024)
A Blockchain-Enabled Framework for Storage and Retrieval of Social Data
by: Parab, Aishwarya, et al.
Published: (2025)
by: Parab, Aishwarya, et al.
Published: (2025)
On Fault Tolerance of Data Storage Systems: A Holistic Perspective
by: Zheng, Mai, et al.
Published: (2025)
by: Zheng, Mai, et al.
Published: (2025)
Exploiting the Uncertainty of the Longest Paths: Response Time Analysis for Probabilistic DAG Tasks
by: Gao, Yiyang, et al.
Published: (2025)
by: Gao, Yiyang, et al.
Published: (2025)
TokenRing: An Efficient Parallelism Framework for Infinite-Context LLMs via Bidirectional Communication
by: Wang, Zongwu, et al.
Published: (2024)
by: Wang, Zongwu, et al.
Published: (2024)
FairBatching: Fairness-Aware Batch Formation for LLM Inference
by: Lyu, Hongtao, et al.
Published: (2025)
by: Lyu, Hongtao, et al.
Published: (2025)
Parallel Writing of Nested Data in Columnar Formats
by: Hahnfeld, Jonas, et al.
Published: (2024)
by: Hahnfeld, Jonas, et al.
Published: (2024)
Gathering Teams of Bounded Memory Agents on a Line
by: Gao, Younan, et al.
Published: (2025)
by: Gao, Younan, et al.
Published: (2025)
Generalized Data Placement Strategies for Racetrack Memories
by: Khan, Asif Ali, et al.
Published: (2019)
by: Khan, Asif Ali, et al.
Published: (2019)
Personalized Federated Learning on Data with Dynamic Heterogeneity under Limited Storage
by: Tan, Sixing, et al.
Published: (2024)
by: Tan, Sixing, et al.
Published: (2024)
SkimROOT: Accelerating LHC Data Filtering with Near-Storage Processing
by: Batsoyol, Narangerelt, et al.
Published: (2025)
by: Batsoyol, Narangerelt, et al.
Published: (2025)
Thread and Data Mapping in Software Transactional Memory: An Overview
by: Pasqualin, Douglas Pereira, et al.
Published: (2022)
by: Pasqualin, Douglas Pereira, et al.
Published: (2022)
SWARM: Replicating Shared Disaggregated-Memory Data in No Time
by: Murat, Antoine, et al.
Published: (2024)
by: Murat, Antoine, et al.
Published: (2024)
LOG.io: Unified Rollback Recovery and Data Lineage Capture for Distributed Data Pipelines
by: Simon, Eric, et al.
Published: (2025)
by: Simon, Eric, et al.
Published: (2025)
MemFine: Memory-Aware Fine-Grained Scheduling for MoE Training
by: Zhao, Lu, et al.
Published: (2025)
by: Zhao, Lu, et al.
Published: (2025)
Porting HPC Applications to AMD Instinct$^\text{TM}$ MI300A Using Unified Memory and OpenMP
by: Tandon, Suyash, et al.
Published: (2024)
by: Tandon, Suyash, et al.
Published: (2024)
GMLake: Efficient and Transparent GPU Memory Defragmentation for Large-scale DNN Training with Virtual Memory Stitching
by: Guo, Cong, et al.
Published: (2024)
by: Guo, Cong, et al.
Published: (2024)
ML-based Modeling to Predict I/O Performance on Different Storage Sub-systems
by: Xu, Yiheng, et al.
Published: (2023)
by: Xu, Yiheng, et al.
Published: (2023)
Similar Items
-
PIM-SHERPA: Software Method for On-device LLM Inference by Resolving PIM Memory Attribute and Layout Inconsistencies
by: Lee, Sunjung, et al.
Published: (2026) -
UMDAM: A Unified Data Layout and DRAM Address Mapping for Heterogenous NPU-PIM
by: Huang, Hai
Published: (2025) -
HyperOffload: Graph-Driven Hierarchical Memory Management for Large Language Models on SuperNode Architectures
by: Liu, Fangxin, et al.
Published: (2026) -
Performance Models for a Two-tiered Storage System
by: Sasidharan, Aparna, et al.
Published: (2025) -
Do We Need Tensor Cores for Stencil Computations?
by: Gu, Qiqi, et al.
Published: (2026)