CXL Topology-Aware and Expander-Driven Prefetching: Unlocking SSD Performance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Oh, Dongsuk, Kwon, Miryeong, Kim, Jiseon, Na, Eunjee, Moon, Junseok, Choi, Hyunkyu, Jang, Seonghyeon, Choi, Hanjin, Jung, Hongjoo, Lee, Sangwon, Jung, Myoungsoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ScalePool: Hybrid XLink-CXL Fabric for Composable Resource Disaggregation in Unified Scale-up Domains
von: Woo, Hyein, et al.
Veröffentlicht: (2025)
von: Woo, Hyein, et al.
Veröffentlicht: (2025)
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
From Block to Byte: Transforming PCIe SSDs with CXL Memory Protocol and Instruction Annotation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN Performance
von: Kang, Seungkwan, et al.
Veröffentlicht: (2026)
von: Kang, Seungkwan, et al.
Veröffentlicht: (2026)
Compute Can't Handle the Truth: Why Communication Tax Prioritizes Memory and Interconnects in Modern AI Infrastructure
von: Jung, Myoungsoo
Veröffentlicht: (2025)
von: Jung, Myoungsoo
Veröffentlicht: (2025)
A Full-System Simulation Framework for CXL-Based SSD Memory System
von: Wang, Yaohui, et al.
Veröffentlicht: (2025)
von: Wang, Yaohui, et al.
Veröffentlicht: (2025)
Low-overhead General-purpose Near-Data Processing in CXL Memory Expanders
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
MPI-over-CXL: Enhancing Communication Efficiency in Distributed HPC Systems
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
SkyByte: Architecting an Efficient Memory-Semantic CXL-based SSD with OS and Hardware Co-design
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
CXLRAMSim v1.0: System-Level Exploration of CXL Memory Expander Cards
von: Pathak, Karan, et al.
Veröffentlicht: (2026)
von: Pathak, Karan, et al.
Veröffentlicht: (2026)
SSD Offloading for LLM Mixture-of-Experts Weights Considered Harmful in Energy Efficiency
von: Kyung, Kwanhee, et al.
Veröffentlicht: (2025)
von: Kyung, Kwanhee, et al.
Veröffentlicht: (2025)
Cosmos: A CXL-Based Full In-Memory System for Approximate Nearest Neighbor Search
von: Ko, Seoyoung, et al.
Veröffentlicht: (2025)
von: Ko, Seoyoung, et al.
Veröffentlicht: (2025)
TRACE: Unlocking Effective CXL Bandwidth via Lossless Compression and Precision Scaling
von: Xie, Rui, et al.
Veröffentlicht: (2025)
von: Xie, Rui, et al.
Veröffentlicht: (2025)
Octopus: Enhancing CXL Memory Pods via Sparse Topology
von: Zhong, Yuhong, et al.
Veröffentlicht: (2025)
von: Zhong, Yuhong, et al.
Veröffentlicht: (2025)
Pickle Prefetcher: Programmable and Scalable Last-Level Cache Prefetcher
von: Nguyen, Hoa, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoa, et al.
Veröffentlicht: (2025)
Integrating Prefetcher Selection with Dynamic Request Allocation Improves Prefetching Efficiency
von: Li, Mengming, et al.
Veröffentlicht: (2025)
von: Li, Mengming, et al.
Veröffentlicht: (2025)
Block-SSD: A New Block-Based Blocking SSD Architecture
von: Wong, Ryan, et al.
Veröffentlicht: (2024)
von: Wong, Ryan, et al.
Veröffentlicht: (2024)
Profile-Guided Temporal Prefetching
von: Li, Mengming, et al.
Veröffentlicht: (2025)
von: Li, Mengming, et al.
Veröffentlicht: (2025)
Performance Characterizations and Usage Guidelines of Samsung CXL Memory Module Hybrid Prototype
von: Zeng, Jianping, et al.
Veröffentlicht: (2025)
von: Zeng, Jianping, et al.
Veröffentlicht: (2025)
EONSim: An NPU Simulator for On-Chip Memory and Embedding Vector Operations
von: Choi, Sangun, et al.
Veröffentlicht: (2025)
von: Choi, Sangun, et al.
Veröffentlicht: (2025)
A Host-SSD Collaborative Write Accelerator for LSM-Tree-Based Key-Value Stores
von: Kim, KiHwan, et al.
Veröffentlicht: (2024)
von: Kim, KiHwan, et al.
Veröffentlicht: (2024)
A Case for Kolmogorov-Arnold Networks in Prefetching: Towards Low-Latency, Generalizable ML-Based Prefetchers
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
The Case for Persistent CXL switches
von: Hadi, Khan Shaikhul, et al.
Veröffentlicht: (2025)
von: Hadi, Khan Shaikhul, et al.
Veröffentlicht: (2025)
Scalable Processing-Near-Memory for 1M-Token LLM Inference: CXL-Enabled KV-Cache Management Beyond GPU Limits
von: Kim, Dowon, et al.
Veröffentlicht: (2025)
von: Kim, Dowon, et al.
Veröffentlicht: (2025)
Exploring DRAM Cache Prefetching for Pooled Memory
von: Tirumalasetty, Chandrahas, et al.
Veröffentlicht: (2024)
von: Tirumalasetty, Chandrahas, et al.
Veröffentlicht: (2024)
Agile TLB Prefetching and Prediction Replacement Policy
von: Mersha, Melkamu, et al.
Veröffentlicht: (2024)
von: Mersha, Melkamu, et al.
Veröffentlicht: (2024)
SPPAM: Signature Pattern Prediction and Access-Map Prefetcher
von: Merrell, Maccoy, et al.
Veröffentlicht: (2026)
von: Merrell, Maccoy, et al.
Veröffentlicht: (2026)
ORAP: Optimized Row Access Prefetching for Rowhammer-mitigated Memory
von: Merrell, Maccoy, et al.
Veröffentlicht: (2026)
von: Merrell, Maccoy, et al.
Veröffentlicht: (2026)
ICP: Exploiting Instruction Correlation for Prefetching Irregular Memory Accesses
von: Li, Mengming, et al.
Veröffentlicht: (2026)
von: Li, Mengming, et al.
Veröffentlicht: (2026)
CXL-DMSim: A Full-System CXL Disaggregated Memory Simulator With Comprehensive Silicon Validation
von: Wang, Yanjing, et al.
Veröffentlicht: (2024)
von: Wang, Yanjing, et al.
Veröffentlicht: (2024)
LMB: Augmenting PCIe Devices with CXL-Linked Memory Buffer
von: Wang, Jiapin, et al.
Veröffentlicht: (2024)
von: Wang, Jiapin, et al.
Veröffentlicht: (2024)
A Novel Extensible Simulation Framework for CXL-Enabled Systems
von: An, Yuda, et al.
Veröffentlicht: (2024)
von: An, Yuda, et al.
Veröffentlicht: (2024)
Enabling Efficient Transaction Processing on CXL-Based Memory Sharing
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
TTP: A Hardware-Efficient Design for Precise Prefetching in Ray Tracing
von: Tozlu, Yavuz Selim, et al.
Veröffentlicht: (2026)
von: Tozlu, Yavuz Selim, et al.
Veröffentlicht: (2026)
Triangel: A High-Performance, Accurate, Timely On-Chip Temporal Prefetcher
von: Ainsworth, Sam, et al.
Veröffentlicht: (2024)
von: Ainsworth, Sam, et al.
Veröffentlicht: (2024)
Enhancing Instruction Prefetching via Cache and TLB Management
von: Jamet, Alexandre Valentin, et al.
Veröffentlicht: (2026)
von: Jamet, Alexandre Valentin, et al.
Veröffentlicht: (2026)
CiFHER: A Chiplet-Based FHE Accelerator with a Resizable Structure
von: Kim, Sangpyo, et al.
Veröffentlicht: (2023)
von: Kim, Sangpyo, et al.
Veröffentlicht: (2023)
Gaze into the Pattern: Characterizing Spatial Patterns with Internal Temporal Correlations for Hardware Prefetching
von: Chen, Zixiao, et al.
Veröffentlicht: (2024)
von: Chen, Zixiao, et al.
Veröffentlicht: (2024)
CXL-Interference: Analysis and Characterization in Modern Computer Systems
von: Mao, Shunyu, et al.
Veröffentlicht: (2024)
von: Mao, Shunyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ScalePool: Hybrid XLink-CXL Fabric for Composable Resource Disaggregation in Unified Scale-up Domains
von: Woo, Hyein, et al.
Veröffentlicht: (2025) -
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025) -
From Block to Byte: Transforming PCIe SSDs with CXL Memory Protocol and Instruction Annotation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025) -
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025) -
AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN Performance
von: Kang, Seungkwan, et al.
Veröffentlicht: (2026)