Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics
Fuente:
arXiv
Salvato in:
| Autori principali: | Yuan, Yichao, Iyer, Advait, Ma, Lin, Talati, Nishil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MoE-Lens: Towards the Hardware Limit of High-Throughput MoE LLM Serving Under Resource Constraints
di: Yuan, Yichao, et al.
Pubblicazione: (2025)
di: Yuan, Yichao, et al.
Pubblicazione: (2025)
Mayura: Exploiting Similarities in Motifs for Temporal Co-Mining
di: Singapuram, Sanjay Sri Vallabh, et al.
Pubblicazione: (2025)
di: Singapuram, Sanjay Sri Vallabh, et al.
Pubblicazione: (2025)
Agentic AI Workload Characteristics
di: Yuan, Yichao, et al.
Pubblicazione: (2026)
di: Yuan, Yichao, et al.
Pubblicazione: (2026)
KAIROS: Stateful, Context-Aware Power-Efficient Agentic Inference Serving
di: Yuan, Yichao, et al.
Pubblicazione: (2026)
di: Yuan, Yichao, et al.
Pubblicazione: (2026)
MoDM: Efficient Serving for Image Generation via Mixture-of-Diffusion Models
di: Xia, Yuchen, et al.
Pubblicazione: (2025)
di: Xia, Yuchen, et al.
Pubblicazione: (2025)
BlazingAML: High-Throughput Anti-Money Laundering (AML) via Multi-Stage Graph Mining
di: Ye, Haojie, et al.
Pubblicazione: (2026)
di: Ye, Haojie, et al.
Pubblicazione: (2026)
Cuckoo-GPU: Accelerating Cuckoo Filters on Modern GPUs
di: Dortmann, Tim, et al.
Pubblicazione: (2026)
di: Dortmann, Tim, et al.
Pubblicazione: (2026)
PolarStore: High-Performance Data Compression for Large-Scale Cloud-Native Databases
di: Hu, Qingda, et al.
Pubblicazione: (2025)
di: Hu, Qingda, et al.
Pubblicazione: (2025)
Distributed Indexing Schemes for k-Dominant Skyline Analytics on Uncertain Edge-IoT Data
di: Lai, Chuan-Chi, et al.
Pubblicazione: (2023)
di: Lai, Chuan-Chi, et al.
Pubblicazione: (2023)
ZKProphet: Understanding Performance of Zero-Knowledge Proofs on GPUs
di: Verma, Tarunesh, et al.
Pubblicazione: (2025)
di: Verma, Tarunesh, et al.
Pubblicazione: (2025)
Exploring Novel Data Storage Approaches for Large-Scale Numerical Weather Prediction
di: Gil, Nicolau Manubens
Pubblicazione: (2026)
di: Gil, Nicolau Manubens
Pubblicazione: (2026)
Terabyte-Scale Analytics in the Blink of an Eye
di: Wu, Bowen, et al.
Pubblicazione: (2025)
di: Wu, Bowen, et al.
Pubblicazione: (2025)
PIMDAL: Mitigating the Memory Bottleneck in Data Analytics using a Real Processing-in-Memory System
di: Frouzakis, Manos, et al.
Pubblicazione: (2025)
di: Frouzakis, Manos, et al.
Pubblicazione: (2025)
Employ SmartNICs' Data Path Accelerators for Ordered Key-Value Stores
di: Schimmelpfennig, Frederic, et al.
Pubblicazione: (2026)
di: Schimmelpfennig, Frederic, et al.
Pubblicazione: (2026)
Data Caching for Enterprise-Grade Petabyte-Scale OLAP
di: Tang, Chunxu, et al.
Pubblicazione: (2024)
di: Tang, Chunxu, et al.
Pubblicazione: (2024)
A Unified Ontology for Scalable Knowledge Graph-Driven Operational Data Analytics in High-Performance Computing Systems
di: Khan, Junaid Ahmed, et al.
Pubblicazione: (2025)
di: Khan, Junaid Ahmed, et al.
Pubblicazione: (2025)
CUTTANA: Scalable Graph Partitioning for Faster Distributed Graph Databases and Analytics
di: Hajidehi, Milad Rezaei, et al.
Pubblicazione: (2023)
di: Hajidehi, Milad Rezaei, et al.
Pubblicazione: (2023)
FluxSieve: Unifying Streaming and Analytical Data Planes for Scalable Cloud Observability
di: Vogel, Adriano, et al.
Pubblicazione: (2026)
di: Vogel, Adriano, et al.
Pubblicazione: (2026)
OASIS: Object-based Analytics Storage for Intelligent SQL Query Offloading in Scientific Tabular Workloads
di: Hwang, Soon, et al.
Pubblicazione: (2025)
di: Hwang, Soon, et al.
Pubblicazione: (2025)
DEX: Scalable Range Indexing on Disaggregated Memory [Extended Version]
di: Lu, Baotong, et al.
Pubblicazione: (2024)
di: Lu, Baotong, et al.
Pubblicazione: (2024)
CIDER: Boosting Memory-Disaggregated Key-Value Stores with Pessimistic Synchronization
di: Du, Yuxuan, et al.
Pubblicazione: (2026)
di: Du, Yuxuan, et al.
Pubblicazione: (2026)
XMiner: Efficient Directed Subgraph Matching with Pattern Reduction
di: Yuan, Pingpeng, et al.
Pubblicazione: (2024)
di: Yuan, Pingpeng, et al.
Pubblicazione: (2024)
Theseus: A Distributed and Scalable GPU-Accelerated Query Processing Platform Optimized for Efficient Data Movement
di: Aramburú, Felipe, et al.
Pubblicazione: (2025)
di: Aramburú, Felipe, et al.
Pubblicazione: (2025)
Next Generation Cloud-native In-Memory Stores: From Redis to Valkey and Beyond
di: Rosensch"old, Carl-Johan Fauvelle Munck af, et al.
Pubblicazione: (2025)
di: Rosensch"old, Carl-Johan Fauvelle Munck af, et al.
Pubblicazione: (2025)
GPU-RMQ: Accelerating Range Minimum Queries on Modern GPUs
di: Kreis, Lara, et al.
Pubblicazione: (2026)
di: Kreis, Lara, et al.
Pubblicazione: (2026)
LiveData -- A Worldwide Data Mesh for Stratified Data
di: Bocca, Simone, et al.
Pubblicazione: (2024)
di: Bocca, Simone, et al.
Pubblicazione: (2024)
Parallel R-tree-based Spatial Query Processing on a Commercial Processing-in-Memory System
di: Jannat, Tasmia, et al.
Pubblicazione: (2026)
di: Jannat, Tasmia, et al.
Pubblicazione: (2026)
Fine-Grained Modeling and Optimization for Intelligent Resource Management in Big Data Processing
di: Lyu, Chenghao, et al.
Pubblicazione: (2022)
di: Lyu, Chenghao, et al.
Pubblicazione: (2022)
NMP-PaK: Near-Memory Processing Acceleration of Scalable De Novo Genome Assembly
di: Kim, Heewoo, et al.
Pubblicazione: (2025)
di: Kim, Heewoo, et al.
Pubblicazione: (2025)
Data Dams: A Novel Framework for Regulating and Managing Data Flow in Large-Scale Systems
di: Bouke, Mohamed Aly, et al.
Pubblicazione: (2025)
di: Bouke, Mohamed Aly, et al.
Pubblicazione: (2025)
AQUA: Network-Accelerated Memory Offloading for LLMs in Scale-Up GPU Domains
di: Kumar, Abhishek Vijaya, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek Vijaya, et al.
Pubblicazione: (2024)
Snowpark: Performant, Secure, User-Friendly Data Engineering and AI/ML Next To Your Data
di: Baker, Brandon, et al.
Pubblicazione: (2025)
di: Baker, Brandon, et al.
Pubblicazione: (2025)
Multi-Relational Algebra for Multi-Granular Data Analytics
di: Wu, Xi, et al.
Pubblicazione: (2023)
di: Wu, Xi, et al.
Pubblicazione: (2023)
LatentBox: Storing AI-Generated Images at Scale via a Latent-First Design
di: Wang, Zirui, et al.
Pubblicazione: (2026)
di: Wang, Zirui, et al.
Pubblicazione: (2026)
Floating-Point Data Transformation for Lossless Compression
di: Jamalidinan, Samirasadat, et al.
Pubblicazione: (2025)
di: Jamalidinan, Samirasadat, et al.
Pubblicazione: (2025)
Towards Serverless Processing of Spatiotemporal Big Data Queries
di: Baumann, Diana, et al.
Pubblicazione: (2025)
di: Baumann, Diana, et al.
Pubblicazione: (2025)
Demystifying Object-based Big Data Storage Systems
di: Mondal, Anindita Sarkar, et al.
Pubblicazione: (2024)
di: Mondal, Anindita Sarkar, et al.
Pubblicazione: (2024)
EdgeMiner: Distributed Process Mining at the Data Sources
di: Andersen, Julia, et al.
Pubblicazione: (2024)
di: Andersen, Julia, et al.
Pubblicazione: (2024)
Data Trading and Monetization: Challenges and Open Research Directions
di: Ramadan, Qusai, et al.
Pubblicazione: (2024)
di: Ramadan, Qusai, et al.
Pubblicazione: (2024)
LCP: Enhancing Scientific Data Management with Lossy Compression for Particles
di: Zhang, Longtao, et al.
Pubblicazione: (2024)
di: Zhang, Longtao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MoE-Lens: Towards the Hardware Limit of High-Throughput MoE LLM Serving Under Resource Constraints
di: Yuan, Yichao, et al.
Pubblicazione: (2025) -
Mayura: Exploiting Similarities in Motifs for Temporal Co-Mining
di: Singapuram, Sanjay Sri Vallabh, et al.
Pubblicazione: (2025) -
Agentic AI Workload Characteristics
di: Yuan, Yichao, et al.
Pubblicazione: (2026) -
KAIROS: Stateful, Context-Aware Power-Efficient Agentic Inference Serving
di: Yuan, Yichao, et al.
Pubblicazione: (2026) -
MoDM: Efficient Serving for Image Generation via Mixture-of-Diffusion Models
di: Xia, Yuchen, et al.
Pubblicazione: (2025)