PystachIO: Efficient Distributed GPU Query Processing with PyTorch over Fast Networks & Fast Storage
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Jigao, Boeschen, Nils, El-Hindi, Muhammad, Binnig, Carsten |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do GPUs Really Need New Tabular File Formats?
by: Luo, Jigao, et al.
Published: (2026)
by: Luo, Jigao, et al.
Published: (2026)
Benchmarking Analytical Query Processing in Intel SGXv2
by: Lutsch, Adrian, et al.
Published: (2024)
by: Lutsch, Adrian, et al.
Published: (2024)
Efficient Learned Query Execution over Text and Tables [Technical Report]
by: Urban, Matthias, et al.
Published: (2024)
by: Urban, Matthias, et al.
Published: (2024)
High-Performance DBMSs with io_uring: When and How to use it
by: Jasny, Matthias, et al.
Published: (2025)
by: Jasny, Matthias, et al.
Published: (2025)
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning
by: Hu, Weihua, et al.
Published: (2024)
by: Hu, Weihua, et al.
Published: (2024)
Redbench: Workload Synthesis From Cloud Traces
by: Wehrstein, Johannes, et al.
Published: (2025)
by: Wehrstein, Johannes, et al.
Published: (2025)
JOB-Complex: A Challenging Benchmark for Traditional & Learned Query Optimization
by: Wehrstein, Johannes, et al.
Published: (2025)
by: Wehrstein, Johannes, et al.
Published: (2025)
How Good are Learned Cost Models, Really? Insights from Query Optimization Tasks
by: Heinrich, Roman, et al.
Published: (2025)
by: Heinrich, Roman, et al.
Published: (2025)
PDSP-Bench: A Benchmarking System for Parallel and Distributed Stream Processing
by: Agnihotri, Pratyush, et al.
Published: (2025)
by: Agnihotri, Pratyush, et al.
Published: (2025)
PyTorch-IE: Fast and Reproducible Prototyping for Information Extraction
by: Binder, Arne, et al.
Published: (2024)
by: Binder, Arne, et al.
Published: (2024)
PairwiseHist: Fast, Accurate and Space-Efficient Approximate Query Processing with Data Compression
by: Hurst, Aaron, et al.
Published: (2024)
by: Hurst, Aaron, et al.
Published: (2024)
The Stretto Execution Engine for LLM-Augmented Data Systems
by: Sanmartino, Gabriele, et al.
Published: (2026)
by: Sanmartino, Gabriele, et al.
Published: (2026)
GRACEFUL: A Learned Cost Estimator For UDFs
by: Wehrstein, Johannes, et al.
Published: (2025)
by: Wehrstein, Johannes, et al.
Published: (2025)
Bespoke OLAP: Synthesizing Workload-Specific One-size-fits-one Database Engines
by: Wehrstein, Johannes, et al.
Published: (2026)
by: Wehrstein, Johannes, et al.
Published: (2026)
Distributed Processing of kNN Queries over Moving Objects on Dynamic Road Networks
by: Tao, Mingjin, et al.
Published: (2025)
by: Tao, Mingjin, et al.
Published: (2025)
TorchFX: A modern approach to Audio DSP with PyTorch and GPU acceleration
by: Spanio, Matteo, et al.
Published: (2025)
by: Spanio, Matteo, et al.
Published: (2025)
Towards a Multimodal Stream Processing System
by: Santos, Uélison Jean Lopes dos, et al.
Published: (2025)
by: Santos, Uélison Jean Lopes dos, et al.
Published: (2025)
Data Path Fusion in GPU for Analytical Query Processing
by: Ozawa, Tsuyoshi, et al.
Published: (2026)
by: Ozawa, Tsuyoshi, et al.
Published: (2026)
DMFF in PyTorch backend
by: Zhu, Jia-Xin, et al.
Published: (2026)
by: Zhu, Jia-Xin, et al.
Published: (2026)
ADMP in PyTorch backend
by: Zhu, Jia-Xin, et al.
Published: (2026)
by: Zhu, Jia-Xin, et al.
Published: (2026)
Reflex: Faster Secure Collaborative Analytics via Controlled Intermediate Result Size Disclosure
by: Gu, Long, et al.
Published: (2025)
by: Gu, Long, et al.
Published: (2025)
SoftDTW-CUDA-Torch: Memory-Efficient GPU-Accelerated Soft Dynamic Time Warping for PyTorch
by: Weber, Ron Shapira, et al.
Published: (2026)
by: Weber, Ron Shapira, et al.
Published: (2026)
Indexing Join Inputs for Fast Queries and Maintenance
by: Lyu, Wenhui, et al.
Published: (2025)
by: Lyu, Wenhui, et al.
Published: (2025)
HL-index: Fast Reachability Query in Hypergraphs
by: Xie, Peiting, et al.
Published: (2025)
by: Xie, Peiting, et al.
Published: (2025)
Opening The Black-Box: Explaining Learned Cost Models For Databases
by: Heinrich, Roman, et al.
Published: (2025)
by: Heinrich, Roman, et al.
Published: (2025)
GTS: GPU-based Tree Index for Fast Similarity Search
by: Zhu, Yifan, et al.
Published: (2024)
by: Zhu, Yifan, et al.
Published: (2024)
Unveiling Challenges for LLMs in Enterprise Data Engineering
by: Bodensohn, Jan-Micha, et al.
Published: (2025)
by: Bodensohn, Jan-Micha, et al.
Published: (2025)
PyPOD-GP: Using PyTorch for Accelerated Chip-Level Thermal Simulation of the GPU
by: He, Neil, et al.
Published: (2024)
by: He, Neil, et al.
Published: (2024)
Milliscale: Fast Commit on Low-Latency Object Storage
by: Zhou, Jiatang, et al.
Published: (2026)
by: Zhou, Jiatang, et al.
Published: (2026)
LTNtorch: PyTorch Implementation of Logic Tensor Networks
by: Carraro, Tommaso, et al.
Published: (2024)
by: Carraro, Tommaso, et al.
Published: (2024)
RadixGraph: A Fast, Space-Optimized Data Structure for Dynamic Graph Storage (Extended Version)
by: Xie, Haoxuan, et al.
Published: (2026)
by: Xie, Haoxuan, et al.
Published: (2026)
Theseus: A Distributed and Scalable GPU-Accelerated Query Processing Platform Optimized for Efficient Data Movement
by: Aramburú, Felipe, et al.
Published: (2025)
by: Aramburú, Felipe, et al.
Published: (2025)
Torch Geometric Pool: the PyTorch library for pooling in Graph Neural Networks
by: Abate, Carlo, et al.
Published: (2025)
by: Abate, Carlo, et al.
Published: (2025)
LHGstore: An In-Memory Learned Graph Storage for Fast Updates and Analytics
by: Qiao, Pengpeng, et al.
Published: (2026)
by: Qiao, Pengpeng, et al.
Published: (2026)
ITR: Grammar-based Graph Compression Supporting Fast Triple Queries
by: Adler, Enno, et al.
Published: (2023)
by: Adler, Enno, et al.
Published: (2023)
TorchSim: An efficient atomistic simulation engine in PyTorch
by: Cohen, Orion, et al.
Published: (2025)
by: Cohen, Orion, et al.
Published: (2025)
Fast-Vollib: A Fast Implied Volatility Library for Pythonwith PyTorch, JAX, and CUDA Fused-Kernel Backends
by: Saqur, Raeid
Published: (2026)
by: Saqur, Raeid
Published: (2026)
COSTREAM: Learned Cost Models for Operator Placement in Edge-Cloud Environments
by: Heinrich, Roman, et al.
Published: (2024)
by: Heinrich, Roman, et al.
Published: (2024)
RapidStore: An Efficient Dynamic Graph Storage System for Concurrent Queries
by: Hao, Chiyu, et al.
Published: (2025)
by: Hao, Chiyu, et al.
Published: (2025)
CardBench: A Benchmark for Learned Cardinality Estimation in Relational Databases
by: Chronis, Yannis, et al.
Published: (2024)
by: Chronis, Yannis, et al.
Published: (2024)
Similar Items
-
Do GPUs Really Need New Tabular File Formats?
by: Luo, Jigao, et al.
Published: (2026) -
Benchmarking Analytical Query Processing in Intel SGXv2
by: Lutsch, Adrian, et al.
Published: (2024) -
Efficient Learned Query Execution over Text and Tables [Technical Report]
by: Urban, Matthias, et al.
Published: (2024) -
High-Performance DBMSs with io_uring: When and How to use it
by: Jasny, Matthias, et al.
Published: (2025) -
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning
by: Hu, Weihua, et al.
Published: (2024)