Delta Tensor: Efficient Vector and Tensor Storage in Delta Lake
Fuente:
arXiv
Guardado en:
| Autores principales: | Bao, Zhiwei, Liao-Liao, Liu, Wu, Zhiyu, Zhou, Yifan, Fan, Dan, Aibin, Michal, Coady, Yvonne, Brownsword, Andrew |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Delta Fair Sharing: Performance Isolation for Multi-Tenant Storage Systems
por: Griggs, Tyler, et al.
Publicado: (2026)
por: Griggs, Tyler, et al.
Publicado: (2026)
SwitchDelta: Asynchronous Metadata Updating for Distributed Storage with In-Network Data Visibility
por: Li, Junru, et al.
Publicado: (2025)
por: Li, Junru, et al.
Publicado: (2025)
Flock: A Low-Cost Streaming Query Engine on FaaS Platforms
por: Liao, Gang, et al.
Publicado: (2023)
por: Liao, Gang, et al.
Publicado: (2023)
ZipLLM: Efficient LLM Storage via Model-Aware Synergistic Data Deduplication and Compression
por: Wang, Zirui, et al.
Publicado: (2025)
por: Wang, Zirui, et al.
Publicado: (2025)
Keigo: Co-designing Log-Structured Merge Key-Value Stores with a Non-Volatile, Concurrency-aware Storage Hierarchy (Extended Version)
por: Adão, Rúben, et al.
Publicado: (2025)
por: Adão, Rúben, et al.
Publicado: (2025)
SFVInt: Simple, Fast and Generic Variable-Length Integer Decoding using Bit Manipulation Instructions
por: Liao, Gang, et al.
Publicado: (2024)
por: Liao, Gang, et al.
Publicado: (2024)
Demystifying Object-based Big Data Storage Systems
por: Mondal, Anindita Sarkar, et al.
Publicado: (2024)
por: Mondal, Anindita Sarkar, et al.
Publicado: (2024)
OffloadFS: Leveraging Disaggregated Storage for Computation Offloading
por: Moon, Sungho, et al.
Publicado: (2026)
por: Moon, Sungho, et al.
Publicado: (2026)
SQUASH: Serverless and Distributed Quantization-based Attributed Vector Similarity Search
por: Oakley, Joe, et al.
Publicado: (2025)
por: Oakley, Joe, et al.
Publicado: (2025)
Exploring Distributed Vector Databases Performance on HPC Platforms: A Study with Qdrant
por: Ockerman, Seth, et al.
Publicado: (2025)
por: Ockerman, Seth, et al.
Publicado: (2025)
IDSS, a Novel P2P Relational Data Storage Service
por: Cafaro, Massimo, et al.
Publicado: (2025)
por: Cafaro, Massimo, et al.
Publicado: (2025)
Exploring Novel Data Storage Approaches for Large-Scale Numerical Weather Prediction
por: Gil, Nicolau Manubens
Publicado: (2026)
por: Gil, Nicolau Manubens
Publicado: (2026)
Data Backup System with No Impact on Business Processing Utilizing Storage and Container Technologies
por: Watanabe, Satoru
Publicado: (2024)
por: Watanabe, Satoru
Publicado: (2024)
Context Lake: A System Class Defined by Decision Coherence
por: Jiang, Xiaowei
Publicado: (2026)
por: Jiang, Xiaowei
Publicado: (2026)
MSF-Model: Queuing-Based Analysis and Prediction of Metastable Failures in Replicated Storage Systems
por: Habibi, Farzad, et al.
Publicado: (2023)
por: Habibi, Farzad, et al.
Publicado: (2023)
OASIS: Object-based Analytics Storage for Intelligent SQL Query Offloading in Scientific Tabular Workloads
por: Hwang, Soon, et al.
Publicado: (2025)
por: Hwang, Soon, et al.
Publicado: (2025)
A Chunked-Object Pattern for Multi-Region Large Payload Storage in Managed NoSQL Databases
por: Chinthareddy, Manideep Reddy
Publicado: (2025)
por: Chinthareddy, Manideep Reddy
Publicado: (2025)
StreamShield: A Production-Proven Resiliency Solution for Apache Flink at ByteDance
por: Fang, Yong, et al.
Publicado: (2026)
por: Fang, Yong, et al.
Publicado: (2026)
DIMS: Distributed Index for Similarity Search in Metric Spaces
por: Zhu, Yifan, et al.
Publicado: (2024)
por: Zhu, Yifan, et al.
Publicado: (2024)
Scalable Graph Indexing using GPUs for Approximate Nearest Neighbor Search
por: Li, Zhonggen, et al.
Publicado: (2025)
por: Li, Zhonggen, et al.
Publicado: (2025)
Ira: Efficient Transaction Replay for Distributed Systems
por: Bhat, Adithya, et al.
Publicado: (2026)
por: Bhat, Adithya, et al.
Publicado: (2026)
XMiner: Efficient Directed Subgraph Matching with Pattern Reduction
por: Yuan, Pingpeng, et al.
Publicado: (2024)
por: Yuan, Pingpeng, et al.
Publicado: (2024)
PolarStore: High-Performance Data Compression for Large-Scale Cloud-Native Databases
por: Hu, Qingda, et al.
Publicado: (2025)
por: Hu, Qingda, et al.
Publicado: (2025)
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory
por: Cheng, Rongxin, et al.
Publicado: (2024)
por: Cheng, Rongxin, et al.
Publicado: (2024)
ACGraph: An Efficient Asynchronous Out-of-Core Graph Processing Framework
por: Chen, Dechuang, et al.
Publicado: (2025)
por: Chen, Dechuang, et al.
Publicado: (2025)
Learning from the Past: Adaptive Parallelism Tuning for Stream Processing Systems
por: Han, Yuxing, et al.
Publicado: (2025)
por: Han, Yuxing, et al.
Publicado: (2025)
Efficient Fault Tolerance for Pipelined Query Engines via Write-ahead Lineage
por: Wang, Ziheng, et al.
Publicado: (2024)
por: Wang, Ziheng, et al.
Publicado: (2024)
CheetahGIS: Architecting a Scalable and Efficient Streaming Spatial Query Processing System
por: Cao, Jiaping, et al.
Publicado: (2025)
por: Cao, Jiaping, et al.
Publicado: (2025)
Fine-Grained Modeling and Optimization for Intelligent Resource Management in Big Data Processing
por: Lyu, Chenghao, et al.
Publicado: (2022)
por: Lyu, Chenghao, et al.
Publicado: (2022)
Efficient Candidate-Free R-S Set Similarity Joins with Filter-and-Verification Trees on MapReduce
por: Feng, Yuhong, et al.
Publicado: (2025)
por: Feng, Yuhong, et al.
Publicado: (2025)
Minimizing Communication for Parallel Symmetric Tensor Times Same Vector Computation
por: Daas, Hussam Al, et al.
Publicado: (2025)
por: Daas, Hussam Al, et al.
Publicado: (2025)
LARK -- Linearizability Algorithms for Replicated Keys in Aerospike
por: Goodng, Andrew, et al.
Publicado: (2025)
por: Goodng, Andrew, et al.
Publicado: (2025)
CIDER: Boosting Memory-Disaggregated Key-Value Stores with Pessimistic Synchronization
por: Du, Yuxuan, et al.
Publicado: (2026)
por: Du, Yuxuan, et al.
Publicado: (2026)
Scalable, reproducible, and cost-effective processing of large-scale medical imaging datasets
por: Kim, Michael E., et al.
Publicado: (2024)
por: Kim, Michael E., et al.
Publicado: (2024)
Curator: Efficient Indexing for Multi-Tenant Vector Databases
por: Jin, Yicheng, et al.
Publicado: (2024)
por: Jin, Yicheng, et al.
Publicado: (2024)
Data Caching for Enterprise-Grade Petabyte-Scale OLAP
por: Tang, Chunxu, et al.
Publicado: (2024)
por: Tang, Chunxu, et al.
Publicado: (2024)
SIVF: GPU-Resident IVF Index for Streaming Vector Search
por: Zhao, Dongfang
Publicado: (2026)
por: Zhao, Dongfang
Publicado: (2026)
xNVMe: Unleashing Storage Hardware-Software Co-design
por: Lund, Simon A. F., et al.
Publicado: (2024)
por: Lund, Simon A. F., et al.
Publicado: (2024)
Kairos: Efficient Temporal Graph Analytics on a Single Machine
por: da Trindade, Joana M. F., et al.
Publicado: (2024)
por: da Trindade, Joana M. F., et al.
Publicado: (2024)
vTensor: Flexible Virtual Tensor Management for Efficient LLM Serving
por: Xu, Jiale, et al.
Publicado: (2024)
por: Xu, Jiale, et al.
Publicado: (2024)
Ejemplares similares
-
Delta Fair Sharing: Performance Isolation for Multi-Tenant Storage Systems
por: Griggs, Tyler, et al.
Publicado: (2026) -
SwitchDelta: Asynchronous Metadata Updating for Distributed Storage with In-Network Data Visibility
por: Li, Junru, et al.
Publicado: (2025) -
Flock: A Low-Cost Streaming Query Engine on FaaS Platforms
por: Liao, Gang, et al.
Publicado: (2023) -
ZipLLM: Efficient LLM Storage via Model-Aware Synergistic Data Deduplication and Compression
por: Wang, Zirui, et al.
Publicado: (2025) -
Keigo: Co-designing Log-Structured Merge Key-Value Stores with a Non-Volatile, Concurrency-aware Storage Hierarchy (Extended Version)
por: Adão, Rúben, et al.
Publicado: (2025)