Bauplan: zero-copy, scale-up FaaS for data pipelines
Fuente:
arXiv
Saved in:
| Main Authors: | Tagliabue, Jacopo, Caraza-Harter, Tyler, Greco, Ciro |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zerrow: True Zero-Copy Arrow Pipelines in Bauplan
by: Dai, Yifan, et al.
Published: (2025)
by: Dai, Yifan, et al.
Published: (2025)
Reproducible data science over data lakes: replayable data pipelines with Bauplan and Nessie
by: Tagliabue, Jacopo, et al.
Published: (2024)
by: Tagliabue, Jacopo, et al.
Published: (2024)
FaaS and Furious: abstractions and differential caching for efficient data pre-processing
by: Tagliabue, Jacopo, et al.
Published: (2024)
by: Tagliabue, Jacopo, et al.
Published: (2024)
Eudoxia: a FaaS scheduling simulator for the composable lakehouse
by: Srivastava, Tapan, et al.
Published: (2025)
by: Srivastava, Tapan, et al.
Published: (2025)
Everything You Always Wanted to Know About Storage Compressibility of Pre-Trained ML Models but Were Afraid to Ask
by: Su, Zhaoyuan, et al.
Published: (2024)
by: Su, Zhaoyuan, et al.
Published: (2024)
Safe, Untrusted, "Proof-Carrying" AI Agents: toward the agentic lakehouse
by: Tagliabue, Jacopo, et al.
Published: (2025)
by: Tagliabue, Jacopo, et al.
Published: (2025)
Asynchronous I/O -- With Great Power Comes Great Responsibility
by: Pestka, Constantin, et al.
Published: (2024)
by: Pestka, Constantin, et al.
Published: (2024)
Design and Reliability of a User Space Write-Ahead Log in Rust
by: Pellegatti, Vitor K. F., et al.
Published: (2025)
by: Pellegatti, Vitor K. F., et al.
Published: (2025)
Taking the Leap: Efficient and Reliable Fine-Grained NUMA Migration in User-space
by: Schuhknecht, Felix, et al.
Published: (2026)
by: Schuhknecht, Felix, et al.
Published: (2026)
Virtual-Memory Assisted Buffer Management In Tiered Memory
by: Rayhan, Yeasir, et al.
Published: (2026)
by: Rayhan, Yeasir, et al.
Published: (2026)
PipeANN-Filter: An Efficient Filtered Vector Search System on SSD
by: Guo, Hao, et al.
Published: (2026)
by: Guo, Hao, et al.
Published: (2026)
Decoupling Vector Data and Index Storage for Space Efficiency
by: Ren, Yuanming, et al.
Published: (2026)
by: Ren, Yuanming, et al.
Published: (2026)
GateANN: I/O-Efficient Filtered Vector Search on SSDs
by: Lee, Nakyung, et al.
Published: (2026)
by: Lee, Nakyung, et al.
Published: (2026)
FusionANNS: An Efficient CPU/GPU Cooperative Processing Architecture for Billion-scale Approximate Nearest Neighbor Search
by: Tian, Bing, et al.
Published: (2024)
by: Tian, Bing, et al.
Published: (2024)
ROSfs: A User-Level File System for ROS
by: Xu, Zijun, et al.
Published: (2024)
by: Xu, Zijun, et al.
Published: (2024)
Trustworthy AI in the Agentic Lakehouse: from Concurrency to Governance
by: Tagliabue, Jacopo, et al.
Published: (2025)
by: Tagliabue, Jacopo, et al.
Published: (2025)
sqlelf: a SQL-centric Approach to ELF Analysis
by: Zakaria, Farid, et al.
Published: (2024)
by: Zakaria, Farid, et al.
Published: (2024)
Herding LLaMaS: Using LLMs as an OS Module
by: Kamath, Aditya K, et al.
Published: (2024)
by: Kamath, Aditya K, et al.
Published: (2024)
Reinforcement Learning for Dynamic Memory Allocation
by: Lim, Arisrei, et al.
Published: (2024)
by: Lim, Arisrei, et al.
Published: (2024)
Energy-Efficient Computation with DVFS using Deep Reinforcement Learning for Multi-Task Systems in Edge Computing
by: Li, Xinyi, et al.
Published: (2024)
by: Li, Xinyi, et al.
Published: (2024)
vAttention: Dynamic Memory Management for Serving LLMs without PagedAttention
by: Prabhu, Ramya, et al.
Published: (2024)
by: Prabhu, Ramya, et al.
Published: (2024)
MaLV-OS: Rethinking the Operating System Architecture for Machine Learning in Virtualized Clouds
by: Bitchebe, Stella, et al.
Published: (2025)
by: Bitchebe, Stella, et al.
Published: (2025)
Crash-Consistent Checkpointing for AI Training on macOS/APFS
by: Jeon, Juha
Published: (2025)
by: Jeon, Juha
Published: (2025)
Puzzle: Scheduling Multiple Deep Learning Models on Mobile Device with Heterogeneous Processors
by: Kang, Duseok, et al.
Published: (2025)
by: Kang, Duseok, et al.
Published: (2025)
PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU
by: Song, Yixin, et al.
Published: (2023)
by: Song, Yixin, et al.
Published: (2023)
Machine Learning (ML) library in Linux kernel
by: Dubeyko, Viacheslav
Published: (2026)
by: Dubeyko, Viacheslav
Published: (2026)
LithOS: An Operating System for Efficient Machine Learning on GPUs
by: Coppock, Patrick H., et al.
Published: (2025)
by: Coppock, Patrick H., et al.
Published: (2025)
Accelerated Training on Low-Power Edge Devices
by: Ahmed, Mohamed Aboelenien, et al.
Published: (2025)
by: Ahmed, Mohamed Aboelenien, et al.
Published: (2025)
TempoNet: Slack-Quantized Transformer-Guided Reinforcement Scheduler for Adaptive Deadline-Centric Real-Time Dispatchs
by: Fu, Rong, et al.
Published: (2026)
by: Fu, Rong, et al.
Published: (2026)
xNVMe: Unleashing Storage Hardware-Software Co-design
by: Lund, Simon A. F., et al.
Published: (2024)
by: Lund, Simon A. F., et al.
Published: (2024)
Thallus: An RDMA-based Columnar Data Transport Protocol
by: Chakraborty, Jayjeet, et al.
Published: (2024)
by: Chakraborty, Jayjeet, et al.
Published: (2024)
SteelDB: Diagnosing Kernel-Space Bottlenecks in Cloud OLTP Databases
by: Kondo, Mitsumasa
Published: (2026)
by: Kondo, Mitsumasa
Published: (2026)
EROICA: Online Performance Troubleshooting for Large-scale Model Training
by: Guan, Yu, et al.
Published: (2025)
by: Guan, Yu, et al.
Published: (2025)
When eBPF Meets Machine Learning: On-the-fly OS Kernel Compartmentalization
by: Wang, Zicheng, et al.
Published: (2024)
by: Wang, Zicheng, et al.
Published: (2024)
An Integrated Artificial Intelligence Operating System for Advanced Low-Altitude Aviation Applications
by: Tan, Minzhe, et al.
Published: (2024)
by: Tan, Minzhe, et al.
Published: (2024)
FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation
by: Wu, Yinpeng, et al.
Published: (2026)
by: Wu, Yinpeng, et al.
Published: (2026)
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
by: Abhyankar, Reyna, et al.
Published: (2025)
by: Abhyankar, Reyna, et al.
Published: (2025)
AdaptCache: KV Cache Native Storage Hierarchy for Low-Delay and High-Quality Language Model Serving
by: Feng, Shaoting, et al.
Published: (2025)
by: Feng, Shaoting, et al.
Published: (2025)
Semantic Scheduling for LLM Inference
by: Hua, Wenyue, et al.
Published: (2025)
by: Hua, Wenyue, et al.
Published: (2025)
Selective KV-Cache Sharing to Mitigate Timing Side-Channels in LLM Inference
by: Chu, Kexin, et al.
Published: (2025)
by: Chu, Kexin, et al.
Published: (2025)
Similar Items
-
Zerrow: True Zero-Copy Arrow Pipelines in Bauplan
by: Dai, Yifan, et al.
Published: (2025) -
Reproducible data science over data lakes: replayable data pipelines with Bauplan and Nessie
by: Tagliabue, Jacopo, et al.
Published: (2024) -
FaaS and Furious: abstractions and differential caching for efficient data pre-processing
by: Tagliabue, Jacopo, et al.
Published: (2024) -
Eudoxia: a FaaS scheduling simulator for the composable lakehouse
by: Srivastava, Tapan, et al.
Published: (2025) -
Everything You Always Wanted to Know About Storage Compressibility of Pre-Trained ML Models but Were Afraid to Ask
by: Su, Zhaoyuan, et al.
Published: (2024)