Characterizing and Fixing Silent Data Loss in Spark-on-AWS-Lambda with Open Table Formats
Fuente:
arXiv
Salvato in:
| Autore principale: | Gandla, Srujan Kumar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Alea-BFT: Practical Asynchronous Byzantine Fault Tolerance
di: Antunes, Diogo S., et al.
Pubblicazione: (2024)
di: Antunes, Diogo S., et al.
Pubblicazione: (2024)
DPDPU: Data Processing with DPUs
di: Hu, Jiasheng, et al.
Pubblicazione: (2024)
di: Hu, Jiasheng, et al.
Pubblicazione: (2024)
DDS: DPU-optimized Disaggregated Storage [Extended Report]
di: Zhang, Qizhen, et al.
Pubblicazione: (2024)
di: Zhang, Qizhen, et al.
Pubblicazione: (2024)
Valori: A Deterministic Memory Substrate for AI Systems
di: Gudur, Varshith
Pubblicazione: (2025)
di: Gudur, Varshith
Pubblicazione: (2025)
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
di: Melnychuk, Anton, et al.
Pubblicazione: (2025)
di: Melnychuk, Anton, et al.
Pubblicazione: (2025)
From Consensus to Chaos: A Vulnerability Assessment of the RAFT Algorithm
di: Afifi, Tamer, et al.
Pubblicazione: (2026)
di: Afifi, Tamer, et al.
Pubblicazione: (2026)
dpBento: Benchmarking DPUs for Data Processing
di: Hu, Jiasheng, et al.
Pubblicazione: (2025)
di: Hu, Jiasheng, et al.
Pubblicazione: (2025)
Data Race Satisfiability on Array Elements
di: Shim, Junhyung, et al.
Pubblicazione: (2025)
di: Shim, Junhyung, et al.
Pubblicazione: (2025)
Combining Serverless and High-Performance Computing Paradigms to support ML Data-Intensive Applications
di: Staylor, Mills, et al.
Pubblicazione: (2025)
di: Staylor, Mills, et al.
Pubblicazione: (2025)
Design and Implementation of an Analysis Pipeline for Heterogeneous Data
di: Sarker, Arup Kumar, et al.
Pubblicazione: (2024)
di: Sarker, Arup Kumar, et al.
Pubblicazione: (2024)
Deep RC: A Scalable Data Engineering and Deep Learning Pipeline
di: Sarker, Arup Kumar, et al.
Pubblicazione: (2025)
di: Sarker, Arup Kumar, et al.
Pubblicazione: (2025)
Shipwright: Proving liveness of distributed systems with Byzantine participants
di: Leung, Derek, et al.
Pubblicazione: (2025)
di: Leung, Derek, et al.
Pubblicazione: (2025)
Theseus: A Distributed and Scalable GPU-Accelerated Query Processing Platform Optimized for Efficient Data Movement
di: Aramburú, Felipe, et al.
Pubblicazione: (2025)
di: Aramburú, Felipe, et al.
Pubblicazione: (2025)
Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems
di: Sharma, Aasish Kumar, et al.
Pubblicazione: (2026)
di: Sharma, Aasish Kumar, et al.
Pubblicazione: (2026)
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
di: Pathak, Prashant Kumar, et al.
Pubblicazione: (2026)
di: Pathak, Prashant Kumar, et al.
Pubblicazione: (2026)
Trident: Adaptive Scheduling for Heterogeneous Multimodal Data Pipelines
di: Pan, Ding, et al.
Pubblicazione: (2026)
di: Pan, Ding, et al.
Pubblicazione: (2026)
Fifty Years of Transaction Processing Research (extended)
di: Bernstein, Philip A.
Pubblicazione: (2026)
di: Bernstein, Philip A.
Pubblicazione: (2026)
Laminar: A Probe-First Scheduling Paradigm with Deterministic Runtime Survival
di: Chu, Zhengyan
Pubblicazione: (2026)
di: Chu, Zhengyan
Pubblicazione: (2026)
Rank-Aware Resource Scheduling for Tightly-Coupled MPI Workloads on Kubernetes
di: Xie, Tianfang
Pubblicazione: (2026)
di: Xie, Tianfang
Pubblicazione: (2026)
Verifying In-Network Computing Systems for Design Risks
di: Bai, Tianyu, et al.
Pubblicazione: (2026)
di: Bai, Tianyu, et al.
Pubblicazione: (2026)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
di: Guan, Bo
Pubblicazione: (2026)
di: Guan, Bo
Pubblicazione: (2026)
Trustworthiness in Digital Twin Systems: Systematic Review and Research Horizons
di: Lam, Chi Fai David, et al.
Pubblicazione: (2026)
di: Lam, Chi Fai David, et al.
Pubblicazione: (2026)
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
di: Kang, Daemyung, et al.
Pubblicazione: (2026)
di: Kang, Daemyung, et al.
Pubblicazione: (2026)
Fast Truncated SVD of Sparse and Dense Matrices on Graphics Processors
di: Tomas, Andres E., et al.
Pubblicazione: (2024)
di: Tomas, Andres E., et al.
Pubblicazione: (2024)
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
Cloud Uptime Archive: Open-Access Availability Data of Web, Cloud, and Gaming Services
di: Talluri, Sacheendra, et al.
Pubblicazione: (2025)
di: Talluri, Sacheendra, et al.
Pubblicazione: (2025)
Rebooting Microreboot: Architectural Support for Safe, Parallel Recovery in Microservice Systems
di: Bindschaedler, Laurent
Pubblicazione: (2026)
di: Bindschaedler, Laurent
Pubblicazione: (2026)
HashKitty: Distributed Password Analysis
di: Antunes, Pedro, et al.
Pubblicazione: (2025)
di: Antunes, Pedro, et al.
Pubblicazione: (2025)
A Survey on Transactional Stream Processing
di: Zhang, Shuhao, et al.
Pubblicazione: (2022)
di: Zhang, Shuhao, et al.
Pubblicazione: (2022)
GPU-Augmented OLAP Execution Engine: GPU Offloading
di: Chang, Ilsun
Pubblicazione: (2025)
di: Chang, Ilsun
Pubblicazione: (2025)
FCDP: Fully Cached Data Parallel for Communication-Avoiding Large-Scale Training
di: Park, Gyeongseo, et al.
Pubblicazione: (2026)
di: Park, Gyeongseo, et al.
Pubblicazione: (2026)
push0: Scalable and Fault-Tolerant Orchestration for Zero-Knowledge Proof Generation
di: Ahmadvand, Mohsen, et al.
Pubblicazione: (2026)
di: Ahmadvand, Mohsen, et al.
Pubblicazione: (2026)
Privacy-Aware Split Inference with Speculative Decoding for Large Language Models over Wide-Area Networks
di: Cunningham, Michael
Pubblicazione: (2026)
di: Cunningham, Michael
Pubblicazione: (2026)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
di: Li, Zhifei, et al.
Pubblicazione: (2026)
di: Li, Zhifei, et al.
Pubblicazione: (2026)
Distributed Recoverable Sketches (Extended Version)
di: Cohen, Diana, et al.
Pubblicazione: (2025)
di: Cohen, Diana, et al.
Pubblicazione: (2025)
Intersections of Web3 and AI -- View in 2024
di: Hyland-Wood, David, et al.
Pubblicazione: (2024)
di: Hyland-Wood, David, et al.
Pubblicazione: (2024)
Artifact Evaluation for Distributed Systems: Current Practices and Beyond
di: Sedghpour, Mohammad Reza Saleh, et al.
Pubblicazione: (2024)
di: Sedghpour, Mohammad Reza Saleh, et al.
Pubblicazione: (2024)
NotebookOS: A Replicated Notebook Platform for Interactive Training with On-Demand GPUs
di: Carver, Benjamin, et al.
Pubblicazione: (2025)
di: Carver, Benjamin, et al.
Pubblicazione: (2025)
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
di: Stokely, Murray, et al.
Pubblicazione: (2025)
di: Stokely, Murray, et al.
Pubblicazione: (2025)
Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters
di: Da, Wei, et al.
Pubblicazione: (2025)
di: Da, Wei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Alea-BFT: Practical Asynchronous Byzantine Fault Tolerance
di: Antunes, Diogo S., et al.
Pubblicazione: (2024) -
DPDPU: Data Processing with DPUs
di: Hu, Jiasheng, et al.
Pubblicazione: (2024) -
DDS: DPU-optimized Disaggregated Storage [Extended Report]
di: Zhang, Qizhen, et al.
Pubblicazione: (2024) -
Valori: A Deterministic Memory Substrate for AI Systems
di: Gudur, Varshith
Pubblicazione: (2025) -
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
di: Melnychuk, Anton, et al.
Pubblicazione: (2025)