Flash-Fusion: Enabling Expressive, Low-Latency Queries on IoT Sensor Streams with LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Patherya, Kausar, Dhekne, Ashutosh, Romero, Francisco |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Probabilistic Top-k Dominating Query Monitoring over Multiple Uncertain IoT Data Streams in Edge Computing Environments
por: Lai, Chuan-Chi, et al.
Publicado: (2019)
por: Lai, Chuan-Chi, et al.
Publicado: (2019)
Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling
por: Luo, Zizhang, et al.
Publicado: (2026)
por: Luo, Zizhang, et al.
Publicado: (2026)
Flock: A Low-Cost Streaming Query Engine on FaaS Platforms
por: Liao, Gang, et al.
Publicado: (2023)
por: Liao, Gang, et al.
Publicado: (2023)
FedMon: Federated eBPF Monitoring for Distributed Anomaly Detection in Multi-Cluster Cloud Environments
por: Zehra, Sehar, et al.
Publicado: (2025)
por: Zehra, Sehar, et al.
Publicado: (2025)
DAGER: Exact Gradient Inversion for Large Language Models
por: Petrov, Ivo, et al.
Publicado: (2024)
por: Petrov, Ivo, et al.
Publicado: (2024)
Cost Trade-offs of Reasoning and Non-Reasoning Large Language Models in Text-to-SQL
por: Deochake, Saurabh, et al.
Publicado: (2025)
por: Deochake, Saurabh, et al.
Publicado: (2025)
Training LLMs on HPC Systems: Best Practices from the OpenGPT-X Project
por: Penke, Carolin, et al.
Publicado: (2025)
por: Penke, Carolin, et al.
Publicado: (2025)
CooperLLM: Cloud-Edge-End Cooperative Federated Fine-tuning for LLMs via ZOO-based Gradient Correction
por: Sun, He, et al.
Publicado: (2026)
por: Sun, He, et al.
Publicado: (2026)
A Context-Aware Knowledge Graph Platform for Stream Processing in Industrial IoT
por: Sciarroni, Monica Marconi, et al.
Publicado: (2026)
por: Sciarroni, Monica Marconi, et al.
Publicado: (2026)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
por: Li, Zonghang, et al.
Publicado: (2025)
por: Li, Zonghang, et al.
Publicado: (2025)
Distributed Indexing Schemes for k-Dominant Skyline Analytics on Uncertain Edge-IoT Data
por: Lai, Chuan-Chi, et al.
Publicado: (2023)
por: Lai, Chuan-Chi, et al.
Publicado: (2023)
CheetahGIS: Architecting a Scalable and Efficient Streaming Spatial Query Processing System
por: Cao, Jiaping, et al.
Publicado: (2025)
por: Cao, Jiaping, et al.
Publicado: (2025)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
por: Chen, Mu-Chi, et al.
Publicado: (2025)
por: Chen, Mu-Chi, et al.
Publicado: (2025)
MAS-Attention: Memory-Aware Stream Processing for Attention Acceleration on Resource-Constrained Edge Devices
por: Shakerdargah, Mohammadali, et al.
Publicado: (2024)
por: Shakerdargah, Mohammadali, et al.
Publicado: (2024)
AAFLOW: Scalable Patterns for Agentic AI Workflows
por: Sarker, Arup Kumar, et al.
Publicado: (2026)
por: Sarker, Arup Kumar, et al.
Publicado: (2026)
Comparative Analysis of Large Language Model Inference Serving Systems: A Performance Study of vLLM and HuggingFace TGI
por: Kolluru, Saicharan
Publicado: (2025)
por: Kolluru, Saicharan
Publicado: (2025)
AIvailable: A Software-Defined Architecture for LLM-as-a-Service on Heterogeneous and Legacy GPUs
por: Antunes, Pedro, et al.
Publicado: (2025)
por: Antunes, Pedro, et al.
Publicado: (2025)
POD-Attention: Unlocking Full Prefill-Decode Overlap for Faster LLM Inference
por: Kamath, Aditya K, et al.
Publicado: (2024)
por: Kamath, Aditya K, et al.
Publicado: (2024)
Reputation-based partition scheme for IoT security
por: Chen, Zhikui, et al.
Publicado: (2025)
por: Chen, Zhikui, et al.
Publicado: (2025)
Benchmarking Federated Learning for Throughput Prediction in 5G Live Streaming Applications
por: Dutta, Yuvraj, et al.
Publicado: (2025)
por: Dutta, Yuvraj, et al.
Publicado: (2025)
Streaming SQL Multi-Way Join Method for Long State Streams
por: Hu, Jinlong, et al.
Publicado: (2024)
por: Hu, Jinlong, et al.
Publicado: (2024)
Decentralized Stratified Sampling for Low-Latency Approximate Geospatial Data Stream Processing in Edge-Cloud Architectures
por: Jawarneh, Isam Mashhour Al, et al.
Publicado: (2026)
por: Jawarneh, Isam Mashhour Al, et al.
Publicado: (2026)
Worldwide Federated Training of Language Models
por: Iacob, Alex, et al.
Publicado: (2024)
por: Iacob, Alex, et al.
Publicado: (2024)
A Survey on Parallel Text Generation: From Parallel Decoding to Diffusion Language Models
por: Zhang, Lingzhe, et al.
Publicado: (2025)
por: Zhang, Lingzhe, et al.
Publicado: (2025)
Rotary GPU: Exploring Local Execution Paths for Large Mixture-of-Experts Models Under Limited GPU Memory
por: Jo, Myeong Jun
Publicado: (2026)
por: Jo, Myeong Jun
Publicado: (2026)
Architecture-Aware LLM Inference Optimization on AMD Instinct GPUs: A Comprehensive Benchmark and Deployment Study
por: Georgiou, Athos
Publicado: (2026)
por: Georgiou, Athos
Publicado: (2026)
GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration
por: Sarker, Yeahia, et al.
Publicado: (2026)
por: Sarker, Yeahia, et al.
Publicado: (2026)
Runtime-optimized Multi-way Stream Join Operator for Large-scale Streaming data
por: Hu, Jinlong, et al.
Publicado: (2024)
por: Hu, Jinlong, et al.
Publicado: (2024)
Parameter-Efficient and Personalized Federated Training of Generative Models at the Edge
por: Khan, Kabir, et al.
Publicado: (2025)
por: Khan, Kabir, et al.
Publicado: (2025)
Streaming CityJSON datasets
por: Ledoux, Hugo, et al.
Publicado: (2024)
por: Ledoux, Hugo, et al.
Publicado: (2024)
Batch Query Processing and Optimization for Agentic Workflows
por: Shen, Junyi, et al.
Publicado: (2025)
por: Shen, Junyi, et al.
Publicado: (2025)
Towards Serverless Processing of Spatiotemporal Big Data Queries
por: Baumann, Diana, et al.
Publicado: (2025)
por: Baumann, Diana, et al.
Publicado: (2025)
Optimizing Distributed Protocols with Query Rewrites [Technical Report]
por: Chu, David, et al.
Publicado: (2024)
por: Chu, David, et al.
Publicado: (2024)
Styx: Transactional Stateful Functions on Streaming Dataflows
por: Psarakis, Kyriakos, et al.
Publicado: (2023)
por: Psarakis, Kyriakos, et al.
Publicado: (2023)
CheckMate: Evaluating Checkpointing Protocols for Streaming Dataflows
por: Siachamis, George, et al.
Publicado: (2024)
por: Siachamis, George, et al.
Publicado: (2024)
Odyssey: An End-to-End System for Pareto-Optimal Serverless Query Processing
por: Jesalpura, Shyam, et al.
Publicado: (2025)
por: Jesalpura, Shyam, et al.
Publicado: (2025)
AgileDART: An Agile and Scalable Edge Stream Processing Engine
por: Ching, Cheng-Wei, et al.
Publicado: (2024)
por: Ching, Cheng-Wei, et al.
Publicado: (2024)
A Novel Approach to Translate Structural Aggregation Queries to MapReduce Code
por: Abdelmoniem, Ahmed M., et al.
Publicado: (2025)
por: Abdelmoniem, Ahmed M., et al.
Publicado: (2025)
Distributed Continuous Range-Skyline Query Monitoring over the Internet of Mobile Things
por: Lai, Chuan-Chi, et al.
Publicado: (2019)
por: Lai, Chuan-Chi, et al.
Publicado: (2019)
Efficient Fault Tolerance for Pipelined Query Engines via Write-ahead Lineage
por: Wang, Ziheng, et al.
Publicado: (2024)
por: Wang, Ziheng, et al.
Publicado: (2024)
Ejemplares similares
-
Probabilistic Top-k Dominating Query Monitoring over Multiple Uncertain IoT Data Streams in Edge Computing Environments
por: Lai, Chuan-Chi, et al.
Publicado: (2019) -
Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling
por: Luo, Zizhang, et al.
Publicado: (2026) -
Flock: A Low-Cost Streaming Query Engine on FaaS Platforms
por: Liao, Gang, et al.
Publicado: (2023) -
FedMon: Federated eBPF Monitoring for Distributed Anomaly Detection in Multi-Cluster Cloud Environments
por: Zehra, Sehar, et al.
Publicado: (2025) -
DAGER: Exact Gradient Inversion for Large Language Models
por: Petrov, Ivo, et al.
Publicado: (2024)