FlowMesh: A Service Fabric for Composable LLM Workflows
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Junyi, Wadlom, Noppanat, Zhou, Lingfeng, Wang, Dequan, Miao, Xu, Fang, Lei, Lu, Yao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Batch Query Processing and Optimization for Agentic Workflows
by: Shen, Junyi, et al.
Published: (2025)
by: Shen, Junyi, et al.
Published: (2025)
HexAGenT: Efficient Agentic LLM Serving via Workflow- and Heterogeneity-Aware Scheduling
by: Peng, You, et al.
Published: (2026)
by: Peng, You, et al.
Published: (2026)
An Ecosystem of Services for FAIR Computational Workflows
by: Wilkinson, Sean R., et al.
Published: (2025)
by: Wilkinson, Sean R., et al.
Published: (2025)
ScalePool: Hybrid XLink-CXL Fabric for Composable Resource Disaggregation in Unified Scale-up Domains
by: Woo, Hyein, et al.
Published: (2025)
by: Woo, Hyein, et al.
Published: (2025)
Flow-Bench: A Dataset for Computational Workflow Anomaly Detection
by: Papadimitriou, George, et al.
Published: (2023)
by: Papadimitriou, George, et al.
Published: (2023)
ESG: Pipeline-Conscious Efficient Scheduling of DNN Workflows on Serverless Platforms with Shareable GPUs
by: Hui, Xinning, et al.
Published: (2024)
by: Hui, Xinning, et al.
Published: (2024)
FATE: Future-State-Aware Scheduling for Heterogeneous LLM Workflows
by: Huang, Zirui, et al.
Published: (2026)
by: Huang, Zirui, et al.
Published: (2026)
Workflow as a Service Broker in Cloud Environment: A Systematic Mapping Study
by: Abrishami, Saeid, et al.
Published: (2025)
by: Abrishami, Saeid, et al.
Published: (2025)
A Decentralized Microservice Scheduling Approach Using Service Mesh in Cloud-Edge Systems
by: Wen, Yangyang, et al.
Published: (2025)
by: Wen, Yangyang, et al.
Published: (2025)
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
by: Pan, Zaifeng, et al.
Published: (2025)
by: Pan, Zaifeng, et al.
Published: (2025)
AgentX: Towards Orchestrating Robust Agentic Workflow Patterns with FaaS-hosted MCP Services
by: Tokal, Shiva Sai Krishna Anand, et al.
Published: (2025)
by: Tokal, Shiva Sai Krishna Anand, et al.
Published: (2025)
SpecBranch: Speculative Decoding via Hybrid Drafting and Rollback-Aware Branch Parallelism
by: Shen, Yuhao, et al.
Published: (2025)
by: Shen, Yuhao, et al.
Published: (2025)
Workflow Mini-Apps: Portable, Scalable, Tunable & Faithful Representations of Scientific Workflows
by: Kilic, Ozgur Ozan, et al.
Published: (2024)
by: Kilic, Ozgur Ozan, et al.
Published: (2024)
CacheFlow: Efficient LLM Serving with 3D-Parallel KV Cache Restoration
by: Nian, Sean, et al.
Published: (2026)
by: Nian, Sean, et al.
Published: (2026)
DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training
by: Wang, Zhixin, et al.
Published: (2025)
by: Wang, Zhixin, et al.
Published: (2025)
Parallax: Efficient LLM Inference Service over Decentralized Environment
by: Tong, Chris, et al.
Published: (2025)
by: Tong, Chris, et al.
Published: (2025)
Workflows Community Summit 2024: Future Trends and Challenges in Scientific Workflows
by: da Silva, Rafael Ferreira, et al.
Published: (2024)
by: da Silva, Rafael Ferreira, et al.
Published: (2024)
exa-AMD: A Scalable Workflow for Accelerating AI-Assisted Materials Discovery and Design
by: Moraru, Maxim, et al.
Published: (2025)
by: Moraru, Maxim, et al.
Published: (2025)
SeaLLM: Service-Aware and Latency-Optimized Resource Sharing for Large Language Model Inference
by: Zhao, Yihao, et al.
Published: (2025)
by: Zhao, Yihao, et al.
Published: (2025)
Performance Modeling and Evaluation of Hyperledger Fabric: An Analysis Based on Transaction Flow and Endorsement Policies
by: Melo, Carlos, et al.
Published: (2025)
by: Melo, Carlos, et al.
Published: (2025)
OnePiece: A Large-Scale Distributed Inference System with RDMA for Complex AI-Generated Content (AIGC) Workflows
by: Chen, June, et al.
Published: (2026)
by: Chen, June, et al.
Published: (2026)
WWW.Serve: Interconnecting Global LLM Services through Decentralization
by: Wang, Huanyu, et al.
Published: (2026)
by: Wang, Huanyu, et al.
Published: (2026)
WOW: Workflow-Aware Data Movement and Task Scheduling for Dynamic Scientific Workflows
by: Lehmann, Fabian, et al.
Published: (2025)
by: Lehmann, Fabian, et al.
Published: (2025)
Composing Distributed Computations Through Task and Kernel Fusion
by: Yadav, Rohan, et al.
Published: (2024)
by: Yadav, Rohan, et al.
Published: (2024)
Amoeba: Runtime Tensor Parallel Transformation for LLM Inference Services
by: Chen, Haoyu, et al.
Published: (2025)
by: Chen, Haoyu, et al.
Published: (2025)
BLOCKS: Blockchain-supported Cross-Silo Knowledge Sharing for Efficient LLM Services
by: Zhou, Zhaojiacheng, et al.
Published: (2025)
by: Zhou, Zhaojiacheng, et al.
Published: (2025)
Improving the End-to-End Efficiency of Offline Inference for Multi-LLM Applications Based on Sampling and Simulation
by: Fang, Jingzhi, et al.
Published: (2025)
by: Fang, Jingzhi, et al.
Published: (2025)
Pythia: Exploiting Workflow Predictability for Efficient Agent-Native LLM Serving
by: Yu, Shan, et al.
Published: (2026)
by: Yu, Shan, et al.
Published: (2026)
Ten Pillars for Data Meshes
by: Grossman, Robert L., et al.
Published: (2024)
by: Grossman, Robert L., et al.
Published: (2024)
DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers
by: Maurya, Avinash, et al.
Published: (2026)
by: Maurya, Avinash, et al.
Published: (2026)
Towards Advancing Research with Workflows: A perspective from the Workflows Community Summit -- Amsterdam, 2025
by: Bonati, Irene, et al.
Published: (2026)
by: Bonati, Irene, et al.
Published: (2026)
Proceedings of 3rd Workshop on Heterogeneous Composable and Disaggregated Systems
by: Pinto, Christian, et al.
Published: (2024)
by: Pinto, Christian, et al.
Published: (2024)
CARISMA: CAR-Integrated Service Mesh Architecture
by: Klein, Kevin, et al.
Published: (2024)
by: Klein, Kevin, et al.
Published: (2024)
AI-coupled HPC Workflow Applications, Middleware and Performance
by: Brewer, Wes, et al.
Published: (2024)
by: Brewer, Wes, et al.
Published: (2024)
FlowWalker: A Memory-efficient and High-performance GPU-based Dynamic Graph Random Walk Framework
by: Mei, Junyi, et al.
Published: (2024)
by: Mei, Junyi, et al.
Published: (2024)
A Terminology for Scientific Workflow Systems
by: Suter, Frédéric, et al.
Published: (2025)
by: Suter, Frédéric, et al.
Published: (2025)
Demystifying Cost-Efficiency in LLM Serving over Heterogeneous GPUs
by: Jiang, Youhe, et al.
Published: (2025)
by: Jiang, Youhe, et al.
Published: (2025)
Joint$λ$: Orchestrating Serverless Workflows on Jointcloud FaaS Systems
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
QoSFlow: Ensuring Service Quality of Distributed Workflows Using Interpretable Sensitivity Models
by: Rashid, Md Hasanur, et al.
Published: (2026)
by: Rashid, Md Hasanur, et al.
Published: (2026)
Managing Federated Learning on Decentralized Infrastructures as a Reputation-based Collaborative Workflow
by: Wang, Yuandou, et al.
Published: (2025)
by: Wang, Yuandou, et al.
Published: (2025)
Similar Items
-
Batch Query Processing and Optimization for Agentic Workflows
by: Shen, Junyi, et al.
Published: (2025) -
HexAGenT: Efficient Agentic LLM Serving via Workflow- and Heterogeneity-Aware Scheduling
by: Peng, You, et al.
Published: (2026) -
An Ecosystem of Services for FAIR Computational Workflows
by: Wilkinson, Sean R., et al.
Published: (2025) -
ScalePool: Hybrid XLink-CXL Fabric for Composable Resource Disaggregation in Unified Scale-up Domains
by: Woo, Hyein, et al.
Published: (2025) -
Flow-Bench: A Dataset for Computational Workflow Anomaly Detection
by: Papadimitriou, George, et al.
Published: (2023)