GraphFlow: A Graph-Based Workflow Management for Efficient LLM-Agent Serving
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Ao, Yang, Shangpeng, Chen, Fahao, Xu, Tianheng, Li, Peng, Su, Zhou |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency
por: Li, Ruixiao, et al.
Publicado: (2025)
por: Li, Ruixiao, et al.
Publicado: (2025)
Scaling Graph Chain-of-Thought Reasoning: A Multi-Agent Framework with Efficient LLM Serving
por: Huan, Chengying, et al.
Publicado: (2025)
por: Huan, Chengying, et al.
Publicado: (2025)
Efficient Serving for Dynamic Agent Workflows with Prediction-based KV-Cache Management
por: Zheng, Haoyu, et al.
Publicado: (2026)
por: Zheng, Haoyu, et al.
Publicado: (2026)
GraphFlow: An Architecture for Formally Verifiable Visual Workflows Enabling Reliable Agentic AI Automation
por: Morris V, Drewry H., et al.
Publicado: (2026)
por: Morris V, Drewry H., et al.
Publicado: (2026)
GraphMaster: Automated Graph Synthesis via LLM Agents in Data-Limited Environments
por: Du, Enjun, et al.
Publicado: (2025)
por: Du, Enjun, et al.
Publicado: (2025)
BayesFlow: A Probability Inference Framework for Meta-Agent Assisted Workflow Generation
por: Yuan, Bo, et al.
Publicado: (2026)
por: Yuan, Bo, et al.
Publicado: (2026)
When LLM Agents Meet Graph Optimization: An Automated Data Quality Improvement Approach
por: Zhang, Zhihan, et al.
Publicado: (2025)
por: Zhang, Zhihan, et al.
Publicado: (2025)
Parrot: Efficient Serving of LLM-based Applications with Semantic Variable
por: Lin, Chaofan, et al.
Publicado: (2024)
por: Lin, Chaofan, et al.
Publicado: (2024)
Graph Agent Network: Empowering Nodes with Inference Capabilities for Adversarial Resilience
por: Liu, Ao, et al.
Publicado: (2023)
por: Liu, Ao, et al.
Publicado: (2023)
Can Graph Learning Improve Planning in LLM-based Agents?
por: Wu, Xixi, et al.
Publicado: (2024)
por: Wu, Xixi, et al.
Publicado: (2024)
FlowPrecision: Advancing FPGA-Based Real-Time Fluid Flow Estimation with Linear Quantization
por: Ling, Tianheng, et al.
Publicado: (2024)
por: Ling, Tianheng, et al.
Publicado: (2024)
vTensor: Flexible Virtual Tensor Management for Efficient LLM Serving
por: Xu, Jiale, et al.
Publicado: (2024)
por: Xu, Jiale, et al.
Publicado: (2024)
Grimm: A Plug-and-Play Perturbation Rectifier for Graph Neural Networks Defending against Poisoning Attacks
por: Liu, Ao, et al.
Publicado: (2024)
por: Liu, Ao, et al.
Publicado: (2024)
Synergizing LLM Agents and Knowledge Graph for Socioeconomic Prediction in LBSN
por: Zhou, Zhilun, et al.
Publicado: (2024)
por: Zhou, Zhilun, et al.
Publicado: (2024)
Efficient Curvature-aware Graph Network
por: Fei, Chaoqun, et al.
Publicado: (2025)
por: Fei, Chaoqun, et al.
Publicado: (2025)
The Role of Visualization in LLM-Assisted Knowledge Graph Systems: Effects on User Trust, Exploration, and Workflows
por: Li, Harry, et al.
Publicado: (2025)
por: Li, Harry, et al.
Publicado: (2025)
Graph Augmentation for Cross Graph Domain Generalization
por: Chen, Guanzi, et al.
Publicado: (2025)
por: Chen, Guanzi, et al.
Publicado: (2025)
Autellix: An Efficient Serving Engine for LLM Agents as General Programs
por: Luo, Michael, et al.
Publicado: (2025)
por: Luo, Michael, et al.
Publicado: (2025)
Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
por: Zhao, Yilong, et al.
Publicado: (2023)
por: Zhao, Yilong, et al.
Publicado: (2023)
Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection
por: Huang, Tairan, et al.
Publicado: (2026)
por: Huang, Tairan, et al.
Publicado: (2026)
Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start
por: Liu, Xueshen, et al.
Publicado: (2026)
por: Liu, Xueshen, et al.
Publicado: (2026)
SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation
por: Liu, Hao, et al.
Publicado: (2026)
por: Liu, Hao, et al.
Publicado: (2026)
Multi-Domain Riemannian Graph Gluing for Building Graph Foundation Models
por: Sun, Li, et al.
Publicado: (2026)
por: Sun, Li, et al.
Publicado: (2026)
LLM4GNAS: A Large Language Model Based Toolkit for Graph Neural Architecture Search
por: Gao, Yang, et al.
Publicado: (2025)
por: Gao, Yang, et al.
Publicado: (2025)
Mell: Memory-Efficient Large Language Model Serving via Multi-GPU KV Cache Management
por: Qianli, Liu, et al.
Publicado: (2025)
por: Qianli, Liu, et al.
Publicado: (2025)
Combinatorial Optimization with Automated Graph Neural Networks
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
AgentKit: Structured LLM Reasoning with Dynamic Graphs
por: Wu, Yue, et al.
Publicado: (2024)
por: Wu, Yue, et al.
Publicado: (2024)
Toward General and Robust LLM-enhanced Text-attributed Graph Learning
por: Zhang, Zihao, et al.
Publicado: (2025)
por: Zhang, Zihao, et al.
Publicado: (2025)
Heterophilous Distribution Propagation for Graph Neural Networks
por: Zheng, Zhuonan, et al.
Publicado: (2024)
por: Zheng, Zhuonan, et al.
Publicado: (2024)
Walk Wisely on Graph: Knowledge Graph Reasoning with Dual Agents via Efficient Guidance-Exploration
por: Wang, Zijian, et al.
Publicado: (2024)
por: Wang, Zijian, et al.
Publicado: (2024)
MorphServe: Efficient and Workload-Aware LLM Serving via Runtime Quantized Layer Swapping and KV Cache Resizing
por: Su, Zhaoyuan, et al.
Publicado: (2025)
por: Su, Zhaoyuan, et al.
Publicado: (2025)
Khan-GCL: Kolmogorov-Arnold Network Based Graph Contrastive Learning with Hard Negatives
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation
por: Feng, Tao, et al.
Publicado: (2025)
por: Feng, Tao, et al.
Publicado: (2025)
Equivariant Efficient Joint Discrete and Continuous MeanFlow for Molecular Graph Generation
por: Xu, Rongjian, et al.
Publicado: (2026)
por: Xu, Rongjian, et al.
Publicado: (2026)
Preble: Efficient Distributed Prompt Scheduling for LLM Serving
por: Srivatsa, Vikranth, et al.
Publicado: (2024)
por: Srivatsa, Vikranth, et al.
Publicado: (2024)
AutoFlow: Automated Workflow Generation for Large Language Model Agents
por: Li, Zelong, et al.
Publicado: (2024)
por: Li, Zelong, et al.
Publicado: (2024)
Network Distance Based on Laplacian Flows on Graphs
por: Bao, Dianbin, et al.
Publicado: (2018)
por: Bao, Dianbin, et al.
Publicado: (2018)
OpenGLT: A Comprehensive Benchmark of Graph Neural Networks for Graph-Level Tasks
por: Li, Haoyang, et al.
Publicado: (2025)
por: Li, Haoyang, et al.
Publicado: (2025)
Simplifying Graph Kernels for Efficient
por: Wang, Lin, et al.
Publicado: (2025)
por: Wang, Lin, et al.
Publicado: (2025)
The CAP Principle for LLM Serving: A Survey of Long-Context Large Language Model Serving
por: Zeng, Pai, et al.
Publicado: (2024)
por: Zeng, Pai, et al.
Publicado: (2024)
Ejemplares similares
-
Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency
por: Li, Ruixiao, et al.
Publicado: (2025) -
Scaling Graph Chain-of-Thought Reasoning: A Multi-Agent Framework with Efficient LLM Serving
por: Huan, Chengying, et al.
Publicado: (2025) -
Efficient Serving for Dynamic Agent Workflows with Prediction-based KV-Cache Management
por: Zheng, Haoyu, et al.
Publicado: (2026) -
GraphFlow: An Architecture for Formally Verifiable Visual Workflows Enabling Reliable Agentic AI Automation
por: Morris V, Drewry H., et al.
Publicado: (2026) -
GraphMaster: Automated Graph Synthesis via LLM Agents in Data-Limited Environments
por: Du, Enjun, et al.
Publicado: (2025)