FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Zhengding, Lu, Mingge, Wang, Zhen, Ruan, Jixuan, Chen, Chang, Pan, Zaifeng, Guan, Yue, Wang, Ruiyi, Yu, Zhongkai, Zhang, Chao, Ding, Yufei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory Management
by: Pan, Zaifeng, et al.
Published: (2026)
by: Pan, Zaifeng, et al.
Published: (2026)
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
by: Pan, Zaifeng, et al.
Published: (2025)
by: Pan, Zaifeng, et al.
Published: (2025)
Syncopate: Efficient Multi-GPU AI Kernels via Automatic Chunk-Centric Compute-Communication Overlap
by: Qiang, Xinwei, et al.
Published: (2026)
by: Qiang, Xinwei, et al.
Published: (2026)
OServe: Accelerating LLM Serving via Spatial-Temporal Workload Orchestration
by: Jiang, Youhe, et al.
Published: (2026)
by: Jiang, Youhe, et al.
Published: (2026)
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage
by: Lin, Junqing, et al.
Published: (2025)
by: Lin, Junqing, et al.
Published: (2025)
AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving
by: Yu, Zhongkai, et al.
Published: (2026)
by: Yu, Zhongkai, et al.
Published: (2026)
Federated Semi-Supervised and Semi-Asynchronous Learning for Anomaly Detection in IoT Networks
by: Zhai, Wenbin, et al.
Published: (2023)
by: Zhai, Wenbin, et al.
Published: (2023)
Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference
by: Yu, Zhongkai, et al.
Published: (2025)
by: Yu, Zhongkai, et al.
Published: (2025)
AsyncSparse: Accelerating Sparse Matrix-Matrix Multiplication on Asynchronous GPU Architectures
by: Liu, Jie, et al.
Published: (2026)
by: Liu, Jie, et al.
Published: (2026)
KPerfIR: Towards an Open and Compiler-centric Ecosystem for GPU Kernel Performance Tooling on Modern AI Workloads
by: Guan, Yue, et al.
Published: (2025)
by: Guan, Yue, et al.
Published: (2025)
Energy-aware Incremental OTA Update for Flash-based Batteryless IoT Devices
by: Wei, Wei, et al.
Published: (2024)
by: Wei, Wei, et al.
Published: (2024)
Prioritized-MVBA: A New Approach to Design an Optimal Asynchronous Byzantine Agreement Protocol
by: Sony, Nasit S, et al.
Published: (2024)
by: Sony, Nasit S, et al.
Published: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
by: Liang, Antian, et al.
Published: (2025)
by: Liang, Antian, et al.
Published: (2025)
iDDS: Intelligent Distributed Dispatch and Scheduling for Workflow Orchestration
by: Guan, Wen, et al.
Published: (2025)
by: Guan, Wen, et al.
Published: (2025)
Warp-STAR: High-performance, Differentiable GPU-Accelerated Static Timing Analysis through Warp-oriented Parallel Orchestration
by: Huang, En-Ming, et al.
Published: (2026)
by: Huang, En-Ming, et al.
Published: (2026)
Falcon: Advancing Asynchronous BFT Consensus for Lower Latency and Enhanced Throughput
by: Dai, Xiaohai, et al.
Published: (2025)
by: Dai, Xiaohai, et al.
Published: (2025)
Joint$λ$: Orchestrating Serverless Workflows on Jointcloud FaaS Systems
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
Exploiting Multicast for Accelerating Collective Communication
by: Xu, Chao, et al.
Published: (2026)
by: Xu, Chao, et al.
Published: (2026)
Self-Evolving Distributed Memory Architecture for Scalable AI Systems
by: Li, Zixuan, et al.
Published: (2026)
by: Li, Zixuan, et al.
Published: (2026)
Asynchronous Checkpoint for Eventually Consistent Databases
by: Ravishankar, Raaghav, et al.
Published: (2025)
by: Ravishankar, Raaghav, et al.
Published: (2025)
Asynchronous Latency and Fast Atomic Snapshot
by: Bezerra, João Paulo, et al.
Published: (2024)
by: Bezerra, João Paulo, et al.
Published: (2024)
Byzantine Consensus in the Random Asynchronous Model
by: Danezis, George, et al.
Published: (2025)
by: Danezis, George, et al.
Published: (2025)
DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training
by: Hu, Tianhao, et al.
Published: (2026)
by: Hu, Tianhao, et al.
Published: (2026)
Accelerating Compound LLM Training Workloads with Maestro
by: Yuan, Xiulong, et al.
Published: (2026)
by: Yuan, Xiulong, et al.
Published: (2026)
Flash-KMeans: Fast and Memory-Efficient Exact K-Means
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
StarTrail: Concentric Ring Sequence Parallelism for Efficient Near-Infinite-Context Transformer Model Training
by: Liu, Ziming, et al.
Published: (2024)
by: Liu, Ziming, et al.
Published: (2024)
Lemonshark: Asynchronous DAG-BFT With Early Finality
by: Hu, Michael Yiqing, et al.
Published: (2026)
by: Hu, Michael Yiqing, et al.
Published: (2026)
PALE: Partially Asynchronous Agile Leader Election
by: Sidik, Bronislav, et al.
Published: (2018)
by: Sidik, Bronislav, et al.
Published: (2018)
From Symmetric to Asymmetric Asynchronous Byzantine Consensus
by: Cachin, Christian, et al.
Published: (2020)
by: Cachin, Christian, et al.
Published: (2020)
Asynchronous Secure Federated Learning with Byzantine aggregators
by: Del Pozzo, Antonella, et al.
Published: (2026)
by: Del Pozzo, Antonella, et al.
Published: (2026)
Orchestrating the Execution of Serverless Functions in Hybrid Clouds
by: Peri, Aristotelis, et al.
Published: (2024)
by: Peri, Aristotelis, et al.
Published: (2024)
Synergizing Monetization, Orchestration, and Semantics in Computing Continuum
by: Dehury, Chinmaya Kumar, et al.
Published: (2025)
by: Dehury, Chinmaya Kumar, et al.
Published: (2025)
Unleashing Efficient Asynchronous RL Post-Training via Staleness-Constrained Rollout Coordination
by: Li, Haoyang, et al.
Published: (2026)
by: Li, Haoyang, et al.
Published: (2026)
Boosting Scientific Error-Bounded Lossy Compression through Optimized Synergistic Lossy-Lossless Orchestration
by: Wu, Shixun, et al.
Published: (2025)
by: Wu, Shixun, et al.
Published: (2025)
ACGraph: An Efficient Asynchronous Out-of-Core Graph Processing Framework
by: Chen, Dechuang, et al.
Published: (2025)
by: Chen, Dechuang, et al.
Published: (2025)
FlashSketch: Sketch-Kernel Co-Design for Fast Sparse Sketching on GPUs
by: Dwaraknath, Rajat Vadiraj, et al.
Published: (2026)
by: Dwaraknath, Rajat Vadiraj, et al.
Published: (2026)
Consensus Through Knot Discovery in Asynchronous Dynamic Networks
by: Bricker, Rachel, et al.
Published: (2024)
by: Bricker, Rachel, et al.
Published: (2024)
Optimal Uniform Circle Formation by Asynchronous Luminous Robots
by: Feletti, Caterina, et al.
Published: (2024)
by: Feletti, Caterina, et al.
Published: (2024)
AGILE: Lightweight and Efficient Asynchronous GPU-SSD Integration
by: Yang, Zhuoping, et al.
Published: (2025)
by: Yang, Zhuoping, et al.
Published: (2025)
Examining MPI and its Extensions for Asynchronous Multithreaded Communication
by: Yan, Jiakun, et al.
Published: (2025)
by: Yan, Jiakun, et al.
Published: (2025)
Similar Items
-
ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory Management
by: Pan, Zaifeng, et al.
Published: (2026) -
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
by: Pan, Zaifeng, et al.
Published: (2025) -
Syncopate: Efficient Multi-GPU AI Kernels via Automatic Chunk-Centric Compute-Communication Overlap
by: Qiang, Xinwei, et al.
Published: (2026) -
OServe: Accelerating LLM Serving via Spatial-Temporal Workload Orchestration
by: Jiang, Youhe, et al.
Published: (2026) -
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage
by: Lin, Junqing, et al.
Published: (2025)