Self-Evolving Distributed Memory Architecture for Scalable AI Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Zixuan, Wang, Chuanzhen, Sun, Haotian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Autopoiesis: A Self-Evolving System Paradigm for LLM Serving Under Runtime Dynamics
por: Jiang, Youhe, et al.
Publicado: (2026)
por: Jiang, Youhe, et al.
Publicado: (2026)
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda
por: Xu, Minxian, et al.
Publicado: (2026)
por: Xu, Minxian, et al.
Publicado: (2026)
Joint Temporal-Structural Representation Learning for Distributed Fault Discrimination in Microservice Architectures
por: Xue, Yihan, et al.
Publicado: (2026)
por: Xue, Yihan, et al.
Publicado: (2026)
Automatic BLAS Offloading on Unified Memory Architecture: A Study on NVIDIA Grace-Hopper
por: Li, Junjie, et al.
Publicado: (2024)
por: Li, Junjie, et al.
Publicado: (2024)
sVIRGO: A Scalable Virtual Tree Hierarchical Framework for Distributed Systems
por: Huang, Lican
Publicado: (2026)
por: Huang, Lican
Publicado: (2026)
TD-Orch: Scalable Load-Balancing for Distributed Systems with Applications to Graph Processing
por: Zhao, Yiwei, et al.
Publicado: (2025)
por: Zhao, Yiwei, et al.
Publicado: (2025)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
por: Lu, Zhengxian, et al.
Publicado: (2024)
por: Lu, Zhengxian, et al.
Publicado: (2024)
AI-Driven Health Monitoring of Distributed Computing Architecture: Insights from XGBoost and SHAP
por: Sun, Xiaoxuan, et al.
Publicado: (2024)
por: Sun, Xiaoxuan, et al.
Publicado: (2024)
DiFache: Efficient and Scalable Caching on Disaggregated Memory using Decentralized Coherence
por: Zhang, Hanze, et al.
Publicado: (2025)
por: Zhang, Hanze, et al.
Publicado: (2025)
CondenseGraph: Communication-Efficient Distributed GNN Training via On-the-Fly Graph Condensation
por: Zhang, Zizhao, et al.
Publicado: (2026)
por: Zhang, Zizhao, et al.
Publicado: (2026)
Pilotfish: Distributed Execution for Scalable Blockchains
por: Kniep, Quentin, et al.
Publicado: (2024)
por: Kniep, Quentin, et al.
Publicado: (2024)
Performance Trade-offs of High Order Meshless Approximation on Distributed Memory Systems
por: Vehovar, Jon, et al.
Publicado: (2025)
por: Vehovar, Jon, et al.
Publicado: (2025)
Towards Efficient and Scalable Distributed Vector Search with RDMA
por: Zhi, Xiangyu, et al.
Publicado: (2025)
por: Zhi, Xiangyu, et al.
Publicado: (2025)
Scalable Distributed Vector Search via Accuracy Preserving Index Construction
por: Xu, Yuming, et al.
Publicado: (2025)
por: Xu, Yuming, et al.
Publicado: (2025)
Distributed Renaming with Subquadratic Bits via Scalable Committee Election
por: Bai, Sirui, et al.
Publicado: (2026)
por: Bai, Sirui, et al.
Publicado: (2026)
Triton-distributed: Programming Overlapping Kernels on Distributed AI Systems with the Triton Compiler
por: Zheng, Size, et al.
Publicado: (2025)
por: Zheng, Size, et al.
Publicado: (2025)
Distributed Resource Selection for Self-Organising Cloud-Edge Systems
por: Renau, Quentin, et al.
Publicado: (2025)
por: Renau, Quentin, et al.
Publicado: (2025)
Scalable Systems and Software Architectures for High-Performance Computing on cloud platforms
por: Ramesh, Risshab Srinivas
Publicado: (2024)
por: Ramesh, Risshab Srinivas
Publicado: (2024)
Beluga: A CXL-Based Memory Architecture for Scalable and Efficient LLM KVCache Management
por: Yang, Xinjun, et al.
Publicado: (2025)
por: Yang, Xinjun, et al.
Publicado: (2025)
Memory-aware Adaptive Scheduling of Scientific Workflows on Heterogeneous Architectures
por: Kulagina, Svetlana, et al.
Publicado: (2025)
por: Kulagina, Svetlana, et al.
Publicado: (2025)
exa-AMD: A Scalable Workflow for Accelerating AI-Assisted Materials Discovery and Design
por: Moraru, Maxim, et al.
Publicado: (2025)
por: Moraru, Maxim, et al.
Publicado: (2025)
Kairos: A Scalable Serving System for Physical AI
por: Dai, Yinwei, et al.
Publicado: (2026)
por: Dai, Yinwei, et al.
Publicado: (2026)
BlockAMC: Scalable In-Memory Analog Matrix Computing for Solving Linear Systems
por: Pan, Lunshuai, et al.
Publicado: (2024)
por: Pan, Lunshuai, et al.
Publicado: (2024)
Beyond A Single AI Cluster: A Survey of Decentralized LLM Training
por: Dong, Haotian, et al.
Publicado: (2025)
por: Dong, Haotian, et al.
Publicado: (2025)
HybridTier: an Adaptive and Lightweight CXL-Memory Tiering System
por: Song, Kevin, et al.
Publicado: (2023)
por: Song, Kevin, et al.
Publicado: (2023)
Optimizing Memory Allocation in Distributed Clusters with Predictive Modeling
por: Bader, Jonathan, et al.
Publicado: (2026)
por: Bader, Jonathan, et al.
Publicado: (2026)
Enabling Scientific Workflow Scheduling Research in Non-Uniform Memory Access Architectures
por: Vivas, Aurelio, et al.
Publicado: (2025)
por: Vivas, Aurelio, et al.
Publicado: (2025)
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
por: Fan, Yuankai, et al.
Publicado: (2025)
por: Fan, Yuankai, et al.
Publicado: (2025)
Solutions for Distributed Memory Access Mechanism on HPC Clusters
por: Meizner, Jan, et al.
Publicado: (2025)
por: Meizner, Jan, et al.
Publicado: (2025)
Daedalus: Self-Adaptive Horizontal Autoscaling for Resource Efficiency of Distributed Stream Processing Systems
por: Pfister, Benjamin J. J., et al.
Publicado: (2024)
por: Pfister, Benjamin J. J., et al.
Publicado: (2024)
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
por: Yang, Shaofeng, et al.
Publicado: (2026)
por: Yang, Shaofeng, et al.
Publicado: (2026)
DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training
por: Wang, Zhixin, et al.
Publicado: (2025)
por: Wang, Zhixin, et al.
Publicado: (2025)
FLARE: A Dataflow-Aware and Scalable Hardware Architecture for Neural-Hybrid Scientific Lossy Compression
por: Jia, Wenqi, et al.
Publicado: (2025)
por: Jia, Wenqi, et al.
Publicado: (2025)
DEX: Scalable Range Indexing on Disaggregated Memory [Extended Version]
por: Lu, Baotong, et al.
Publicado: (2024)
por: Lu, Baotong, et al.
Publicado: (2024)
A Scalable Clustered Architecture for Cyber-Physical Systems
por: Cabral, Bernardo
Publicado: (2024)
por: Cabral, Bernardo
Publicado: (2024)
Lattica: A Decentralized Cross-NAT Communication Framework for Scalable AI Inference and Training
por: Yang, Ween, et al.
Publicado: (2025)
por: Yang, Ween, et al.
Publicado: (2025)
Distributed Log-driven Anomaly Detection System based on Evolving Decision Making
por: Tan, Zhuoran, et al.
Publicado: (2025)
por: Tan, Zhuoran, et al.
Publicado: (2025)
Data Augmentation and Convolutional Network Architecture Influence on Distributed Learning
por: Jansen, Victor Forattini, et al.
Publicado: (2026)
por: Jansen, Victor Forattini, et al.
Publicado: (2026)
Akita: A High Usability Simulation Framework for Computer Architecture
por: Jannat, Sabila Al, et al.
Publicado: (2026)
por: Jannat, Sabila Al, et al.
Publicado: (2026)
Efficient and Portable Support for Overdecomposition on Distributed Memory GPGPU Platforms
por: Bhosale, Aditya, et al.
Publicado: (2026)
por: Bhosale, Aditya, et al.
Publicado: (2026)
Ejemplares similares
-
Autopoiesis: A Self-Evolving System Paradigm for LLM Serving Under Runtime Dynamics
por: Jiang, Youhe, et al.
Publicado: (2026) -
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda
por: Xu, Minxian, et al.
Publicado: (2026) -
Joint Temporal-Structural Representation Learning for Distributed Fault Discrimination in Microservice Architectures
por: Xue, Yihan, et al.
Publicado: (2026) -
Automatic BLAS Offloading on Unified Memory Architecture: A Study on NVIDIA Grace-Hopper
por: Li, Junjie, et al.
Publicado: (2024) -
sVIRGO: A Scalable Virtual Tree Hierarchical Framework for Distributed Systems
por: Huang, Lican
Publicado: (2026)