Saved in:
| Main Authors: | Zhang, Bingbing, Lin, Ziyu, Su, Yingxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.08215 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Cloud-Based Spatio-Temporal GNN-Transformer Hybrid Model for Traffic Flow Forecasting with External Feature Integration
by: Zheng, Zhuo, et al.
Published: (2025)
by: Zheng, Zhuo, et al.
Published: (2025)
VQ-LLM: High-performance Code Generation for Vector Quantization Augmented LLM Inference
by: Liu, Zihan, et al.
Published: (2025)
by: Liu, Zihan, et al.
Published: (2025)
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
by: Bitan, Tomer, et al.
Published: (2025)
by: Bitan, Tomer, et al.
Published: (2025)
RCOMPSs: A Scalable Runtime System for R Code Execution on Manycore Systems
by: Zhang, Xiran, et al.
Published: (2025)
by: Zhang, Xiran, et al.
Published: (2025)
Exploring Influence Factors on LLM Suitability for No-Code Development of End User IoT Applications
by: Wang, Minghe, et al.
Published: (2025)
by: Wang, Minghe, et al.
Published: (2025)
Cooperative Gradient Coding
by: Weng, Shudi, et al.
Published: (2025)
by: Weng, Shudi, et al.
Published: (2025)
Code once, Run Green: Automated Green Code Translation in Serverless Computing
by: Werner, Sebastian, et al.
Published: (2025)
by: Werner, Sebastian, et al.
Published: (2025)
CoCoI: Distributed Coded Inference System for Straggler Mitigation
by: Liu, Xing, et al.
Published: (2025)
by: Liu, Xing, et al.
Published: (2025)
ACC Saturator: Automatic Kernel Optimization for Directive-Based GPU Code
by: Matsumura, Kazuaki, et al.
Published: (2023)
by: Matsumura, Kazuaki, et al.
Published: (2023)
WWW.Serve: Interconnecting Global LLM Services through Decentralization
by: Wang, Huanyu, et al.
Published: (2026)
by: Wang, Huanyu, et al.
Published: (2026)
The Design and Implementation of a High-Performance Log-Structured RAID System for ZNS SSDs
by: Li, Jinhong, et al.
Published: (2024)
by: Li, Jinhong, et al.
Published: (2024)
Taming the Memory Footprint Crisis: System Design for Production Diffusion LLM Serving
by: Fan, Jiakun, et al.
Published: (2025)
by: Fan, Jiakun, et al.
Published: (2025)
ACE-Sync: An Adaptive Cloud-Edge Synchronization Framework for Communication-Efficient Large-Scale Distributed Model Training
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
Federated Neural Radiance Field for Distributed Intelligence
by: Zhang, Yintian, et al.
Published: (2024)
by: Zhang, Yintian, et al.
Published: (2024)
TSUE: A Two-Stage Data Update Method for an Erasure Coded Cluster File System
by: Wei, Zheng, et al.
Published: (2025)
by: Wei, Zheng, et al.
Published: (2025)
On Similarity of Computational Kernels in our Codes and Proxies
by: McKinsey, Michael, et al.
Published: (2026)
by: McKinsey, Michael, et al.
Published: (2026)
Biased Compression in Gradient Coding for Distributed Learning
by: Li, Chengxi, et al.
Published: (2026)
by: Li, Chengxi, et al.
Published: (2026)
ResiHP: Taming LLM Training Failures with Dynamic Hybrid Parallelism
by: Ma, Tenghui, et al.
Published: (2026)
by: Ma, Tenghui, et al.
Published: (2026)
DSPE: Profit Maximization in Edge-Cloud Storage System using Dynamic Space Partitioning with Erasure Code
by: Roy, Shubhradeep, et al.
Published: (2025)
by: Roy, Shubhradeep, et al.
Published: (2025)
HybridFlow: Resource-Adaptive Subtask Routing for Efficient Edge-Cloud LLM Inference
by: Dong, Jiangwen, et al.
Published: (2025)
by: Dong, Jiangwen, et al.
Published: (2025)
Compositional Design, Implementation, and Verification of Swarms (Technical Report)
by: Furbach, Florian, et al.
Published: (2026)
by: Furbach, Florian, et al.
Published: (2026)
New Wide Locally Recoverable Codes with Unified Locality
by: Xu, Liangliang, et al.
Published: (2025)
by: Xu, Liangliang, et al.
Published: (2025)
A High-throughput and Secure Coded Blockchain for IoT
by: Taherpour, Amirhossein, et al.
Published: (2023)
by: Taherpour, Amirhossein, et al.
Published: (2023)
StarDist: A Code Generator for Distributed Graph Algorithms
by: Nandy, Barenya Kumar, et al.
Published: (2025)
by: Nandy, Barenya Kumar, et al.
Published: (2025)
Code Generation for a Variety of Accelerators for a Graph DSL
by: Kumar, Ashwina, et al.
Published: (2024)
by: Kumar, Ashwina, et al.
Published: (2024)
Analytical Performance Estimation during Code Generation on Modern GPUs
by: Ernst, Dominik, et al.
Published: (2022)
by: Ernst, Dominik, et al.
Published: (2022)
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
by: Daiß, Gregor, et al.
Published: (2024)
by: Daiß, Gregor, et al.
Published: (2024)
QiMeng-Kernel: Macro-Thinking Micro-Coding Paradigm for LLM-Based High-Performance GPU Kernel Generation
by: Zhu, Xinguo, et al.
Published: (2025)
by: Zhu, Xinguo, et al.
Published: (2025)
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
by: García-Raigada, Ricard S., et al.
Published: (2026)
by: García-Raigada, Ricard S., et al.
Published: (2026)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
by: Lin, Mao, et al.
Published: (2026)
by: Lin, Mao, et al.
Published: (2026)
LLM-Driven Intent-Based Privacy-Aware Orchestration Across the Cloud-Edge Continuum
by: Su, Zijie, et al.
Published: (2026)
by: Su, Zijie, et al.
Published: (2026)
BlueBottle: Fast and Robust Blockchains through Subsystem Specialization
by: Vos, Preston Vander, et al.
Published: (2025)
by: Vos, Preston Vander, et al.
Published: (2025)
Privacy-Preserving Coding Schemes for Multi-Access Distributed Computing Models
by: Sasi, Shanuja
Published: (2026)
by: Sasi, Shanuja
Published: (2026)
NCCLZ: Compression-Enabled GPU Collectives with Decoupled Quantization and Entropy Coding
by: Wang, Jiamin, et al.
Published: (2026)
by: Wang, Jiamin, et al.
Published: (2026)
Offline Energy-Optimal LLM Serving: Workload-Based Energy Models for LLM Inference on Heterogeneous Systems
by: Wilkins, Grant, et al.
Published: (2024)
by: Wilkins, Grant, et al.
Published: (2024)
Efficient LLM Inference with Activation Checkpointing and Hybrid Caching
by: Lee, Sanghyeon, et al.
Published: (2025)
by: Lee, Sanghyeon, et al.
Published: (2025)
FleetOpt: Analytical Fleet Provisioning for LLM Inference with Compress-and-Route as Implementation Mechanism
by: Chen, Huamin, et al.
Published: (2026)
by: Chen, Huamin, et al.
Published: (2026)
MixServe: An Automatic Distributed Serving System for MoE Models with Hybrid Parallelism Based on Fused Communication Algorithm
by: Zhou, Bowen, et al.
Published: (2026)
by: Zhou, Bowen, et al.
Published: (2026)
LLM4FaaS: No-Code Application Development using LLMs and FaaS
by: Wang, Minghe, et al.
Published: (2025)
by: Wang, Minghe, et al.
Published: (2025)
Approximated Coded Computing: Towards Fast, Private and Secure Distributed Machine Learning
by: Qiu, Houming, et al.
Published: (2024)
by: Qiu, Houming, et al.
Published: (2024)
Similar Items
-
A Cloud-Based Spatio-Temporal GNN-Transformer Hybrid Model for Traffic Flow Forecasting with External Feature Integration
by: Zheng, Zhuo, et al.
Published: (2025) -
VQ-LLM: High-performance Code Generation for Vector Quantization Augmented LLM Inference
by: Liu, Zihan, et al.
Published: (2025) -
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
by: Bitan, Tomer, et al.
Published: (2025) -
RCOMPSs: A Scalable Runtime System for R Code Execution on Manycore Systems
by: Zhang, Xiran, et al.
Published: (2025) -
Exploring Influence Factors on LLM Suitability for No-Code Development of End User IoT Applications
by: Wang, Minghe, et al.
Published: (2025)