Reproduction Research of FSA-Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ludolf, Joshua, Reyna-Hernandez, Yesmin, Trevino, Matthew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference
von: Nguyen, An Xuan
Veröffentlicht: (2026)
von: Nguyen, An Xuan
Veröffentlicht: (2026)
Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative Models
von: Niu, Ziru, et al.
Veröffentlicht: (2025)
von: Niu, Ziru, et al.
Veröffentlicht: (2025)
FSA: An Alternative Efficient Implementation of Native Sparse Attention Kernel
von: Yan, Ran, et al.
Veröffentlicht: (2025)
von: Yan, Ran, et al.
Veröffentlicht: (2025)
A Comprehensive Survey on Orbital Edge Computing: Systems, Applications, and Algorithms
von: Wu, Changhao, et al.
Veröffentlicht: (2023)
von: Wu, Changhao, et al.
Veröffentlicht: (2023)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
DDS: DPU-optimized Disaggregated Storage [Extended Report]
von: Zhang, Qizhen, et al.
Veröffentlicht: (2024)
von: Zhang, Qizhen, et al.
Veröffentlicht: (2024)
Enhancing Cluster Resilience: LLM-agent Based Autonomous Intelligent Cluster Diagnosis System and Evaluation Framework
von: Shi, Honghao, et al.
Veröffentlicht: (2024)
von: Shi, Honghao, et al.
Veröffentlicht: (2024)
Flexible Swapping for the Cloud
von: Pandurov, Milan, et al.
Veröffentlicht: (2024)
von: Pandurov, Milan, et al.
Veröffentlicht: (2024)
Rendezvous and Merging for Two Metamorphic Robotic Systems without Global Compass
von: Yamada, Ryonosuke, et al.
Veröffentlicht: (2024)
von: Yamada, Ryonosuke, et al.
Veröffentlicht: (2024)
SCION: Size-aware Policy Orchestration for Nonstationary Object Caches (Long Paper Version)
von: Wang, Qizhi
Veröffentlicht: (2026)
von: Wang, Qizhi
Veröffentlicht: (2026)
Characterising resource management performance in Kubernetes
von: Medel, Víctor, et al.
Veröffentlicht: (2024)
von: Medel, Víctor, et al.
Veröffentlicht: (2024)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
von: Melnychuk, Anton, et al.
Veröffentlicht: (2025)
von: Melnychuk, Anton, et al.
Veröffentlicht: (2025)
T3: Transparent Tracking & Triggering for Fine-grained Overlap of Compute & Collectives
von: Pati, Suchita, et al.
Veröffentlicht: (2024)
von: Pati, Suchita, et al.
Veröffentlicht: (2024)
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
An Empirical Study of the Impact of Federated Learning on Machine Learning Model Accuracy
von: Yang, Haotian, et al.
Veröffentlicht: (2025)
von: Yang, Haotian, et al.
Veröffentlicht: (2025)
Readout-Side Bypass for Residual Hybrid Quantum-Classical Models
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
Melding the Serverless Control Plane with the Conventional Cluster Manager for Speed and Resource Efficiency
von: Kondrashov, Leonid, et al.
Veröffentlicht: (2025)
von: Kondrashov, Leonid, et al.
Veröffentlicht: (2025)
TStore: Rethinking AI Model Hub with Tensor-Centric Compression
von: Lan, Tingfeng, et al.
Veröffentlicht: (2026)
von: Lan, Tingfeng, et al.
Veröffentlicht: (2026)
Model-driven development of data intensive applications over cloud resources
von: Tolosana-Calasanz, Rafael, et al.
Veröffentlicht: (2024)
von: Tolosana-Calasanz, Rafael, et al.
Veröffentlicht: (2024)
The $qs$ Inequality: Quantifying the Double Penalty of Mixture-of-Experts at Inference
von: Adhinarayanan, Vignesh, et al.
Veröffentlicht: (2026)
von: Adhinarayanan, Vignesh, et al.
Veröffentlicht: (2026)
TPFL: A Trustworthy Personalized Federated Learning Framework via Subjective Logic
von: Chen, Jinqian, et al.
Veröffentlicht: (2024)
von: Chen, Jinqian, et al.
Veröffentlicht: (2024)
Communication-Efficient Diffusion Strategy for Performance Improvement of Federated Learning with Non-IID Data
von: Ahn, Seyoung, et al.
Veröffentlicht: (2022)
von: Ahn, Seyoung, et al.
Veröffentlicht: (2022)
Intersections of Web3 and AI -- View in 2024
von: Hyland-Wood, David, et al.
Veröffentlicht: (2024)
von: Hyland-Wood, David, et al.
Veröffentlicht: (2024)
Artifact Evaluation for Distributed Systems: Current Practices and Beyond
von: Sedghpour, Mohammad Reza Saleh, et al.
Veröffentlicht: (2024)
von: Sedghpour, Mohammad Reza Saleh, et al.
Veröffentlicht: (2024)
Generic Multicast (Extended Version)
von: Bolina, José Augusto, et al.
Veröffentlicht: (2024)
von: Bolina, José Augusto, et al.
Veröffentlicht: (2024)
SkyNomad: On Using Multi-Region Spot Instances to Minimize AI Batch Job Cost
von: Li, Zhifei, et al.
Veröffentlicht: (2026)
von: Li, Zhifei, et al.
Veröffentlicht: (2026)
Distributed Recoverable Sketches (Extended Version)
von: Cohen, Diana, et al.
Veröffentlicht: (2025)
von: Cohen, Diana, et al.
Veröffentlicht: (2025)
NotebookOS: A Replicated Notebook Platform for Interactive Training with On-Demand GPUs
von: Carver, Benjamin, et al.
Veröffentlicht: (2025)
von: Carver, Benjamin, et al.
Veröffentlicht: (2025)
Using a Market Economy to Provision Compute Resources Across Planet-wide Clusters
von: Stokely, Murray, et al.
Veröffentlicht: (2025)
von: Stokely, Murray, et al.
Veröffentlicht: (2025)
Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters
von: Da, Wei, et al.
Veröffentlicht: (2025)
von: Da, Wei, et al.
Veröffentlicht: (2025)
FCDP: Fully Cached Data Parallel for Communication-Avoiding Large-Scale Training
von: Park, Gyeongseo, et al.
Veröffentlicht: (2026)
von: Park, Gyeongseo, et al.
Veröffentlicht: (2026)
Sky$^ε$-Tree: Embracing the Batch Updates of B$^ε$-trees through Access Port Parallelism on Skyrmion Racetrack Memory
von: Tsai, Yu-Shiang, et al.
Veröffentlicht: (2024)
von: Tsai, Yu-Shiang, et al.
Veröffentlicht: (2024)
dpBento: Benchmarking DPUs for Data Processing
von: Hu, Jiasheng, et al.
Veröffentlicht: (2025)
von: Hu, Jiasheng, et al.
Veröffentlicht: (2025)
New Solutions Based on the Generalized Eigenvalue Problem for the Data Collaboration Analysis
von: Kawakami, Yuta, et al.
Veröffentlicht: (2024)
von: Kawakami, Yuta, et al.
Veröffentlicht: (2024)
FlashSparse: Minimizing Computation Redundancy for Fast Sparse Matrix Multiplications on Tensor Cores
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
POD-Attention: Unlocking Full Prefill-Decode Overlap for Faster LLM Inference
von: Kamath, Aditya K, et al.
Veröffentlicht: (2024)
von: Kamath, Aditya K, et al.
Veröffentlicht: (2024)
Near-Optimal Sparse Allreduce for Distributed Deep Learning
von: Li, Shigang, et al.
Veröffentlicht: (2022)
von: Li, Shigang, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025) -
Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference
von: Nguyen, An Xuan
Veröffentlicht: (2026) -
Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative Models
von: Niu, Ziru, et al.
Veröffentlicht: (2025) -
FSA: An Alternative Efficient Implementation of Native Sparse Attention Kernel
von: Yan, Ran, et al.
Veröffentlicht: (2025) -
A Comprehensive Survey on Orbital Edge Computing: Systems, Applications, and Algorithms
von: Wu, Changhao, et al.
Veröffentlicht: (2023)