Distributed quantum computing with black-box subroutines
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, X., Liu, Y. -D., Shi, S., Wang, Y. -J., Wang, D. -S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Residue Number System (RNS) based Distributed Quantum Addition
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2024)
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2024)
Atomique: A Quantum Compiler for Reconfigurable Neutral Atom Arrays
von: Wang, Hanrui, et al.
Veröffentlicht: (2023)
von: Wang, Hanrui, et al.
Veröffentlicht: (2023)
ML-QLS: Multilevel Quantum Layout Synthesis
von: Lin, Wan-Hsuan, et al.
Veröffentlicht: (2024)
von: Lin, Wan-Hsuan, et al.
Veröffentlicht: (2024)
Modeling Short-Range Microwave Networks to Scale Superconducting Quantum Computation
von: LaRacuente, Nicholas, et al.
Veröffentlicht: (2022)
von: LaRacuente, Nicholas, et al.
Veröffentlicht: (2022)
Residue Number System (RNS) based Distributed Quantum Multiplication
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2025)
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2025)
Evaluation of computational and energy performance in matrix multiplication algorithms on CPU and GPU using MKL, cuBLAS and SYCL
von: Torres, L. A., et al.
Veröffentlicht: (2024)
von: Torres, L. A., et al.
Veröffentlicht: (2024)
Multi-GPU-Enabled Hybrid Quantum-Classical Workflow in Quantum-HPC Middleware: Applications in Quantum Simulations
von: Chen, Kuan-Cheng, et al.
Veröffentlicht: (2024)
von: Chen, Kuan-Cheng, et al.
Veröffentlicht: (2024)
DeepStack: Scalable and Accurate Design Space Exploration for Distributed 3D-Stacked AI Accelerators
von: Mo, Zhiwen, et al.
Veröffentlicht: (2026)
von: Mo, Zhiwen, et al.
Veröffentlicht: (2026)
CCSS: Hardware-Accelerated RTL Simulation with Fast Combinational Logic Computing and Sequential Logic Synchronization
von: Feng, Weigang, et al.
Veröffentlicht: (2025)
von: Feng, Weigang, et al.
Veröffentlicht: (2025)
Workload-Aware Hardware Accelerator Mining for Distributed Deep Learning Training
von: Adnan, Muhammad, et al.
Veröffentlicht: (2024)
von: Adnan, Muhammad, et al.
Veröffentlicht: (2024)
DUET: Disaggregated Hybrid Mamba-Transformer LLMs with Prefill and Decode-Specific Packages
von: Kanani, Alish, et al.
Veröffentlicht: (2026)
von: Kanani, Alish, et al.
Veröffentlicht: (2026)
Architecting Distributed Quantum Computers: Design Insights from Resource Estimation
von: Filippov, Dmitry, et al.
Veröffentlicht: (2025)
von: Filippov, Dmitry, et al.
Veröffentlicht: (2025)
GreenLLM: Disaggregating Large Language Model Serving on Heterogeneous GPUs for Lower Carbon Emissions
von: Shi, Tianyao, et al.
Veröffentlicht: (2024)
von: Shi, Tianyao, et al.
Veröffentlicht: (2024)
Infinite-LLM: Efficient LLM Service for Long Context with DistAttention and Distributed KVCache
von: Lin, Bin, et al.
Veröffentlicht: (2024)
von: Lin, Bin, et al.
Veröffentlicht: (2024)
HieraSparse: Hierarchical Semi-Structured Sparse KV Attention
von: Wang, Haoxuan, et al.
Veröffentlicht: (2026)
von: Wang, Haoxuan, et al.
Veröffentlicht: (2026)
Navigating the Landscape of Distributed File Systems: Architectures, Implementations, and Considerations
von: Pan, Xueting, et al.
Veröffentlicht: (2024)
von: Pan, Xueting, et al.
Veröffentlicht: (2024)
Optimizing Distributed ML Communication with Fused Computation-Collective Operations
von: Punniyamurthy, Kishore, et al.
Veröffentlicht: (2023)
von: Punniyamurthy, Kishore, et al.
Veröffentlicht: (2023)
PAM: Processing Across Memory Hierarchy for Efficient KV-centric LLM Serving System
von: Liu, Lian, et al.
Veröffentlicht: (2026)
von: Liu, Lian, et al.
Veröffentlicht: (2026)
DCRA: A Distributed Chiplet-based Reconfigurable Architecture for Irregular Applications
von: Orenes-Vera, Marcelo, et al.
Veröffentlicht: (2023)
von: Orenes-Vera, Marcelo, et al.
Veröffentlicht: (2023)
TAPA-CS: Enabling Scalable Accelerator Design on Distributed HBM-FPGAs
von: Prakriya, Neha, et al.
Veröffentlicht: (2023)
von: Prakriya, Neha, et al.
Veröffentlicht: (2023)
Automated Deep Neural Network Inference Partitioning for Distributed Embedded Systems
von: Kreß, Fabian, et al.
Veröffentlicht: (2024)
von: Kreß, Fabian, et al.
Veröffentlicht: (2024)
MLDSE: Scaling Design Space Exploration Infrastructure for Multi-Level Hardware
von: Qu, Huanyu, et al.
Veröffentlicht: (2025)
von: Qu, Huanyu, et al.
Veröffentlicht: (2025)
Enabling Time-Aware Priority Traffic Management over Distributed FPGA Nodes
von: Scionti, Alberto, et al.
Veröffentlicht: (2025)
von: Scionti, Alberto, et al.
Veröffentlicht: (2025)
Torrent: A Distributed DMA for Efficient and Flexible Point-to-Multipoint Data Movement
von: Deng, Yunhao, et al.
Veröffentlicht: (2025)
von: Deng, Yunhao, et al.
Veröffentlicht: (2025)
Exploring the Efficiency of 3D-Stacked AI Chip Architecture for LLM Inference with Voxel
von: Liu, Yiqi, et al.
Veröffentlicht: (2026)
von: Liu, Yiqi, et al.
Veröffentlicht: (2026)
Adaptive KV Cache Reuse for Fast Long-Context LLM Serving
von: li, Fei, et al.
Veröffentlicht: (2026)
von: li, Fei, et al.
Veröffentlicht: (2026)
Deep Learning and Machine Learning with GPGPU and CUDA: Unlocking the Power of Parallel Computing
von: Li, Ming, et al.
Veröffentlicht: (2024)
von: Li, Ming, et al.
Veröffentlicht: (2024)
SpArch: Efficient Architecture for Sparse Matrix Multiplication
von: Zhang, Zhekai, et al.
Veröffentlicht: (2020)
von: Zhang, Zhekai, et al.
Veröffentlicht: (2020)
XDMA: A Distributed, Extensible DMA Architecture for Layout-Flexible Data Movements in Heterogeneous Multi-Accelerator SoCs
von: Kong, Fanchen, et al.
Veröffentlicht: (2025)
von: Kong, Fanchen, et al.
Veröffentlicht: (2025)
Survey of Disaggregated Memory: Cross-layer Technique Insights for Next-Generation Datacenters
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
Improving Multi-Instance GPU Efficiency via Sub-Entry Sharing TLB Design
von: Li, Bingyao, et al.
Veröffentlicht: (2024)
von: Li, Bingyao, et al.
Veröffentlicht: (2024)
Understanding Bottlenecks for Efficiently Serving LLM Inference With KV Offloading
von: Meng, William, et al.
Veröffentlicht: (2025)
von: Meng, William, et al.
Veröffentlicht: (2025)
Adaptive Multi-Objective Tiered Storage Configuration for KV Cache in LLM Service
von: Zheng, Xianzhe, et al.
Veröffentlicht: (2026)
von: Zheng, Xianzhe, et al.
Veröffentlicht: (2026)
FLEX: Leveraging FPGA-CPU Synergy for Mixed-Cell-Height Legalization Acceleration
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
von: Liu, Xingyu, et al.
Veröffentlicht: (2025)
FlexVector: A SpMM Vector Processor with Flexible VRF for GCNs on Varying-Sparsity Graphs
von: Li, Bohan, et al.
Veröffentlicht: (2026)
von: Li, Bohan, et al.
Veröffentlicht: (2026)
MANOJAVAM: A Scalable, Unified FPGA Accelerator for Matrix Multiplication and Singular Value Decomposition in Principal Component Analysis
von: Ramasubramanian, Srivaths, et al.
Veröffentlicht: (2026)
von: Ramasubramanian, Srivaths, et al.
Veröffentlicht: (2026)
CMDS: Cross-layer Dataflow Optimization for DNN Accelerators Exploiting Multi-bank Memories
von: Shi, Man, et al.
Veröffentlicht: (2024)
von: Shi, Man, et al.
Veröffentlicht: (2024)
How Fast Can Graph Computations Go on Fine-grained Parallel Architectures
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
Revisiting Computational Storage for Data Integrity and Security
von: Shi, Chao, et al.
Veröffentlicht: (2025)
von: Shi, Chao, et al.
Veröffentlicht: (2025)
UniFormer: Unified and Efficient Transformer for Reasoning Across General and Custom Computing
von: Ran, Zhuoheng, et al.
Veröffentlicht: (2025)
von: Ran, Zhuoheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Residue Number System (RNS) based Distributed Quantum Addition
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2024) -
Atomique: A Quantum Compiler for Reconfigurable Neutral Atom Arrays
von: Wang, Hanrui, et al.
Veröffentlicht: (2023) -
ML-QLS: Multilevel Quantum Layout Synthesis
von: Lin, Wan-Hsuan, et al.
Veröffentlicht: (2024) -
Modeling Short-Range Microwave Networks to Scale Superconducting Quantum Computation
von: LaRacuente, Nicholas, et al.
Veröffentlicht: (2022) -
Residue Number System (RNS) based Distributed Quantum Multiplication
von: Gaur, Bhaskar, et al.
Veröffentlicht: (2025)