Validating quantum-supremacy experiments with exact and fast tensor network contraction
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Yong, Chen, Yaojian, Guo, Chu, Song, Jiawei, Shi, Xinmin, Gan, Lin, Wu, Wenzhao, Wu, Wei, Fu, Haohuan, Liu, Xin, Chen, Dexun, Zhao, Zhifeng, Yang, Guangwen, Gao, Jiangang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SW-TNC : Reaching the Most Complex Random Quantum Circuit via Tensor Network Contraction
por: Chen, Yaojian, et al.
Publicado: (2025)
por: Chen, Yaojian, et al.
Publicado: (2025)
FastMPS: Revisit Data Parallel in Large-scale Matrix Product State Sampling
por: Chen, Yaojian, et al.
Publicado: (2025)
por: Chen, Yaojian, et al.
Publicado: (2025)
GenTT: Generate Vectorized Codes for General Tensor Permutation
por: Chen, Yaojian, et al.
Publicado: (2025)
por: Chen, Yaojian, et al.
Publicado: (2025)
DiT-HC: Enabling Efficient Training of Visual Generation Model DiT on HPC-oriented CPU Cluster
por: Zhang, Jinxiao, et al.
Publicado: (2026)
por: Zhang, Jinxiao, et al.
Publicado: (2026)
SageSched: Efficient LLM Scheduling Confronting Demand Uncertainty and Hybridity
por: Gan, Zhenghao, et al.
Publicado: (2026)
por: Gan, Zhenghao, et al.
Publicado: (2026)
Kilometer-Level Coupled Modeling Using 40 Million Cores: An Eight-Year Journey of Model Development
por: Duan, Xiaohui, et al.
Publicado: (2024)
por: Duan, Xiaohui, et al.
Publicado: (2024)
MMStencil: Optimizing High-order Stencils on Multicore CPU using Matrix Unit
por: Wang, Yinuo, et al.
Publicado: (2025)
por: Wang, Yinuo, et al.
Publicado: (2025)
Space-efficient population protocols for exact majority on general graphs
por: Rybicki, Joel, et al.
Publicado: (2025)
por: Rybicki, Joel, et al.
Publicado: (2025)
Efficient Column-Wise N:M Pruning on RISC-V CPU
por: Chu, Chi-Wei, et al.
Publicado: (2025)
por: Chu, Chi-Wei, et al.
Publicado: (2025)
SP-MoE: Speculative Decoding and Prefetching for Accelerating MoE-based Model Inference
por: Chen, Liangkun, et al.
Publicado: (2025)
por: Chen, Liangkun, et al.
Publicado: (2025)
Edge-Cloud Collaborative Pothole Detection via Onboard Event Screening and Federated Temporal Segmentation
por: Wu, Yingjie, et al.
Publicado: (2026)
por: Wu, Yingjie, et al.
Publicado: (2026)
RingAda: Pipelining Large Model Fine-Tuning on Edge Devices with Scheduled Layer Unfreezing
por: Li, Liang, et al.
Publicado: (2025)
por: Li, Liang, et al.
Publicado: (2025)
On-the-fly Communication-and-Computing to Enable Representation Learning for Distributed Point Clouds
por: Chen, Xu, et al.
Publicado: (2024)
por: Chen, Xu, et al.
Publicado: (2024)
PVU: Design and Implementation of a Posit Vector Arithmetic Unit (PVU) for Enhanced Floating-Point Computing in Edge and AI Applications
por: Wu, Xinyu, et al.
Publicado: (2025)
por: Wu, Xinyu, et al.
Publicado: (2025)
Fantasy: Efficient Large-scale Vector Search on GPU Clusters with GPUDirect Async
por: Liu, Yi, et al.
Publicado: (2025)
por: Liu, Yi, et al.
Publicado: (2025)
Error-controlled Progressive Retrieval of Scientific Data under Derivable Quantities of Interest
por: Wu, Xuan, et al.
Publicado: (2024)
por: Wu, Xuan, et al.
Publicado: (2024)
It Takes Two to Tango: Serverless Workflow Serving via Bilaterally Engaged Resource Adaptation
por: Wu, Jing, et al.
Publicado: (2025)
por: Wu, Jing, et al.
Publicado: (2025)
S-HPLB: Efficient LLM Attention Serving via Sparsity-Aware Head Parallelism Load Balance
por: Liu, Di, et al.
Publicado: (2026)
por: Liu, Di, et al.
Publicado: (2026)
HydraInfer: Hybrid Disaggregated Scheduling for Multimodal Large Language Model Serving
por: Dong, Xianzhe, et al.
Publicado: (2025)
por: Dong, Xianzhe, et al.
Publicado: (2025)
InstCache: A Predictive Cache for LLM Serving
por: Zou, Longwei, et al.
Publicado: (2024)
por: Zou, Longwei, et al.
Publicado: (2024)
Memory-Efficient Split Federated Learning for LLM Fine-Tuning on Heterogeneous Mobile Devices
por: Chen, Xiaopei, et al.
Publicado: (2025)
por: Chen, Xiaopei, et al.
Publicado: (2025)
PS-WL: A Probability-Sensitive Wear Leveling scheme for SSD array scaling
por: Xu, Shuhang, et al.
Publicado: (2025)
por: Xu, Shuhang, et al.
Publicado: (2025)
OOCO: Latency-disaggregated Architecture for Online-Offline Co-locate LLM Serving
por: Wu, Siyu, et al.
Publicado: (2025)
por: Wu, Siyu, et al.
Publicado: (2025)
CloudQC: A Network-aware Framework for Multi-tenant Distributed Quantum Computing
por: Zhou, Ruilin, et al.
Publicado: (2025)
por: Zhou, Ruilin, et al.
Publicado: (2025)
Mosaic: Towards Efficient Training of Multimodal Models with Spatial Resource Multiplexing
por: Wang, Yanbo, et al.
Publicado: (2026)
por: Wang, Yanbo, et al.
Publicado: (2026)
Lifting to tensors when compiling scientific computing workloads for AI Engines
por: Brown, Nick, et al.
Publicado: (2026)
por: Brown, Nick, et al.
Publicado: (2026)
Towards Exascale Computation for Turbomachinery Flows
por: Fu, Yuhang, et al.
Publicado: (2023)
por: Fu, Yuhang, et al.
Publicado: (2023)
The 1/W Law: An Analytical Study of Context-Length Routing Topology and GPU Generation Gains for LLM Inference Energy Efficiency
por: Chen, Huamin, et al.
Publicado: (2026)
por: Chen, Huamin, et al.
Publicado: (2026)
FleetOpt: Analytical Fleet Provisioning for LLM Inference with Compress-and-Route as Implementation Mechanism
por: Chen, Huamin, et al.
Publicado: (2026)
por: Chen, Huamin, et al.
Publicado: (2026)
inference-fleet-sim: A Queueing-Theory-Grounded Fleet Capacity Planner for LLM Inference
por: Chen, Huamin, et al.
Publicado: (2026)
por: Chen, Huamin, et al.
Publicado: (2026)
DynaShard: Secure and Adaptive Blockchain Sharding Protocol with Hybrid Consensus and Dynamic Shard Management
por: Liu, Ao, et al.
Publicado: (2024)
por: Liu, Ao, et al.
Publicado: (2024)
TurboFFT: A High-Performance Fast Fourier Transform with Fault Tolerance on GPU
por: Wu, Shixun, et al.
Publicado: (2024)
por: Wu, Shixun, et al.
Publicado: (2024)
TurboFFT: Co-Designed High-Performance and Fault-Tolerant Fast Fourier Transform on GPUs
por: Wu, Shixun, et al.
Publicado: (2024)
por: Wu, Shixun, et al.
Publicado: (2024)
FaaSTube: Optimizing GPU-oriented Data Transfer for Serverless Computing
por: Wu, Hao, et al.
Publicado: (2024)
por: Wu, Hao, et al.
Publicado: (2024)
Communication-Efficient Model Aggregation with Layer Divergence Feedback in Federated Learning
por: Wang, Liwei, et al.
Publicado: (2024)
por: Wang, Liwei, et al.
Publicado: (2024)
NetSenseML: Network-Adaptive Compression for Efficient Distributed Machine Learning
por: Wang, Yisu, et al.
Publicado: (2025)
por: Wang, Yisu, et al.
Publicado: (2025)
Argus: Token Aware Distributed LLM Inference Optimization
por: Wu, Panlong, et al.
Publicado: (2025)
por: Wu, Panlong, et al.
Publicado: (2025)
DGRO: Diameter-Guided Ring Optimization for Integrated Research Infrastructure Membership
por: Wu, Shixun, et al.
Publicado: (2024)
por: Wu, Shixun, et al.
Publicado: (2024)
SuperBench: Improving Cloud AI Infrastructure Reliability with Proactive Validation
por: Xiong, Yifan, et al.
Publicado: (2024)
por: Xiong, Yifan, et al.
Publicado: (2024)
A Preliminary Study on Accelerating Simulation Optimization with GPU Implementation
por: He, Jinghai, et al.
Publicado: (2024)
por: He, Jinghai, et al.
Publicado: (2024)
Ejemplares similares
-
SW-TNC : Reaching the Most Complex Random Quantum Circuit via Tensor Network Contraction
por: Chen, Yaojian, et al.
Publicado: (2025) -
FastMPS: Revisit Data Parallel in Large-scale Matrix Product State Sampling
por: Chen, Yaojian, et al.
Publicado: (2025) -
GenTT: Generate Vectorized Codes for General Tensor Permutation
por: Chen, Yaojian, et al.
Publicado: (2025) -
DiT-HC: Enabling Efficient Training of Visual Generation Model DiT on HPC-oriented CPU Cluster
por: Zhang, Jinxiao, et al.
Publicado: (2026) -
SageSched: Efficient LLM Scheduling Confronting Demand Uncertainty and Hybridity
por: Gan, Zhenghao, et al.
Publicado: (2026)