Switchboard: An Open-Source Framework for Modular Simulation of Large Hardware Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Herbst, Steven, Moroze, Noah, Iglesias, Edgar, Olofsson, Andreas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Evaluation and Comparison of GPU Hardware and Solver Libraries for Accelerating the OPM Flow Reservoir Simulator
por: Qiu, Tong Dong, et al.
Publicado: (2023)
por: Qiu, Tong Dong, et al.
Publicado: (2023)
CCSS: Hardware-Accelerated RTL Simulation with Fast Combinational Logic Computing and Sequential Logic Synchronization
por: Feng, Weigang, et al.
Publicado: (2025)
por: Feng, Weigang, et al.
Publicado: (2025)
Muchisim: A Simulation Framework for Design Exploration of Multi-Chip Manycore Systems
por: Orenes-Vera, Marcelo, et al.
Publicado: (2023)
por: Orenes-Vera, Marcelo, et al.
Publicado: (2023)
RISC-V Word-Size Modular Instructions for Residue Number Systems
por: Didier, Laurent-Stéphane, et al.
Publicado: (2024)
por: Didier, Laurent-Stéphane, et al.
Publicado: (2024)
Taming Offload Overheads in a Massively Parallel Open-Source RISC-V MPSoC: Analysis and Optimization
por: Colagrande, Luca, et al.
Publicado: (2025)
por: Colagrande, Luca, et al.
Publicado: (2025)
MoE-Hub: Taming Software Complexity for Seamless MoE Overlap with Hardware-Accelerated Communication on Multi-GPU Systems
por: Zhou, Zhuoshan, et al.
Publicado: (2026)
por: Zhou, Zhuoshan, et al.
Publicado: (2026)
Workload-Aware Hardware Accelerator Mining for Distributed Deep Learning Training
por: Adnan, Muhammad, et al.
Publicado: (2024)
por: Adnan, Muhammad, et al.
Publicado: (2024)
Tascade: Hardware Support for Atomic-free, Asynchronous and Efficient Reduction Trees
por: Orenes-Vera, Marcelo, et al.
Publicado: (2023)
por: Orenes-Vera, Marcelo, et al.
Publicado: (2023)
MLDSE: Scaling Design Space Exploration Infrastructure for Multi-Level Hardware
por: Qu, Huanyu, et al.
Publicado: (2025)
por: Qu, Huanyu, et al.
Publicado: (2025)
Next-generation Probabilistic Computing Hardware with 3D MOSAICs, Illusion Scale-up, and Co-design
por: Srimani, Tathagata, et al.
Publicado: (2024)
por: Srimani, Tathagata, et al.
Publicado: (2024)
TeraPool-SDR: An 1.89TOPS 1024 RV-Cores 4MiB Shared-L1 Cluster for Next-Generation Open-Source Software-Defined Radios
por: Zhang, Yichao, et al.
Publicado: (2024)
por: Zhang, Yichao, et al.
Publicado: (2024)
TT-Edge: A Hardware-Software Co-Design for Energy-Efficient Tensor-Train Decomposition on Edge AI
por: Kwak, Hyunseok, et al.
Publicado: (2025)
por: Kwak, Hyunseok, et al.
Publicado: (2025)
Parendi: Thousand-Way Parallel RTL Simulation
por: Emami, Mahyar, et al.
Publicado: (2024)
por: Emami, Mahyar, et al.
Publicado: (2024)
RapidOMS: FPGA-based Open Modification Spectral Library Searching with HD Computing
por: Pinge, Sumukh, et al.
Publicado: (2024)
por: Pinge, Sumukh, et al.
Publicado: (2024)
Simopt-Power: Leveraging Simulation Metadata for Low-Power Design Synthesis
por: Wadhwa, Eashan, et al.
Publicado: (2025)
por: Wadhwa, Eashan, et al.
Publicado: (2025)
Multi-Partner Project: Multi-GPU Performance Portability Analysis for CFD Simulations at Scale
por: Eleftherakis, Panagiotis-Eleftherios, et al.
Publicado: (2026)
por: Eleftherakis, Panagiotis-Eleftherios, et al.
Publicado: (2026)
Leveraging SIMD for Accelerating Large-number Arithmetic
por: Das, Subhrajit, et al.
Publicado: (2026)
por: Das, Subhrajit, et al.
Publicado: (2026)
NetSmith: An Optimization Framework for Machine-Discovered Network Topologies
por: Green, Conor, et al.
Publicado: (2024)
por: Green, Conor, et al.
Publicado: (2024)
COMET: A Framework for Modeling Compound Operation Dataflows with Explicit Collectives
por: Negi, Shubham, et al.
Publicado: (2025)
por: Negi, Shubham, et al.
Publicado: (2025)
Pooling Engram Conditional Memory in Large Language Models using CXL
por: Ma, Ruiyang, et al.
Publicado: (2026)
por: Ma, Ruiyang, et al.
Publicado: (2026)
DP-HLS: A High-Level Synthesis Framework for Accelerating Dynamic Programming Algorithms in Bioinformatics
por: Cao, Yingqi, et al.
Publicado: (2024)
por: Cao, Yingqi, et al.
Publicado: (2024)
PID-Comm: A Fast and Flexible Collective Communication Framework for Commodity Processing-in-DIMM Devices
por: Noh, Si Ung, et al.
Publicado: (2024)
por: Noh, Si Ung, et al.
Publicado: (2024)
Accelerating Triangle Counting with Real Processing-in-Memory Systems
por: Asquini, Lorenzo, et al.
Publicado: (2025)
por: Asquini, Lorenzo, et al.
Publicado: (2025)
Navigating the Landscape of Distributed File Systems: Architectures, Implementations, and Considerations
por: Pan, Xueting, et al.
Publicado: (2024)
por: Pan, Xueting, et al.
Publicado: (2024)
Evaluating Rapid Makespan Predictions for Heterogeneous Systems with Programmable Logic
por: Wilhelm, Martin, et al.
Publicado: (2025)
por: Wilhelm, Martin, et al.
Publicado: (2025)
Accelerating Data Chunking in Deduplication Systems using Vector Instructions
por: Udayashankar, Sreeharsha, et al.
Publicado: (2025)
por: Udayashankar, Sreeharsha, et al.
Publicado: (2025)
Data-aware Dynamic Execution of Irregular Workloads on Heterogeneous Systems
por: Bai, Zhenyu, et al.
Publicado: (2025)
por: Bai, Zhenyu, et al.
Publicado: (2025)
GreenLLM: Disaggregating Large Language Model Serving on Heterogeneous GPUs for Lower Carbon Emissions
por: Shi, Tianyao, et al.
Publicado: (2024)
por: Shi, Tianyao, et al.
Publicado: (2024)
A Lightweight High-Throughput Collective-Capable NoC for Large-Scale ML Accelerators
por: Colagrande, Luca, et al.
Publicado: (2026)
por: Colagrande, Luca, et al.
Publicado: (2026)
Cache Your Prompt When It's Green: Carbon-Aware Caching for Large Language Model Serving
por: Tian, Yuyang, et al.
Publicado: (2025)
por: Tian, Yuyang, et al.
Publicado: (2025)
Automated Deep Neural Network Inference Partitioning for Distributed Embedded Systems
por: Kreß, Fabian, et al.
Publicado: (2024)
por: Kreß, Fabian, et al.
Publicado: (2024)
SLIM: A Heterogeneous Accelerator for Edge Inference of Sparse Large Language Model via Adaptive Thresholding
por: Xu, Weihong, et al.
Publicado: (2025)
por: Xu, Weihong, et al.
Publicado: (2025)
EPOCH: Enabling Preemption Operation for Context Saving in Heterogeneous FPGA Systems
por: Malik, Arsalan Ali, et al.
Publicado: (2025)
por: Malik, Arsalan Ali, et al.
Publicado: (2025)
New Tools, Programming Models, and System Support for Processing-in-Memory Architectures
por: Oliveira, Geraldo F.
Publicado: (2025)
por: Oliveira, Geraldo F.
Publicado: (2025)
CMDS: Cross-layer Dataflow Optimization for DNN Accelerators Exploiting Multi-bank Memories
por: Shi, Man, et al.
Publicado: (2024)
por: Shi, Man, et al.
Publicado: (2024)
BlockAMC: Scalable In-Memory Analog Matrix Computing for Solving Linear Systems
por: Pan, Lunshuai, et al.
Publicado: (2024)
por: Pan, Lunshuai, et al.
Publicado: (2024)
Enhancing Regression Models for Complex Systems Using Evolutionary Techniques for Feature Engineering
por: Arroba, Patricia, et al.
Publicado: (2024)
por: Arroba, Patricia, et al.
Publicado: (2024)
PAM: Processing Across Memory Hierarchy for Efficient KV-centric LLM Serving System
por: Liu, Lian, et al.
Publicado: (2026)
por: Liu, Lian, et al.
Publicado: (2026)
Towards Compute-Aware In-Switch Computing for LLMs Tensor-Parallelism on Multi-GPU Systems
por: Zhang, Chen, et al.
Publicado: (2026)
por: Zhang, Chen, et al.
Publicado: (2026)
FlexStep: Enabling Flexible Error Detection in Multi/Many-core Real-time Systems
por: Wang, Tinglue, et al.
Publicado: (2025)
por: Wang, Tinglue, et al.
Publicado: (2025)
Ejemplares similares
-
An Evaluation and Comparison of GPU Hardware and Solver Libraries for Accelerating the OPM Flow Reservoir Simulator
por: Qiu, Tong Dong, et al.
Publicado: (2023) -
CCSS: Hardware-Accelerated RTL Simulation with Fast Combinational Logic Computing and Sequential Logic Synchronization
por: Feng, Weigang, et al.
Publicado: (2025) -
Muchisim: A Simulation Framework for Design Exploration of Multi-Chip Manycore Systems
por: Orenes-Vera, Marcelo, et al.
Publicado: (2023) -
RISC-V Word-Size Modular Instructions for Residue Number Systems
por: Didier, Laurent-Stéphane, et al.
Publicado: (2024) -
Taming Offload Overheads in a Massively Parallel Open-Source RISC-V MPSoC: Analysis and Optimization
por: Colagrande, Luca, et al.
Publicado: (2025)