Mestra: Exploring Migration on Virtualized CGRAs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kyriazis, Agamemnon, Miliadis, Panagiotis, Theodoropoulos, Dimitris, Koziris, Nectarios, Pnevmatikatos, Dionisios |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A WASM-Subset Stack Architecture for Low-cost FPGAs using Open-Source EDA Flows
von: Chakrabarti, Aradhya
Veröffentlicht: (2025)
von: Chakrabarti, Aradhya
Veröffentlicht: (2025)
Design and Implementation of a RISC-V SoC with Custom DSP Accelerators for Edge Computing
von: Yadav, Priyanshu
Veröffentlicht: (2025)
von: Yadav, Priyanshu
Veröffentlicht: (2025)
RayFlex: An Open-Source RTL Implementation of the Hardware Ray Tracer Datapath
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
SISA: A Scale-In Systolic Array for GEMM Acceleration
von: Altamura, Luigi, et al.
Veröffentlicht: (2026)
von: Altamura, Luigi, et al.
Veröffentlicht: (2026)
FREESS: An Educational Simulator of a RISC-V-Inspired Superscalar Processor Based on Tomasulo's Algorithm
von: Giorgi, Roberto
Veröffentlicht: (2025)
von: Giorgi, Roberto
Veröffentlicht: (2025)
SambaNova SN40L: Scaling the AI Memory Wall with Dataflow and Composition of Experts
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
Aurora: Architecting Argonne's First Exascale Supercomputer for Accelerated Scientific Discovery
von: Allcock, William E., et al.
Veröffentlicht: (2025)
von: Allcock, William E., et al.
Veröffentlicht: (2025)
ASTER: Attention-based Spiking Transformer Engine for Event-driven Reasoning
von: Das, Tamoghno, et al.
Veröffentlicht: (2025)
von: Das, Tamoghno, et al.
Veröffentlicht: (2025)
MEDEA: A Design-Time Multi-Objective Manager for Energy-Efficient DNN Inference on Heterogeneous Ultra-Low Power Platforms
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
Exploring the Design Space for Message-Driven Systems for Dynamic Graph Processing using CCA
von: Chandio, Bibrak Qamar, et al.
Veröffentlicht: (2024)
von: Chandio, Bibrak Qamar, et al.
Veröffentlicht: (2024)
Fast and Practical Strassen's Matrix Multiplication using FPGAs
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
Make LLM Inference Affordable to Everyone: Augmenting GPU Memory with NDP-DIMM
von: Liu, Lian, et al.
Veröffentlicht: (2025)
von: Liu, Lian, et al.
Veröffentlicht: (2025)
Application-Driven Exascale: The JUPITER Benchmark Suite
von: Herten, Andreas, et al.
Veröffentlicht: (2024)
von: Herten, Andreas, et al.
Veröffentlicht: (2024)
SynapticCore-X: A Modular Neural Processing Architecture for Low-Cost FPGA Acceleration
von: Parameshwara, Arya
Veröffentlicht: (2025)
von: Parameshwara, Arya
Veröffentlicht: (2025)
Ember: A Compiler for Efficient Embedding Operations on Decoupled Access-Execute Architectures
von: Siracusa, Marco, et al.
Veröffentlicht: (2025)
von: Siracusa, Marco, et al.
Veröffentlicht: (2025)
FlexiBit: Fully Flexible Precision Bit-parallel Accelerator Architecture for Arbitrary Mixed Precision AI
von: Tahmasebi, Faraz, et al.
Veröffentlicht: (2024)
von: Tahmasebi, Faraz, et al.
Veröffentlicht: (2024)
Multi-diseases detection with memristive system on chip
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Streamlining SIMD ISA Extensions with Takum Arithmetic: A Case Study on Intel AVX10.2
von: Hunhold, Laslo
Veröffentlicht: (2025)
von: Hunhold, Laslo
Veröffentlicht: (2025)
DSPE: An Energy-Efficient Edge Processor for DeepSeek Inference with MerkleTree-based Incremental Pruning, Multi-Stage Boothing Lookup and Dynamic Adaptive Posit Processing
von: Zhang, Yuhan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuhan, et al.
Veröffentlicht: (2026)
FPGA-Accelerated RISC-V ISA Extensions for Efficient Neural Network Inference on Edge Devices
von: Parameshwara, Arya, et al.
Veröffentlicht: (2025)
von: Parameshwara, Arya, et al.
Veröffentlicht: (2025)
SparseZipper: Enhancing Matrix Extensions to Accelerate SpGEMM on CPUs
von: Ta, Tuan, et al.
Veröffentlicht: (2025)
von: Ta, Tuan, et al.
Veröffentlicht: (2025)
DARE: An Irregularity-Tolerant Matrix Processing Unit with a Densifying ISA and Filtered Runahead Execution
von: Yang, Xin, et al.
Veröffentlicht: (2025)
von: Yang, Xin, et al.
Veröffentlicht: (2025)
Taming Wild Branches: Overcoming Hard-to-Predict Branches using the Bullseye Predictor
von: Behrendt, Emet, et al.
Veröffentlicht: (2025)
von: Behrendt, Emet, et al.
Veröffentlicht: (2025)
OpenEye: A Scalable Open-Source Hardware Accelerator for DNNs
von: Lebold, Denis, et al.
Veröffentlicht: (2026)
von: Lebold, Denis, et al.
Veröffentlicht: (2026)
How long can you sleep? Idle Time System Inefficiencies and Opportunities
von: Antoniou, Georgia, et al.
Veröffentlicht: (2025)
von: Antoniou, Georgia, et al.
Veröffentlicht: (2025)
A Comparative Analysis of ARM and x86-64 Laptop-Class Processors: Architecture, Assembly-Level Performance, and Energy Efficiency
von: Özyılmaz, Mustafa Mert
Veröffentlicht: (2026)
von: Özyılmaz, Mustafa Mert
Veröffentlicht: (2026)
Exploring GPU-to-GPU Communication: Insights into Supercomputer Interconnects
von: De Sensi, Daniele, et al.
Veröffentlicht: (2024)
von: De Sensi, Daniele, et al.
Veröffentlicht: (2024)
Architecting Long-Context LLM Acceleration with Packing-Prefetch Scheduler and Ultra-Large Capacity On-Chip Memories
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
HERMES: High-Performance RISC-V Memory Hierarchy for ML Workloads
von: Suryadevara, Pranav
Veröffentlicht: (2025)
von: Suryadevara, Pranav
Veröffentlicht: (2025)
Improved Prefetching Techniques for Linked Data Structures
von: Maruszewski, Nikola Vuk
Veröffentlicht: (2025)
von: Maruszewski, Nikola Vuk
Veröffentlicht: (2025)
RV-IM100: Quantifying ISA Extension, Datapath Width, and Pipeline Depth Trade-offs in RISC-V Microarchitectures
von: Kang, Hyunwoo
Veröffentlicht: (2026)
von: Kang, Hyunwoo
Veröffentlicht: (2026)
basic_RV32s: An Open-Source Microarchitectural Roadmap for RISC-V RV32I
von: Kang, Hyun Woo, et al.
Veröffentlicht: (2025)
von: Kang, Hyun Woo, et al.
Veröffentlicht: (2025)
Systolic Arrays and Structured Pruning Co-design for Efficient Transformers in Edge Systems
von: Palacios, Pedro, et al.
Veröffentlicht: (2024)
von: Palacios, Pedro, et al.
Veröffentlicht: (2024)
GPU-centric Communication Schemes for HPC and ML Applications
von: Namashivayam, Naveen
Veröffentlicht: (2025)
von: Namashivayam, Naveen
Veröffentlicht: (2025)
FREESS: A Web-Based Educational Simulator for a RISC-V-Inspired Superscalar Processor with Tomasulo-Style Dynamic Scheduling
von: Giorgi, Roberto, et al.
Veröffentlicht: (2026)
von: Giorgi, Roberto, et al.
Veröffentlicht: (2026)
Biological Intuition on Digital Hardware: An RTL Implementation of Poisson-Encoded SNNs for Static Image Classification
von: Das, Debabrata, et al.
Veröffentlicht: (2026)
von: Das, Debabrata, et al.
Veröffentlicht: (2026)
Guess-Verify-Refine: Data-Aware Top-K for Sparse-Attention Decoding on Blackwell via Temporal Correlation
von: Cheng, Long, et al.
Veröffentlicht: (2026)
von: Cheng, Long, et al.
Veröffentlicht: (2026)
pLUTo: Enabling Massively Parallel Computation in DRAM via Lookup Tables
von: Ferreira, João Dinis, et al.
Veröffentlicht: (2021)
von: Ferreira, João Dinis, et al.
Veröffentlicht: (2021)
A Compilation Framework for Quantum Circuits with Mid-Circuit Measurement Error Awareness
von: Zhong, Ming, et al.
Veröffentlicht: (2025)
von: Zhong, Ming, et al.
Veröffentlicht: (2025)
Tekum: Balanced Ternary Tapered Precision Real Arithmetic
von: Hunhold, Laslo
Veröffentlicht: (2025)
von: Hunhold, Laslo
Veröffentlicht: (2025)
Ähnliche Einträge
-
A WASM-Subset Stack Architecture for Low-cost FPGAs using Open-Source EDA Flows
von: Chakrabarti, Aradhya
Veröffentlicht: (2025) -
Design and Implementation of a RISC-V SoC with Custom DSP Accelerators for Edge Computing
von: Yadav, Priyanshu
Veröffentlicht: (2025) -
RayFlex: An Open-Source RTL Implementation of the Hardware Ray Tracer Datapath
von: Shen, Fangjia, et al.
Veröffentlicht: (2024) -
SISA: A Scale-In Systolic Array for GEMM Acceleration
von: Altamura, Luigi, et al.
Veröffentlicht: (2026) -
FREESS: An Educational Simulator of a RISC-V-Inspired Superscalar Processor Based on Tomasulo's Algorithm
von: Giorgi, Roberto
Veröffentlicht: (2025)