To Overlay or to Customize? Revisiting Architectural Choices in Heterogeneous Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xingzhen, Ji, Shixin, Dong, Zheng, Zhou, Peipei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DORA: Dataflow-Instruction Orchestration Architecture for DNN Acceleration
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026)
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026)
FILCO: Flexible Composing Architecture with Real-Time Reconfigurability for DNN Acceleration
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026)
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026)
PHAROS: Pipelined Heterogeneous Accelerators for Real-time Safety-critical Systems With Deadline Compliance
von: Ji, Shixin, et al.
Veröffentlicht: (2026)
von: Ji, Shixin, et al.
Veröffentlicht: (2026)
Challenges and Opportunities to Enable Large-Scale Computing via Heterogeneous Chiplets
von: Yang, Zhuoping, et al.
Veröffentlicht: (2023)
von: Yang, Zhuoping, et al.
Veröffentlicht: (2023)
SSR: Spatial Sequential Hybrid Architecture for Latency Throughput Tradeoff in Transformer Acceleration
von: Zhuang, Jinming, et al.
Veröffentlicht: (2024)
von: Zhuang, Jinming, et al.
Veröffentlicht: (2024)
FireFly-T: High-Throughput Sparsity Exploitation for Spiking Transformer Acceleration with Dual-Engine Overlay Architecture
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
The Survey of Chiplet-based Integrated Architecture: An EDA perspective
von: Chen, Shixin, et al.
Veröffentlicht: (2024)
von: Chen, Shixin, et al.
Veröffentlicht: (2024)
IMAGine: An In-Memory Accelerated GEMV Engine Overlay
von: Kabir, MD Arafat, et al.
Veröffentlicht: (2024)
von: Kabir, MD Arafat, et al.
Veröffentlicht: (2024)
Revisit Choice Network for Synthesis and Technology Mapping
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
SkipOPU: An FPGA-based Overlay Processor for Large Language Models with Dynamically Allocated Computation
von: He, Zicheng, et al.
Veröffentlicht: (2026)
von: He, Zicheng, et al.
Veröffentlicht: (2026)
Mixed Structural Choice Operator: Enhancing Technology Mapping with Heterogeneous Representations
von: Hu, Zhang, et al.
Veröffentlicht: (2025)
von: Hu, Zhang, et al.
Veröffentlicht: (2025)
Towards Generalized On-Chip Communication for Programmable Accelerators in Heterogeneous Architectures
von: Zuckerman, Joseph, et al.
Veröffentlicht: (2024)
von: Zuckerman, Joseph, et al.
Veröffentlicht: (2024)
LoopLynx: A Scalable Dataflow Architecture for Efficient LLM Inference
von: Zheng, Jianing, et al.
Veröffentlicht: (2025)
von: Zheng, Jianing, et al.
Veröffentlicht: (2025)
A Novel Cost-Effective MIMO Architecture with Ray Antenna Array for Enhanced Wireless Communication Performance
von: Dong, Zhenjun, et al.
Veröffentlicht: (2025)
von: Dong, Zhenjun, et al.
Veröffentlicht: (2025)
Optimized Memory System Architecture for VESA VDC-M Decoder with Multi-Slice Support
von: Yang, Hannah, et al.
Veröffentlicht: (2025)
von: Yang, Hannah, et al.
Veröffentlicht: (2025)
Hemlet: A Heterogeneous Compute-in-Memory Chiplet Architecture for Vision Transformers with Group-Level Parallelism
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
A System Architecture for Low Latency Multiprogramming Quantum Computing
von: Zhao, Yilun, et al.
Veröffentlicht: (2026)
von: Zhao, Yilun, et al.
Veröffentlicht: (2026)
Realizing Hardware-Optimized General Tree-Based Data Structures for Heterogeneous System Classes
von: Biebert, Daniel, et al.
Veröffentlicht: (2025)
von: Biebert, Daniel, et al.
Veröffentlicht: (2025)
THERMOS: Thermally-Aware Multi-Objective Scheduling of AI Workloads on Heterogeneous Multi-Chiplet PIM Architectures
von: Kanani, Alish, et al.
Veröffentlicht: (2025)
von: Kanani, Alish, et al.
Veröffentlicht: (2025)
LlamaF: An Efficient Llama2 Architecture Accelerator on Embedded FPGAs
von: Xu, Han, et al.
Veröffentlicht: (2024)
von: Xu, Han, et al.
Veröffentlicht: (2024)
Cohet: A CXL-Driven Coherent Heterogeneous Computing Framework with Hardware-Calibrated Full-System Simulation
von: Wang, Yanjing, et al.
Veröffentlicht: (2025)
von: Wang, Yanjing, et al.
Veröffentlicht: (2025)
MEEK: Re-thinking Heterogeneous Parallel Error Detection Architecture for Real-World OoO Superscalar Processors
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
MARVEL: An End-to-End Framework for Generating Model-Class Aware Custom RISC-V Extensions for Lightweight AI
von: M, Ajay Kumar, et al.
Veröffentlicht: (2025)
von: M, Ajay Kumar, et al.
Veröffentlicht: (2025)
Hermes: A Unified High-Performance NTT Architecture with Hybrid Dataflow
von: Gu, Hang, et al.
Veröffentlicht: (2026)
von: Gu, Hang, et al.
Veröffentlicht: (2026)
CAT: Customized Transformer Accelerator Framework on Versal ACAP
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
Evolution, Challenges, and Optimization in Computer Architecture: The Role of Reconfigurable Systems
von: Ederhion, Jefferson, et al.
Veröffentlicht: (2024)
von: Ederhion, Jefferson, et al.
Veröffentlicht: (2024)
Full System Architecture Modeling for Wearable Egocentric Contextual AI
von: Lee, Vincent T., et al.
Veröffentlicht: (2025)
von: Lee, Vincent T., et al.
Veröffentlicht: (2025)
FPGA-based Emulation and Device-Side Management for CXL-based Memory Tiering Systems
von: Chen, Yiqi, et al.
Veröffentlicht: (2025)
von: Chen, Yiqi, et al.
Veröffentlicht: (2025)
UpANNS: Enhancing Billion-Scale ANNS Efficiency with Real-World PIM Architecture
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
DreamRAM: A Fine-Grained Configurable Design Space Modeling Tool for Custom 3D Die-Stacked DRAM
von: Cai, Victor, et al.
Veröffentlicht: (2025)
von: Cai, Victor, et al.
Veröffentlicht: (2025)
Fast Generation of Custom Floating-Point Spatial Filters on FPGAs
von: Campos, Nelson, et al.
Veröffentlicht: (2024)
von: Campos, Nelson, et al.
Veröffentlicht: (2024)
Be CIM or Be Memory: A Dual-mode-aware DNN Compiler for CIM Accelerators
von: Zhao, Shixin, et al.
Veröffentlicht: (2025)
von: Zhao, Shixin, et al.
Veröffentlicht: (2025)
SLDB: An End-To-End Heterogeneous System-on-Chip Benchmark Suite for LLM-Aided Design
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
A Protocol-Independent Transport Architecture
von: Mohammadtaheri, Kimiya, et al.
Veröffentlicht: (2026)
von: Mohammadtaheri, Kimiya, et al.
Veröffentlicht: (2026)
FLASH-FHE: A Heterogeneous Architecture for Fully Homomorphic Encryption Acceleration
von: Zhang, Junxue, et al.
Veröffentlicht: (2025)
von: Zhang, Junxue, et al.
Veröffentlicht: (2025)
Automatic Microarchitecture-Aware Custom Instruction Design for RISC-V Processors
von: Rezunov, Evgenii, et al.
Veröffentlicht: (2025)
von: Rezunov, Evgenii, et al.
Veröffentlicht: (2025)
AGON: Automated Design Framework for Customizing Processors from ISA Documents
von: Li, Chongxiao, et al.
Veröffentlicht: (2024)
von: Li, Chongxiao, et al.
Veröffentlicht: (2024)
SpeedLLM: An FPGA Co-design of Large Language Model Inference Accelerator
von: Wang, Peipei, et al.
Veröffentlicht: (2025)
von: Wang, Peipei, et al.
Veröffentlicht: (2025)
HE^2: A Communication-Light Heterogeneous Architecture for Efficient Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2026)
von: Shi, Shangyi, et al.
Veröffentlicht: (2026)
AXI-REALM: Safe, Modular and Lightweight Traffic Monitoring and Regulation for Heterogeneous Mixed-Criticality Systems
von: Benz, Thomas, et al.
Veröffentlicht: (2025)
von: Benz, Thomas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DORA: Dataflow-Instruction Orchestration Architecture for DNN Acceleration
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026) -
FILCO: Flexible Composing Architecture with Real-Time Reconfigurability for DNN Acceleration
von: Chen, Xingzhen, et al.
Veröffentlicht: (2026) -
PHAROS: Pipelined Heterogeneous Accelerators for Real-time Safety-critical Systems With Deadline Compliance
von: Ji, Shixin, et al.
Veröffentlicht: (2026) -
Challenges and Opportunities to Enable Large-Scale Computing via Heterogeneous Chiplets
von: Yang, Zhuoping, et al.
Veröffentlicht: (2023) -
SSR: Spatial Sequential Hybrid Architecture for Latency Throughput Tradeoff in Transformer Acceleration
von: Zhuang, Jinming, et al.
Veröffentlicht: (2024)