ODIN-Based CPU-GPU Architecture with Replay-Driven Simulation and Emulation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Dorairaj, Nij, Chatterjee, Debabrata, Wang, Hong, Jiang, Hong, Saxena, Alankar, Koker, Altug, Lim, Thiam Ern, Teoh, Cathrane, Loo, Chuan Yin, Shomar, Bishara, Lester, Anthony
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912971192532992
author Dorairaj, Nij
Chatterjee, Debabrata
Wang, Hong
Jiang, Hong
Saxena, Alankar
Koker, Altug
Lim, Thiam Ern
Teoh, Cathrane
Loo, Chuan Yin
Shomar, Bishara
Lester, Anthony
author_facet Dorairaj, Nij
Chatterjee, Debabrata
Wang, Hong
Jiang, Hong
Saxena, Alankar
Koker, Altug
Lim, Thiam Ern
Teoh, Cathrane
Loo, Chuan Yin
Shomar, Bishara
Lester, Anthony
contents Integration of CPU and GPU technologies is a key enabler for modern AI and graphics workloads, combining control-oriented processing with massive parallel compute capability. As systems evolve toward chiplet-based architectures, pre-silicon validation of tightly coupled CPU-GPU subsystems becomes increasingly challenging due to complex validation framework setup, large design scale, high concurrency, non-deterministic execution, and intricate protocol interactions at chiplet boundaries, often resulting in long integration cycles. This paper presents a replay-driven validation methodology developed during the integration of a CPU subsystem, multiple Xe GPU cores, and a configurable Network-on-Chip (NoC) within a foundational SoC building block targeting the ODIN integrated chiplet architecture. By leveraging deterministic waveform capture and replay across both simulation and emulation using a single design database, complex GPU workloads and protocol sequences can be reproduced reliably at the system level. This approach significantly accelerates debug, improves integration confidence, and enables end-to-end system boot and workload execution within a single quarter, demonstrating the effectiveness of replay-based validation as a scalable methodology for chiplet-based systems.
format Preprint
id arxiv_https___arxiv_org_abs_2603_16812
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle ODIN-Based CPU-GPU Architecture with Replay-Driven Simulation and Emulation
Dorairaj, Nij
Chatterjee, Debabrata
Wang, Hong
Jiang, Hong
Saxena, Alankar
Koker, Altug
Lim, Thiam Ern
Teoh, Cathrane
Loo, Chuan Yin
Shomar, Bishara
Lester, Anthony
Distributed, Parallel, and Cluster Computing
Artificial Intelligence
Hardware Architecture
Integration of CPU and GPU technologies is a key enabler for modern AI and graphics workloads, combining control-oriented processing with massive parallel compute capability. As systems evolve toward chiplet-based architectures, pre-silicon validation of tightly coupled CPU-GPU subsystems becomes increasingly challenging due to complex validation framework setup, large design scale, high concurrency, non-deterministic execution, and intricate protocol interactions at chiplet boundaries, often resulting in long integration cycles. This paper presents a replay-driven validation methodology developed during the integration of a CPU subsystem, multiple Xe GPU cores, and a configurable Network-on-Chip (NoC) within a foundational SoC building block targeting the ODIN integrated chiplet architecture. By leveraging deterministic waveform capture and replay across both simulation and emulation using a single design database, complex GPU workloads and protocol sequences can be reproduced reliably at the system level. This approach significantly accelerates debug, improves integration confidence, and enables end-to-end system boot and workload execution within a single quarter, demonstrating the effectiveness of replay-based validation as a scalable methodology for chiplet-based systems.
title ODIN-Based CPU-GPU Architecture with Replay-Driven Simulation and Emulation
topic Distributed, Parallel, and Cluster Computing
Artificial Intelligence
Hardware Architecture
url https://arxiv.org/abs/2603.16812