Design of a GPU with Heterogeneous Cores for Graphics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tomás, Aurora, Aragón, Juan Luis, Parcerisa, Joan Manuel, González, Antonio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WaSP: Warp Scheduling to Mimic Prefetching in Graphics Workloads
von: Joseph, Diya, et al.
Veröffentlicht: (2024)
von: Joseph, Diya, et al.
Veröffentlicht: (2024)
e-GPU: An Open-Source and Configurable RISC-V Graphic Processing Unit for TinyAI Applications
von: Machetti, Simone, et al.
Veröffentlicht: (2025)
von: Machetti, Simone, et al.
Veröffentlicht: (2025)
Analyzing Modern NVIDIA GPU cores
von: Huerta, Rodrigo, et al.
Veröffentlicht: (2025)
von: Huerta, Rodrigo, et al.
Veröffentlicht: (2025)
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
Lightweight Congruence Profiling for Early Design Exploration of Heterogeneous FPGAs
von: Boston, Allen, et al.
Veröffentlicht: (2025)
von: Boston, Allen, et al.
Veröffentlicht: (2025)
RTGPU: Real-Time Computing with Graphics Processing Units
von: Gheibi-Fetrat, Atiyeh, et al.
Veröffentlicht: (2025)
von: Gheibi-Fetrat, Atiyeh, et al.
Veröffentlicht: (2025)
Confidential Computing on Heterogeneous CPU-GPU Systems: Survey and Future Directions
von: Wang, Qifan, et al.
Veröffentlicht: (2024)
von: Wang, Qifan, et al.
Veröffentlicht: (2024)
Designing Reconfigurable Interconnection Network of Heterogeneous Chiplets Using Kalman Filter
von: Biglari, Siamak, et al.
Veröffentlicht: (2024)
von: Biglari, Siamak, et al.
Veröffentlicht: (2024)
MemExplorer: Navigating the Heterogeneous Memory Design Space for Agentic Inference NPUs
von: Wu, Haoran, et al.
Veröffentlicht: (2026)
von: Wu, Haoran, et al.
Veröffentlicht: (2026)
Stream: Design Space Exploration of Layer-Fused DNNs on Heterogeneous Dataflow Accelerators
von: Symons, Arne, et al.
Veröffentlicht: (2022)
von: Symons, Arne, et al.
Veröffentlicht: (2022)
RoboGPU: Accelerating GPU Collision Detection for Robotics
von: Liu, Lufei, et al.
Veröffentlicht: (2026)
von: Liu, Lufei, et al.
Veröffentlicht: (2026)
Benchmarking and Dissecting the Nvidia Hopper GPU Architecture
von: Luo, Weile, et al.
Veröffentlicht: (2024)
von: Luo, Weile, et al.
Veröffentlicht: (2024)
COOK Access Control on an embedded Volta GPU
von: Lesage, Benjamin, et al.
Veröffentlicht: (2024)
von: Lesage, Benjamin, et al.
Veröffentlicht: (2024)
Designing Approximate Arithmetic Circuits with Combined Error Constraints
von: Češka, Milan, et al.
Veröffentlicht: (2022)
von: Češka, Milan, et al.
Veröffentlicht: (2022)
EdgeMM: Multi-Core CPU with Heterogeneous AI-Extension and Activation-aware Weight Pruning for Multimodal LLMs at Edge
von: Bai, Kangbo, et al.
Veröffentlicht: (2025)
von: Bai, Kangbo, et al.
Veröffentlicht: (2025)
SLDB: An End-To-End Heterogeneous System-on-Chip Benchmark Suite for LLM-Aided Design
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
CuLifter: Lifting GPU Binaries to Typed IR
von: Zhao, Jisheng, et al.
Veröffentlicht: (2026)
von: Zhao, Jisheng, et al.
Veröffentlicht: (2026)
Multiport Support for Vortex OpenGPU Memory Hierarchy
von: Shin, Injae, et al.
Veröffentlicht: (2025)
von: Shin, Injae, et al.
Veröffentlicht: (2025)
Siracusa: A 16 nm Heterogenous RISC-V SoC for Extended Reality with At-MRAM Neural Engine
von: Prasad, Arpan Suravi, et al.
Veröffentlicht: (2023)
von: Prasad, Arpan Suravi, et al.
Veröffentlicht: (2023)
Empirical Measurements of AI Training Power Demand on a GPU-Accelerated Node
von: Latif, Imran, et al.
Veröffentlicht: (2024)
von: Latif, Imran, et al.
Veröffentlicht: (2024)
A Prototype-Based Framework to Design Scalable Heterogeneous SoCs with Fine-Grained DFS
von: Montanaro, Gabriele, et al.
Veröffentlicht: (2024)
von: Montanaro, Gabriele, et al.
Veröffentlicht: (2024)
Thermal Analysis for NVIDIA GTX480 Fermi GPU Architecture
von: Nagendra, Savinay
Veröffentlicht: (2024)
von: Nagendra, Savinay
Veröffentlicht: (2024)
CMD: A Cache-assisted GPU Memory Deduplication Architecture
von: Zhao, Wei, et al.
Veröffentlicht: (2024)
von: Zhao, Wei, et al.
Veröffentlicht: (2024)
GAP-LA: GPU-Accelerated Performance-Driven Layer Assignment
von: Zhao, Chunyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Chunyuan, et al.
Veröffentlicht: (2025)
The Anatomy of Silent Data Corruption: GPU Error Pattern Study and Modeling Guidance
von: Tung, Chung-Hsuan, et al.
Veröffentlicht: (2026)
von: Tung, Chung-Hsuan, et al.
Veröffentlicht: (2026)
EnergAIzer: Fast and Accurate GPU Power Estimation Framework for AI Workloads
von: Lee, Kyungmi, et al.
Veröffentlicht: (2026)
von: Lee, Kyungmi, et al.
Veröffentlicht: (2026)
Towards Performance-Aware Allocation for Accelerated Machine Learning on GPU-SSD Systems
von: Gundawar, Ayush, et al.
Veröffentlicht: (2024)
von: Gundawar, Ayush, et al.
Veröffentlicht: (2024)
Parallelizing a modern GPU simulator
von: Huerta, Rodrigo, et al.
Veröffentlicht: (2025)
von: Huerta, Rodrigo, et al.
Veröffentlicht: (2025)
Scaling Photonic Tensor Cores with Unary and Homodyne Designs
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments
von: Guan, Yue, et al.
Veröffentlicht: (2026)
von: Guan, Yue, et al.
Veröffentlicht: (2026)
Edge GPU Aware Multiple AI Model Pipeline for Accelerated MRI Reconstruction and Analysis
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
Hardware vs. Software Implementation of Warp-Level Features in Vortex RISC-V GPU
von: Pu, Huanzhi, et al.
Veröffentlicht: (2025)
von: Pu, Huanzhi, et al.
Veröffentlicht: (2025)
GPU-Accelerated Simulated Oscillator Ising/Potts Machine Solving Combinatorial Optimization Problems
von: Gonul, Yilmaz Ege, et al.
Veröffentlicht: (2025)
von: Gonul, Yilmaz Ege, et al.
Veröffentlicht: (2025)
A RISC-V Multicore and GPU SoC Platform with a Qualifiable Software Stack for Safety Critical Systems
von: Bonet, Marc Solé i, et al.
Veröffentlicht: (2025)
von: Bonet, Marc Solé i, et al.
Veröffentlicht: (2025)
LLM-PRISM: Characterizing Silent Data Corruption from Permanent GPU Faults in LLM Training
von: Tyagi, Abhishek, et al.
Veröffentlicht: (2026)
von: Tyagi, Abhishek, et al.
Veröffentlicht: (2026)
Evaluation of GPU Video Encoder for Low-Latency Real-Time 4K UHD Encoding
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation
von: Zhou, Yaoyun, et al.
Veröffentlicht: (2025)
von: Zhou, Yaoyun, et al.
Veröffentlicht: (2025)
A Comparison of the Cerebras Wafer-Scale Integration Technology with Nvidia GPU-based Systems for Artificial Intelligence
von: Kundu, Yudhishthira, et al.
Veröffentlicht: (2025)
von: Kundu, Yudhishthira, et al.
Veröffentlicht: (2025)
PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference
von: Gu, Yufeng, et al.
Veröffentlicht: (2025)
von: Gu, Yufeng, et al.
Veröffentlicht: (2025)
Invited Paper: FEMU: An Open-Source and Configurable Emulation Framework for Prototyping TinyAI Heterogeneous Systems
von: Machetti, Simone, et al.
Veröffentlicht: (2025)
von: Machetti, Simone, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
WaSP: Warp Scheduling to Mimic Prefetching in Graphics Workloads
von: Joseph, Diya, et al.
Veröffentlicht: (2024) -
e-GPU: An Open-Source and Configurable RISC-V Graphic Processing Unit for TinyAI Applications
von: Machetti, Simone, et al.
Veröffentlicht: (2025) -
Analyzing Modern NVIDIA GPU cores
von: Huerta, Rodrigo, et al.
Veröffentlicht: (2025) -
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025) -
Lightweight Congruence Profiling for Early Design Exploration of Heterogeneous FPGAs
von: Boston, Allen, et al.
Veröffentlicht: (2025)