AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN Performance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kang, Seungkwan, Lee, Seungjun, Gouk, Donghyun, Kwon, Miryeong, Choi, Hyunkyu, Jang, Junhyeok, Lee, Sangwon, Choi, Huiwon, Zhang, Jie, Choi, Wonil, Kandemir, Mahmut Taylan, Jung, Myoungsoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Block to Byte: Transforming PCIe SSDs with CXL Memory Protocol and Instruction Annotation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025)
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025)
CXL Topology-Aware and Expander-Driven Prefetching: Unlocking SSD Performance
von: Oh, Dongsuk, et al.
Veröffentlicht: (2025)
von: Oh, Dongsuk, et al.
Veröffentlicht: (2025)
Diagnosing FP4 inference: a layer-wise and block-wise sensitivity analysis of NVFP4 and MXFP4
von: Cim, Musa, et al.
Veröffentlicht: (2026)
von: Cim, Musa, et al.
Veröffentlicht: (2026)
Mamba-X: An End-to-End Vision Mamba Accelerator for Edge Computing Devices
von: Yoon, Dongho, et al.
Veröffentlicht: (2025)
von: Yoon, Dongho, et al.
Veröffentlicht: (2025)
ApproxGNN: A Pretrained GNN for Parameter Prediction in Design Space Exploration for Approximate Computing
von: Vlcek, Ondrej, et al.
Veröffentlicht: (2025)
von: Vlcek, Ondrej, et al.
Veröffentlicht: (2025)
MCMComm: Hardware-Software Co-Optimization for End-to-End Communication in Multi-Chip-Modules
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
HSCO-Bench: An Agent-Driven End-to-End Hardware-Software Co-design Benchmark for Systems-on-Chip
von: Tsai, Pei-Huan, et al.
Veröffentlicht: (2026)
von: Tsai, Pei-Huan, et al.
Veröffentlicht: (2026)
SLDB: An End-To-End Heterogeneous System-on-Chip Benchmark Suite for LLM-Aided Design
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
von: Alvanaki, Elisavet Lydia, et al.
Veröffentlicht: (2025)
elasticAI.explorer: Towards a Unified End-to-End Framework for Hardware-Aware Neural Architecture Search
von: Maman, Natalie, et al.
Veröffentlicht: (2026)
von: Maman, Natalie, et al.
Veröffentlicht: (2026)
HARP: Hadamard-Domain Write-and-Verify for Noise-Robust RRAM Programming
von: Choi, Ilhuan, et al.
Veröffentlicht: (2026)
von: Choi, Ilhuan, et al.
Veröffentlicht: (2026)
CMAX-CAMEL: A Coarse-to-Fine Adaptive, Memory-Efficient, and Low-Power Edge Processor for Contrast Maximization
von: Min, Kyeongpil, et al.
Veröffentlicht: (2026)
von: Min, Kyeongpil, et al.
Veröffentlicht: (2026)
Accelerating GNN Training through Locality-aware Dropout and Merge
von: Sun, Gongjian, et al.
Veröffentlicht: (2025)
von: Sun, Gongjian, et al.
Veröffentlicht: (2025)
ApproxPilot: A GNN-based Accelerator Approximation Framework
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
A Host-SSD Collaborative Write Accelerator for LSM-Tree-Based Key-Value Stores
von: Kim, KiHwan, et al.
Veröffentlicht: (2024)
von: Kim, KiHwan, et al.
Veröffentlicht: (2024)
Low Latency GNN Accelerator for Quantum Error Correction
von: Cicero, Alessio, et al.
Veröffentlicht: (2026)
von: Cicero, Alessio, et al.
Veröffentlicht: (2026)
Hardware-Efficient Softmax and Layer Normalization with Guaranteed Normalization for Edge Devices
von: Choi, Dawon, et al.
Veröffentlicht: (2026)
von: Choi, Dawon, et al.
Veröffentlicht: (2026)
TriGen: NPU Architecture for End-to-End Acceleration of Large Language Models based on SW-HW Co-Design
von: Lee, Jonghun, et al.
Veröffentlicht: (2026)
von: Lee, Jonghun, et al.
Veröffentlicht: (2026)
Salient Store: Enabling Smart Storage for Continuous Learning Edge Servers
von: Mishra, Cyan Subhra, et al.
Veröffentlicht: (2024)
von: Mishra, Cyan Subhra, et al.
Veröffentlicht: (2024)
Piccolo: Large-Scale Graph Processing with Fine-Grained In-Memory Scatter-Gather
von: Shin, Changmin, et al.
Veröffentlicht: (2025)
von: Shin, Changmin, et al.
Veröffentlicht: (2025)
Efficient yet Accurate End-to-End SC Accelerator Design
von: Li, Meng, et al.
Veröffentlicht: (2024)
von: Li, Meng, et al.
Veröffentlicht: (2024)
PIMCOMP: An End-to-End DNN Compiler for Processing-In-Memory Accelerators
von: Sun, Xiaotian, et al.
Veröffentlicht: (2024)
von: Sun, Xiaotian, et al.
Veröffentlicht: (2024)
FIXME: Towards End-to-End Benchmarking of LLM-Aided Design Verification
von: Wan, Gwok-Waa, et al.
Veröffentlicht: (2025)
von: Wan, Gwok-Waa, et al.
Veröffentlicht: (2025)
Voyager: An End-to-End Framework for Design-Space Exploration and Generation of DNN Accelerators
von: Prabhu, Kartik, et al.
Veröffentlicht: (2025)
von: Prabhu, Kartik, et al.
Veröffentlicht: (2025)
EONSim: An NPU Simulator for On-Chip Memory and Embedding Vector Operations
von: Choi, Sangun, et al.
Veröffentlicht: (2025)
von: Choi, Sangun, et al.
Veröffentlicht: (2025)
FuncGNN: Learning Functional Semantics of Logic Circuits with Graph Neural Networks
von: Zhao, Qiyun
Veröffentlicht: (2025)
von: Zhao, Qiyun
Veröffentlicht: (2025)
DCI: A Coordinated Allocation and Filling Workload-Aware Dual-Cache Allocation GNN Inference Acceleration System
von: Luo, Yi, et al.
Veröffentlicht: (2025)
von: Luo, Yi, et al.
Veröffentlicht: (2025)
GenPairX: A Hardware-Algorithm Co-Designed Accelerator for Paired-End Read Mapping
von: Eudine, Julien, et al.
Veröffentlicht: (2026)
von: Eudine, Julien, et al.
Veröffentlicht: (2026)
SA-Kura: An Energy-Efficient Systolic Array Accelerator for Locally-Coupled Kuramoto Drift in Diffusion Sampling
von: Jin, Jeongmin, et al.
Veröffentlicht: (2026)
von: Jin, Jeongmin, et al.
Veröffentlicht: (2026)
Towards an End-To-End System for Real-Time Gesture Recognition from Surface Vibrations
von: Hettstedt, Florian, et al.
Veröffentlicht: (2026)
von: Hettstedt, Florian, et al.
Veröffentlicht: (2026)
MemIntelli: A Generic End-to-End Simulation Framework for Memristive Intelligent Computing
von: Zhou, Houji, et al.
Veröffentlicht: (2025)
von: Zhou, Houji, et al.
Veröffentlicht: (2025)
FASE: FPGA-Assisted Syscall Emulation for Rapid End-to-End Processor Performance Validation
von: Meng, Chengzhen, et al.
Veröffentlicht: (2025)
von: Meng, Chengzhen, et al.
Veröffentlicht: (2025)
System-Level Design Space Exploration for High-Level Synthesis under End-to-End Latency Constraints
von: Liao, Yuchao, et al.
Veröffentlicht: (2024)
von: Liao, Yuchao, et al.
Veröffentlicht: (2024)
T-MAN: Enabling End-to-End Low-Bit LLM Inference on NPUs via Unified Table Lookup
von: Wei, Jianyu, et al.
Veröffentlicht: (2025)
von: Wei, Jianyu, et al.
Veröffentlicht: (2025)
EasyDRAM: An FPGA-based Infrastructure for Fast and Accurate End-to-End Evaluation of Emerging DRAM Techniques
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2025)
CIMR-V: An End-to-End SRAM-based CIM Accelerator with RISC-V for AI Edge Device
von: and, Yan-Cheng Guo, et al.
Veröffentlicht: (2025)
von: and, Yan-Cheng Guo, et al.
Veröffentlicht: (2025)
SPADE: Sparse Pillar-based 3D Object Detection Accelerator for Autonomous Driving
von: Lee, Minjae, et al.
Veröffentlicht: (2023)
von: Lee, Minjae, et al.
Veröffentlicht: (2023)
Analyzing Performance Characteristics of PostgreSQL and MariaDB on NVMeVirt
von: Han, Juhee, et al.
Veröffentlicht: (2024)
von: Han, Juhee, et al.
Veröffentlicht: (2024)
End-to-End Transformer Acceleration Through Processing-in-Memory Architectures
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Block to Byte: Transforming PCIe SSDs with CXL Memory Protocol and Instruction Annotation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025) -
CXL-GPU: Pushing GPU Memory Boundaries with the Integration of CXL Technologies
von: Gouk, Donghyun, et al.
Veröffentlicht: (2025) -
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
von: Kwon, Miryeong, et al.
Veröffentlicht: (2025) -
CXL Topology-Aware and Expander-Driven Prefetching: Unlocking SSD Performance
von: Oh, Dongsuk, et al.
Veröffentlicht: (2025) -
Diagnosing FP4 inference: a layer-wise and block-wise sensitivity analysis of NVFP4 and MXFP4
von: Cim, Musa, et al.
Veröffentlicht: (2026)