Hardware Acceleration of Kolmogorov-Arnold Network (KAN) in Large-Scale Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Wei-Hsing, Jia, Jianwei, Kong, Yuyao, Waqar, Faaiq, Wen, Tai-Hao, Chang, Meng-Fan, Yu, Shimeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hardware Acceleration of Kolmogorov-Arnold Network (KAN) for Lightweight Edge Inference
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2024)
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2024)
A3D-MoE: Acceleration of Large Language Models with Mixture of Experts via 3D Heterogeneous Integration
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
KAN-SAs: Efficient Acceleration of Kolmogorov-Arnold Networks on Systolic Arrays
von: Errabii, Sohaib, et al.
Veröffentlicht: (2025)
von: Errabii, Sohaib, et al.
Veröffentlicht: (2025)
BiKA: Kolmogorov-Arnold-Network-inspired Ultra Lightweight Neural Network Hardware Accelerator
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
3DGauCIM: Accelerating Static/Dynamic 3D Gaussian Splatting via Digital CIM for High Frame Rate Real-Time Edge Rendering
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)
Scalable Digital Compute-in-Memory Ising Machines for Robustness Verification of Binary Neural Networks
von: Vadlamani, Madhav, et al.
Veröffentlicht: (2026)
von: Vadlamani, Madhav, et al.
Veröffentlicht: (2026)
Architecting Long-Context LLM Acceleration with Packing-Prefetch Scheduler and Ultra-Large Capacity On-Chip Memories
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
Exploring the Limitations of Kolmogorov-Arnold Networks in Classification: Insights to Software Training and Hardware Implementation
von: Tran, Van Duy, et al.
Veröffentlicht: (2024)
von: Tran, Van Duy, et al.
Veröffentlicht: (2024)
CMOS+X: Stacking Persistent Embedded Memories based on Oxide Transistors upon GPGPU Platforms
von: Waqar, Faaiq, et al.
Veröffentlicht: (2025)
von: Waqar, Faaiq, et al.
Veröffentlicht: (2025)
A Case for Kolmogorov-Arnold Networks in Prefetching: Towards Low-Latency, Generalizable ML-Based Prefetchers
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
HAVEN: High-Bandwidth Flash Augmented Vector Engine for Large-Scale Approximate Nearest-Neighbor Search Acceleration
von: Hsu, Po-Kai, et al.
Veröffentlicht: (2026)
von: Hsu, Po-Kai, et al.
Veröffentlicht: (2026)
KANtize: Exploring Low-bit Quantization of Kolmogorov-Arnold Networks for Efficient Inference
von: Errabii, Sohaib, et al.
Veröffentlicht: (2026)
von: Errabii, Sohaib, et al.
Veröffentlicht: (2026)
Hardware Efficient Accelerator for Spiking Transformer With Reconfigurable Parallel Time Step Computing
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
FETTA: Flexible and Efficient Hardware Accelerator for Tensorized Neural Network Training
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
Large Language Model Inference Acceleration: A Comprehensive Hardware Perspective
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
Enabling Efficient Hardware Acceleration of Hybrid Vision Transformer (ViT) Networks at the Edge
von: Dumoulin, Joren, et al.
Veröffentlicht: (2025)
von: Dumoulin, Joren, et al.
Veröffentlicht: (2025)
SD-Acc: Accelerating Stable Diffusion through Phase-aware Sampling and Hardware Co-Optimizations
von: Wang, Zhican, et al.
Veröffentlicht: (2025)
von: Wang, Zhican, et al.
Veröffentlicht: (2025)
Hardware-Aware Neural Network Compilation with Learned Optimization: A RISC-V Accelerator Approach
von: Ganti, Ravindra, et al.
Veröffentlicht: (2025)
von: Ganti, Ravindra, et al.
Veröffentlicht: (2025)
Bombyx: OpenCilk Compilation for FPGA Hardware Acceleration
von: Shahawy, Mohamed, et al.
Veröffentlicht: (2025)
von: Shahawy, Mohamed, et al.
Veröffentlicht: (2025)
An Efficient Sparse Hardware Accelerator for Spike-Driven Transformer
von: Li, Zhengke, et al.
Veröffentlicht: (2025)
von: Li, Zhengke, et al.
Veröffentlicht: (2025)
System-Technology Co-Optimization of Bitline Routing and Bonding Pathways in Monolithic 3D DRAM Architectures
von: Lee, Kiseok, et al.
Veröffentlicht: (2026)
von: Lee, Kiseok, et al.
Veröffentlicht: (2026)
ChatNeuroSim: An LLM Agent Framework for Automated Compute-in-Memory Accelerator Deployment and Optimization
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2026)
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2026)
GSIM: Accelerating RTL Simulation for Large-Scale Designs
von: Chen, Lu, et al.
Veröffentlicht: (2025)
von: Chen, Lu, et al.
Veröffentlicht: (2025)
Overcoming Quadratic Hardware Scaling for a Fully Connected Digital Oscillatory Neural Network
von: Haverkort, Bram, et al.
Veröffentlicht: (2025)
von: Haverkort, Bram, et al.
Veröffentlicht: (2025)
Cross-Layer Design of Vector-Symbolic Computing: Bridging Cognition and Brain-Inspired Hardware Acceleration
von: Du, Shuting, et al.
Veröffentlicht: (2025)
von: Du, Shuting, et al.
Veröffentlicht: (2025)
HyDRA: Deadline and Reuse-Aware Cacheability for Hardware Accelerators
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
Energy-Efficient Hardware Acceleration of Whisper ASR on a CGLA
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
Hardware Acceleration in Portable MRIs: State of the Art and Future Prospects
von: Habsi, Omar Al, et al.
Veröffentlicht: (2025)
von: Habsi, Omar Al, et al.
Veröffentlicht: (2025)
Xpikeformer: Hybrid Analog-Digital Hardware Acceleration for Spiking Transformers
von: Song, Zihang, et al.
Veröffentlicht: (2024)
von: Song, Zihang, et al.
Veröffentlicht: (2024)
Realizing Hardware-Optimized General Tree-Based Data Structures for Heterogeneous System Classes
von: Biebert, Daniel, et al.
Veröffentlicht: (2025)
von: Biebert, Daniel, et al.
Veröffentlicht: (2025)
Proxima: Near-storage Acceleration for Graph-based Approximate Nearest Neighbor Search in 3D NAND
von: Xu, Weihong, et al.
Veröffentlicht: (2023)
von: Xu, Weihong, et al.
Veröffentlicht: (2023)
Ultrafast On-chip Online Learning via Spline Locality in Kolmogorov-Arnold Networks
von: Hoang, Duc, et al.
Veröffentlicht: (2026)
von: Hoang, Duc, et al.
Veröffentlicht: (2026)
Memory-Guided Unified Hardware Accelerator for Mixed-Precision Scientific Computing
von: Wang, Chuanzhen, et al.
Veröffentlicht: (2026)
von: Wang, Chuanzhen, et al.
Veröffentlicht: (2026)
TurboFuzz: FPGA Accelerated Hardware Fuzzing for Processor Agile Verification
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
HFRWKV: A High-Performance Fully On-Chip Hardware Accelerator for RWKV
von: Shijie, Liu, et al.
Veröffentlicht: (2026)
von: Shijie, Liu, et al.
Veröffentlicht: (2026)
SeDA: Secure and Efficient DNN Accelerators with Hardware/Software Synergy
von: Xuan, Wei, et al.
Veröffentlicht: (2025)
von: Xuan, Wei, et al.
Veröffentlicht: (2025)
Manticore: Hardware-Accelerated RTL Simulation with Static Bulk-Synchronous Parallelism
von: Emami, Mahyar, et al.
Veröffentlicht: (2023)
von: Emami, Mahyar, et al.
Veröffentlicht: (2023)
Design and Analysis of Approximate Hardware Accelerators for VVC Intra Angular Prediction
von: de Fraga, Lucas M. Leipnitz, et al.
Veröffentlicht: (2025)
von: de Fraga, Lucas M. Leipnitz, et al.
Veröffentlicht: (2025)
NeuroSim V1.5: Improved Software Backbone for Benchmarking Compute-in-Memory Accelerators with Device and Circuit-level Non-idealities
von: Read, James, et al.
Veröffentlicht: (2025)
von: Read, James, et al.
Veröffentlicht: (2025)
Hardware Accelerators for Autonomous Cars: A Review
von: Islayem, Ruba, et al.
Veröffentlicht: (2024)
von: Islayem, Ruba, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hardware Acceleration of Kolmogorov-Arnold Network (KAN) for Lightweight Edge Inference
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2024) -
A3D-MoE: Acceleration of Large Language Models with Mixture of Experts via 3D Heterogeneous Integration
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025) -
KAN-SAs: Efficient Acceleration of Kolmogorov-Arnold Networks on Systolic Arrays
von: Errabii, Sohaib, et al.
Veröffentlicht: (2025) -
BiKA: Kolmogorov-Arnold-Network-inspired Ultra Lightweight Neural Network Hardware Accelerator
von: Liu, Yuhao, et al.
Veröffentlicht: (2026) -
3DGauCIM: Accelerating Static/Dynamic 3D Gaussian Splatting via Digital CIM for High Frame Rate Real-Time Edge Rendering
von: Huang, Wei-Hsing, et al.
Veröffentlicht: (2025)