IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
Fuente:
arXiv
Saved in:
| Main Authors: | Seo, Minseok, Nguyen, Xuan Truong, Hwang, Seok Joong, Kwon, Yongkee, Kim, Guhyun, Park, Chanwook, Kim, Ilkon, Park, Jaehan, Kim, Jeongbin, Shin, Woojae, Won, Jongsoon, Choi, Haerang, Kim, Kyuyoung, Kwon, Daehan, Jeong, Chunseok, Lee, Sangheon, Choi, Yongseok, Byun, Wooseok, Baek, Seungcheol, Lee, Hyuk-Jae, Kim, John |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PIMphony: Overcoming Bandwidth and Capacity Inefficiency in PIM-based Long-Context LLM Inference System
by: Kwon, Hyucksung, et al.
Published: (2024)
by: Kwon, Hyucksung, et al.
Published: (2024)
Darwin: A DRAM-based Multi-level Processing-in-Memory Architecture for Data Analytics
by: Kim, Donghyuk, et al.
Published: (2023)
by: Kim, Donghyuk, et al.
Published: (2023)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
by: Heo, Guseul, et al.
Published: (2024)
by: Heo, Guseul, et al.
Published: (2024)
CryptoGuard: Lightweight Hybrid Detection and Response to Host-based Cryptojackers in Linux Cloud Environments
by: Park, Gyeonghoon, et al.
Published: (2025)
by: Park, Gyeonghoon, et al.
Published: (2025)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
by: Park, Seungcheol, et al.
Published: (2025)
by: Park, Seungcheol, et al.
Published: (2025)
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
by: Hong, Junguk, et al.
Published: (2026)
by: Hong, Junguk, et al.
Published: (2026)
VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?
by: Kim, Minkyu, et al.
Published: (2026)
by: Kim, Minkyu, et al.
Published: (2026)
Improving Cyclability and Structural Stability of Co‐Free Layered Cathode by Controlling Porosity and Cracks in Secondary Particles for Low‐Cost and High‐Energy LIBs
by: Myungeun Choi, et al.
Published: (2025)
by: Myungeun Choi, et al.
Published: (2025)
Floquet Chern Insulators and Radiation-Induced Zero Resistance in Irradiated Graphene
by: Kim, Youngjae, et al.
Published: (2025)
by: Kim, Youngjae, et al.
Published: (2025)
Double-Side Polarization and Beamforming Alignment in Polarization Reconfigurable MISO System with Deep Neural Networks
by: Oh, Seungcheol, et al.
Published: (2024)
by: Oh, Seungcheol, et al.
Published: (2024)
Polarization Reconfigurable Transmit-Receive Beam Alignment with Interpretable Transformer
by: Oh, Seungcheol, et al.
Published: (2025)
by: Oh, Seungcheol, et al.
Published: (2025)
Bond-Strength-Based Understanding of Oxygen Vacancy Migration Barriers in Rutile Oxides
by: Kim, Inseo, et al.
Published: (2026)
by: Kim, Inseo, et al.
Published: (2026)
Adversarial Reinforcement Learning Framework for ESP Cheater Simulation
by: Park, Inkyu, et al.
Published: (2025)
by: Park, Inkyu, et al.
Published: (2025)
Woojae Kim
by: Woojae Kim
Published: (2025)
by: Woojae Kim
Published: (2025)
Woojae Kim
by: Woojae Kim
Published: (2025)
by: Woojae Kim
Published: (2025)
Acceleration of Grokking in Learning Arithmetic Operations via Kolmogorov-Arnold Representation
by: Park, Yeachan, et al.
Published: (2024)
by: Park, Yeachan, et al.
Published: (2024)
LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
by: Park, Gunho, et al.
Published: (2022)
by: Park, Gunho, et al.
Published: (2022)
From Tokens to Photons: Test-Time Physical Prompting for Vision-Language Models
by: Im, Boyeong, et al.
Published: (2025)
by: Im, Boyeong, et al.
Published: (2025)
Mixed Non-linear Quantization for Vision Transformers
by: Kim, Gihwan, et al.
Published: (2024)
by: Kim, Gihwan, et al.
Published: (2024)
BIPED: Pedagogically Informed Tutoring System for ESL Education
by: Kwon, Soonwoo, et al.
Published: (2024)
by: Kwon, Soonwoo, et al.
Published: (2024)
Lightweight Wasserstein Audio-Visual Model for Unified Speech Enhancement and Separation
by: Park, Jisoo, et al.
Published: (2025)
by: Park, Jisoo, et al.
Published: (2025)
Empowering Personalized Learning through a Conversation-based Tutoring System with Student Modeling
by: Park, Minju, et al.
Published: (2024)
by: Park, Minju, et al.
Published: (2024)
Controllable 3D Molecular Generation for Structure-Based Drug Design Through Bayesian Flow Networks and Gradient Integration
by: Choi, Seungyeon, et al.
Published: (2025)
by: Choi, Seungyeon, et al.
Published: (2025)
Unraveling reaction discrepancy and electrolyte stabilizing effects of auto‐oxygenated porphyrin catalysts in lithium–oxygen and lithium–air cells
by: Boran Kim, et al.
Published: (2024)
by: Boran Kim, et al.
Published: (2024)
PIM-SHERPA: Software Method for On-device LLM Inference by Resolving PIM Memory Attribute and Layout Inconsistencies
by: Lee, Sunjung, et al.
Published: (2026)
by: Lee, Sunjung, et al.
Published: (2026)
Sheaf Graph Neural Networks via PAC-Bayes Spectral Optimization
by: Choi, Yoonhyuk, et al.
Published: (2025)
by: Choi, Yoonhyuk, et al.
Published: (2025)
Adaptive Branch Specialization in Spectral-Spatial Graph Neural Networks for Certified Robustness
by: Choi, Yoonhyuk, et al.
Published: (2025)
by: Choi, Yoonhyuk, et al.
Published: (2025)
User Experience in Metaverse Libraries: Lessons from Four Cases
by: Yumi Kim, et al.
Published: (2025)
by: Yumi Kim, et al.
Published: (2025)
READRetro: natural product biosynthesis predicting with retrieval‐augmented dual‐view retrosynthesis
by: Taein Kim, et al.
Published: (2024)
by: Taein Kim, et al.
Published: (2024)
MAGE: All-[MASK] Block Already Knows Where to Look in Diffusion LLM
by: Kwon, Omin, et al.
Published: (2026)
by: Kwon, Omin, et al.
Published: (2026)
Personalized Language Models via Privacy-Preserving Evolutionary Model Merging
by: Kim, Kyuyoung, et al.
Published: (2025)
by: Kim, Kyuyoung, et al.
Published: (2025)
REDUCED GRAPHITE OXIDE-INDIUM TIN OXIDE COMPOSITES FOR TRANSPARENT ELECTRODE USING SOLUTION PROCESS
by: K. S. Choi, et al.
Published: (2011)
by: K. S. Choi, et al.
Published: (2011)
The Iterative Chainlet Partitioning Algorithm for the Traveling Salesman Problem with Drone and Neural Acceleration
by: Lee, Jae Hyeok, et al.
Published: (2025)
by: Lee, Jae Hyeok, et al.
Published: (2025)
ABM-LoRA: Activation Boundary Matching for Fast Convergence in Low-Rank Adaptation
by: Lee, Dongha, et al.
Published: (2025)
by: Lee, Dongha, et al.
Published: (2025)
Presidential agendas in statutory interpretation: A case study of the Ministry of Government Legislation of Korea (MGLK)
by: Hyang‐mi Kim, et al.
Published: (2024)
by: Hyang‐mi Kim, et al.
Published: (2024)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
Crystallographic defects in Weyl semimetal LaAlGe
by: Kim, Inseo, et al.
Published: (2024)
by: Kim, Inseo, et al.
Published: (2024)
Implementation of Magic State Injection within Heavy-Hexagon Architecture
by: Kim, Hansol, et al.
Published: (2024)
by: Kim, Hansol, et al.
Published: (2024)
Holi-DETR: Holistic Fashion Item Detection Leveraging Contextual Information
by: Kwon, Youngchae, et al.
Published: (2025)
by: Kwon, Youngchae, et al.
Published: (2025)
Item Region-based Style Classification Network (IRSN): A Fashion Style Classifier Based on Domain Knowledge of Fashion Experts
by: Choi, Jinyoung, et al.
Published: (2025)
by: Choi, Jinyoung, et al.
Published: (2025)
Similar Items
-
PIMphony: Overcoming Bandwidth and Capacity Inefficiency in PIM-based Long-Context LLM Inference System
by: Kwon, Hyucksung, et al.
Published: (2024) -
Darwin: A DRAM-based Multi-level Processing-in-Memory Architecture for Data Analytics
by: Kim, Donghyuk, et al.
Published: (2023) -
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
by: Heo, Guseul, et al.
Published: (2024) -
CryptoGuard: Lightweight Hybrid Detection and Response to Host-based Cryptojackers in Linux Cloud Environments
by: Park, Gyeonghoon, et al.
Published: (2025) -
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
by: Park, Seungcheol, et al.
Published: (2025)