RACAM: Enhancing DRAM with Reuse-Aware Computation and Automated Mapping for ML Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Siyuan, Hu, Jiajun, Ryoo, Jeeho, Arora, Aman, John, Lizy Kurian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A comparative study on power delivery aspects of compute-in/near-memory approaches using DRAM
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
GAMA: High-Performance GEMM Acceleration on AMD Versal ML-Optimized AI Engines
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2025)
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2025)
Regular-Dead on Arrival: Characterizing and Protecting Against Dead-Entry TLB Misses in GPU Microarchitectures
von: Anik, Shafayat Mowla, et al.
Veröffentlicht: (2026)
von: Anik, Shafayat Mowla, et al.
Veröffentlicht: (2026)
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
von: Hong, Junguk, et al.
Veröffentlicht: (2026)
von: Hong, Junguk, et al.
Veröffentlicht: (2026)
CarbonSet: A Dataset to Analyze Trends and Benchmark the Sustainability of CPUs and GPUs
von: Hu, Jiajun, et al.
Veröffentlicht: (2025)
von: Hu, Jiajun, et al.
Veröffentlicht: (2025)
Understanding Inference-Time Token Allocation and Coverage Limits in Agentic Hardware Verification
von: Patel, Vihaan, et al.
Veröffentlicht: (2026)
von: Patel, Vihaan, et al.
Veröffentlicht: (2026)
Shared-PIM: Enabling Concurrent Computation and Data Flow for Faster Processing-in-DRAM
von: Mamdouh, Ahmed, et al.
Veröffentlicht: (2024)
von: Mamdouh, Ahmed, et al.
Veröffentlicht: (2024)
Accelerating CRONet on AMD Versal AIE-ML Engines
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2026)
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2026)
GreenFPGA: Evaluating FPGAs as Environmentally Sustainable Computing Solutions
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2023)
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2023)
Shifting in-DRAM
von: Tegge, William C., et al.
Veröffentlicht: (2026)
von: Tegge, William C., et al.
Veröffentlicht: (2026)
GenDRAM:Hardware-Software Co-Design of General Platform in DRAM
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2026)
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2026)
CarbonPATH: Carbon-aware pathfinding and architecture optimization for chiplet-based AI systems
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2026)
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2026)
Sectored DRAM: A Practical Energy-Efficient and High-Performance Fine-Grained DRAM Architecture
von: Olgun, Ataberk, et al.
Veröffentlicht: (2022)
von: Olgun, Ataberk, et al.
Veröffentlicht: (2022)
Field-Programmable Gate Array Architecture for Deep Learning: Survey & Future Directions
von: Boutros, Andrew, et al.
Veröffentlicht: (2024)
von: Boutros, Andrew, et al.
Veröffentlicht: (2024)
EasyDRAM: An FPGA-based Infrastructure for Fast and Accurate End-to-End Evaluation of Emerging DRAM Techniques
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Canpolat, Oğuzhan, et al.
Veröffentlicht: (2025)
Sudoku: Decomposing DRAM Address Mapping into Component Functions
von: Wi, Minbok, et al.
Veröffentlicht: (2025)
von: Wi, Minbok, et al.
Veröffentlicht: (2025)
CHICO-Agent: An LLM Agent for the Cross-layer Optimization of 2.5D and 3D Chiplet-based Systems
von: Wu, Qihang, et al.
Veröffentlicht: (2026)
von: Wu, Qihang, et al.
Veröffentlicht: (2026)
Exploring DRAM Cache Prefetching for Pooled Memory
von: Tirumalasetty, Chandrahas, et al.
Veröffentlicht: (2024)
von: Tirumalasetty, Chandrahas, et al.
Veröffentlicht: (2024)
TDRAM: Tag-enhanced DRAM for Efficient Caching
von: Babaie, Maryam, et al.
Veröffentlicht: (2024)
von: Babaie, Maryam, et al.
Veröffentlicht: (2024)
A complete discussion on fully reconfigurable, digital, scalable, graph and sparsity-aware near-memory accelerator for graph neural networks
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
HyDRA: Deadline and Reuse-Aware Cacheability for Hardware Accelerators
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
ATiM: Autotuning Tensor Programs for Processing-in-DRAM
von: Shin, Yongwon, et al.
Veröffentlicht: (2024)
von: Shin, Yongwon, et al.
Veröffentlicht: (2024)
A comprehensive study on ILP acceleration accounting for sparsity, area, energy, data movement using near-memory architecture
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026)
SoMa: Identifying, Exploring, and Understanding the DRAM Communication Scheduling Space for DNN Accelerators
von: Cai, Jingwei, et al.
Veröffentlicht: (2025)
von: Cai, Jingwei, et al.
Veröffentlicht: (2025)
Evaluating Computing Platforms for Sustainability: A Comparative Analysis of FPGAs against ASICs, GPUs, and CPUs
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2026)
von: Sudarshan, Chetan Choppali, et al.
Veröffentlicht: (2026)
Rethinking the Producer-Consumer Relationship in Modern DRAM-Based Systems
von: Patel, Minesh, et al.
Veröffentlicht: (2024)
von: Patel, Minesh, et al.
Veröffentlicht: (2024)
Bandwidth-Effective DRAM Cache for GPUs with Storage-Class Memory
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
HLSFactory: A Framework Empowering High-Level Synthesis Datasets for Machine Learning and Beyond
von: Abi-Karam, Stefan, et al.
Veröffentlicht: (2024)
von: Abi-Karam, Stefan, et al.
Veröffentlicht: (2024)
DRAM-Profiler: An Experimental DRAM RowHammer Vulnerability Profiling Mechanism
von: Zhou, Ranyang, et al.
Veröffentlicht: (2024)
von: Zhou, Ranyang, et al.
Veröffentlicht: (2024)
Sangam: Chiplet-Based DRAM-PIM Accelerator with CXL Integration for LLM Inferencing
von: Kiyawat, Khyati, et al.
Veröffentlicht: (2025)
von: Kiyawat, Khyati, et al.
Veröffentlicht: (2025)
GEM3D CIM General Purpose Matrix Computation Using 3D Integrated SRAM eDRAM Hybrid Compute In Memory on Memory Architecture
von: Chakraborty, Subhradip, et al.
Veröffentlicht: (2026)
von: Chakraborty, Subhradip, et al.
Veröffentlicht: (2026)
Membrane: Accelerating Database Analytics with Bank-Level DRAM-PIM Filtering
von: Shekar, Akhil, et al.
Veröffentlicht: (2025)
von: Shekar, Akhil, et al.
Veröffentlicht: (2025)
Corrigendum to: A Systematic Study of DDR4 DRAM Faults in the Field
von: Beigi, Majed Valad, et al.
Veröffentlicht: (2024)
von: Beigi, Majed Valad, et al.
Veröffentlicht: (2024)
A Full-Stack Performance Evaluation Infrastructure for 3D-DRAM-based LLM Accelerators
von: Li, Cong, et al.
Veröffentlicht: (2026)
von: Li, Cong, et al.
Veröffentlicht: (2026)
Efficient Approaches for GEMM Acceleration on Leading AI-Optimized FPGAs
von: Taka, Endri, et al.
Veröffentlicht: (2024)
von: Taka, Endri, et al.
Veröffentlicht: (2024)
DPUConfig: Optimizing ML Inference in FPGAs Using Reinforcement Learning
von: Patras, Alexandros, et al.
Veröffentlicht: (2026)
von: Patras, Alexandros, et al.
Veröffentlicht: (2026)
Self-Managing DRAM: A Low-Cost Framework for Enabling Autonomous and Efficient in-DRAM Operations
von: Hassan, Hasan, et al.
Veröffentlicht: (2022)
von: Hassan, Hasan, et al.
Veröffentlicht: (2022)
A Logic-Reuse Approach to Nibble-based Multiplier Design for Low Power Vector Computing
von: Chowdhury, Md Rownak Hossain, et al.
Veröffentlicht: (2026)
von: Chowdhury, Md Rownak Hossain, et al.
Veröffentlicht: (2026)
PIM-FW: Hardware-Software Co-Design of All-pairs Shortest Paths in DRAM
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2025)
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2025)
Hardware-Software Co-design for 3D-DRAM-based LLM Serving Accelerator
von: Li, Cong, et al.
Veröffentlicht: (2026)
von: Li, Cong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A comparative study on power delivery aspects of compute-in/near-memory approaches using DRAM
von: Raman, Siddhartha Raman Sundara, et al.
Veröffentlicht: (2026) -
GAMA: High-Performance GEMM Acceleration on AMD Versal ML-Optimized AI Engines
von: Mhatre, Kaustubh, et al.
Veröffentlicht: (2025) -
Regular-Dead on Arrival: Characterizing and Protecting Against Dead-Entry TLB Misses in GPU Microarchitectures
von: Anik, Shafayat Mowla, et al.
Veröffentlicht: (2026) -
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
von: Hong, Junguk, et al.
Veröffentlicht: (2026) -
CarbonSet: A Dataset to Analyze Trends and Benchmark the Sustainability of CPUs and GPUs
von: Hu, Jiajun, et al.
Veröffentlicht: (2025)