Enhancing Scalability in Recommender Systems through Lottery Ticket Hypothesis and Knowledge Distillation-based Neural Network Pruning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | R, Rajaram, Bharadhwaj, Manoj, VS, Vasan, Pervin, Nargis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RAG-Enhanced Kernel-Based Heuristic Synthesis (RKHS): A Structured Methodology Using Large Language Models for Hardware Design
von: Ahir, Shiva, et al.
Veröffentlicht: (2026)
von: Ahir, Shiva, et al.
Veröffentlicht: (2026)
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
Pushing the Performance Envelope of DNN-based Recommendation Systems Inference on GPUs
von: Jain, Rishabh, et al.
Veröffentlicht: (2024)
von: Jain, Rishabh, et al.
Veröffentlicht: (2024)
PIFS-Rec: Process-In-Fabric-Switch for Large-Scale Recommendation System Inferences
von: Huo, Pingyi, et al.
Veröffentlicht: (2024)
von: Huo, Pingyi, et al.
Veröffentlicht: (2024)
Digital Signal Processing from Classical Coherent Systems to Continuous-Variable QKD: A Review of Cross-Domain Techniques, Applications, and Challenges
von: de Sousa, Davi Juvêncio Gomes, et al.
Veröffentlicht: (2025)
von: de Sousa, Davi Juvêncio Gomes, et al.
Veröffentlicht: (2025)
Accelerating Recommendation System Training by Leveraging Popular Choices
von: Adnan, Muhammad, et al.
Veröffentlicht: (2021)
von: Adnan, Muhammad, et al.
Veröffentlicht: (2021)
D2S-FLOW: Automated Parameter Extraction from Datasheets for SPICE Model Generation Using Large Language Models
von: Chen, Hong Cai, et al.
Veröffentlicht: (2025)
von: Chen, Hong Cai, et al.
Veröffentlicht: (2025)
The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency
von: Geens, Robin, et al.
Veröffentlicht: (2026)
von: Geens, Robin, et al.
Veröffentlicht: (2026)
Real-Time Adaptive Neural Network on FPGA: Enhancing Adaptability through Dynamic Classifier Selection
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
CIMPool: Scalable Neural Network Acceleration for Compute-In-Memory using Weight Pools
von: Li, Shurui, et al.
Veröffentlicht: (2025)
von: Li, Shurui, et al.
Veröffentlicht: (2025)
KLiNQ: Knowledge Distillation-Assisted Lightweight Neural Network for Qubit Readout on FPGA
von: Guo, Xiaorang, et al.
Veröffentlicht: (2025)
von: Guo, Xiaorang, et al.
Veröffentlicht: (2025)
Table-Lookup MAC: Scalable Processing of Quantised Neural Networks in FPGA Soft Logic
von: Gerlinghoff, Daniel, et al.
Veröffentlicht: (2024)
von: Gerlinghoff, Daniel, et al.
Veröffentlicht: (2024)
Towards High-Performance Network Coding: FPGA Acceleration With Bounded-value Generators
von: Qing, Jiaxin, et al.
Veröffentlicht: (2025)
von: Qing, Jiaxin, et al.
Veröffentlicht: (2025)
Enhancing LUT-based Deep Neural Networks Inference through Architecture and Connectivity Optimization
von: Lou, Binglei, et al.
Veröffentlicht: (2026)
von: Lou, Binglei, et al.
Veröffentlicht: (2026)
Unrolled and Pipelined Decoders based on Look-Up Tables for Polar Codes
von: Giard, Pascal, et al.
Veröffentlicht: (2023)
von: Giard, Pascal, et al.
Veröffentlicht: (2023)
Lottery BP: Unlocking Quantum Error Decoding at Scale
von: Zhu, Yanzhang, et al.
Veröffentlicht: (2026)
von: Zhu, Yanzhang, et al.
Veröffentlicht: (2026)
LRSCwait: Enabling Scalable and Efficient Synchronization in Manycore Systems through Polling-Free and Retry-Free Operation
von: Riedel, Samuel, et al.
Veröffentlicht: (2024)
von: Riedel, Samuel, et al.
Veröffentlicht: (2024)
Accelerating Retrieval-Augmented Generation
von: Quinn, Derrick, et al.
Veröffentlicht: (2024)
von: Quinn, Derrick, et al.
Veröffentlicht: (2024)
FPGA Resource-aware Structured Pruning for Real-Time Neural Networks
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2023)
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2023)
Look-Up Table based Neural Network Hardware
von: Sen, Ovishake, et al.
Veröffentlicht: (2024)
von: Sen, Ovishake, et al.
Veröffentlicht: (2024)
Bit-Flip Fault Attack: Crushing Graph Neural Networks via Gradual Bit Search
von: Abharian, Sanaz Kazemi, et al.
Veröffentlicht: (2025)
von: Abharian, Sanaz Kazemi, et al.
Veröffentlicht: (2025)
PDF: PUF-based DNN Fingerprinting for Knowledge Distillation Traceability
von: Lyu, Ning, et al.
Veröffentlicht: (2026)
von: Lyu, Ning, et al.
Veröffentlicht: (2026)
A Survey on LUT-based Deep Neural Networks Implemented in FPGAs
von: Guo, Zeyu
Veröffentlicht: (2025)
von: Guo, Zeyu
Veröffentlicht: (2025)
ProactivePIM: Accelerating Weight-Sharing Embedding Layer with PIM for Scalable Recommendation System
von: Kim, Youngsuk, et al.
Veröffentlicht: (2024)
von: Kim, Youngsuk, et al.
Veröffentlicht: (2024)
Hecaton: Training Large Language Models with Scalable Chiplet Systems
von: Huang, Zongle, et al.
Veröffentlicht: (2024)
von: Huang, Zongle, et al.
Veröffentlicht: (2024)
E-Commerce Product Recommendation System based on ML Algorithms
von: Haque, Md. Zahurul
Veröffentlicht: (2024)
von: Haque, Md. Zahurul
Veröffentlicht: (2024)
Aquas: Enhancing Domain Specialization through Holistic Hardware-Software Co-Optimization based on MLIR
von: Zou, Yuyang, et al.
Veröffentlicht: (2025)
von: Zou, Yuyang, et al.
Veröffentlicht: (2025)
SLTarch: Towards Scalable Point-Based Neural Rendering by Taming Workload Imbalance and Memory Irregularity
von: Li, Xingyang, et al.
Veröffentlicht: (2025)
von: Li, Xingyang, et al.
Veröffentlicht: (2025)
AutoRAC: Automated Processing-in-Memory Accelerator Design for Recommender Systems
von: Cheng, Feng, et al.
Veröffentlicht: (2025)
von: Cheng, Feng, et al.
Veröffentlicht: (2025)
GEN-Graph: Heterogeneous PIM Accelerator for General Computational Patterns in Graph-based Dynamic Programming
von: Chen, Yanru, et al.
Veröffentlicht: (2026)
von: Chen, Yanru, et al.
Veröffentlicht: (2026)
Voxel-CIM: An Efficient Compute-in-Memory Accelerator for Voxel-based Point Cloud Neural Networks
von: Lin, Xipeng, et al.
Veröffentlicht: (2024)
von: Lin, Xipeng, et al.
Veröffentlicht: (2024)
Unraveling codes: fast, robust, beyond-bound error correction for DRAM
von: Hamburg, Mike, et al.
Veröffentlicht: (2024)
von: Hamburg, Mike, et al.
Veröffentlicht: (2024)
Codes for Metastability-Containing Addition
von: Bund, Johannes, et al.
Veröffentlicht: (2026)
von: Bund, Johannes, et al.
Veröffentlicht: (2026)
Non-Binary LDPC Arithmetic Error Correction For Processing-in-Memory
von: Shi, Daijing, et al.
Veröffentlicht: (2025)
von: Shi, Daijing, et al.
Veröffentlicht: (2025)
Towards Reliable Systems: A Scalable Approach to AXI4 Transaction Monitoring
von: Liang, Chaoqun, et al.
Veröffentlicht: (2025)
von: Liang, Chaoqun, et al.
Veröffentlicht: (2025)
SCRec: A Scalable Computational Storage System with Statistical Sharding and Tensor-train Decomposition for Recommendation Models
von: Yang, Jinho, et al.
Veröffentlicht: (2025)
von: Yang, Jinho, et al.
Veröffentlicht: (2025)
FsimNNs: An Open-Source Graph Neural Network Platform for SEU Simulation-based Fault Injection
von: Lu, Li, et al.
Veröffentlicht: (2025)
von: Lu, Li, et al.
Veröffentlicht: (2025)
A Case for Kolmogorov-Arnold Networks in Prefetching: Towards Low-Latency, Generalizable ML-Based Prefetchers
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
von: Kulkarni, Dhruv, et al.
Veröffentlicht: (2025)
RecFlash: Fast Recommendation System on In-Storage Computing with Frequency-Based Data Mapping
von: Baik, Jangho, et al.
Veröffentlicht: (2026)
von: Baik, Jangho, et al.
Veröffentlicht: (2026)
Scalable and Efficient Intra- and Inter-node Interconnection Networks for Post-Exascale Supercomputers and Data centers
von: Tarraga-Moreno, Joaquin, et al.
Veröffentlicht: (2025)
von: Tarraga-Moreno, Joaquin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RAG-Enhanced Kernel-Based Heuristic Synthesis (RKHS): A Structured Methodology Using Large Language Models for Hardware Design
von: Ahir, Shiva, et al.
Veröffentlicht: (2026) -
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026) -
Pushing the Performance Envelope of DNN-based Recommendation Systems Inference on GPUs
von: Jain, Rishabh, et al.
Veröffentlicht: (2024) -
PIFS-Rec: Process-In-Fabric-Switch for Large-Scale Recommendation System Inferences
von: Huo, Pingyi, et al.
Veröffentlicht: (2024) -
Digital Signal Processing from Classical Coherent Systems to Continuous-Variable QKD: A Review of Cross-Domain Techniques, Applications, and Challenges
von: de Sousa, Davi Juvêncio Gomes, et al.
Veröffentlicht: (2025)