GPU-accelerated Multi-relational Parallel Graph Retrieval for Web-scale Recommendations
Fuente:
arXiv
Salvato in:
| Autori principali: | Guo, Zhuoning, Chen, Guangxing, Gao, Qian, Liao, Xiaochao, Zheng, Jianjia, Shen, Lu, Liu, Hao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Model-agnostic Strategy to Mitigate Embedding Degradation in Personalized Federated Recommendation
di: Shen, Jiakui, et al.
Pubblicazione: (2025)
di: Shen, Jiakui, et al.
Pubblicazione: (2025)
G-Meta: Distributed Meta Learning in GPU Clusters for Large-Scale Recommender Systems
di: Xiao, Youshao, et al.
Pubblicazione: (2024)
di: Xiao, Youshao, et al.
Pubblicazione: (2024)
A Big Data Architecture for Early Identification and Categorization of Dark Web Sites
di: Pastor-Galindo, Javier, et al.
Pubblicazione: (2024)
di: Pastor-Galindo, Javier, et al.
Pubblicazione: (2024)
SIVF: GPU-Resident IVF Index for Streaming Vector Search
di: Zhao, Dongfang
Pubblicazione: (2026)
di: Zhao, Dongfang
Pubblicazione: (2026)
C-FedRAG: A Confidential Federated Retrieval-Augmented Generation System
di: Addison, Parker, et al.
Pubblicazione: (2024)
di: Addison, Parker, et al.
Pubblicazione: (2024)
HiFGL: A Hierarchical Framework for Cross-silo Cross-device Federated Graph Learning
di: Guo, Zhuoning, et al.
Pubblicazione: (2024)
di: Guo, Zhuoning, et al.
Pubblicazione: (2024)
Far From Sight, Far From Mind: Inverse Distance Weighting for Graph Federated Recommendation
di: Khouas, Aymen Rayane, et al.
Pubblicazione: (2025)
di: Khouas, Aymen Rayane, et al.
Pubblicazione: (2025)
EXaCTz: Guaranteed Extremum Graph and Contour Tree Preservation for Distributed- and GPU-Parallel Lossy Compression
di: Li, Yuxiao, et al.
Pubblicazione: (2026)
di: Li, Yuxiao, et al.
Pubblicazione: (2026)
Disaggregated Multi-Tower: Topology-aware Modeling Technique for Efficient Large-Scale Recommendation
di: Luo, Liang, et al.
Pubblicazione: (2024)
di: Luo, Liang, et al.
Pubblicazione: (2024)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
di: Mo, Zizhao, et al.
Pubblicazione: (2025)
di: Mo, Zizhao, et al.
Pubblicazione: (2025)
RecIS: Sparse to Dense, A Unified Training Framework for Recommendation Models
di: Zong, Hua, et al.
Pubblicazione: (2025)
di: Zong, Hua, et al.
Pubblicazione: (2025)
Fantasy: Efficient Large-scale Vector Search on GPU Clusters with GPUDirect Async
di: Liu, Yi, et al.
Pubblicazione: (2025)
di: Liu, Yi, et al.
Pubblicazione: (2025)
A GPU accelerated mixed-precision Smoothed Particle Hydrodynamics framework with cell-based relative coordinates
di: Mao, Zirui, et al.
Pubblicazione: (2023)
di: Mao, Zirui, et al.
Pubblicazione: (2023)
PiPNN: Ultra-Scalable Graph-Based Nearest Neighbor Indexing
di: Rubel, Tobias, et al.
Pubblicazione: (2026)
di: Rubel, Tobias, et al.
Pubblicazione: (2026)
A Context-Aware Knowledge Graph Platform for Stream Processing in Industrial IoT
di: Sciarroni, Monica Marconi, et al.
Pubblicazione: (2026)
di: Sciarroni, Monica Marconi, et al.
Pubblicazione: (2026)
Intelligent Model Update Strategy for Sequential Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2023)
di: Lv, Zheqi, et al.
Pubblicazione: (2023)
Combining GPU and CPU for accelerating evolutionary computing workloads
di: Eynaliyev, Rustam, et al.
Pubblicazione: (2025)
di: Eynaliyev, Rustam, et al.
Pubblicazione: (2025)
DIET: Customized Slimming for Incompatible Networks in Sequential Recommendation
di: Fu, Kairui, et al.
Pubblicazione: (2024)
di: Fu, Kairui, et al.
Pubblicazione: (2024)
SiPipe: Bridging the CPU-GPU Utilization Gap for Efficient Pipeline-Parallel LLM Inference
di: He, Yongchao, et al.
Pubblicazione: (2025)
di: He, Yongchao, et al.
Pubblicazione: (2025)
A Large-Scale Web Search Dataset for Federated Online Learning to Rank
di: Gregoriadis, Marcel, et al.
Pubblicazione: (2025)
di: Gregoriadis, Marcel, et al.
Pubblicazione: (2025)
From Data to Decisions: The Transformational Power of Machine Learning in Business Recommendations
di: Gangadharan, Kapilya, et al.
Pubblicazione: (2024)
di: Gangadharan, Kapilya, et al.
Pubblicazione: (2024)
Exact Nearest-Neighbor Search on Energy-Efficient FPGA Devices
di: Dazzi, Patrizio, et al.
Pubblicazione: (2025)
di: Dazzi, Patrizio, et al.
Pubblicazione: (2025)
An OPC UA-based industrial Big Data architecture
di: Hirsch, Eduard, et al.
Pubblicazione: (2023)
di: Hirsch, Eduard, et al.
Pubblicazione: (2023)
GRNND: A GPU-Parallel Relative NN-Descent Algorithm for Efficient Approximate Nearest Neighbor Graph Construction
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Forward Once for All: Structural Parameterized Adaptation for Efficient Cloud-coordinated On-device Recommendation
di: Fu, Kairui, et al.
Pubblicazione: (2025)
di: Fu, Kairui, et al.
Pubblicazione: (2025)
One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving
di: Yu, Wenjun, et al.
Pubblicazione: (2026)
di: Yu, Wenjun, et al.
Pubblicazione: (2026)
Efficient Federated Search for Retrieval-Augmented Generation using Lightweight Routing
di: Dhasade, Akash, et al.
Pubblicazione: (2025)
di: Dhasade, Akash, et al.
Pubblicazione: (2025)
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory
di: Cheng, Rongxin, et al.
Pubblicazione: (2024)
di: Cheng, Rongxin, et al.
Pubblicazione: (2024)
GMLake: Efficient and Transparent GPU Memory Defragmentation for Large-scale DNN Training with Virtual Memory Stitching
di: Guo, Cong, et al.
Pubblicazione: (2024)
di: Guo, Cong, et al.
Pubblicazione: (2024)
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
di: Medeiros, Daniel, et al.
Pubblicazione: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
di: Liang, Antian, et al.
Pubblicazione: (2025)
di: Liang, Antian, et al.
Pubblicazione: (2025)
Warp-STAR: High-performance, Differentiable GPU-Accelerated Static Timing Analysis through Warp-oriented Parallel Orchestration
di: Huang, En-Ming, et al.
Pubblicazione: (2026)
di: Huang, En-Ming, et al.
Pubblicazione: (2026)
Parallel Order-Based Core Maintenance in Dynamic Graphs
di: Guo, Bin, et al.
Pubblicazione: (2022)
di: Guo, Bin, et al.
Pubblicazione: (2022)
Accelerating Biclique Counting on GPU
di: Qiu, Linshan, et al.
Pubblicazione: (2024)
di: Qiu, Linshan, et al.
Pubblicazione: (2024)
ElasticRec: A Microservice-based Model Serving Architecture Enabling Elastic Resource Scaling for Recommendation Models
di: Choi, Yujeong, et al.
Pubblicazione: (2024)
di: Choi, Yujeong, et al.
Pubblicazione: (2024)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
di: Yi, Xinyao
Pubblicazione: (2024)
di: Yi, Xinyao
Pubblicazione: (2024)
GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
di: Yi, Xinyao, et al.
Pubblicazione: (2024)
di: Yi, Xinyao, et al.
Pubblicazione: (2024)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
di: Knorr, Fabian, et al.
Pubblicazione: (2025)
di: Knorr, Fabian, et al.
Pubblicazione: (2025)
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
di: Li, Zhonggen, et al.
Pubblicazione: (2025)
di: Li, Zhonggen, et al.
Pubblicazione: (2025)
Performance Evaluation of LLMs in Automated RDF Knowledge Graph Generation
di: Martin, Ioana Ramona, et al.
Pubblicazione: (2026)
di: Martin, Ioana Ramona, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Model-agnostic Strategy to Mitigate Embedding Degradation in Personalized Federated Recommendation
di: Shen, Jiakui, et al.
Pubblicazione: (2025) -
G-Meta: Distributed Meta Learning in GPU Clusters for Large-Scale Recommender Systems
di: Xiao, Youshao, et al.
Pubblicazione: (2024) -
A Big Data Architecture for Early Identification and Categorization of Dark Web Sites
di: Pastor-Galindo, Javier, et al.
Pubblicazione: (2024) -
SIVF: GPU-Resident IVF Index for Streaming Vector Search
di: Zhao, Dongfang
Pubblicazione: (2026) -
C-FedRAG: A Confidential Federated Retrieval-Augmented Generation System
di: Addison, Parker, et al.
Pubblicazione: (2024)