G-Meta: Distributed Meta Learning in GPU Clusters for Large-Scale Recommender Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Youshao, Zhao, Shangchun, Zhou, Zhenglei, Huan, Zhaoxin, Ju, Lin, Zhang, Xiaolu, Wang, Lin, Zhou, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AntDT: A Self-Adaptive Distributed Training Framework for Leader and Straggler Nodes
by: Xiao, Youshao, et al.
Published: (2024)
by: Xiao, Youshao, et al.
Published: (2024)
AntBatchInfer: Elastic Batch Inference in the Kubernetes Cluster
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
ERCache: An Efficient and Reliable Caching Framework for Large-Scale User Representations in Meta's Ads System
by: Zhou, Fang, et al.
Published: (2024)
by: Zhou, Fang, et al.
Published: (2024)
SIVF: GPU-Resident IVF Index for Streaming Vector Search
by: Zhao, Dongfang
Published: (2026)
by: Zhao, Dongfang
Published: (2026)
Disaggregated Multi-Tower: Topology-aware Modeling Technique for Efficient Large-Scale Recommendation
by: Luo, Liang, et al.
Published: (2024)
by: Luo, Liang, et al.
Published: (2024)
GPU-accelerated Multi-relational Parallel Graph Retrieval for Web-scale Recommendations
by: Guo, Zhuoning, et al.
Published: (2025)
by: Guo, Zhuoning, et al.
Published: (2025)
A Model-agnostic Strategy to Mitigate Embedding Degradation in Personalized Federated Recommendation
by: Shen, Jiakui, et al.
Published: (2025)
by: Shen, Jiakui, et al.
Published: (2025)
Data Dams: A Novel Framework for Regulating and Managing Data Flow in Large-Scale Systems
by: Bouke, Mohamed Aly, et al.
Published: (2025)
by: Bouke, Mohamed Aly, et al.
Published: (2025)
FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost
by: Feng, Chenhao, et al.
Published: (2026)
by: Feng, Chenhao, et al.
Published: (2026)
Poplar: Efficient Scaling of Distributed DNN Training on Heterogeneous GPU Clusters
by: Zhang, WenZheng, et al.
Published: (2024)
by: Zhang, WenZheng, et al.
Published: (2024)
PIFS-Rec: Process-In-Fabric-Switch for Large-Scale Recommendation System Inferences
by: Huo, Pingyi, et al.
Published: (2024)
by: Huo, Pingyi, et al.
Published: (2024)
One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving
by: Yu, Wenjun, et al.
Published: (2026)
by: Yu, Wenjun, et al.
Published: (2026)
PRISM: Dynamic Primitive-Based Forecasting for Large-Scale GPU Cluster Workloads
by: Wu, Xin, et al.
Published: (2026)
by: Wu, Xin, et al.
Published: (2026)
RecIS: Sparse to Dense, A Unified Training Framework for Recommendation Models
by: Zong, Hua, et al.
Published: (2025)
by: Zong, Hua, et al.
Published: (2025)
Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
by: Lv, Zheqi, et al.
Published: (2025)
by: Lv, Zheqi, et al.
Published: (2025)
C-FedRAG: A Confidential Federated Retrieval-Augmented Generation System
by: Addison, Parker, et al.
Published: (2024)
by: Addison, Parker, et al.
Published: (2024)
ElasticRec: A Microservice-based Model Serving Architecture Enabling Elastic Resource Scaling for Recommendation Models
by: Choi, Yujeong, et al.
Published: (2024)
by: Choi, Yujeong, et al.
Published: (2024)
A Large-Scale Web Search Dataset for Federated Online Learning to Rank
by: Gregoriadis, Marcel, et al.
Published: (2025)
by: Gregoriadis, Marcel, et al.
Published: (2025)
Frenzy: A Memory-Aware Serverless LLM Training System for Heterogeneous GPU Clusters
by: Chang, Zihan, et al.
Published: (2024)
by: Chang, Zihan, et al.
Published: (2024)
FLAME: A Serving System Optimized for Large-Scale Generative Recommendation with Efficiency
by: Guo, Xianwen, et al.
Published: (2025)
by: Guo, Xianwen, et al.
Published: (2025)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
by: Mo, Zizhao, et al.
Published: (2025)
by: Mo, Zizhao, et al.
Published: (2025)
HAP: SPMD DNN Training on Heterogeneous GPU Clusters with Automated Program Synthesis
by: Zhang, Shiwei, et al.
Published: (2024)
by: Zhang, Shiwei, et al.
Published: (2024)
Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics
by: Yuan, Yichao, et al.
Published: (2025)
by: Yuan, Yichao, et al.
Published: (2025)
Intelligent Model Update Strategy for Sequential Recommendation
by: Lv, Zheqi, et al.
Published: (2023)
by: Lv, Zheqi, et al.
Published: (2023)
gZCCL: Compression-Accelerated Collective Communication Framework for GPU Clusters
by: Huang, Jiajun, et al.
Published: (2023)
by: Huang, Jiajun, et al.
Published: (2023)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
by: Lee, Munkyu, et al.
Published: (2024)
by: Lee, Munkyu, et al.
Published: (2024)
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory
by: Cheng, Rongxin, et al.
Published: (2024)
by: Cheng, Rongxin, et al.
Published: (2024)
DIET: Customized Slimming for Incompatible Networks in Sequential Recommendation
by: Fu, Kairui, et al.
Published: (2024)
by: Fu, Kairui, et al.
Published: (2024)
HETHUB: A Distributed Training System with Heterogeneous Cluster for Large-Scale Models
by: Xu, Si, et al.
Published: (2024)
by: Xu, Si, et al.
Published: (2024)
ColonyOS -- A Meta-Operating System for Distributed Computing Across Heterogeneous Platform
by: Kristiansson, Johan
Published: (2024)
by: Kristiansson, Johan
Published: (2024)
An Efficient, Reliable and Observable Collective Communication Library in Large-scale GPU Training Clusters
by: Zhang, Mingjun, et al.
Published: (2025)
by: Zhang, Mingjun, et al.
Published: (2025)
Semantic-Aware Scheduling for GPU Clusters with Large Language Models
by: Wang, Zerui, et al.
Published: (2025)
by: Wang, Zerui, et al.
Published: (2025)
From Data to Decisions: The Transformational Power of Machine Learning in Business Recommendations
by: Gangadharan, Kapilya, et al.
Published: (2024)
by: Gangadharan, Kapilya, et al.
Published: (2024)
GPU-Accelerated Distributed QAOA on Large-scale HPC Ecosystems
by: Xu, Zhihao, et al.
Published: (2025)
by: Xu, Zhihao, et al.
Published: (2025)
Fantasy: Efficient Large-scale Vector Search on GPU Clusters with GPUDirect Async
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
OnePiece: A Large-Scale Distributed Inference System with RDMA for Complex AI-Generated Content (AIGC) Workflows
by: Chen, June, et al.
Published: (2026)
by: Chen, June, et al.
Published: (2026)
Exact Nearest-Neighbor Search on Energy-Efficient FPGA Devices
by: Dazzi, Patrizio, et al.
Published: (2025)
by: Dazzi, Patrizio, et al.
Published: (2025)
A Big Data Architecture for Early Identification and Categorization of Dark Web Sites
by: Pastor-Galindo, Javier, et al.
Published: (2024)
by: Pastor-Galindo, Javier, et al.
Published: (2024)
An OPC UA-based industrial Big Data architecture
by: Hirsch, Eduard, et al.
Published: (2023)
by: Hirsch, Eduard, et al.
Published: (2023)
Predictable LLM Serving on GPU Clusters
by: Darzi, Erfan, et al.
Published: (2025)
by: Darzi, Erfan, et al.
Published: (2025)
Similar Items
-
AntDT: A Self-Adaptive Distributed Training Framework for Leader and Straggler Nodes
by: Xiao, Youshao, et al.
Published: (2024) -
AntBatchInfer: Elastic Batch Inference in the Kubernetes Cluster
by: Li, Siyuan, et al.
Published: (2024) -
ERCache: An Efficient and Reliable Caching Framework for Large-Scale User Representations in Meta's Ads System
by: Zhou, Fang, et al.
Published: (2024) -
SIVF: GPU-Resident IVF Index for Streaming Vector Search
by: Zhao, Dongfang
Published: (2026) -
Disaggregated Multi-Tower: Topology-aware Modeling Technique for Efficient Large-Scale Recommendation
by: Luo, Liang, et al.
Published: (2024)