Billion-Scale Graph Foundation Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bechler-Speicher, Maya, Gottlieb, Yoel, Isakov, Andrey, Abensur, David, Tavory, Ami, Haimovich, Daniel, Guy, Ido, Weinsberg, Udi
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917518594015232
author Bechler-Speicher, Maya
Gottlieb, Yoel
Isakov, Andrey
Abensur, David
Tavory, Ami
Haimovich, Daniel
Guy, Ido
Weinsberg, Udi
author_facet Bechler-Speicher, Maya
Gottlieb, Yoel
Isakov, Andrey
Abensur, David
Tavory, Ami
Haimovich, Daniel
Guy, Ido
Weinsberg, Udi
contents Graph-structured data underpins many critical applications. While foundation models have transformed language and vision via large-scale pretraining and lightweight adaptation, extending this paradigm to general, real-world graphs is challenging. In this work, we present Graph Billion-Foundation-Fusion (GraphBFF): an end-to-end recipe for building billion-parameter Graph Foundation Models (GFMs) for large-scale heterogeneous graphs. Central to the recipe is the GraphBFF Transformer, a flexible and scalable architecture designed for practical billion-scale GFMs. Using the GraphBFF, we present neural scaling laws for heterogeneous graphs and show that loss decreases predictably as either model capacity or training data scales, depending on which factor is the bottleneck. The GraphBFF framework provides concrete methodologies for data batching, pretraining, and fine-tuning for building GFMs at scale. We demonstrate the effectiveness of the framework over a real-world billion-scale graph, with an evaluation of a billion-parameter GraphBFF Transformer following the proposed recipe. Across ten diverse, real-world downstream tasks on graphs unseen during training, spanning node- and link-level classification and regression, GraphBFF consistently outperforms baselines, with large margins of up to 31 PRAUC points, including in few-shot settings. Finally, we discuss key challenges and open opportunities for making GFMs a practical and principled foundation for graph learning at industrial scale.
format Preprint
id arxiv_https___arxiv_org_abs_2602_04768
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Billion-Scale Graph Foundation Models
Bechler-Speicher, Maya
Gottlieb, Yoel
Isakov, Andrey
Abensur, David
Tavory, Ami
Haimovich, Daniel
Guy, Ido
Weinsberg, Udi
Machine Learning
Artificial Intelligence
Graph-structured data underpins many critical applications. While foundation models have transformed language and vision via large-scale pretraining and lightweight adaptation, extending this paradigm to general, real-world graphs is challenging. In this work, we present Graph Billion-Foundation-Fusion (GraphBFF): an end-to-end recipe for building billion-parameter Graph Foundation Models (GFMs) for large-scale heterogeneous graphs. Central to the recipe is the GraphBFF Transformer, a flexible and scalable architecture designed for practical billion-scale GFMs. Using the GraphBFF, we present neural scaling laws for heterogeneous graphs and show that loss decreases predictably as either model capacity or training data scales, depending on which factor is the bottleneck. The GraphBFF framework provides concrete methodologies for data batching, pretraining, and fine-tuning for building GFMs at scale. We demonstrate the effectiveness of the framework over a real-world billion-scale graph, with an evaluation of a billion-parameter GraphBFF Transformer following the proposed recipe. Across ten diverse, real-world downstream tasks on graphs unseen during training, spanning node- and link-level classification and regression, GraphBFF consistently outperforms baselines, with large margins of up to 31 PRAUC points, including in few-shot settings. Finally, we discuss key challenges and open opportunities for making GFMs a practical and principled foundation for graph learning at industrial scale.
title Billion-Scale Graph Foundation Models
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2602.04768