When Do Graph Foundation Models Transfer? A Data-Centric Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Jiajun, Chen, Ying, Wang, Peihao, He, Yixuan, Li, Pan, Akella, Aditya, Wang, Zhangyang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Why Neural Network Can Discover Symbolic Structures with Gradient-based Training: An Algebraic and Geometric Foundation for Neurosymbolic Reasoning
by: Wang, Peihao, et al.
Published: (2025)
by: Wang, Peihao, et al.
Published: (2025)
Understanding and Mitigating Bottlenecks of State Space Models through the Lens of Recency and Over-smoothing
by: Wang, Peihao, et al.
Published: (2024)
by: Wang, Peihao, et al.
Published: (2024)
Graph-KV: Breaking Sequence via Injecting Structural Biases into Large Language Models
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Read-ME: Refactorizing LLMs as Router-Decoupled Mixture of Experts with System Co-Design
by: Cai, Ruisi, et al.
Published: (2024)
by: Cai, Ruisi, et al.
Published: (2024)
Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding
by: Zhu, Jiajun, et al.
Published: (2025)
by: Zhu, Jiajun, et al.
Published: (2025)
On-the-Fly Adaptive Distillation of Transformer to Dual-State Linear Attention
by: Ro, Yeonju, et al.
Published: (2025)
by: Ro, Yeonju, et al.
Published: (2025)
Generalization Error Analysis for Sparse Mixture-of-Experts: A Preliminary Study
by: Zhao, Jinze, et al.
Published: (2024)
by: Zhao, Jinze, et al.
Published: (2024)
Polynomial Width is Sufficient for Set Representation with High-dimensional Features
by: Wang, Peihao, et al.
Published: (2023)
by: Wang, Peihao, et al.
Published: (2023)
Position: Weight Space Should Be a First-Class Generative AI Modality
by: Wang, Zhangyang, et al.
Published: (2026)
by: Wang, Zhangyang, et al.
Published: (2026)
Improving the Throughput of Diffusion-based Large Language Models via a Training-Free Confidence-Aware Calibration
by: Shen, Jucheng, et al.
Published: (2025)
by: Shen, Jucheng, et al.
Published: (2025)
$\nabla$-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space
by: Wang, Peihao, et al.
Published: (2026)
by: Wang, Peihao, et al.
Published: (2026)
HALoS: Hierarchical Asynchronous Local SGD over Slow Networks for Geo-Distributed Large Language Model Training
by: Kim, Geon-Woo, et al.
Published: (2025)
by: Kim, Geon-Woo, et al.
Published: (2025)
Comparative Analysis of Time Series Foundation Models for Demographic Forecasting: Enhancing Predictive Accuracy in US Population Dynamics
by: Akella, Aditya, et al.
Published: (2025)
by: Akella, Aditya, et al.
Published: (2025)
Towards Graph Foundation Models: A Transferability Perspective
by: Wang, Yuxiang, et al.
Published: (2025)
by: Wang, Yuxiang, et al.
Published: (2025)
Sparse Mixture-of-Experts for Compositional Generalization: Empirical Evidence and Theoretical Foundations of Optimal Sparsity
by: Zhao, Jinze, et al.
Published: (2024)
by: Zhao, Jinze, et al.
Published: (2024)
GraphCLIP: Enhancing Transferability in Graph Foundation Models for Text-Attributed Graphs
by: Zhu, Yun, et al.
Published: (2024)
by: Zhu, Yun, et al.
Published: (2024)
Out-of-Distribution Generalization in Graph Foundation Models
by: Li, Haoyang, et al.
Published: (2026)
by: Li, Haoyang, et al.
Published: (2026)
Griffin: Towards a Graph-Centric Relational Database Foundation Model
by: Wang, Yanbo, et al.
Published: (2025)
by: Wang, Yanbo, et al.
Published: (2025)
Structure-Centric Graph Foundation Model via Geometric Bases
by: He, Xiaodong, et al.
Published: (2026)
by: He, Xiaodong, et al.
Published: (2026)
Meta ControlNet: Enhancing Task Adaptation via Meta Learning
by: Yang, Junjie, et al.
Published: (2023)
by: Yang, Junjie, et al.
Published: (2023)
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search
by: Robertson, John T., et al.
Published: (2026)
by: Robertson, John T., et al.
Published: (2026)
Towards Understanding Sensitive and Decisive Patterns in Explainable AI: A Case Study of Model Interpretation in Geometric Deep Learning
by: Zhu, Jiajun, et al.
Published: (2024)
by: Zhu, Jiajun, et al.
Published: (2024)
Data-Centric Foundation Models in Computational Healthcare: A Survey
by: Zhang, Yunkun, et al.
Published: (2024)
by: Zhang, Yunkun, et al.
Published: (2024)
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025)
by: Akella, Aditya
Published: (2025)
Foundations and Frontiers of Graph Learning Theory
by: Huang, Yu, et al.
Published: (2024)
by: Huang, Yu, et al.
Published: (2024)
CrossHGL: A Text-Free Foundation Model for Cross-Domain Heterogeneous Graph Learning
by: Chen, Xuanze, et al.
Published: (2026)
by: Chen, Xuanze, et al.
Published: (2026)
TCGU: Data-centric Graph Unlearning based on Transferable Condensation
by: Li, Fan, et al.
Published: (2024)
by: Li, Fan, et al.
Published: (2024)
A Survey of Generalization of Graph Anomaly Detection: From Transfer Learning to Foundation Models
by: Pan, Junjun, et al.
Published: (2025)
by: Pan, Junjun, et al.
Published: (2025)
Towards Graph Foundation Models: Training on Knowledge Graphs Enables Transferability to General Graphs
by: Wang, Kai, et al.
Published: (2024)
by: Wang, Kai, et al.
Published: (2024)
Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning
by: Wang, Peihao, et al.
Published: (2026)
by: Wang, Peihao, et al.
Published: (2026)
Bridging OOD Detection and Generalization: A Graph-Theoretic View
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
by: Bhatt, Neel P., et al.
Published: (2024)
by: Bhatt, Neel P., et al.
Published: (2024)
Data Wrangling Task Automation Using Code-Generating Language Models
by: Akella, Ashlesha, et al.
Published: (2025)
by: Akella, Ashlesha, et al.
Published: (2025)
MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D World
by: Hong, Yining, et al.
Published: (2024)
by: Hong, Yining, et al.
Published: (2024)
COMBA: Cross Batch Aggregation for Learning Large Graphs with Context Gating State Space Models
by: Shen, Jiajun, et al.
Published: (2026)
by: Shen, Jiajun, et al.
Published: (2026)
Learning Shortest Paths When Data is Scarce
by: Matsypura, Dmytro, et al.
Published: (2026)
by: Matsypura, Dmytro, et al.
Published: (2026)
GFT: Graph Foundation Model with Transferable Tree Vocabulary
by: Wang, Zehong, et al.
Published: (2024)
by: Wang, Zehong, et al.
Published: (2024)
Transferable and Forecastable User Targeting Foundation Model
by: Dou, Bin, et al.
Published: (2024)
by: Dou, Bin, et al.
Published: (2024)
Relation-Aware Graph Foundation Model
by: Yu, Jianxiang, et al.
Published: (2025)
by: Yu, Jianxiang, et al.
Published: (2025)
GLIP-OOD: Zero-Shot Graph OOD Detection with Graph Foundation Model
by: Xu, Haoyan, et al.
Published: (2025)
by: Xu, Haoyan, et al.
Published: (2025)
Similar Items
-
Why Neural Network Can Discover Symbolic Structures with Gradient-based Training: An Algebraic and Geometric Foundation for Neurosymbolic Reasoning
by: Wang, Peihao, et al.
Published: (2025) -
Understanding and Mitigating Bottlenecks of State Space Models through the Lens of Recency and Over-smoothing
by: Wang, Peihao, et al.
Published: (2024) -
Graph-KV: Breaking Sequence via Injecting Structural Biases into Large Language Models
by: Wang, Haoyu, et al.
Published: (2025) -
Read-ME: Refactorizing LLMs as Router-Decoupled Mixture of Experts with System Co-Design
by: Cai, Ruisi, et al.
Published: (2024) -
Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding
by: Zhu, Jiajun, et al.
Published: (2025)