Relational Learning in Pre-Trained Models: A Theory from Hypergraph Recovery Perspective
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Yang, Fang, Cong, Lin, Zhouchen, Liu, Bing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DIGIC: Domain Generalizable Imitation Learning by Causal Discovery
di: Chen, Yang, et al.
Pubblicazione: (2024)
di: Chen, Yang, et al.
Pubblicazione: (2024)
Conda: Column-Normalized Adam for Training Large Language Models Faster
di: Wang, Junjie, et al.
Pubblicazione: (2025)
di: Wang, Junjie, et al.
Pubblicazione: (2025)
On the Limitations and Capabilities of Position Embeddings for Length Generalization
di: Chen, Yang, et al.
Pubblicazione: (2025)
di: Chen, Yang, et al.
Pubblicazione: (2025)
Improving Model Representation and Reducing KV Cache via Skip Connections with First Value Heads
di: Wu, Zhoutong, et al.
Pubblicazione: (2025)
di: Wu, Zhoutong, et al.
Pubblicazione: (2025)
Attention Sinks Induce Gradient Sinks: Massive Activations as Gradient Regulators in Transformers
di: Chen, Yihong, et al.
Pubblicazione: (2026)
di: Chen, Yihong, et al.
Pubblicazione: (2026)
Training-Free Message Passing for Learning on Hypergraphs
di: Tang, Bohan, et al.
Pubblicazione: (2024)
di: Tang, Bohan, et al.
Pubblicazione: (2024)
On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery
di: Liu, Renpu, et al.
Pubblicazione: (2024)
di: Liu, Renpu, et al.
Pubblicazione: (2024)
Quadratic Direct Forecast for Training Multi-Step Time-Series Forecast Models
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
Link Prediction with Relational Hypergraphs
di: Huang, Xingyue, et al.
Pubblicazione: (2024)
di: Huang, Xingyue, et al.
Pubblicazione: (2024)
Hyperbolic Hypergraph Neural Networks for Multi-Relational Knowledge Hypergraph Representation
di: Li, Mengfan, et al.
Pubblicazione: (2024)
di: Li, Mengfan, et al.
Pubblicazione: (2024)
Generalist++: A Meta-learning Framework for Mitigating Trade-off in Adversarial Training
di: Wang, Yisen, et al.
Pubblicazione: (2025)
di: Wang, Yisen, et al.
Pubblicazione: (2025)
Contrastive Language-Image Pre-Training Model based Semantic Communication Performance Optimization
di: Yang, Shaoran, et al.
Pubblicazione: (2025)
di: Yang, Shaoran, et al.
Pubblicazione: (2025)
Enhanced Atrial Fibrillation Prediction in ESUS Patients with Hypergraph-based Pre-training
di: Xie, Yuzhang, et al.
Pubblicazione: (2026)
di: Xie, Yuzhang, et al.
Pubblicazione: (2026)
Protecting Copyright of Medical Pre-trained Language Models: Training-Free Backdoor Model Watermarking
di: Kong, Cong, et al.
Pubblicazione: (2024)
di: Kong, Cong, et al.
Pubblicazione: (2024)
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models at the Edge
di: Hu, Senkang, et al.
Pubblicazione: (2025)
di: Hu, Senkang, et al.
Pubblicazione: (2025)
Proximity Matters: Local Proximity Enhanced Balancing for Treatment Effect Estimation
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
PTMs-TSCIL Pre-Trained Models Based Class-Incremental Learning
di: Wu, Yuanlong, et al.
Pubblicazione: (2025)
di: Wu, Yuanlong, et al.
Pubblicazione: (2025)
CyclicFL: A Cyclic Model Pre-Training Approach to Efficient Federated Learning
di: Zhang, Pengyu, et al.
Pubblicazione: (2023)
di: Zhang, Pengyu, et al.
Pubblicazione: (2023)
Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning Models
di: Javanmard, Adel, et al.
Pubblicazione: (2026)
di: Javanmard, Adel, et al.
Pubblicazione: (2026)
FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models
di: Liu, Junkang, et al.
Pubblicazione: (2025)
di: Liu, Junkang, et al.
Pubblicazione: (2025)
A Survey on Time-Series Pre-Trained Models
di: Ma, Qianli, et al.
Pubblicazione: (2023)
di: Ma, Qianli, et al.
Pubblicazione: (2023)
HypergraphFormer: Learning Hypergraphs from LLMs for Editable Floor Plan Generation
di: Klimenko, Nikita, et al.
Pubblicazione: (2026)
di: Klimenko, Nikita, et al.
Pubblicazione: (2026)
ADORA: Training Reasoning Models with Dynamic Advantage Estimation on Reinforcement Learning
di: Ren, Qingnan, et al.
Pubblicazione: (2026)
di: Ren, Qingnan, et al.
Pubblicazione: (2026)
On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective
di: Zhou, Zhi, et al.
Pubblicazione: (2026)
di: Zhou, Zhi, et al.
Pubblicazione: (2026)
Online Pseudo-Zeroth-Order Training of Neuromorphic Spiking Neural Networks
di: Xiao, Mingqing, et al.
Pubblicazione: (2024)
di: Xiao, Mingqing, et al.
Pubblicazione: (2024)
Simple Convergence Proof of Adam From a Sign-like Descent Perspective
di: Peng, Hanyang, et al.
Pubblicazione: (2025)
di: Peng, Hanyang, et al.
Pubblicazione: (2025)
On the Surprising Efficacy of Distillation as an Alternative to Pre-Training Small Models
di: Farhat, Sean, et al.
Pubblicazione: (2024)
di: Farhat, Sean, et al.
Pubblicazione: (2024)
Implicit Hypergraph Neural Networks: A Stable Framework for Higher-Order Relational Learning with Provable Guarantees
di: Li, Xiaoyu, et al.
Pubblicazione: (2025)
di: Li, Xiaoyu, et al.
Pubblicazione: (2025)
Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning
di: Piccoli, Elia, et al.
Pubblicazione: (2025)
di: Piccoli, Elia, et al.
Pubblicazione: (2025)
Reinforcement Learning on Pre-Training Data
di: Li, Siheng, et al.
Pubblicazione: (2025)
di: Li, Siheng, et al.
Pubblicazione: (2025)
How Particle System Theory Enhances Hypergraph Message Passing
di: Ma, Yixuan, et al.
Pubblicazione: (2025)
di: Ma, Yixuan, et al.
Pubblicazione: (2025)
Ensemble of Pre-Trained Models for Long-Tailed Trajectory Prediction
di: Thuremella, Divya, et al.
Pubblicazione: (2025)
di: Thuremella, Divya, et al.
Pubblicazione: (2025)
CHGNN: A Semi-Supervised Contrastive Hypergraph Learning Network
di: Song, Yumeng, et al.
Pubblicazione: (2023)
di: Song, Yumeng, et al.
Pubblicazione: (2023)
SeqFusion: Sequential Fusion of Pre-Trained Models for Zero-Shot Time-Series Forecasting
di: Huang, Ting-Ji, et al.
Pubblicazione: (2025)
di: Huang, Ting-Ji, et al.
Pubblicazione: (2025)
Single Parent Family: A Spectrum of Family Members from a Single Pre-Trained Foundation Model
di: Hajimolahoseini, Habib, et al.
Pubblicazione: (2024)
di: Hajimolahoseini, Habib, et al.
Pubblicazione: (2024)
Toward a Graph Foundation Model: Pre-Training Transformers With Random Walks
di: Tang, Ziyuan, et al.
Pubblicazione: (2025)
di: Tang, Ziyuan, et al.
Pubblicazione: (2025)
MaGNet: A Mamba Dual-Hypergraph Network for Stock Prediction via Temporal-Causal and Global Relational Learning
di: Tan, Peilin, et al.
Pubblicazione: (2025)
di: Tan, Peilin, et al.
Pubblicazione: (2025)
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
di: Qing, Yunpeng, et al.
Pubblicazione: (2024)
di: Qing, Yunpeng, et al.
Pubblicazione: (2024)
A Survey of Few-Shot Learning on Graphs: from Meta-Learning to Pre-Training and Prompt Learning
di: Yu, Xingtong, et al.
Pubblicazione: (2024)
di: Yu, Xingtong, et al.
Pubblicazione: (2024)
Enhancing Pre-Trained Model-Based Class-Incremental Learning through Neural Collapse
di: He, Kun, et al.
Pubblicazione: (2025)
di: He, Kun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DIGIC: Domain Generalizable Imitation Learning by Causal Discovery
di: Chen, Yang, et al.
Pubblicazione: (2024) -
Conda: Column-Normalized Adam for Training Large Language Models Faster
di: Wang, Junjie, et al.
Pubblicazione: (2025) -
On the Limitations and Capabilities of Position Embeddings for Length Generalization
di: Chen, Yang, et al.
Pubblicazione: (2025) -
Improving Model Representation and Reducing KV Cache via Skip Connections with First Value Heads
di: Wu, Zhoutong, et al.
Pubblicazione: (2025) -
Attention Sinks Induce Gradient Sinks: Massive Activations as Gradient Regulators in Transformers
di: Chen, Yihong, et al.
Pubblicazione: (2026)