Utilizing Strategic Pre-training to Reduce Overfitting: Baguan -- A Pre-trained Weather Forecasting Model
Fuente:
arXiv
Saved in:
| Main Authors: | Niu, Peisong, Ma, Ziqing, Zhou, Tian, Chen, Weiqi, Shen, Lefei, Jin, Rong, Sun, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Integrating Weather Foundation Model and Satellite to Enable Fine-Grained Solar Irradiance Forecasting
by: Ma, Ziqing, et al.
Published: (2026)
by: Ma, Ziqing, et al.
Published: (2026)
Baguan-TS: A Sequence-Native In-Context Learning Model for Time Series Forecasting with Covariates
by: Yang, Linxiao, et al.
Published: (2026)
by: Yang, Linxiao, et al.
Published: (2026)
Skillful Kilometer-Scale Regional Weather Forecasting via Global and Regional Coupling
by: Chen, Weiqi, et al.
Published: (2026)
by: Chen, Weiqi, et al.
Published: (2026)
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
by: Song, Weixi, et al.
Published: (2023)
by: Song, Weixi, et al.
Published: (2023)
Enhancing AI-Based Tropical Cyclone Track and Intensity Forecasting via Systematic Bias Correction
by: Niu, Peisong, et al.
Published: (2026)
by: Niu, Peisong, et al.
Published: (2026)
Mitigating Time Discretization Challenges with WeatherODE: A Sandwich Physics-Driven Neural ODE for Weather Forecasting
by: Liu, Peiyuan, et al.
Published: (2024)
by: Liu, Peiyuan, et al.
Published: (2024)
Machine Unlearning of Pre-trained Large Language Models
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
Centroid-centered Modeling for Efficient Vision Transformer Pre-training
by: Yan, Xin, et al.
Published: (2023)
by: Yan, Xin, et al.
Published: (2023)
Aligning Instruction Tuning with Pre-training
by: Liang, Yiming, et al.
Published: (2025)
by: Liang, Yiming, et al.
Published: (2025)
Can Pre-trained Language Models Understand Chinese Humor?
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
Pre-train and Fine-tune: Recommenders as Large Models
by: Jiang, Zhenhao, et al.
Published: (2025)
by: Jiang, Zhenhao, et al.
Published: (2025)
Pre-training Generative Recommender with Multi-Identifier Item Tokenization
by: Zheng, Bowen, et al.
Published: (2025)
by: Zheng, Bowen, et al.
Published: (2025)
Pre-trained Vision-Language Models Learn Discoverable Visual Concepts
by: Zang, Yuan, et al.
Published: (2024)
by: Zang, Yuan, et al.
Published: (2024)
Target Concept Tuning Improves Extreme Weather Forecasting
by: Ren, Shijie, et al.
Published: (2026)
by: Ren, Shijie, et al.
Published: (2026)
A Pre-trained Reaction Embedding Descriptor Capturing Bond Transformation Patterns
by: Liu, Weiqi, et al.
Published: (2026)
by: Liu, Weiqi, et al.
Published: (2026)
Pre-training for Recommendation Unlearning
by: Chen, Guoxuan, et al.
Published: (2025)
by: Chen, Guoxuan, et al.
Published: (2025)
Understanding the Role of Textual Prompts in LLM for Time Series Forecasting: an Adapter View
by: Niu, Peisong, et al.
Published: (2023)
by: Niu, Peisong, et al.
Published: (2023)
Forecast-PEFT: Parameter-Efficient Fine-Tuning for Pre-trained Motion Forecasting Models
by: Wang, Jifeng, et al.
Published: (2024)
by: Wang, Jifeng, et al.
Published: (2024)
Exploring the Benefit of Activation Sparsity in Pre-training
by: Zhang, Zhengyan, et al.
Published: (2024)
by: Zhang, Zhengyan, et al.
Published: (2024)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
by: Chen, Tong, et al.
Published: (2025)
by: Chen, Tong, et al.
Published: (2025)
Making Pre-trained Language Models Better Continual Few-Shot Relation Extractors
by: Ma, Shengkun, et al.
Published: (2024)
by: Ma, Shengkun, et al.
Published: (2024)
Pre-training Auto-regressive Robotic Models with 4D Representations
by: Niu, Dantong, et al.
Published: (2025)
by: Niu, Dantong, et al.
Published: (2025)
Structure-aware Fine-tuning for Code Pre-trained Models
by: Wu, Jiayi, et al.
Published: (2024)
by: Wu, Jiayi, et al.
Published: (2024)
FusionSF: Fuse Heterogeneous Modalities in a Vector Quantized Framework for Robust Solar Power Forecasting
by: Ma, Ziqing, et al.
Published: (2024)
by: Ma, Ziqing, et al.
Published: (2024)
RADAR: Revealing Asymmetric Development of Abilities in MLLM Pre-training
by: Nie, Yunshuang, et al.
Published: (2026)
by: Nie, Yunshuang, et al.
Published: (2026)
Pre-training Limited Memory Language Models with Internal and External Knowledge
by: Zhao, Linxi, et al.
Published: (2025)
by: Zhao, Linxi, et al.
Published: (2025)
Graph Generative Pre-trained Transformer
by: Chen, Xiaohui, et al.
Published: (2025)
by: Chen, Xiaohui, et al.
Published: (2025)
Probing Language Models for Pre-training Data Detection
by: Liu, Zhenhua, et al.
Published: (2024)
by: Liu, Zhenhua, et al.
Published: (2024)
Demystifying Manifold Constraints in LLM Pre-training
by: An, Kang, et al.
Published: (2026)
by: An, Kang, et al.
Published: (2026)
Towards Efficient Pre-training: Exploring FP4 Precision in Large Language Models
by: Zhou, Jiecheng, et al.
Published: (2025)
by: Zhou, Jiecheng, et al.
Published: (2025)
QKCV Attention: Enhancing Time Series Forecasting with Static Categorical Embeddings for Both Lightweight and Pre-trained Foundation Models
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Do Pre-trained Vision-Language Models Encode Object States?
by: Newman, Kaleb, et al.
Published: (2024)
by: Newman, Kaleb, et al.
Published: (2024)
TCGPN: Temporal-Correlation Graph Pre-trained Network for Stock Forecasting
by: Yan, Wenbo, et al.
Published: (2024)
by: Yan, Wenbo, et al.
Published: (2024)
Unlocking the Power of Spatial and Temporal Information in Medical Multimodal Pre-training
by: Yang, Jinxia, et al.
Published: (2024)
by: Yang, Jinxia, et al.
Published: (2024)
Exploring the Innovation Opportunities for Pre-trained Models
by: Park, Minjung, et al.
Published: (2025)
by: Park, Minjung, et al.
Published: (2025)
Urban Region Pre-training and Prompting: A Graph-based Approach
by: Jin, Jiahui, et al.
Published: (2024)
by: Jin, Jiahui, et al.
Published: (2024)
Pre-training with Fractional Denoising to Enhance Molecular Property Prediction
by: Ni, Yuyan, et al.
Published: (2024)
by: Ni, Yuyan, et al.
Published: (2024)
Multimodal 3D Genome Pre-training
by: Yang, Minghao, et al.
Published: (2025)
by: Yang, Minghao, et al.
Published: (2025)
The Future of Large Language Model Pre-training is Federated
by: Sani, Lorenzo, et al.
Published: (2024)
by: Sani, Lorenzo, et al.
Published: (2024)
Pre-training on High Definition X-ray Images: An Experimental Study
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
Similar Items
-
Integrating Weather Foundation Model and Satellite to Enable Fine-Grained Solar Irradiance Forecasting
by: Ma, Ziqing, et al.
Published: (2026) -
Baguan-TS: A Sequence-Native In-Context Learning Model for Time Series Forecasting with Covariates
by: Yang, Linxiao, et al.
Published: (2026) -
Skillful Kilometer-Scale Regional Weather Forecasting via Global and Regional Coupling
by: Chen, Weiqi, et al.
Published: (2026) -
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
by: Song, Weixi, et al.
Published: (2023) -
Enhancing AI-Based Tropical Cyclone Track and Intensity Forecasting via Systematic Bias Correction
by: Niu, Peisong, et al.
Published: (2026)