Domain Generalization Guided by Large-Scale Pre-Trained Priors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zongbin, Pan, Bin, Shen, Shiyu, Shi, Tianyang, Shi, Zhenwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Domain Agnostic Conditional Invariant Predictions for Domain Generalization
von: Wang, Zongbin, et al.
Veröffentlicht: (2024)
von: Wang, Zongbin, et al.
Veröffentlicht: (2024)
Preserving Domain Generalization in Fine-Tuning via Joint Parameter Selection
von: Pan, Bin, et al.
Veröffentlicht: (2025)
von: Pan, Bin, et al.
Veröffentlicht: (2025)
Be Bayesian by Attachments to Catch More Uncertainty
von: Shen, Shiyu, et al.
Veröffentlicht: (2023)
von: Shen, Shiyu, et al.
Veröffentlicht: (2023)
A Pretrained Probabilistic Transformer for City-Scale Traffic Volume Prediction
von: Shen, Shiyu, et al.
Veröffentlicht: (2025)
von: Shen, Shiyu, et al.
Veröffentlicht: (2025)
Hyperspectral Image Generation with Unmixing Guided Diffusion Model
von: Shen, Shiyu, et al.
Veröffentlicht: (2025)
von: Shen, Shiyu, et al.
Veröffentlicht: (2025)
Source-Free Domain Adaptation Guided by Vision and Vision-Language Pre-Training
von: Zhang, Wenyu, et al.
Veröffentlicht: (2024)
von: Zhang, Wenyu, et al.
Veröffentlicht: (2024)
IMITATE: Clinical Prior Guided Hierarchical Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
von: He, Haoran, et al.
Veröffentlicht: (2024)
von: He, Haoran, et al.
Veröffentlicht: (2024)
Relational In-Context Learning via Synthetic Pre-training with Structural Prior
von: Wang, Yanbo, et al.
Veröffentlicht: (2026)
von: Wang, Yanbo, et al.
Veröffentlicht: (2026)
Mixture-of-Experts Operator Transformer for Large-Scale PDE Pre-Training
von: Wang, Hong, et al.
Veröffentlicht: (2025)
von: Wang, Hong, et al.
Veröffentlicht: (2025)
Joint Selection for Large-Scale Pre-Training Data via Policy Gradient-based Mask Learning
von: Fan, Ziqing, et al.
Veröffentlicht: (2025)
von: Fan, Ziqing, et al.
Veröffentlicht: (2025)
Prior-Guided Multi-Omic Transformers for Single-Cell Gene Regulatory Network Inference
von: Xu, Tianyang, et al.
Veröffentlicht: (2026)
von: Xu, Tianyang, et al.
Veröffentlicht: (2026)
LaMM: Semi-Supervised Pre-Training of Large-Scale Materials Models
von: Oyama, Yosuke, et al.
Veröffentlicht: (2025)
von: Oyama, Yosuke, et al.
Veröffentlicht: (2025)
Rethinking Local Learning: A Cheaper and Faster Recipe for LLM Post-Training
von: Shi, Hengyu, et al.
Veröffentlicht: (2026)
von: Shi, Hengyu, et al.
Veröffentlicht: (2026)
Using Scaling Laws for Data Source Utility Estimation in Domain-Specific Pre-Training
von: Ostapenko, Oleksiy, et al.
Veröffentlicht: (2025)
von: Ostapenko, Oleksiy, et al.
Veröffentlicht: (2025)
Scaling Law for Large-Scale Pre-Training Using Chaotic Time Series and Predictability in Financial Time Series
von: Takemoto, Yuki
Veröffentlicht: (2025)
von: Takemoto, Yuki
Veröffentlicht: (2025)
Multi-Scale Heterogeneous Text-Attributed Graph Datasets From Diverse Domains
von: Liu, Yunhui, et al.
Veröffentlicht: (2024)
von: Liu, Yunhui, et al.
Veröffentlicht: (2024)
A Large Scale Heterogeneous Treatment Effect Estimation Framework and Its Applications of Users' Journey at Snap
von: Pan, Jing, et al.
Veröffentlicht: (2025)
von: Pan, Jing, et al.
Veröffentlicht: (2025)
Stochastic Normalized Gradient Descent with Momentum for Large-Batch Training
von: Zhao, Shen-Yi, et al.
Veröffentlicht: (2020)
von: Zhao, Shen-Yi, et al.
Veröffentlicht: (2020)
Hybrid Attribution Priors for Explainable and Robust Model Training
von: Zhang, Zhuoran, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuoran, et al.
Veröffentlicht: (2025)
GraphControl: Adding Conditional Control to Universal Graph Pre-trained Models for Graph Domain Transfer Learning
von: Zhu, Yun, et al.
Veröffentlicht: (2023)
von: Zhu, Yun, et al.
Veröffentlicht: (2023)
Mixture-of-Channels: Exploiting Sparse FFNs for Efficient LLMs Pre-Training and Inference
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
SVFit: Parameter-Efficient Fine-Tuning of Large Pre-Trained Models Using Singular Values
von: Sun, Chengwei, et al.
Veröffentlicht: (2024)
von: Sun, Chengwei, et al.
Veröffentlicht: (2024)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-Training
von: Hao, Zhongkai, et al.
Veröffentlicht: (2024)
von: Hao, Zhongkai, et al.
Veröffentlicht: (2024)
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
Pushing the Limits of All-Atom Geometric Graph Neural Networks: Pre-Training, Scaling and Zero-Shot Transfer
von: Pengmei, Zihan, et al.
Veröffentlicht: (2024)
von: Pengmei, Zihan, et al.
Veröffentlicht: (2024)
SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
von: Zhang, Xiaojiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaojiang, et al.
Veröffentlicht: (2025)
Learning Multimodal Latent Generative Models with Energy-Based Prior
von: Yuan, Shiyu, et al.
Veröffentlicht: (2024)
von: Yuan, Shiyu, et al.
Veröffentlicht: (2024)
Domain Generalization Under Posterior Drift
von: Zhu, Yilun, et al.
Veröffentlicht: (2025)
von: Zhu, Yilun, et al.
Veröffentlicht: (2025)
Gradient-Guided Annealing for Domain Generalization
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2025)
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2025)
Domain-Adaptive Continued Pre-Training of Small Language Models
von: Faroz, Salman
Veröffentlicht: (2025)
von: Faroz, Salman
Veröffentlicht: (2025)
GaLore 2: Large-Scale LLM Pre-Training by Gradient Low-Rank Projection
von: Su, DiJia, et al.
Veröffentlicht: (2025)
von: Su, DiJia, et al.
Veröffentlicht: (2025)
Threshold-Guided Optimization for Visual Generative Models
von: Bai, Jinbin, et al.
Veröffentlicht: (2026)
von: Bai, Jinbin, et al.
Veröffentlicht: (2026)
Provable Target Sample Complexity Improvements as Pre-Trained Models Scale
von: Fukuchi, Kazuto, et al.
Veröffentlicht: (2026)
von: Fukuchi, Kazuto, et al.
Veröffentlicht: (2026)
Structural Priors and Modular Adapters in the Composable Fine-Tuning Algorithm of Large-Scale Models
von: Wang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Wang, Yuxiao, et al.
Veröffentlicht: (2025)
PTPP-Aware Adaptation Scaling Laws: Predicting Domain-Adaptation Performance at Unseen Pre-Training Budgets
von: Goffinet, Etienne, et al.
Veröffentlicht: (2025)
von: Goffinet, Etienne, et al.
Veröffentlicht: (2025)
TimeHF: Billion-Scale Time Series Models Guided by Human Feedback
von: Qi, Yongzhi, et al.
Veröffentlicht: (2025)
von: Qi, Yongzhi, et al.
Veröffentlicht: (2025)
ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment
von: Li, Xiuyu, et al.
Veröffentlicht: (2026)
von: Li, Xiuyu, et al.
Veröffentlicht: (2026)
DACP: Domain-Adaptive Continual Pre-Training of Large Language Models for Phone Conversation Summarization
von: Fu, Xue-Yong, et al.
Veröffentlicht: (2025)
von: Fu, Xue-Yong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Domain Agnostic Conditional Invariant Predictions for Domain Generalization
von: Wang, Zongbin, et al.
Veröffentlicht: (2024) -
Preserving Domain Generalization in Fine-Tuning via Joint Parameter Selection
von: Pan, Bin, et al.
Veröffentlicht: (2025) -
Be Bayesian by Attachments to Catch More Uncertainty
von: Shen, Shiyu, et al.
Veröffentlicht: (2023) -
A Pretrained Probabilistic Transformer for City-Scale Traffic Volume Prediction
von: Shen, Shiyu, et al.
Veröffentlicht: (2025) -
Hyperspectral Image Generation with Unmixing Guided Diffusion Model
von: Shen, Shiyu, et al.
Veröffentlicht: (2025)