Efficient Generative Model Training via Embedded Representation Warmup
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Deyuan, Sun, Peng, Li, Xufeng, Lin, Tao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training
by: Lin, Jinhong, et al.
Published: (2026)
by: Lin, Jinhong, et al.
Published: (2026)
Test-Time Warmup for Multimodal Large Language Models
by: Rajaneesh, Nikita, et al.
Published: (2025)
by: Rajaneesh, Nikita, et al.
Published: (2025)
Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler
by: Dremov, Aleksandr, et al.
Published: (2025)
by: Dremov, Aleksandr, et al.
Published: (2025)
Efficiency for Free: Ideal Data Are Transportable Representations
by: Sun, Peng, et al.
Published: (2024)
by: Sun, Peng, et al.
Published: (2024)
Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching
by: Haidari, Rasched, et al.
Published: (2026)
by: Haidari, Rasched, et al.
Published: (2026)
FOAM: Blocked State Folding for Memory-Efficient LLM Training
by: Wen, Ziqing, et al.
Published: (2025)
by: Wen, Ziqing, et al.
Published: (2025)
Humanoid-inspired Causal Representation Learning for Domain Generalization
by: Tao, Ze, et al.
Published: (2025)
by: Tao, Ze, et al.
Published: (2025)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
by: Deiseroth, Björn, et al.
Published: (2024)
by: Deiseroth, Björn, et al.
Published: (2024)
A Time-Series Foundation Model by Universal Delay Embedding
by: Wang, Zijian, et al.
Published: (2025)
by: Wang, Zijian, et al.
Published: (2025)
Unified Continuous Generative Models
by: Sun, Peng, et al.
Published: (2025)
by: Sun, Peng, et al.
Published: (2025)
Efficient Process Reward Model Training via Active Learning
by: Duan, Keyu, et al.
Published: (2025)
by: Duan, Keyu, et al.
Published: (2025)
PSNE: Efficient Spectral Sparsification Algorithms for Scaling Network Embedding
by: Lin, Longlong, et al.
Published: (2024)
by: Lin, Longlong, et al.
Published: (2024)
Auxiliary Reward Generation with Transition Distance Representation Learning
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
GRE^2-MDCL: Graph Representation Embedding Enhanced via Multidimensional Contrastive Learning
by: Fan, Kaizhe, et al.
Published: (2024)
by: Fan, Kaizhe, et al.
Published: (2024)
Towards Robust Multi-Modal Reasoning via Model Selection
by: Liu, Xiangyan, et al.
Published: (2023)
by: Liu, Xiangyan, et al.
Published: (2023)
ELAS: Efficient Pre-Training of Low-Rank Large Language Models via 2:4 Activation Sparsity
by: Li, Jiaxi, et al.
Published: (2026)
by: Li, Jiaxi, et al.
Published: (2026)
LinkedIn Post Embeddings: Industrial Scale Embedding Generation and Usage across LinkedIn
by: Ramanujam, Sudarshan Srinivasa, et al.
Published: (2024)
by: Ramanujam, Sudarshan Srinivasa, et al.
Published: (2024)
Joint Design of Protein Surface and Structure Using a Diffusion Bridge Model
by: Li, Guanlue, et al.
Published: (2025)
by: Li, Guanlue, et al.
Published: (2025)
Post-Training Quantization of OpenPangu Models for Efficient Deployment on Atlas A2
by: Luo, Yilun, et al.
Published: (2025)
by: Luo, Yilun, et al.
Published: (2025)
Collaborative Unlabeled Data Optimization
by: Shang, Xinyi, et al.
Published: (2025)
by: Shang, Xinyi, et al.
Published: (2025)
Embedding Reliability Verification Constraints into Generation Expansion Planning
by: Liu, Peng, et al.
Published: (2025)
by: Liu, Peng, et al.
Published: (2025)
Scalable Numerical Embeddings for Multivariate Time Series: Enhancing Healthcare Data Representation Learning
by: Huang, Chun-Kai, et al.
Published: (2024)
by: Huang, Chun-Kai, et al.
Published: (2024)
Echo-LoRA: Parameter-Efficient Fine-Tuning via Cross-Layer Representation Injection
by: Peng, Yihang, et al.
Published: (2026)
by: Peng, Yihang, et al.
Published: (2026)
RAG-GFM: Overcoming In-Memory Bottlenecks in Graph Foundation Models via Retrieval-Augmented Generation
by: Yuan, Haonan, et al.
Published: (2026)
by: Yuan, Haonan, et al.
Published: (2026)
Embedding by Elicitation: Dynamic Representations for Bayesian Optimization of System Prompts
by: Lin, Zhiyuan Jerry, et al.
Published: (2026)
by: Lin, Zhiyuan Jerry, et al.
Published: (2026)
Grouter: Decoupling Routing from Representation for Accelerated MoE Training
by: Xu, Yuqi, et al.
Published: (2026)
by: Xu, Yuqi, et al.
Published: (2026)
Efficiently Aligning Draft Models via Parameter- and Data-Efficient Adaptation
by: Lin, Luxi, et al.
Published: (2026)
by: Lin, Luxi, et al.
Published: (2026)
ACT-JEPA: Novel Joint-Embedding Predictive Architecture for Efficient Policy Representation Learning
by: Vujinovic, Aleksandar, et al.
Published: (2025)
by: Vujinovic, Aleksandar, et al.
Published: (2025)
Pilot selection in the era of Virtual reality: algorithms for accurate and interpretable machine learning models
by: Ke, Luoma, et al.
Published: (2025)
by: Ke, Luoma, et al.
Published: (2025)
Benchmarking Pretrained Molecular Embedding Models For Molecular Representation Learning
by: Praski, Mateusz, et al.
Published: (2025)
by: Praski, Mateusz, et al.
Published: (2025)
Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models
by: Lee, Inhyeok, et al.
Published: (2026)
by: Lee, Inhyeok, et al.
Published: (2026)
A Survey on Memory-Efficient Transformer-Based Model Training in AI for Science
by: Tian, Kaiyuan, et al.
Published: (2025)
by: Tian, Kaiyuan, et al.
Published: (2025)
GWT: Scalable Optimizer State Compression for Large Language Model Training
by: Wen, Ziqing, et al.
Published: (2025)
by: Wen, Ziqing, et al.
Published: (2025)
EvoLlama: Enhancing LLMs' Understanding of Proteins via Multimodal Structure and Sequence Representations
by: Liu, Nuowei, et al.
Published: (2024)
by: Liu, Nuowei, et al.
Published: (2024)
Diffusion Attribution Score: Evaluating Training Data Influence in Diffusion Models
by: Lin, Jinxu, et al.
Published: (2024)
by: Lin, Jinxu, et al.
Published: (2024)
Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning
by: Piccoli, Elia, et al.
Published: (2025)
by: Piccoli, Elia, et al.
Published: (2025)
Learning Reconstructive Embeddings in Reproducing Kernel Hilbert Spaces via the Representer Theorem
by: Feito-Casares, Enrique, et al.
Published: (2026)
by: Feito-Casares, Enrique, et al.
Published: (2026)
Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model
by: Wu, Jiahao, et al.
Published: (2026)
by: Wu, Jiahao, et al.
Published: (2026)
Reinforcement Learning for Scalable Train Timetable Rescheduling with Graph Representation
by: Yue, Peng, et al.
Published: (2024)
by: Yue, Peng, et al.
Published: (2024)
Task-oriented Time Series Imputation Evaluation via Generalized Representers
by: Wang, Zhixian, et al.
Published: (2024)
by: Wang, Zhixian, et al.
Published: (2024)
Similar Items
-
Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training
by: Lin, Jinhong, et al.
Published: (2026) -
Test-Time Warmup for Multimodal Large Language Models
by: Rajaneesh, Nikita, et al.
Published: (2025) -
Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler
by: Dremov, Aleksandr, et al.
Published: (2025) -
Efficiency for Free: Ideal Data Are Transportable Representations
by: Sun, Peng, et al.
Published: (2024) -
Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching
by: Haidari, Rasched, et al.
Published: (2026)