Gespeichert in:
| Hauptverfasser: | Dai, Yijia, Gao, Zhaolin, Sattar, Yahya, Dean, Sarah, Sun, Jennifer J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.07298 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Hidden Markov Models Using Conditional Samples
von: Kakade, Sham M., et al.
Veröffentlicht: (2023)
von: Kakade, Sham M., et al.
Veröffentlicht: (2023)
The Future of Large Language Model Pre-training is Federated
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
Pre-training Limited Memory Language Models with Internal and External Knowledge
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
Machine Unlearning of Pre-trained Large Language Models
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
Accelerating Reinforcement Learning Algorithms Convergence using Pre-trained Large Language Models as Tutors With Advice Reusing
von: Toral, Lukas, et al.
Veröffentlicht: (2025)
von: Toral, Lukas, et al.
Veröffentlicht: (2025)
Ensemble Methods for Sequence Classification with Hidden Markov Models
von: Kawawa-Beaudan, Maxime, et al.
Veröffentlicht: (2024)
von: Kawawa-Beaudan, Maxime, et al.
Veröffentlicht: (2024)
Simple and Scalable Strategies to Continually Pre-train Large Language Models
von: Ibrahim, Adam, et al.
Veröffentlicht: (2024)
von: Ibrahim, Adam, et al.
Veröffentlicht: (2024)
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
von: Song, Weixi, et al.
Veröffentlicht: (2023)
von: Song, Weixi, et al.
Veröffentlicht: (2023)
Utilizing Strategic Pre-training to Reduce Overfitting: Baguan -- A Pre-trained Weather Forecasting Model
von: Niu, Peisong, et al.
Veröffentlicht: (2025)
von: Niu, Peisong, et al.
Veröffentlicht: (2025)
S$^2$ALM: Sequence-Structure Pre-trained Large Language Model for Comprehensive Antibody Representation Learning
von: Yin, Mingze, et al.
Veröffentlicht: (2024)
von: Yin, Mingze, et al.
Veröffentlicht: (2024)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
von: Samragh, Mohammad, et al.
Veröffentlicht: (2024)
von: Samragh, Mohammad, et al.
Veröffentlicht: (2024)
Transfer Learning with Pre-trained Conditional Generative Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2022)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2022)
Learning Linear Dynamics from Bilinear Observations
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
An Efficient Replay for Class-Incremental Learning with Pre-trained Models
von: Yin, Weimin, et al.
Veröffentlicht: (2024)
von: Yin, Weimin, et al.
Veröffentlicht: (2024)
IMU-1: Sample-Efficient Pre-training of Small Language Models
von: Grigorev, George
Veröffentlicht: (2026)
von: Grigorev, George
Veröffentlicht: (2026)
Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models
von: Hao, Yifan, et al.
Veröffentlicht: (2025)
von: Hao, Yifan, et al.
Veröffentlicht: (2025)
Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
Induced Numerical Instability: Hidden Costs in Multimodal Large Language Models
von: Wong, Wai Tuck, et al.
Veröffentlicht: (2026)
von: Wong, Wai Tuck, et al.
Veröffentlicht: (2026)
Probing the Decision Boundaries of In-context Learning in Large Language Models
von: Zhao, Siyan, et al.
Veröffentlicht: (2024)
von: Zhao, Siyan, et al.
Veröffentlicht: (2024)
Revisiting In-context Learning Inference Circuit in Large Language Models
von: Cho, Hakaze, et al.
Veröffentlicht: (2024)
von: Cho, Hakaze, et al.
Veröffentlicht: (2024)
Large Language Models as Markov Chains
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
Investigating Data Contamination for Pre-training Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
Aligning Pre-trained Models for Spoken Language Translation
von: Sedláček, Šimon, et al.
Veröffentlicht: (2024)
von: Sedláček, Šimon, et al.
Veröffentlicht: (2024)
Sequence-to-Sequence Spanish Pre-trained Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2023)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2023)
Towards Efficient Pre-training: Exploring FP4 Precision in Large Language Models
von: Zhou, Jiecheng, et al.
Veröffentlicht: (2025)
von: Zhou, Jiecheng, et al.
Veröffentlicht: (2025)
Representation Learning of Auxiliary Concepts for Improved Student Modeling and Exercise Recommendation
von: Badran, Yahya, et al.
Veröffentlicht: (2025)
von: Badran, Yahya, et al.
Veröffentlicht: (2025)
Pre-trained Vision-Language Models Learn Discoverable Visual Concepts
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
A Vision-Language Pre-training Model-Guided Approach for Mitigating Backdoor Attacks in Federated Learning
von: Gai, Keke, et al.
Veröffentlicht: (2025)
von: Gai, Keke, et al.
Veröffentlicht: (2025)
Reading Your Heart: Learning ECG Words and Sentences via Pre-training ECG Language Model
von: Jin, Jiarui, et al.
Veröffentlicht: (2025)
von: Jin, Jiarui, et al.
Veröffentlicht: (2025)
A Pre-trained Data Deduplication Model based on Active Learning
von: Shi, Haochen, et al.
Veröffentlicht: (2023)
von: Shi, Haochen, et al.
Veröffentlicht: (2023)
Backward-Friendly Optimization: Training Large Language Models with Approximate Gradients under Memory Constraints
von: Yang, Jing, et al.
Veröffentlicht: (2025)
von: Yang, Jing, et al.
Veröffentlicht: (2025)
Embedding Hidden Adversarial Capabilities in Pre-Trained Diffusion Models
von: Beerens, Lucas, et al.
Veröffentlicht: (2025)
von: Beerens, Lucas, et al.
Veröffentlicht: (2025)
FGBERT: Function-Driven Pre-trained Gene Language Model for Metagenomics
von: Duan, ChenRui, et al.
Veröffentlicht: (2024)
von: Duan, ChenRui, et al.
Veröffentlicht: (2024)
Parallel In-context Learning for Large Vision Language Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
Unveiling Hidden Collaboration within Mixture-of-Experts in Large Language Models
von: Tang, Yuanbo, et al.
Veröffentlicht: (2025)
von: Tang, Yuanbo, et al.
Veröffentlicht: (2025)
Distributional Clarity: The Hidden Driver of RL-Friendliness in Large Language Models
von: Sun, Shaoning, et al.
Veröffentlicht: (2026)
von: Sun, Shaoning, et al.
Veröffentlicht: (2026)
Integrating Pre-trained Language Model into Neural Machine Translation
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
Markov Constraint as Large Language Model Surrogate
von: Bonlarron, Alexandre, et al.
Veröffentlicht: (2024)
von: Bonlarron, Alexandre, et al.
Veröffentlicht: (2024)
The Recurrent Sticky Hierarchical Dirichlet Process Hidden Markov Model
von: Słupiński, Mikołaj, et al.
Veröffentlicht: (2024)
von: Słupiński, Mikołaj, et al.
Veröffentlicht: (2024)
Discovering and Reasoning of Causality in the Hidden World with Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Hidden Markov Models Using Conditional Samples
von: Kakade, Sham M., et al.
Veröffentlicht: (2023) -
The Future of Large Language Model Pre-training is Federated
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024) -
Pre-training Limited Memory Language Models with Internal and External Knowledge
von: Zhao, Linxi, et al.
Veröffentlicht: (2025) -
Machine Unlearning of Pre-trained Large Language Models
von: Yao, Jin, et al.
Veröffentlicht: (2024) -
Accelerating Reinforcement Learning Algorithms Convergence using Pre-trained Large Language Models as Tutors With Advice Reusing
von: Toral, Lukas, et al.
Veröffentlicht: (2025)