Gespeichert in:
| Hauptverfasser: | Yao, Xinhao, Qian, Hongjin, Hu, Xiaolin, Xu, Gengze, Liu, Wei, Luan, Jian, Wang, Bin, Liu, Yong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2410.02247 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing In-Context Learning Performance with just SVD-Based Weight Pruning: A Theoretical Perspective
von: Yao, Xinhao, et al.
Veröffentlicht: (2024)
von: Yao, Xinhao, et al.
Veröffentlicht: (2024)
On the Emergence of Weak-to-Strong Generalization: A Bias-Variance Perspective
von: Xu, Gengze, et al.
Veröffentlicht: (2025)
von: Xu, Gengze, et al.
Veröffentlicht: (2025)
PMSS: Pretrained Matrices Skeleton Selection for LLM Fine-tuning
von: Wang, Qibin, et al.
Veröffentlicht: (2024)
von: Wang, Qibin, et al.
Veröffentlicht: (2024)
On the Blessing of Pre-training in Weak-to-Strong Generalization
von: Yao, Wei, et al.
Veröffentlicht: (2026)
von: Yao, Wei, et al.
Veröffentlicht: (2026)
The Capabilities and Limitations of Weak-to-Strong Generalization: Generalization and Calibration
von: Yao, Wei, et al.
Veröffentlicht: (2025)
von: Yao, Wei, et al.
Veröffentlicht: (2025)
DoTA: Weight-Decomposed Tensor Adaptation for Large Language Models
von: Hu, Xiaolin, et al.
Veröffentlicht: (2024)
von: Hu, Xiaolin, et al.
Veröffentlicht: (2024)
On Weak-to-Strong Generalization and f-Divergence
von: Yao, Wei, et al.
Veröffentlicht: (2025)
von: Yao, Wei, et al.
Veröffentlicht: (2025)
Compositional Generalization from Learned Skills via CoT Training: A Theoretical and Structural Analysis for Reasoning
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
The Debate on RLVR Reasoning Capability Boundary: Shrinkage, Expansion, or Both? A Two-Stage Dynamic View
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
Look Within or Look Beyond? A Theoretical Comparison Between Parameter-Efficient and Full Fine-Tuning
von: Liu, Yongkang, et al.
Veröffentlicht: (2025)
von: Liu, Yongkang, et al.
Veröffentlicht: (2025)
Enhancing Reinforcement Learning Fine-Tuning with an Online Refiner
von: Ma, Hao, et al.
Veröffentlicht: (2026)
von: Ma, Hao, et al.
Veröffentlicht: (2026)
Multi-branch of Attention Yields Accurate Results for Tabular Data
von: Li, Xuechen, et al.
Veröffentlicht: (2025)
von: Li, Xuechen, et al.
Veröffentlicht: (2025)
Control Theoretic Approach to Fine-Tuning and Transfer Learning
von: Bayram, Erkan, et al.
Veröffentlicht: (2024)
von: Bayram, Erkan, et al.
Veröffentlicht: (2024)
SASA: Semantic-Aware Contrastive Learning Framework with Separated Attention for Triple Classification
von: Xiaodan, Xu, et al.
Veröffentlicht: (2026)
von: Xiaodan, Xu, et al.
Veröffentlicht: (2026)
LoSiA: Efficient High-Rank Fine-Tuning via Subnet Localization and Optimization
von: Wang, Xujia, et al.
Veröffentlicht: (2025)
von: Wang, Xujia, et al.
Veröffentlicht: (2025)
PSEO: Optimizing Post-hoc Stacking Ensemble Through Hyperparameter Tuning
von: Xu, Beicheng, et al.
Veröffentlicht: (2025)
von: Xu, Beicheng, et al.
Veröffentlicht: (2025)
Mixture of Diverse Size Experts
von: Sun, Manxi, et al.
Veröffentlicht: (2024)
von: Sun, Manxi, et al.
Veröffentlicht: (2024)
Self-Generative Adversarial Fine-Tuning for Large Language Models
von: Wu, Shiguang, et al.
Veröffentlicht: (2026)
von: Wu, Shiguang, et al.
Veröffentlicht: (2026)
Towards Auto-Regressive Next-Token Prediction: In-Context Learning Emerges from Generalization
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
Beyond the Black Box: A Survey on the Theory and Mechanism of Large Language Models
von: Gan, Zeyu, et al.
Veröffentlicht: (2026)
von: Gan, Zeyu, et al.
Veröffentlicht: (2026)
Attention Mechanism, Max-Affine Partition, and Universal Approximation
von: Liu, Hude, et al.
Veröffentlicht: (2025)
von: Liu, Hude, et al.
Veröffentlicht: (2025)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
Fed-pilot: Optimizing LoRA Allocation for Efficient Federated Fine-Tuning with Heterogeneous Clients
von: Zhang, Zikai, et al.
Veröffentlicht: (2024)
von: Zhang, Zikai, et al.
Veröffentlicht: (2024)
Beyond Progress Measures: Theoretical Insights into the Mechanism of Grokking
von: Gu, Zihan, et al.
Veröffentlicht: (2025)
von: Gu, Zihan, et al.
Veröffentlicht: (2025)
Learning to Think: Information-Theoretic Reinforcement Fine-Tuning for LLMs
von: Wang, Jingyao, et al.
Veröffentlicht: (2025)
von: Wang, Jingyao, et al.
Veröffentlicht: (2025)
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning
von: Liu, Yongkang, et al.
Veröffentlicht: (2026)
von: Liu, Yongkang, et al.
Veröffentlicht: (2026)
Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models
von: Lee, Chungpa, et al.
Veröffentlicht: (2026)
von: Lee, Chungpa, et al.
Veröffentlicht: (2026)
Understanding and Preserving Safety in Fine-Tuned LLMs
von: Zhang, Jiawen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiawen, et al.
Veröffentlicht: (2026)
HoPE: A Novel Positional Encoding Without Long-Term Decay for Enhanced Context Awareness and Extrapolation
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
Preserving Domain Generalization in Fine-Tuning via Joint Parameter Selection
von: Pan, Bin, et al.
Veröffentlicht: (2025)
von: Pan, Bin, et al.
Veröffentlicht: (2025)
Information-Theoretic Generalization Bounds for Transductive Learning and its Applications
von: Tang, Huayi, et al.
Veröffentlicht: (2023)
von: Tang, Huayi, et al.
Veröffentlicht: (2023)
Graph Attention is Not Always Beneficial: A Theoretical Analysis of Graph Attention Mechanisms via Contextual Stochastic Block Models
von: Ma, Zhongtian, et al.
Veröffentlicht: (2024)
von: Ma, Zhongtian, et al.
Veröffentlicht: (2024)
Towards a Theoretical Understanding to the Generalization of RLHF
von: Li, Zhaochun, et al.
Veröffentlicht: (2026)
von: Li, Zhaochun, et al.
Veröffentlicht: (2026)
Rethinking Training Dynamics in Scale-wise Autoregressive Generation
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
Invariance Makes LLM Unlearning Resilient Even to Unanticipated Downstream Fine-Tuning
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
STEP: Success-Rate-Aware Trajectory-Efficient Policy Optimization
von: Chen, Yuhan, et al.
Veröffentlicht: (2025)
von: Chen, Yuhan, et al.
Veröffentlicht: (2025)
An Optimization Framework for Differentially Private Sparse Fine-Tuning
von: Makni, Mehdi, et al.
Veröffentlicht: (2025)
von: Makni, Mehdi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Enhancing In-Context Learning Performance with just SVD-Based Weight Pruning: A Theoretical Perspective
von: Yao, Xinhao, et al.
Veröffentlicht: (2024) -
On the Emergence of Weak-to-Strong Generalization: A Bias-Variance Perspective
von: Xu, Gengze, et al.
Veröffentlicht: (2025) -
PMSS: Pretrained Matrices Skeleton Selection for LLM Fine-tuning
von: Wang, Qibin, et al.
Veröffentlicht: (2024) -
On the Blessing of Pre-training in Weak-to-Strong Generalization
von: Yao, Wei, et al.
Veröffentlicht: (2026) -
The Capabilities and Limitations of Weak-to-Strong Generalization: Generalization and Calibration
von: Yao, Wei, et al.
Veröffentlicht: (2025)