Fast and Accurate Probing of In-Training LLMs' Downstream Performances
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Zhichen, Lun, Tianle, Wen, Zhibin, An, Hao, Ou, Yulin, Xu, Jianhui, Zhang, Hao, Fang, Wenyi, Zheng, Yang, Xu, Yang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Scaling Laws for Predicting Downstream Performance in LLMs
por: Chen, Yangyi, et al.
Publicado: (2024)
por: Chen, Yangyi, et al.
Publicado: (2024)
Unveiling Downstream Performance Scaling of LLMs: A Clustering-Based Perspective
por: Xu, Chengyin, et al.
Publicado: (2025)
por: Xu, Chengyin, et al.
Publicado: (2025)
QaRL: Rollout-Aligned Quantization-Aware RL for Fast and Stable Training under Training--Inference Mismatch
por: Gu, Hao, et al.
Publicado: (2026)
por: Gu, Hao, et al.
Publicado: (2026)
Graph VQ-Transformer (GVT): Fast and Accurate Molecular Generation via High-Fidelity Discrete Latents
por: Zheng, Haozhuo, et al.
Publicado: (2025)
por: Zheng, Haozhuo, et al.
Publicado: (2025)
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
por: Chen, Xingwu, et al.
Publicado: (2025)
por: Chen, Xingwu, et al.
Publicado: (2025)
Estimating the Effects of Sample Training Orders for Large Language Models without Retraining
por: Yang, Hao, et al.
Publicado: (2025)
por: Yang, Hao, et al.
Publicado: (2025)
A Theory of Training Profit-Optimal LLMs
por: Hao, Sophie, et al.
Publicado: (2026)
por: Hao, Sophie, et al.
Publicado: (2026)
Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs "Difficult" Downstream Tasks in LLMs
por: Yin, Lu, et al.
Publicado: (2023)
por: Yin, Lu, et al.
Publicado: (2023)
Value-Based Pre-Training with Downstream Feedback
por: Ke, Shuqi, et al.
Publicado: (2026)
por: Ke, Shuqi, et al.
Publicado: (2026)
BWLA: Breaking the Barrier of W1AX Post-Training Quantization for LLMs
por: Zhao, Zhixiong, et al.
Publicado: (2026)
por: Zhao, Zhixiong, et al.
Publicado: (2026)
Vision-Language Model Selection and Reuse for Downstream Adaptation
por: Tan, Hao-Zhe, et al.
Publicado: (2025)
por: Tan, Hao-Zhe, et al.
Publicado: (2025)
Finetune-Informed Pretraining Boosts Downstream Performance
por: Faysal, Atik, et al.
Publicado: (2026)
por: Faysal, Atik, et al.
Publicado: (2026)
CAP: Controllable Alignment Prompting for Unlearning in LLMs
por: Wang, Zhaokun, et al.
Publicado: (2026)
por: Wang, Zhaokun, et al.
Publicado: (2026)
Multi-Value Alignment for LLMs via Value Decorrelation and Extrapolation
por: Xu, Hefei, et al.
Publicado: (2025)
por: Xu, Hefei, et al.
Publicado: (2025)
ENOT: Expectile Regularization for Fast and Accurate Training of Neural Optimal Transport
por: Buzun, Nazar, et al.
Publicado: (2024)
por: Buzun, Nazar, et al.
Publicado: (2024)
Achieving binary weight and activation for LLMs using Post-Training Quantization
por: Song, Siqing, et al.
Publicado: (2025)
por: Song, Siqing, et al.
Publicado: (2025)
EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control
por: Yang, Kai, et al.
Publicado: (2025)
por: Yang, Kai, et al.
Publicado: (2025)
Revisiting Counterfactual Regression through the Lens of Gromov-Wasserstein Information Bottleneck
por: Yang, Hao, et al.
Publicado: (2024)
por: Yang, Hao, et al.
Publicado: (2024)
SliderQuant: Accurate Post-Training Quantization for LLMs
por: Wang, Shigeng, et al.
Publicado: (2026)
por: Wang, Shigeng, et al.
Publicado: (2026)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
por: Li, Miaomiao, et al.
Publicado: (2025)
por: Li, Miaomiao, et al.
Publicado: (2025)
When LLM Meets Time Series: Can LLMs Perform Multi-Step Time Series Reasoning and Inference
por: Ye, Wen, et al.
Publicado: (2025)
por: Ye, Wen, et al.
Publicado: (2025)
Training Overhead Ratio: A Practical Reliability Metric for Large Language Model Training Systems
por: Lu, Ning, et al.
Publicado: (2024)
por: Lu, Ning, et al.
Publicado: (2024)
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
por: Du, Hao, et al.
Publicado: (2025)
por: Du, Hao, et al.
Publicado: (2025)
PolicyEvolve: Evolving Programmatic Policies by LLMs for multi-player games via Population-Based Training
por: Lv, Mingrui, et al.
Publicado: (2025)
por: Lv, Mingrui, et al.
Publicado: (2025)
Entropy-Guided Data-Efficient Training for Multimodal Reasoning Reward Models
por: Yang, Shidong, et al.
Publicado: (2026)
por: Yang, Shidong, et al.
Publicado: (2026)
Teaching LLMs to Abstain via Fine-Grained Semantic Confidence Reward
por: An, Hao, et al.
Publicado: (2025)
por: An, Hao, et al.
Publicado: (2025)
D$^2$Quant: Accurate Low-bit Post-Training Weight Quantization for LLMs
por: Yan, Xianglong, et al.
Publicado: (2026)
por: Yan, Xianglong, et al.
Publicado: (2026)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
por: Xu, Tianyi, et al.
Publicado: (2025)
por: Xu, Tianyi, et al.
Publicado: (2025)
Reinforcement Learning Fine-Tuning Enhances Activation Intensity and Diversity in the Internal Circuitry of LLMs
por: Zhang, Honglin, et al.
Publicado: (2025)
por: Zhang, Honglin, et al.
Publicado: (2025)
PSMGD: Periodic Stochastic Multi-Gradient Descent for Fast Multi-Objective Optimization
por: Xu, Mingjing, et al.
Publicado: (2024)
por: Xu, Mingjing, et al.
Publicado: (2024)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
por: Lin, Zicheng, et al.
Publicado: (2024)
por: Lin, Zicheng, et al.
Publicado: (2024)
Detecting Subtle Differences between Human and Model Languages Using Spectrum of Relative Likelihood
por: Xu, Yang, et al.
Publicado: (2024)
por: Xu, Yang, et al.
Publicado: (2024)
Ignition Phase : Standard Training for Fast Adversarial Robustness
por: Yu-Hang, Wang, et al.
Publicado: (2025)
por: Yu-Hang, Wang, et al.
Publicado: (2025)
ReST-RL: Achieving Accurate Code Reasoning of LLMs with Optimized Self-Training and Decoding
por: Zhoubian, Sining, et al.
Publicado: (2025)
por: Zhoubian, Sining, et al.
Publicado: (2025)
Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing
por: Wang, Xu, et al.
Publicado: (2025)
por: Wang, Xu, et al.
Publicado: (2025)
Erase then Rectify: A Training-Free Parameter Editing Approach for Cost-Effective Graph Unlearning
por: Yang, Zhe-Rui, et al.
Publicado: (2024)
por: Yang, Zhe-Rui, et al.
Publicado: (2024)
Smoothing DiLoCo with Primal Averaging for Faster Training of LLMs
por: Defazio, Aaron, et al.
Publicado: (2025)
por: Defazio, Aaron, et al.
Publicado: (2025)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
por: Liu, Zhichen, et al.
Publicado: (2026)
por: Liu, Zhichen, et al.
Publicado: (2026)
Downstream Task-Oriented Generative Model Selections on Synthetic Data Training for Fraud Detection Models
por: Cheng, Yinan, et al.
Publicado: (2024)
por: Cheng, Yinan, et al.
Publicado: (2024)
FiCoTS: Fine-to-Coarse LLM-Enhanced Hierarchical Cross-Modality Interaction for Time Series Forecasting
por: Lyu, Yafei, et al.
Publicado: (2025)
por: Lyu, Yafei, et al.
Publicado: (2025)
Ejemplares similares
-
Scaling Laws for Predicting Downstream Performance in LLMs
por: Chen, Yangyi, et al.
Publicado: (2024) -
Unveiling Downstream Performance Scaling of LLMs: A Clustering-Based Perspective
por: Xu, Chengyin, et al.
Publicado: (2025) -
QaRL: Rollout-Aligned Quantization-Aware RL for Fast and Stable Training under Training--Inference Mismatch
por: Gu, Hao, et al.
Publicado: (2026) -
Graph VQ-Transformer (GVT): Fast and Accurate Molecular Generation via High-Fidelity Discrete Latents
por: Zheng, Haozhuo, et al.
Publicado: (2025) -
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
por: Chen, Xingwu, et al.
Publicado: (2025)