From Parameters to Performance: A Data-Driven Study on LLM Structure and Development
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Suqing, Li, Zuchao, Shi, Luohe, Du, Bo, Zhao, Hai, Li, Yun, Wang, Qianren |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
by: Song, Weixi, et al.
Published: (2023)
by: Song, Weixi, et al.
Published: (2023)
Faster MoE LLM Inference for Extremely Large Models
by: Yang, Haoqi, et al.
Published: (2025)
by: Yang, Haoqi, et al.
Published: (2025)
CoViPAL: Layer-wise Contextualized Visual Token Pruning for Large Vision-Language Models
by: Tang, Zicong, et al.
Published: (2025)
by: Tang, Zicong, et al.
Published: (2025)
SmallKV: Small Model Assisted Compensation of KV Cache Compression for Efficient LLM Inference
by: Zhao, Yi, et al.
Published: (2025)
by: Zhao, Yi, et al.
Published: (2025)
From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
Ghost in the Transformer: Detecting Model Reuse with Invariant Spectral Signatures
by: Wang, Suqing, et al.
Published: (2025)
by: Wang, Suqing, et al.
Published: (2025)
From Noise to Precision: A Diffusion-Driven Approach to Zero-Inflated Precipitation Prediction
by: Gao, Wentao, et al.
Published: (2025)
by: Gao, Wentao, et al.
Published: (2025)
Developing Distance-Aware, and Evident Uncertainty Quantification in Dynamic Physics-Constrained Neural Networks for Robust Bearing Degradation Estimation
by: Razzaq, Waleed, et al.
Published: (2025)
by: Razzaq, Waleed, et al.
Published: (2025)
BatGPT-Chem: A Foundation Large Model For Retrosynthesis Prediction
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
NanoNet: Parameter-Efficient Learning with Label-Scarce Supervision for Lightweight Text Mining Model
by: Mao, Qianren, et al.
Published: (2026)
by: Mao, Qianren, et al.
Published: (2026)
To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents
by: Shi, Wei, et al.
Published: (2026)
by: Shi, Wei, et al.
Published: (2026)
Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance
by: Wang, Shiqiang, et al.
Published: (2026)
by: Wang, Shiqiang, et al.
Published: (2026)
Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture Search
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
Parameter Importance-Driven Continual Learning for Foundation Models
by: Wang, Lingxiang, et al.
Published: (2025)
by: Wang, Lingxiang, et al.
Published: (2025)
XAI for In-hospital Mortality Prediction via Multimodal ICU Data
by: Li, Xingqiao, et al.
Published: (2023)
by: Li, Xingqiao, et al.
Published: (2023)
Controllable Data Generation by Deep Learning: A Review
by: Wang, Shiyu, et al.
Published: (2022)
by: Wang, Shiyu, et al.
Published: (2022)
Astro: Activation-guided Structured Regularization for Outlier-Robust LLM Post-Training Quantization
by: Chen, Xi, et al.
Published: (2026)
by: Chen, Xi, et al.
Published: (2026)
Symmetry in Neural Network Parameter Spaces
by: Zhao, Bo, et al.
Published: (2025)
by: Zhao, Bo, et al.
Published: (2025)
Can LLM Safety Be Ensured by Constraining Parameter Regions?
by: Li, Zongmin, et al.
Published: (2026)
by: Li, Zongmin, et al.
Published: (2026)
ELLMob: Event-Driven Human Mobility Generation with Self-Aligned LLM Framework
by: Wang, Yusong, et al.
Published: (2026)
by: Wang, Yusong, et al.
Published: (2026)
Expert-Guided LLM Reasoning for Battery Discovery: From AI-Driven Hypothesis to Synthesis and Characterization
by: Liu, Shengchao, et al.
Published: (2025)
by: Liu, Shengchao, et al.
Published: (2025)
Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior
by: Yin, Bo, et al.
Published: (2026)
by: Yin, Bo, et al.
Published: (2026)
From Values to Tokens: An LLM-Driven Framework for Context-aware Time Series Forecasting via Symbolic Discretization
by: Tao, Xiaoyu, et al.
Published: (2025)
by: Tao, Xiaoyu, et al.
Published: (2025)
Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity
by: Li, Bojie
Published: (2026)
by: Li, Bojie
Published: (2026)
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
by: Liu, Siyuan, et al.
Published: (2026)
by: Liu, Siyuan, et al.
Published: (2026)
Quantifying the Capability Boundary of DeepSeek Models: An Application-Driven Performance Analysis
by: Zhao, Kaikai, et al.
Published: (2025)
by: Zhao, Kaikai, et al.
Published: (2025)
WLFM: A Well-Logs Foundation Model for Multi-Task and Cross-Well Geological Interpretation
by: Qi, Zhenyu, et al.
Published: (2025)
by: Qi, Zhenyu, et al.
Published: (2025)
SHARP: Accelerating Language Model Inference by SHaring Adjacent layers with Recovery Parameters
by: Wang, Yiping, et al.
Published: (2025)
by: Wang, Yiping, et al.
Published: (2025)
EDGE: Efficient Data Selection for LLM Agents via Guideline Effectiveness
by: Zhang, Yunxiao, et al.
Published: (2025)
by: Zhang, Yunxiao, et al.
Published: (2025)
Can Past Experience Accelerate LLM Reasoning?
by: Pan, Bo, et al.
Published: (2025)
by: Pan, Bo, et al.
Published: (2025)
Less is More: on the Over-Globalizing Problem in Graph Transformers
by: Xing, Yujie, et al.
Published: (2024)
by: Xing, Yujie, et al.
Published: (2024)
Scaling-Aware Adapter for Structure-Grounded LLM Reasoning
by: Jing, Zihao, et al.
Published: (2026)
by: Jing, Zihao, et al.
Published: (2026)
Predicting LLM Reasoning Performance with Small Proxy Model
by: Koh, Woosung, et al.
Published: (2025)
by: Koh, Woosung, et al.
Published: (2025)
An Optimization Algorithm for Multimodal Data Alignment
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
FLUID: Continuous-Time Hyperconnected Sparse Transformer for Sink-Free Learning
by: Razzaq, Waleed, et al.
Published: (2026)
by: Razzaq, Waleed, et al.
Published: (2026)
Neuronal Stochastic Attention Circuit (NSAC) for Probabilistic Representation Learning
by: Razzaq, Waleed, et al.
Published: (2026)
by: Razzaq, Waleed, et al.
Published: (2026)
A Novel Multimodal RUL Framework for Remaining Useful Life Estimation with Layer-wise Explanations
by: Razzaq, Waleed, et al.
Published: (2025)
by: Razzaq, Waleed, et al.
Published: (2025)
CARLE: A Hybrid Deep-Shallow Learning Framework for Robust and Explainable RUL Estimation of Rolling Element Bearings
by: Razzaq, Waleed, et al.
Published: (2025)
by: Razzaq, Waleed, et al.
Published: (2025)
NaiAD: Initiate Data-Driven Research for LLM Advertising
by: Zhang, Yihang, et al.
Published: (2026)
by: Zhang, Yihang, et al.
Published: (2026)
Efficient CNN-LSTM based Parameter Estimation of Levy Driven Stochastic Differential Equations
by: Li, Shuaiyu, et al.
Published: (2024)
by: Li, Shuaiyu, et al.
Published: (2024)
Similar Items
-
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
by: Song, Weixi, et al.
Published: (2023) -
Faster MoE LLM Inference for Extremely Large Models
by: Yang, Haoqi, et al.
Published: (2025) -
CoViPAL: Layer-wise Contextualized Visual Token Pruning for Large Vision-Language Models
by: Tang, Zicong, et al.
Published: (2025) -
SmallKV: Small Model Assisted Compensation of KV Cache Compression for Efficient LLM Inference
by: Zhao, Yi, et al.
Published: (2025) -
From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning
by: Wang, Xiao, et al.
Published: (2026)