Exploring and Reshaping the Weight Distribution in LLM
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Chunming, Li, Songzhou, Xu, Xu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Analysis on distribution and clustering of weight
by: Ye, Chunming, et al.
Published: (2025)
by: Ye, Chunming, et al.
Published: (2025)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
Softmax Linear Attention: Reclaiming Global Competition
by: Xu, Mingwei, et al.
Published: (2026)
by: Xu, Mingwei, et al.
Published: (2026)
CIKT: A Collaborative and Iterative Knowledge Tracing Framework with Large Language Models
by: Li, Runze, et al.
Published: (2025)
by: Li, Runze, et al.
Published: (2025)
Investigating on RLHF methodology
by: Kutalev, Alexey, et al.
Published: (2024)
by: Kutalev, Alexey, et al.
Published: (2024)
Mixing Times of Glauber Dynamics on Masked Language Models
by: Sana, Suvadip, et al.
Published: (2026)
by: Sana, Suvadip, et al.
Published: (2026)
Exploring the Effectiveness of Instruction Tuning in Biomedical Language Processing
by: Rohanian, Omid, et al.
Published: (2023)
by: Rohanian, Omid, et al.
Published: (2023)
Investigating Distributions of Telecom Adapted Sentence Embeddings for Document Retrieval
by: Roychowdhury, Sujoy, et al.
Published: (2024)
by: Roychowdhury, Sujoy, et al.
Published: (2024)
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation
by: Rouzegar, Hamidreza, et al.
Published: (2024)
by: Rouzegar, Hamidreza, et al.
Published: (2024)
Distractor Injection Attacks on Large Reasoning Models: Characterization and Defense
by: Zhang, Zhehao, et al.
Published: (2025)
by: Zhang, Zhehao, et al.
Published: (2025)
RoleRAG: Enhancing LLM Role-Playing via Graph Guided Retrieval
by: Wang, Yongjie, et al.
Published: (2025)
by: Wang, Yongjie, et al.
Published: (2025)
ProMedTS: A Self-Supervised, Prompt-Guided Multimodal Approach for Integrating Medical Text and Time Series
by: Niu, Shuai, et al.
Published: (2025)
by: Niu, Shuai, et al.
Published: (2025)
A Confidence-Diversity Framework for Calibrating AI Judgement in Accessible Qualitative Coding Tasks
by: Zhao, Zhilong, et al.
Published: (2025)
by: Zhao, Zhilong, et al.
Published: (2025)
Multi-Model Synthetic Training for Mission-Critical Small Language Models
by: Platt, Nolan, et al.
Published: (2025)
by: Platt, Nolan, et al.
Published: (2025)
The Geometry of Persona: Disentangling Personality from Reasoning in Large Language Models
by: Wang, Zhixiang
Published: (2025)
by: Wang, Zhixiang
Published: (2025)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
by: Henry, James
Published: (2026)
by: Henry, James
Published: (2026)
An NLP-Driven Framework for Curriculum-Labor Market Alignment: Schema-Constrained LLM Extraction, ESCO-Anchored Semantic Matching, and Multi-Dimensional Gap Quantification
by: Turaev, Sherzod, et al.
Published: (2026)
by: Turaev, Sherzod, et al.
Published: (2026)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
by: Otal, Hakan T., et al.
Published: (2024)
by: Otal, Hakan T., et al.
Published: (2024)
Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
by: Xu, Beining, et al.
Published: (2025)
by: Xu, Beining, et al.
Published: (2025)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
by: Garg, Saloni, et al.
Published: (2026)
by: Garg, Saloni, et al.
Published: (2026)
PairCFR: Enhancing Model Training on Paired Counterfactually Augmented Data through Contrastive Learning
by: Qiu, Xiaoqi, et al.
Published: (2024)
by: Qiu, Xiaoqi, et al.
Published: (2024)
Copyright Detection in Large Language Models: An Ethical Approach to Generative AI Development
by: Szczecina, David, et al.
Published: (2025)
by: Szczecina, David, et al.
Published: (2025)
Revisiting a Pain in the Neck: Semantic Phrase Processing Benchmark for Language Models
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Empowering Tabular Data Preparation with Language Models: Why and How?
by: Chen, Mengshi, et al.
Published: (2025)
by: Chen, Mengshi, et al.
Published: (2025)
$\text{Memory}^3$: Language Modeling with Explicit Memory
by: Yang, Hongkang, et al.
Published: (2024)
by: Yang, Hongkang, et al.
Published: (2024)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
by: Wang, Yihao, et al.
Published: (2026)
by: Wang, Yihao, et al.
Published: (2026)
CSTRL: Context-Driven Sequential Transfer Learning for Abstractive Radiology Report Summarization
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2025)
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2025)
Fast and Fluent Diffusion Language Models via Convolutional Decoding and Rejective Fine-tuning
by: Seo, Yeongbin, et al.
Published: (2025)
by: Seo, Yeongbin, et al.
Published: (2025)
Aligning Black-box Language Models with Human Judgments
by: Burg, Gerrit J. J. van den, et al.
Published: (2025)
by: Burg, Gerrit J. J. van den, et al.
Published: (2025)
Sparse Logit Sampling: Accelerating Knowledge Distillation in LLMs
by: Anshumann, et al.
Published: (2025)
by: Anshumann, et al.
Published: (2025)
Bielik v3 Small: Technical Report
by: Ociepa, Krzysztof, et al.
Published: (2025)
by: Ociepa, Krzysztof, et al.
Published: (2025)
Recursive Training Loops in LLMs: How training data properties modulate distribution shift in generated data?
by: Kovač, Grgur, et al.
Published: (2025)
by: Kovač, Grgur, et al.
Published: (2025)
ENIGMA: The Geometry of Reasoning and Alignment in Large-Language Models
by: Seneque, Gareth, et al.
Published: (2025)
by: Seneque, Gareth, et al.
Published: (2025)
ATLAS: Constitution-Conditioned Latent Geometry and Redistribution Across Language Models and Neural Perturbation Data
by: Seneque, Gareth, et al.
Published: (2026)
by: Seneque, Gareth, et al.
Published: (2026)
Generative AI for Enhancing Active Learning in Education: A Comparative Study of GPT-3.5 and GPT-4 in Crafting Customized Test Questions
by: Rouzegar, Hamdireza, et al.
Published: (2024)
by: Rouzegar, Hamdireza, et al.
Published: (2024)
Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting
by: Richter-Pechanski, Phillip, et al.
Published: (2024)
by: Richter-Pechanski, Phillip, et al.
Published: (2024)
From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering
by: Santos, José Guilherme Marques dos, et al.
Published: (2026)
by: Santos, José Guilherme Marques dos, et al.
Published: (2026)
KIT-TIP-NLP at MultiPride: Continual Learning with Multilingual Foundation Model
by: HB, Barathi Ganesh, et al.
Published: (2026)
by: HB, Barathi Ganesh, et al.
Published: (2026)
Lightweight Transformers for Clinical Natural Language Processing
by: Rohanian, Omid, et al.
Published: (2023)
by: Rohanian, Omid, et al.
Published: (2023)
TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning
by: Yang, Yujiao
Published: (2026)
by: Yang, Yujiao
Published: (2026)
Similar Items
-
Analysis on distribution and clustering of weight
by: Ye, Chunming, et al.
Published: (2025) -
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026) -
Softmax Linear Attention: Reclaiming Global Competition
by: Xu, Mingwei, et al.
Published: (2026) -
CIKT: A Collaborative and Iterative Knowledge Tracing Framework with Large Language Models
by: Li, Runze, et al.
Published: (2025) -
Investigating on RLHF methodology
by: Kutalev, Alexey, et al.
Published: (2024)