SWSC: Shared Weight for Similar Channel in LLM
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Binrui, Tang, Yongtao, Liu, Xiaodong, Li, Xiaopeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LSAQ: Layer-Specific Adaptive Quantization for Large Language Model Deployment
by: Zeng, Binrui, et al.
Published: (2024)
by: Zeng, Binrui, et al.
Published: (2024)
Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
by: Wu, Di, et al.
Published: (2023)
by: Wu, Di, et al.
Published: (2023)
FlexiGPT: Pruning and Extending Large Language Models with Low-Rank Weight Sharing
by: Smith, James Seale, et al.
Published: (2025)
by: Smith, James Seale, et al.
Published: (2025)
Share Your Attention: Transformer Weight Sharing via Matrix-based Dictionary Learning
by: Zhussip, Magauiya, et al.
Published: (2025)
by: Zhussip, Magauiya, et al.
Published: (2025)
Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
by: Nguyen, Dang, et al.
Published: (2025)
by: Nguyen, Dang, et al.
Published: (2025)
Quantifying True Robustness: Synonymity-Weighted Similarity for Trustworthy XAI Evaluation
by: Burger, Christopher
Published: (2025)
by: Burger, Christopher
Published: (2025)
Variance Control via Weight Rescaling in LLM Pre-training
by: Owen, Louis, et al.
Published: (2025)
by: Owen, Louis, et al.
Published: (2025)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
by: Ren, Jie, et al.
Published: (2025)
by: Ren, Jie, et al.
Published: (2025)
Similar Phrases for Cause of Actions of Civil Cases
by: Huang, Ho-Chien, et al.
Published: (2024)
by: Huang, Ho-Chien, et al.
Published: (2024)
Basis Sharing: Cross-Layer Parameter Sharing for Large Language Model Compression
by: Wang, Jingcun, et al.
Published: (2024)
by: Wang, Jingcun, et al.
Published: (2024)
Not All LLM-Generated Data Are Equal: Rethinking Data Weighting in Text Classification
by: Kuo, Hsun-Yu, et al.
Published: (2024)
by: Kuo, Hsun-Yu, et al.
Published: (2024)
AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization
by: IslamBouli, Beshr, et al.
Published: (2026)
by: IslamBouli, Beshr, et al.
Published: (2026)
WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling
by: Li, Jiacheng, et al.
Published: (2025)
by: Li, Jiacheng, et al.
Published: (2025)
SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning
by: Ai, Zhengyang, et al.
Published: (2026)
by: Ai, Zhengyang, et al.
Published: (2026)
DefSent+: Improving sentence embeddings of language models by projecting definition sentences into a quasi-isotropic or isotropic vector space of unlimited dictionary entries
by: Liu, Xiaodong
Published: (2024)
by: Liu, Xiaodong
Published: (2024)
PolarQuant: Optimal Gaussian Weight Quantization via Hadamard Rotation for LLM Compression
by: Vicentino, Caio
Published: (2026)
by: Vicentino, Caio
Published: (2026)
Weight-of-Thought Reasoning: Exploring Neural Network Weights for Enhanced LLM Reasoning
by: Punjwani, Saif, et al.
Published: (2025)
by: Punjwani, Saif, et al.
Published: (2025)
ResidualTransformer: Residual Low-Rank Learning with Weight-Sharing for Transformer Layers
by: Wang, Yiming, et al.
Published: (2023)
by: Wang, Yiming, et al.
Published: (2023)
Similarity-Distance-Magnitude Activations
by: Schmaltz, Allen
Published: (2025)
by: Schmaltz, Allen
Published: (2025)
LLM Unlearning with LLM Beliefs
by: Li, Kemou, et al.
Published: (2025)
by: Li, Kemou, et al.
Published: (2025)
Mitigating Heterogeneous Token Overfitting in LLM Knowledge Editing
by: Liu, Tianci, et al.
Published: (2025)
by: Liu, Tianci, et al.
Published: (2025)
EMSEdit: Efficient Multi-Step Meta-Learning-based Model Editing
by: Li, Xiaopeng, et al.
Published: (2025)
by: Li, Xiaopeng, et al.
Published: (2025)
Towards Next-Generation LLM Training: From the Data-Centric Perspective
by: Liang, Hao, et al.
Published: (2026)
by: Liang, Hao, et al.
Published: (2026)
In-Context Transfer Learning: Demonstration Synthesis by Transferring Similar Tasks
by: Wang, Dingzirui, et al.
Published: (2024)
by: Wang, Dingzirui, et al.
Published: (2024)
Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets
by: Hsiung, Lei, et al.
Published: (2025)
by: Hsiung, Lei, et al.
Published: (2025)
Similarity-Distance-Magnitude Universal Verification
by: Schmaltz, Allen
Published: (2025)
by: Schmaltz, Allen
Published: (2025)
Explaining Text Similarity in Transformer Models
by: Vasileiou, Alexandros, et al.
Published: (2024)
by: Vasileiou, Alexandros, et al.
Published: (2024)
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
by: Lin, Zihan, et al.
Published: (2026)
by: Lin, Zihan, et al.
Published: (2026)
Predicting Machine Translation Performance on Low-Resource Languages: The Role of Domain Similarity
by: Khiu, Eric, et al.
Published: (2024)
by: Khiu, Eric, et al.
Published: (2024)
Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion
by: Dai, Wei, et al.
Published: (2024)
by: Dai, Wei, et al.
Published: (2024)
Human Texts Are Outliers: Detecting LLM-generated Texts via Out-of-distribution Detection
by: Zeng, Cong, et al.
Published: (2025)
by: Zeng, Cong, et al.
Published: (2025)
SWEA: Updating Factual Knowledge in Large Language Models via Subject Word Embedding Altering
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
ReMix: Reinforcement routing for mixtures of LoRAs in LLM finetuning
by: Qiu, Ruizhong, et al.
Published: (2026)
by: Qiu, Ruizhong, et al.
Published: (2026)
DroidSpeak: KV Cache Sharing for Cross-LLM Communication and Multi-LLM Serving
by: Liu, Yuhan, et al.
Published: (2024)
by: Liu, Yuhan, et al.
Published: (2024)
Efficient Prompt Caching via Embedding Similarity
by: Zhu, Hanlin, et al.
Published: (2024)
by: Zhu, Hanlin, et al.
Published: (2024)
CHESS: Optimizing LLM Inference via Channel-Wise Thresholding and Selective Sparsification
by: He, Junhui, et al.
Published: (2024)
by: He, Junhui, et al.
Published: (2024)
Optimal Transport-Based Token Weighting scheme for Enhanced Preference Optimization
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
Freshness-Aware Prioritized Experience Replay for LLM/VLM Reinforcement Learning
by: Ma, Weiyu, et al.
Published: (2026)
by: Ma, Weiyu, et al.
Published: (2026)
PEANuT: Parameter-Efficient Adaptation with Weight-aware Neural Tweakers
by: Zhong, Yibo, et al.
Published: (2024)
by: Zhong, Yibo, et al.
Published: (2024)
TRACES: Proactive Safety Auditing for Multi-Turn LLM Agents via Trajectory-State Modeling
by: Li, Jiaqian, et al.
Published: (2026)
by: Li, Jiaqian, et al.
Published: (2026)
Similar Items
-
LSAQ: Layer-Specific Adaptive Quantization for Large Language Model Deployment
by: Zeng, Binrui, et al.
Published: (2024) -
Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
by: Wu, Di, et al.
Published: (2023) -
FlexiGPT: Pruning and Extending Large Language Models with Low-Rank Weight Sharing
by: Smith, James Seale, et al.
Published: (2025) -
Share Your Attention: Transformer Weight Sharing via Matrix-based Dictionary Learning
by: Zhussip, Magauiya, et al.
Published: (2025) -
Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
by: Nguyen, Dang, et al.
Published: (2025)