SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Guoxuan, Shi, Han, Li, Jiawei, Gao, Yihang, Ren, Xiaozhe, Chen, Yimeng, Jiang, Xin, Li, Zhenguo, Liu, Weiyang, Huang, Chao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Adjust Softmax
by: Zheng, Chuanyang, et al.
Published: (2025)
by: Zheng, Chuanyang, et al.
Published: (2025)
DAPE: Data-Adaptive Positional Encoding for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024)
by: Zheng, Chuanyang, et al.
Published: (2024)
QuickLLaMA: Query-aware Inference Acceleration for Large Language Models
by: Li, Jingyao, et al.
Published: (2024)
by: Li, Jingyao, et al.
Published: (2024)
DAPE V2: Process Attention Score as Feature Map for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024)
by: Zheng, Chuanyang, et al.
Published: (2024)
Scaling Law for Language Models Training Considering Batch Size
by: Shuai, Xian, et al.
Published: (2024)
by: Shuai, Xian, et al.
Published: (2024)
Efficient Multi-modal Large Language Models via Visual Token Grouping
by: Huang, Minbin, et al.
Published: (2024)
by: Huang, Minbin, et al.
Published: (2024)
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
by: Teng, Yao, et al.
Published: (2025)
by: Teng, Yao, et al.
Published: (2025)
OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner
by: Jiang, Haoyang, et al.
Published: (2026)
by: Jiang, Haoyang, et al.
Published: (2026)
One‐Pot Synthesis of Thiophenoxyimino Vanadium Complexes and Highly Active Catalysis on Ethylene Polymerization
by: Xiaoyu Qi, et al.
Published: (2025)
by: Xiaoyu Qi, et al.
Published: (2025)
StableCodec: Taming One-Step Diffusion for Extreme Image Compression
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
Pre-training for Recommendation Unlearning
by: Chen, Guoxuan, et al.
Published: (2025)
by: Chen, Guoxuan, et al.
Published: (2025)
LightGNN: Simple Graph Neural Network for Recommendation
by: Chen, Guoxuan, et al.
Published: (2025)
by: Chen, Guoxuan, et al.
Published: (2025)
Generative Video Compression with One-Dimensional Latent Representation
by: Zheng, Zihan, et al.
Published: (2026)
by: Zheng, Zihan, et al.
Published: (2026)
Few-Shot Medical Image Segmentation with Large Kernel Attention
by: Wu, Xiaoxiao, et al.
Published: (2024)
by: Wu, Xiaoxiao, et al.
Published: (2024)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
by: Yu, Longhui, et al.
Published: (2023)
by: Yu, Longhui, et al.
Published: (2023)
OMG-Seg: Is One Model Good Enough For All Segmentation?
by: Li, Xiangtai, et al.
Published: (2024)
by: Li, Xiangtai, et al.
Published: (2024)
OneLLM: One Framework to Align All Modalities with Language
by: Han, Jiaming, et al.
Published: (2023)
by: Han, Jiaming, et al.
Published: (2023)
UniPool: A Globally Shared Expert Pool for Mixture-of-Experts
by: Huang, Minbin, et al.
Published: (2026)
by: Huang, Minbin, et al.
Published: (2026)
3One2: One-step Regression Plus One-step Diffusion for One-hot Modulation in Dual-path Video Snapshot Compressive Imaging
by: Wang, Ge, et al.
Published: (2025)
by: Wang, Ge, et al.
Published: (2025)
Answer-Consistent Chain-of-thought Reinforcement Learning For Multi-modal Large Langauge Models
by: Huang, Minbin, et al.
Published: (2025)
by: Huang, Minbin, et al.
Published: (2025)
MARS-Sep: Multimodal-Aligned Reinforced Sound Separation
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
All-in-One Image Compression and Restoration
by: Zeng, Huimin, et al.
Published: (2025)
by: Zeng, Huimin, et al.
Published: (2025)
Is One Score Enough? Rethinking the Evaluation of Sequentially Evolving LLM Memory
by: Dong, Songwei, et al.
Published: (2026)
by: Dong, Songwei, et al.
Published: (2026)
One Step Learning, One Step Review
by: Huang, Xiaolong, et al.
Published: (2024)
by: Huang, Xiaolong, et al.
Published: (2024)
SepPrune: Structured Pruning for Efficient Deep Speech Separation
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Engineering Separated Dual O2 Reduction Cores into One Polymer Framework for Boosting Hydrogen Peroxide Production
by: Mengmeng Fu, et al.
Published: (2025)
by: Mengmeng Fu, et al.
Published: (2025)
Compression-Aware One-Step Diffusion Model for JPEG Artifact Removal
by: Guo, Jinpei, et al.
Published: (2025)
by: Guo, Jinpei, et al.
Published: (2025)
Reinforcement Learning for Chain of Thought Compression with One-Domain-to-All Generalization
by: Li, Hanyu, et al.
Published: (2025)
by: Li, Hanyu, et al.
Published: (2025)
One Signal-Noise Separation based Wiener Filter for Magnetogastrogram
by: Li, Hua
Published: (2023)
by: Li, Hua
Published: (2023)
Self-Retrieval: End-to-End Information Retrieval with One Large Language Model
by: Tang, Qiaoyu, et al.
Published: (2024)
by: Tang, Qiaoyu, et al.
Published: (2024)
OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization Problems
by: Li, Xiaozhe, et al.
Published: (2025)
by: Li, Xiaozhe, et al.
Published: (2025)
Enhancing Shape Perception and Segmentation Consistency for Industrial Image Inspection
by: Mao, Guoxuan, et al.
Published: (2025)
by: Mao, Guoxuan, et al.
Published: (2025)
Compression and Acceleration of Neural Networks for Communications
by: Guo, Jiajia, et al.
Published: (2019)
by: Guo, Jiajia, et al.
Published: (2019)
OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces
by: Li, Xiaozhe, et al.
Published: (2026)
by: Li, Xiaozhe, et al.
Published: (2026)
Oracle Separation Between Quantum Commitments and Quantum One-wayness
by: Bostanci, John, et al.
Published: (2024)
by: Bostanci, John, et al.
Published: (2024)
CoCA: Regaining Safety-awareness of Multimodal Large Language Models with Constitutional Calibration
by: Gao, Jiahui, et al.
Published: (2024)
by: Gao, Jiahui, et al.
Published: (2024)
When One LLM Drools, Multi-LLM Collaboration Rules
by: Feng, Shangbin, et al.
Published: (2025)
by: Feng, Shangbin, et al.
Published: (2025)
One-shot Training for Video Object Segmentation
by: Chen, Baiyu, et al.
Published: (2024)
by: Chen, Baiyu, et al.
Published: (2024)
UniSep: Universal Target Audio Separation with Language Models at Scale
by: Wang, Yuanyuan, et al.
Published: (2025)
by: Wang, Yuanyuan, et al.
Published: (2025)
Adding Additional Control to One-Step Diffusion with Joint Distribution Matching
by: Luo, Yihong, et al.
Published: (2025)
by: Luo, Yihong, et al.
Published: (2025)
Similar Items
-
Self-Adjust Softmax
by: Zheng, Chuanyang, et al.
Published: (2025) -
DAPE: Data-Adaptive Positional Encoding for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024) -
QuickLLaMA: Query-aware Inference Acceleration for Large Language Models
by: Li, Jingyao, et al.
Published: (2024) -
DAPE V2: Process Attention Score as Feature Map for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024) -
Scaling Law for Language Models Training Considering Batch Size
by: Shuai, Xian, et al.
Published: (2024)