Controlling Summarization Length Through EOS Token Weighting
Fuente:
arXiv
Saved in:
| Main Authors: | Belligoli, Zeno, Stergiadis, Emmanouil, Fainman, Eran, Gusev, Ilya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Auto-Regressive Next-Token Predictors are Universal Learners
by: Malach, Eran
Published: (2023)
by: Malach, Eran
Published: (2023)
Prompting and Fine-Tuning of Small LLMs for Length-Controllable Telephone Call Summarization
by: Thulke, David, et al.
Published: (2024)
by: Thulke, David, et al.
Published: (2024)
HotelMatch-LLM: Joint Multi-Task Training of Small and Large Language Models for Efficient Multimodal Hotel Retrieval
by: Askari, Arian, et al.
Published: (2025)
by: Askari, Arian, et al.
Published: (2025)
The Role of Sparsity for Length Generalization in Transformers
by: Golowich, Noah, et al.
Published: (2025)
by: Golowich, Noah, et al.
Published: (2025)
Document Summarization with Conformal Importance Guarantees
by: Kuwahara, Bruce, et al.
Published: (2025)
by: Kuwahara, Bruce, et al.
Published: (2025)
Length-MAX Tokenizer for Language Models
by: Dong, Dong, et al.
Published: (2025)
by: Dong, Dong, et al.
Published: (2025)
Benchmarking Generation and Evaluation Capabilities of Large Language Models for Instruction Controllable Summarization
by: Liu, Yixin, et al.
Published: (2023)
by: Liu, Yixin, et al.
Published: (2023)
PingPong: A Benchmark for Role-Playing Language Models with User Emulation and Multi-Model Evaluation
by: Gusev, Ilya
Published: (2024)
by: Gusev, Ilya
Published: (2024)
Hansel: Output Length Controlling Framework for Large Language Models
by: Song, Seoha, et al.
Published: (2024)
by: Song, Seoha, et al.
Published: (2024)
Mixture of Weight-shared Heterogeneous Group Attention Experts for Dynamic Token-wise KV Optimization
by: Song, Guanghui, et al.
Published: (2025)
by: Song, Guanghui, et al.
Published: (2025)
Weighting What Matters: Boosting Sample Efficiency in Medical Report Generation via Token Reweighting
by: Weers, Alexander, et al.
Published: (2026)
by: Weers, Alexander, et al.
Published: (2026)
TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection
by: Wu, Wei, et al.
Published: (2024)
by: Wu, Wei, et al.
Published: (2024)
Learning to Optimize Multi-Objective Alignment Through Dynamic Reward Weighting
by: Lu, Yining, et al.
Published: (2025)
by: Lu, Yining, et al.
Published: (2025)
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression
by: Elias, Noel, et al.
Published: (2024)
by: Elias, Noel, et al.
Published: (2024)
Dual Process Learning: Controlling Use of In-Context vs. In-Weights Strategies with Weight Forgetting
by: Anand, Suraj, et al.
Published: (2024)
by: Anand, Suraj, et al.
Published: (2024)
Highlight & Summarize: RAG without the jailbreaks
by: Cherubin, Giovanni, et al.
Published: (2025)
by: Cherubin, Giovanni, et al.
Published: (2025)
Logit Reweighting for Topic-Focused Summarization
by: Braun, Joschka, et al.
Published: (2025)
by: Braun, Joschka, et al.
Published: (2025)
Optimal Transport-Based Token Weighting scheme for Enhanced Preference Optimization
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
Query-Focused Extractive Summarization for Sentiment Explanation
by: Moubtahij, Ahmed, et al.
Published: (2025)
by: Moubtahij, Ahmed, et al.
Published: (2025)
BRIDO: Bringing Democratic Order to Abstractive Summarization
by: Lee, Junhyun, et al.
Published: (2025)
by: Lee, Junhyun, et al.
Published: (2025)
PerSEval: Assessing Personalization in Text Summarizers
by: Dasgupta, Sourish, et al.
Published: (2024)
by: Dasgupta, Sourish, et al.
Published: (2024)
DWTSumm: Discrete Wavelet Transform for Document Summarization
by: Salama, Rana, et al.
Published: (2026)
by: Salama, Rana, et al.
Published: (2026)
Improving Variable-Length Generation in Diffusion Language Models via Length Regularization
by: Cheng, Zicong, et al.
Published: (2026)
by: Cheng, Zicong, et al.
Published: (2026)
Improving Faithfulness of Abstractive Summarization by Controlling Confounding Effect of Irrelevant Sentences
by: Ghoshal, Asish, et al.
Published: (2022)
by: Ghoshal, Asish, et al.
Published: (2022)
Guiding Giants: Lightweight Controllers for Weighted Activation Steering in LLMs
by: Hegazy, Amr, et al.
Published: (2025)
by: Hegazy, Amr, et al.
Published: (2025)
Variance Control via Weight Rescaling in LLM Pre-training
by: Owen, Louis, et al.
Published: (2025)
by: Owen, Louis, et al.
Published: (2025)
Multilingual Self-Taught Faithfulness Evaluators
by: Alfano, Carlo, et al.
Published: (2025)
by: Alfano, Carlo, et al.
Published: (2025)
Beyond Multiple Choice: Evaluating Steering Vectors for Summarization
by: Braun, Joschka, et al.
Published: (2025)
by: Braun, Joschka, et al.
Published: (2025)
SUMIE: A Synthetic Benchmark for Incremental Entity Summarization
by: Hwang, Eunjeong, et al.
Published: (2024)
by: Hwang, Eunjeong, et al.
Published: (2024)
Graph Neural Network and NER-Based Text Summarization
by: Khan, Imaad Zaffar, et al.
Published: (2024)
by: Khan, Imaad Zaffar, et al.
Published: (2024)
How to Train Text Summarization Model with Weak Supervisions
by: Wang, Yanbo, et al.
Published: (2024)
by: Wang, Yanbo, et al.
Published: (2024)
Discrete Diffusion Language Model for Efficient Text Summarization
by: Dat, Do Huu, et al.
Published: (2024)
by: Dat, Do Huu, et al.
Published: (2024)
Incremental Extractive Opinion Summarization Using Cover Trees
by: Chowdhury, Somnath Basu Roy, et al.
Published: (2024)
by: Chowdhury, Somnath Basu Roy, et al.
Published: (2024)
On Provable Length and Compositional Generalization
by: Ahuja, Kartik, et al.
Published: (2024)
by: Ahuja, Kartik, et al.
Published: (2024)
DLER: Doing Length pEnalty Right - Incentivizing More Intelligence per Token via Reinforcement Learning
by: Liu, Shih-Yang, et al.
Published: (2025)
by: Liu, Shih-Yang, et al.
Published: (2025)
Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions
by: Sastre, Ignacio, et al.
Published: (2026)
by: Sastre, Ignacio, et al.
Published: (2026)
MOSAIC: Modular Opinion Summarization using Aspect Identification and Clustering
by: Singh, Piyush Kumar, et al.
Published: (2026)
by: Singh, Piyush Kumar, et al.
Published: (2026)
LOCOST: State-Space Models for Long Document Abstractive Summarization
by: Bronnec, Florian Le, et al.
Published: (2024)
by: Bronnec, Florian Le, et al.
Published: (2024)
Event-Keyed Summarization
by: Gantt, William, et al.
Published: (2024)
by: Gantt, William, et al.
Published: (2024)
Token-Weighted RNN-T for Learning from Flawed Data
by: Keren, Gil, et al.
Published: (2024)
by: Keren, Gil, et al.
Published: (2024)
Similar Items
-
Auto-Regressive Next-Token Predictors are Universal Learners
by: Malach, Eran
Published: (2023) -
Prompting and Fine-Tuning of Small LLMs for Length-Controllable Telephone Call Summarization
by: Thulke, David, et al.
Published: (2024) -
HotelMatch-LLM: Joint Multi-Task Training of Small and Large Language Models for Efficient Multimodal Hotel Retrieval
by: Askari, Arian, et al.
Published: (2025) -
The Role of Sparsity for Length Generalization in Transformers
by: Golowich, Noah, et al.
Published: (2025) -
Document Summarization with Conformal Importance Guarantees
by: Kuwahara, Bruce, et al.
Published: (2025)