Hansel: Output Length Controlling Framework for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Seoha, Lee, Junhyun, Ko, Hyeonmok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BRIDO: Bringing Democratic Order to Abstractive Summarization
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
zFLoRA: Zero-Latency Fused Low-Rank Adapters
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
Efficient Compositional Multi-tasking for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
On-device System of Compositional Multi-tasking in Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
FLoRA: Fused forward-backward adapters for parameter efficient fine-tuning and reducing inference-time latencies of LLMs
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling
von: Ding, Ning, et al.
Veröffentlicht: (2026)
von: Ding, Ning, et al.
Veröffentlicht: (2026)
Subgraph-level Universal Prompt Tuning
von: Lee, Junhyun, et al.
Veröffentlicht: (2024)
von: Lee, Junhyun, et al.
Veröffentlicht: (2024)
SLOT: Structuring the Output of Large Language Models
von: Wang, Darren Yow-Bang, et al.
Veröffentlicht: (2025)
von: Wang, Darren Yow-Bang, et al.
Veröffentlicht: (2025)
Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2026)
von: Luo, Feng, et al.
Veröffentlicht: (2026)
AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs
von: Pezeshkpour, Pouya, et al.
Veröffentlicht: (2026)
von: Pezeshkpour, Pouya, et al.
Veröffentlicht: (2026)
An Evaluation on Large Language Model Outputs: Discourse and Memorization
von: de Wynter, Adrian, et al.
Veröffentlicht: (2023)
von: de Wynter, Adrian, et al.
Veröffentlicht: (2023)
Quantifying the Impact of Structured Output Format on Large Language Models through Causal Inference
von: Yuan, Han, et al.
Veröffentlicht: (2025)
von: Yuan, Han, et al.
Veröffentlicht: (2025)
Beyond the Limits: A Survey of Techniques to Extend the Context Length in Large Language Models
von: Wang, Xindi, et al.
Veröffentlicht: (2024)
von: Wang, Xindi, et al.
Veröffentlicht: (2024)
Improving Variable-Length Generation in Diffusion Language Models via Length Regularization
von: Cheng, Zicong, et al.
Veröffentlicht: (2026)
von: Cheng, Zicong, et al.
Veröffentlicht: (2026)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
Adaptive Task Vectors for Large Language Models
von: Kang, Joonseong, et al.
Veröffentlicht: (2025)
von: Kang, Joonseong, et al.
Veröffentlicht: (2025)
DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
von: Jung, Sunghee, et al.
Veröffentlicht: (2025)
von: Jung, Sunghee, et al.
Veröffentlicht: (2025)
Diffusion Language Models Are Natively Length-Aware
von: Rossi, Vittorio, et al.
Veröffentlicht: (2026)
von: Rossi, Vittorio, et al.
Veröffentlicht: (2026)
Confidence Regularized Masked Language Modeling using Text Length
von: Ji, Seunghyun, et al.
Veröffentlicht: (2025)
von: Ji, Seunghyun, et al.
Veröffentlicht: (2025)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
Small Language Models Improve Giants by Rewriting Their Outputs
von: Vernikos, Giorgos, et al.
Veröffentlicht: (2023)
von: Vernikos, Giorgos, et al.
Veröffentlicht: (2023)
Group-Aware Reinforcement Learning for Output Diversity in Large Language Models
von: Anschel, Oron, et al.
Veröffentlicht: (2025)
von: Anschel, Oron, et al.
Veröffentlicht: (2025)
VEHME: A Vision-Language Model For Evaluating Handwritten Mathematics Expressions
von: Nguyen, Thu Phuong, et al.
Veröffentlicht: (2025)
von: Nguyen, Thu Phuong, et al.
Veröffentlicht: (2025)
Phase Transitions in the Output Distribution of Large Language Models
von: Arnold, Julian, et al.
Veröffentlicht: (2024)
von: Arnold, Julian, et al.
Veröffentlicht: (2024)
Controlling Large Language Model with Latent Actions
von: Jia, Chengxing, et al.
Veröffentlicht: (2025)
von: Jia, Chengxing, et al.
Veröffentlicht: (2025)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
von: Krishna, Kundan, et al.
Veröffentlicht: (2024)
von: Krishna, Kundan, et al.
Veröffentlicht: (2024)
Softplus Attention with Re-weighting Boosts Length Extrapolation in Large Language Models
von: Gao, Bo, et al.
Veröffentlicht: (2025)
von: Gao, Bo, et al.
Veröffentlicht: (2025)
A Survey of On-Policy Distillation for Large Language Models
von: Song, Mingyang, et al.
Veröffentlicht: (2026)
von: Song, Mingyang, et al.
Veröffentlicht: (2026)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2024)
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2024)
AXOLOTL: Fairness through Assisted Self-Debiasing of Large Language Model Outputs
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2024)
Lizard: An Efficient Linearization Framework for Large Language Models
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
Selective Generation for Controllable Language Models
von: Lee, Minjae, et al.
Veröffentlicht: (2023)
von: Lee, Minjae, et al.
Veröffentlicht: (2023)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
von: Briesch, Martin, et al.
Veröffentlicht: (2023)
von: Briesch, Martin, et al.
Veröffentlicht: (2023)
Length-MAX Tokenizer for Language Models
von: Dong, Dong, et al.
Veröffentlicht: (2025)
von: Dong, Dong, et al.
Veröffentlicht: (2025)
A Voter-Based Stochastic Rejection-Method Framework for Asymptotically Safe Language Model Outputs
von: Watts, Jake R., et al.
Veröffentlicht: (2024)
von: Watts, Jake R., et al.
Veröffentlicht: (2024)
DistiLLM: Towards Streamlined Distillation for Large Language Models
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
ELMES: An Automated Framework for Evaluating Large Language Models in Educational Scenarios
von: Wei, Shou'ang, et al.
Veröffentlicht: (2025)
von: Wei, Shou'ang, et al.
Veröffentlicht: (2025)
Time Will Tell: Timing Side Channels via Output Token Count in Large Language Models
von: Zhang, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhang, Tianchen, et al.
Veröffentlicht: (2024)
Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs
von: Su, Guinan, et al.
Veröffentlicht: (2026)
von: Su, Guinan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
BRIDO: Bringing Democratic Order to Abstractive Summarization
von: Lee, Junhyun, et al.
Veröffentlicht: (2025) -
zFLoRA: Zero-Latency Fused Low-Rank Adapters
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025) -
Efficient Compositional Multi-tasking for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025) -
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026) -
On-device System of Compositional Multi-tasking in Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)