A Thorough Examination of Decoding Methods in the Era of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Chufan, Yang, Haoran, Cai, Deng, Zhang, Zhisong, Wang, Yifan, Yang, Yujiu, Lam, Wai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LiFi: Lightweight Controlled Text Generation with Fine-Grained Control Codes
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
A Frustratingly Simple Decoding Method for Neural Text Generation
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
LLM2: Let Large Language Models Harness System 2 Reasoning
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
Reasons to Reject? Aligning Language Models with Judgments
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
On the Worst Prompt Performance of Large Language Models
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
InsCL: A Data-efficient Continual Learning Paradigm for Fine-tuning Large Language Models with Instructions
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
InComeS: Integrating Compression and Selection Mechanisms into LLMs for Efficient Model Editing
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
Hint-enhanced In-Context Learning wakes Large Language Models up for knowledge-intensive tasks
von: Wang, Yifan, et al.
Veröffentlicht: (2023)
von: Wang, Yifan, et al.
Veröffentlicht: (2023)
MaskCD: Mitigating LVLM Hallucinations by Image Head Masked Contrastive Decoding
von: Deng, Jingyuan, et al.
Veröffentlicht: (2025)
von: Deng, Jingyuan, et al.
Veröffentlicht: (2025)
Unchosen Experts Can Contribute Too: Unleashing MoE Models' Power by Self-Contrast
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
A Survey on the Honesty of Large Language Models
von: Li, Siheng, et al.
Veröffentlicht: (2024)
von: Li, Siheng, et al.
Veröffentlicht: (2024)
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
Not All Preference Pairs Are Created Equal: A Recipe for Annotation-Efficient Iterative Preference Learning
von: Yang, Sen, et al.
Veröffentlicht: (2024)
von: Yang, Sen, et al.
Veröffentlicht: (2024)
DepWiGNN: A Depth-wise Graph Neural Network for Multi-hop Spatial Reasoning in Text
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
InfiniteICL: Breaking the Limit of Context Window Size via Long Short-term Memory Transformation
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
Consecutive Batch Model Editing with HooK Layers
von: Li, Shuaiyi, et al.
Veröffentlicht: (2024)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2024)
StrategyLLM: Large Language Models as Strategy Generators, Executors, Optimizers, and Evaluators for Problem Solving
von: Gao, Chang, et al.
Veröffentlicht: (2023)
von: Gao, Chang, et al.
Veröffentlicht: (2023)
Large Language Models Can Self-Improve in Long-context Reasoning
von: Li, Siheng, et al.
Veröffentlicht: (2024)
von: Li, Siheng, et al.
Veröffentlicht: (2024)
HoLLMwood: Unleashing the Creativity of Large Language Models in Screenwriting via Role Playing
von: Chen, Jing, et al.
Veröffentlicht: (2024)
von: Chen, Jing, et al.
Veröffentlicht: (2024)
Robust and Minimally Invasive Watermarking for EaaS
von: Wang, Zongqi, et al.
Veröffentlicht: (2024)
von: Wang, Zongqi, et al.
Veröffentlicht: (2024)
Chain-of-Dictionary Prompting Elicits Translation in Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
Atomic Calibration of LLMs in Long-Form Generations
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
Unveiling the Generalization Power of Fine-Tuned Large Language Models
von: Yang, Haoran, et al.
Veröffentlicht: (2024)
von: Yang, Haoran, et al.
Veröffentlicht: (2024)
ChartMimic: Evaluating LMM's Cross-Modal Reasoning Capability via Chart-to-Code Generation
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
SCAN: Structured Capability Assessment and Navigation for LLMs
von: Wang, Zongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zongqi, et al.
Veröffentlicht: (2025)
Can LLMs Act as Historians? Evaluating Historical Research Capabilities of LLMs via the Chinese Imperial Examination
von: Gao, Lirong, et al.
Veröffentlicht: (2026)
von: Gao, Lirong, et al.
Veröffentlicht: (2026)
ContextVis: Envision Contextual Learning and Interaction with Generative Models
von: Shui, Bo, et al.
Veröffentlicht: (2024)
von: Shui, Bo, et al.
Veröffentlicht: (2024)
PTD-SQL: Partitioning and Targeted Drilling with LLMs in Text-to-SQL
von: Luo, Ruilin, et al.
Veröffentlicht: (2024)
von: Luo, Ruilin, et al.
Veröffentlicht: (2024)
Unlocking Multimodal Mathematical Reasoning via Process Reward Model
von: Luo, Ruilin, et al.
Veröffentlicht: (2025)
von: Luo, Ruilin, et al.
Veröffentlicht: (2025)
ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2024)
Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs
von: Gao, Xin, et al.
Veröffentlicht: (2026)
von: Gao, Xin, et al.
Veröffentlicht: (2026)
TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild
von: Li, Huayang, et al.
Veröffentlicht: (2023)
von: Li, Huayang, et al.
Veröffentlicht: (2023)
Stephanie: Step-by-Step Dialogues for Mimicking Human Interactions in Social Conversations
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
Towards Automated Kernel Generation in the Era of LLMs
von: Yu, Yang, et al.
Veröffentlicht: (2026)
von: Yu, Yang, et al.
Veröffentlicht: (2026)
CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges
von: Wang, Zi-Han, et al.
Veröffentlicht: (2026)
von: Wang, Zi-Han, et al.
Veröffentlicht: (2026)
Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers
von: Yu, Guoxin, et al.
Veröffentlicht: (2026)
von: Yu, Guoxin, et al.
Veröffentlicht: (2026)
LoCa: Logit Calibration for Knowledge Distillation
von: Yang, Runming, et al.
Veröffentlicht: (2024)
von: Yang, Runming, et al.
Veröffentlicht: (2024)
Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
Large Search Model: Redefining Search Stack in the Era of LLMs
von: Wang, Liang, et al.
Veröffentlicht: (2023)
von: Wang, Liang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LiFi: Lightweight Controlled Text Generation with Fine-Grained Control Codes
von: Shi, Chufan, et al.
Veröffentlicht: (2024) -
A Frustratingly Simple Decoding Method for Neural Text Generation
von: Yang, Haoran, et al.
Veröffentlicht: (2023) -
LLM2: Let Large Language Models Harness System 2 Reasoning
von: Yang, Cheng, et al.
Veröffentlicht: (2024) -
Reasons to Reject? Aligning Language Models with Judgments
von: Xu, Weiwen, et al.
Veröffentlicht: (2023) -
On the Worst Prompt Performance of Large Language Models
von: Cao, Bowen, et al.
Veröffentlicht: (2024)