Explaining and Improving Contrastive Decoding by Extrapolating the Probabilities of a Huge and Hypothetical LM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Haw-Shiuan, Peng, Nanyun, Bansal, Mohit, Ramakrishna, Anil, Chung, Tagyoung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
REAL Sampling: Boosting Factuality and Diversity of Open-Ended Generation via Asymptotic Entropy
von: Chang, Haw-Shiuan, et al.
Veröffentlicht: (2024)
von: Chang, Haw-Shiuan, et al.
Veröffentlicht: (2024)
CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization
von: Joshi, Brihi, et al.
Veröffentlicht: (2025)
von: Joshi, Brihi, et al.
Veröffentlicht: (2025)
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
von: Ferraz, Thomas Palmeira, et al.
Veröffentlicht: (2024)
von: Ferraz, Thomas Palmeira, et al.
Veröffentlicht: (2024)
FLAMES: Improving LLM Math Reasoning via a Fine-Grained Analysis of the Data Synthesis Pipeline
von: Seegmiller, Parker, et al.
Veröffentlicht: (2025)
von: Seegmiller, Parker, et al.
Veröffentlicht: (2025)
DiNADO: Norm-Disentangled Neurally-Decomposed Oracles for Controlling Language Models
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning
von: Samarinas, Chris, et al.
Veröffentlicht: (2026)
von: Samarinas, Chris, et al.
Veröffentlicht: (2026)
Con-ReCall: Detecting Pre-training Data in LLMs via Contrastive Decoding
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
Model Extrapolation Expedites Alignment
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
Entropy Guided Extrapolative Decoding to Improve Factuality in Large Language Models
von: Das, Souvik, et al.
Veröffentlicht: (2024)
von: Das, Souvik, et al.
Veröffentlicht: (2024)
PROMPT2BOX: Uncovering Entailment Structure among LLM Prompts
von: Bhuiya, Neeladri, et al.
Veröffentlicht: (2026)
von: Bhuiya, Neeladri, et al.
Veröffentlicht: (2026)
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation
von: Hsu, I-Hung, et al.
Veröffentlicht: (2024)
von: Hsu, I-Hung, et al.
Veröffentlicht: (2024)
DeepEdit: Knowledge Editing as Decoding with Constraints
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
QUDSELECT: Selective Decoding for Questions Under Discussion Parsing
von: Suvarna, Ashima, et al.
Veröffentlicht: (2024)
von: Suvarna, Ashima, et al.
Veröffentlicht: (2024)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training
von: Wan, David, et al.
Veröffentlicht: (2024)
von: Wan, David, et al.
Veröffentlicht: (2024)
AdaCAD: Adaptively Decoding to Balance Conflicts between Contextual and Parametric Knowledge
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
Explaining Mixtures of Sources in News Articles
von: Spangher, Alexander, et al.
Veröffentlicht: (2024)
von: Spangher, Alexander, et al.
Veröffentlicht: (2024)
Latent Traits and Cross-Task Transfer: Deconstructing Dataset Interactions in LLM Fine-tuning
von: Krishna, Shambhavi, et al.
Veröffentlicht: (2025)
von: Krishna, Shambhavi, et al.
Veröffentlicht: (2025)
EpiCoDe: Boosting Model Performance Beyond Training with Extrapolation and Contrastive Decoding
von: Tao, Mingxu, et al.
Veröffentlicht: (2025)
von: Tao, Mingxu, et al.
Veröffentlicht: (2025)
Improving Event Definition Following For Zero-Shot Event Detection
von: Cai, Zefan, et al.
Veröffentlicht: (2024)
von: Cai, Zefan, et al.
Veröffentlicht: (2024)
Open-Domain Text Evaluation via Contrastive Distribution Methods
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
Speculative Contrastive Decoding
von: Yuan, Hongyi, et al.
Veröffentlicht: (2023)
von: Yuan, Hongyi, et al.
Veröffentlicht: (2023)
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
Extrapolation Merging: Keep Improving With Extrapolation and Merging
von: Lin, Yiguan, et al.
Veröffentlicht: (2025)
von: Lin, Yiguan, et al.
Veröffentlicht: (2025)
CS4: Measuring the Creativity of Large Language Models Automatically by Controlling the Number of Story-Writing Constraints
von: Atmakuru, Anirudh, et al.
Veröffentlicht: (2024)
von: Atmakuru, Anirudh, et al.
Veröffentlicht: (2024)
Not Every Token Needs Forgetting: Selective Unlearning to Limit Change in Utility in Large Language Model Unlearning
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
von: Phan, Phuc, et al.
Veröffentlicht: (2024)
von: Phan, Phuc, et al.
Veröffentlicht: (2024)
RLCD: Reinforcement Learning from Contrastive Distillation for Language Model Alignment
von: Yang, Kevin, et al.
Veröffentlicht: (2023)
von: Yang, Kevin, et al.
Veröffentlicht: (2023)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
Soft Self-Consistency Improves Language Model Agents
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2023)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2023)
Scientific Discourse Tagging for Evidence Extraction
von: Li, Xiangci, et al.
Veröffentlicht: (2019)
von: Li, Xiangci, et al.
Veröffentlicht: (2019)
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification
von: Li, Xiangci, et al.
Veröffentlicht: (2020)
von: Li, Xiangci, et al.
Veröffentlicht: (2020)
Self-Routing RAG: Binding Selective Retrieval with Knowledge Verbalization
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
MAMM-Refine: A Recipe for Improving Faithfulness in Generation with Multi-Agent Collaboration
von: Wan, David, et al.
Veröffentlicht: (2025)
von: Wan, David, et al.
Veröffentlicht: (2025)
CD4LM: Consistency Distillation and aDaptive Decoding for Diffusion Language Models
von: Liang, Yihao, et al.
Veröffentlicht: (2026)
von: Liang, Yihao, et al.
Veröffentlicht: (2026)
Merging by Matching Models in Task Parameter Subspaces
von: Tam, Derek, et al.
Veröffentlicht: (2023)
von: Tam, Derek, et al.
Veröffentlicht: (2023)
Control Large Language Models via Divide and Conquer
von: Li, Bingxuan, et al.
Veröffentlicht: (2024)
von: Li, Bingxuan, et al.
Veröffentlicht: (2024)
DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
von: Zhou, Yu, et al.
Veröffentlicht: (2025)
von: Zhou, Yu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
REAL Sampling: Boosting Factuality and Diversity of Open-Ended Generation via Asymptotic Entropy
von: Chang, Haw-Shiuan, et al.
Veröffentlicht: (2024) -
CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization
von: Joshi, Brihi, et al.
Veröffentlicht: (2025) -
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
von: Ferraz, Thomas Palmeira, et al.
Veröffentlicht: (2024) -
FLAMES: Improving LLM Math Reasoning via a Fine-Grained Analysis of the Data Synthesis Pipeline
von: Seegmiller, Parker, et al.
Veröffentlicht: (2025) -
DiNADO: Norm-Disentangled Neurally-Decomposed Oracles for Controlling Language Models
von: Lu, Sidi, et al.
Veröffentlicht: (2023)