Confidence-Modulated Speculative Decoding for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Sen, Jaydip, Dasgupta, Subhasis, Waghela, Hetvi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Context-Enhanced Contrastive Search for Improved LLM Text Generation
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
Multi-Amateur Contrastive Decoding for Text Generation
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
Saliency Attention and Semantic Similarity-Driven Adversarial Perturbation
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
Adversarial Text Generation with Dynamic Contextual Perturbation
di: Waghela, Hetvi, et al.
Pubblicazione: (2025)
di: Waghela, Hetvi, et al.
Pubblicazione: (2025)
Advancing Decoding Strategies: Enhancements in Locally Typical Sampling for LLMs
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
Analyzing Consumer Reviews for Understanding Drivers of Hotels Ratings: An Indian Perspective
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
Exploring Sectoral Profitability in the Indian Stock Market Using Deep Learning
di: Sen, Jaydip, et al.
Pubblicazione: (2024)
di: Sen, Jaydip, et al.
Pubblicazione: (2024)
On Speculative Decoding for Multimodal Large Language Models
di: Gagrani, Mukul, et al.
Pubblicazione: (2024)
di: Gagrani, Mukul, et al.
Pubblicazione: (2024)
Fast Large Language Model Collaborative Decoding via Speculation
di: Fu, Jiale, et al.
Pubblicazione: (2025)
di: Fu, Jiale, et al.
Pubblicazione: (2025)
Adversarial Robustness through Dynamic Ensemble Learning
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)
Hierarchical Verification of Speculative Beams for Accelerating LLM Inference
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
Pruning as a Defense: Reducing Memorization in Large Language Models
di: Gupta, Mansi, et al.
Pubblicazione: (2025)
di: Gupta, Mansi, et al.
Pubblicazione: (2025)
Understanding the Impact of News Articles on the Movement of Market Index: A Case on Nifty 50
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
Online Speculative Decoding
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2023)
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2023)
Mixture of Attentions For Speculative Decoding
di: Zimmer, Matthieu, et al.
Pubblicazione: (2024)
di: Zimmer, Matthieu, et al.
Pubblicazione: (2024)
Traversal Verification for Speculative Tree Decoding
di: Weng, Yepeng, et al.
Pubblicazione: (2025)
di: Weng, Yepeng, et al.
Pubblicazione: (2025)
Faster Cascades via Speculative Decoding
di: Narasimhan, Harikrishna, et al.
Pubblicazione: (2024)
di: Narasimhan, Harikrishna, et al.
Pubblicazione: (2024)
A Comparative Study of Hyperparameter Tuning Methods
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
HiSpec: Hierarchical Speculative Decoding for LLMs
di: Kumar, Avinash, et al.
Pubblicazione: (2025)
di: Kumar, Avinash, et al.
Pubblicazione: (2025)
Benchmarking the Energy Savings with Speculative Decoding Strategies
di: Dutta, Rohit, et al.
Pubblicazione: (2026)
di: Dutta, Rohit, et al.
Pubblicazione: (2026)
A Theoretical Perspective for Speculative Decoding Algorithm
di: Yin, Ming, et al.
Pubblicazione: (2024)
di: Yin, Ming, et al.
Pubblicazione: (2024)
Training Domain Draft Models for Speculative Decoding: Best Practices and Insights
di: Hong, Fenglu, et al.
Pubblicazione: (2025)
di: Hong, Fenglu, et al.
Pubblicazione: (2025)
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration
di: Wen, Zhuofan, et al.
Pubblicazione: (2024)
di: Wen, Zhuofan, et al.
Pubblicazione: (2024)
Data Privacy Preservation on the Internet of Things
di: Sen, Jaydip, et al.
Pubblicazione: (2023)
di: Sen, Jaydip, et al.
Pubblicazione: (2023)
Self-Speculative Biased Decoding for Faster Re-Translation
di: Zeng, Linxiao, et al.
Pubblicazione: (2025)
di: Zeng, Linxiao, et al.
Pubblicazione: (2025)
Clover: Regressive Lightweight Speculative Decoding with Sequential Knowledge
di: Xiao, Bin, et al.
Pubblicazione: (2024)
di: Xiao, Bin, et al.
Pubblicazione: (2024)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
di: Goel, Raghavv, et al.
Pubblicazione: (2024)
di: Goel, Raghavv, et al.
Pubblicazione: (2024)
SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration
di: Wen, Zhuofan, et al.
Pubblicazione: (2026)
di: Wen, Zhuofan, et al.
Pubblicazione: (2026)
AdaSPEC: Selective Knowledge Distillation for Efficient Speculative Decoders
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
POSS: Position Specialist Generates Better Draft for Speculative Decoding
di: Huang, Langlin, et al.
Pubblicazione: (2025)
di: Huang, Langlin, et al.
Pubblicazione: (2025)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
di: Zhou, Yongchao, et al.
Pubblicazione: (2023)
di: Zhou, Yongchao, et al.
Pubblicazione: (2023)
Clover-2: Accurate Inference for Regressive Lightweight Speculative Decoding
di: Xiao, Bin, et al.
Pubblicazione: (2024)
di: Xiao, Bin, et al.
Pubblicazione: (2024)
DynaSpec: Context-aware Dynamic Speculative Sampling for Large-Vocabulary Language Models
di: Zhang, Jinbin, et al.
Pubblicazione: (2025)
di: Zhang, Jinbin, et al.
Pubblicazione: (2025)
Large Language Model Confidence Estimation via Black-Box Access
di: Pedapati, Tejaswini, et al.
Pubblicazione: (2024)
di: Pedapati, Tejaswini, et al.
Pubblicazione: (2024)
FR-Spec: Accelerating Large-Vocabulary Language Models via Frequency-Ranked Speculative Sampling
di: Zhao, Weilin, et al.
Pubblicazione: (2025)
di: Zhao, Weilin, et al.
Pubblicazione: (2025)
Spiffy: Multiplying Diffusion LLM Acceleration via Lossless Speculative Decoding
di: Agrawal, Sudhanshu, et al.
Pubblicazione: (2025)
di: Agrawal, Sudhanshu, et al.
Pubblicazione: (2025)
ML-SpecQD: Multi-Level Speculative Decoding with Quantized Drafts
di: Georganas, Evangelos, et al.
Pubblicazione: (2025)
di: Georganas, Evangelos, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Context-Enhanced Contrastive Search for Improved LLM Text Generation
di: Sen, Jaydip, et al.
Pubblicazione: (2025) -
Multi-Amateur Contrastive Decoding for Text Generation
di: Sen, Jaydip, et al.
Pubblicazione: (2025) -
Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent
di: Waghela, Hetvi, et al.
Pubblicazione: (2024) -
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
di: Waghela, Hetvi, et al.
Pubblicazione: (2024) -
Saliency Attention and Semantic Similarity-Driven Adversarial Perturbation
di: Waghela, Hetvi, et al.
Pubblicazione: (2024)