TABED: Test-Time Adaptive Ensemble Drafting for Robust Speculative Decoding in LVLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Minjae, Kang, Wonjun, Ahn, Byeongkeun, Classen, Christian, Galim, Kevin, Oh, Seunghyuk, Yan, Minghao, Koo, Hyung Il, Lee, Kangwook |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Draft-based Approximate Inference for LLMs
von: Galim, Kevin, et al.
Veröffentlicht: (2025)
von: Galim, Kevin, et al.
Veröffentlicht: (2025)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
Parameter-Efficient Fine-Tuning of State Space Models
von: Galim, Kevin, et al.
Veröffentlicht: (2024)
von: Galim, Kevin, et al.
Veröffentlicht: (2024)
Eta Inversion: Designing an Optimal Eta Function for Diffusion-based Real Image Editing
von: Kang, Wonjun, et al.
Veröffentlicht: (2024)
von: Kang, Wonjun, et al.
Veröffentlicht: (2024)
State-offset Tuning: State-based Parameter-Efficient Fine-Tuning for State Space Models
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
Counting Guidance for High Fidelity Text-to-Image Synthesis
von: Kang, Wonjun, et al.
Veröffentlicht: (2023)
von: Kang, Wonjun, et al.
Veröffentlicht: (2023)
Can MLLMs Perform Text-to-Image In-Context Learning?
von: Zeng, Yuchen, et al.
Veröffentlicht: (2024)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2024)
Transformers in the Dark: Navigating Unknown Search Spaces via Bandit Feedback
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data
von: Zeng, Thomas, et al.
Veröffentlicht: (2025)
von: Zeng, Thomas, et al.
Veröffentlicht: (2025)
PEARL: Parallel Speculative Decoding with Adaptive Draft Length
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
Learning to Draft: Adaptive Speculative Decoding with Reinforcement Learning
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
OPT-Tree: Speculative Decoding with Adaptive Draft Tree Structure
von: Wang, Jikai, et al.
Veröffentlicht: (2024)
von: Wang, Jikai, et al.
Veröffentlicht: (2024)
SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation
von: Kang, Jialiang, et al.
Veröffentlicht: (2026)
von: Kang, Jialiang, et al.
Veröffentlicht: (2026)
Draft on the Fly: Adaptive Self-Speculative Decoding using Cosine Similarity
von: Metel, Michael R., et al.
Veröffentlicht: (2024)
von: Metel, Michael R., et al.
Veröffentlicht: (2024)
Decoding Speculative Decoding
von: Yan, Minghao, et al.
Veröffentlicht: (2024)
von: Yan, Minghao, et al.
Veröffentlicht: (2024)
Mamba Drafters for Speculative Decoding
von: Choi, Daewon, et al.
Veröffentlicht: (2025)
von: Choi, Daewon, et al.
Veröffentlicht: (2025)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
Generalizable Prompt Tuning for Audio-Language Models via Semantic Expansion
von: Jang, Jaehyuk, et al.
Veröffentlicht: (2026)
von: Jang, Jaehyuk, et al.
Veröffentlicht: (2026)
Improving Weakly-Supervised Object Localization Using Adversarial Erasing and Pseudo Label
von: Kang, Byeongkeun, et al.
Veröffentlicht: (2024)
von: Kang, Byeongkeun, et al.
Veröffentlicht: (2024)
Generalized Zero-Shot Learning for Point Cloud Segmentation with Evidence-Based Dynamic Calibration
von: Kim, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonseok, et al.
Veröffentlicht: (2025)
ReJump: A Tree-Jump Representation for Analyzing and Improving LLM Reasoning
von: Zeng, Yuchen, et al.
Veröffentlicht: (2025)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2025)
FastEagle: Cascaded Drafting for Accelerating Speculative Decoding
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Cost-Aware Diffusion Draft Trees for Speculative Decoding
von: Zhang, Shuai, et al.
Veröffentlicht: (2026)
von: Zhang, Shuai, et al.
Veröffentlicht: (2026)
Accelerating Speculative Decoding with Block Diffusion Draft Trees
von: Ringel, Liran, et al.
Veröffentlicht: (2026)
von: Ringel, Liran, et al.
Veröffentlicht: (2026)
Stop Jostling: Adaptive Negative Sampling Reduces the Marginalization of Low-Resource Language Tokens by Cross-Entropy Loss
von: Turumtaev, Galim
Veröffentlicht: (2026)
von: Turumtaev, Galim
Veröffentlicht: (2026)
AHASD: Asynchronous Heterogeneous Architecture for LLM Adaptive Drafting Speculative Decoding on Mobile Devices
von: Zirui, Ma, et al.
Veröffentlicht: (2026)
von: Zirui, Ma, et al.
Veröffentlicht: (2026)
AdaEAGLE: Optimizing Speculative Decoding via Explicit Modeling of Adaptive Draft Structures
von: Zhang, Situo, et al.
Veröffentlicht: (2024)
von: Zhang, Situo, et al.
Veröffentlicht: (2024)
OmniDraft: A Cross-vocabulary, Online Adaptive Drafter for On-device Speculative Decoding
von: Ramakrishnan, Ramchalam Kinattinkara, et al.
Veröffentlicht: (2025)
von: Ramakrishnan, Ramchalam Kinattinkara, et al.
Veröffentlicht: (2025)
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
Completely Weakly Supervised Class-Incremental Learning for Semantic Segmentation
von: Kim, David Minkwan, et al.
Veröffentlicht: (2025)
von: Kim, David Minkwan, et al.
Veröffentlicht: (2025)
Enhancing Long-Term Person Re-Identification Using Global, Local Body Part, and Head Streams
von: Thanh, Duy Tran, et al.
Veröffentlicht: (2024)
von: Thanh, Duy Tran, et al.
Veröffentlicht: (2024)
Generalized Class Discovery in Instance Segmentation
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
Unsupervised Contrastive Learning Using Out-Of-Distribution Data for Long-Tailed Dataset
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
TETRIS: Optimal Draft Token Selection for Batch Speculative Decoding
von: Wu, Zhaoxuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhaoxuan, et al.
Veröffentlicht: (2025)
MineDraft: A Framework for Batch Parallel Speculative Decoding
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
When Drafts Evolve: Speculative Decoding Meets Online Learning
von: Qian, Yu-Yang, et al.
Veröffentlicht: (2026)
von: Qian, Yu-Yang, et al.
Veröffentlicht: (2026)
SpecHub: Provable Acceleration to Multi-Draft Speculative Decoding
von: Sun, Ryan, et al.
Veröffentlicht: (2024)
von: Sun, Ryan, et al.
Veröffentlicht: (2024)
Draft, Verify, and Improve: Toward Training-Aware Speculative Decoding
von: Bhansali, Shrenik, et al.
Veröffentlicht: (2025)
von: Bhansali, Shrenik, et al.
Veröffentlicht: (2025)
POSS: Position Specialist Generates Better Draft for Speculative Decoding
von: Huang, Langlin, et al.
Veröffentlicht: (2025)
von: Huang, Langlin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Draft-based Approximate Inference for LLMs
von: Galim, Kevin, et al.
Veröffentlicht: (2025) -
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
von: Kang, Wonjun, et al.
Veröffentlicht: (2025) -
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
von: Kang, Wonjun, et al.
Veröffentlicht: (2025) -
Parameter-Efficient Fine-Tuning of State Space Models
von: Galim, Kevin, et al.
Veröffentlicht: (2024) -
Eta Inversion: Designing an Optimal Eta Function for Diffusion-based Real Image Editing
von: Kang, Wonjun, et al.
Veröffentlicht: (2024)