WGRAMMAR: Leverage Prior Knowledge to Accelerate Structured Decoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Ran, Liu, Xiaoxuan, Ren, Hao, Chen, Gang, Qi, Fanchao, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
von: Zhou, Yiqing, et al.
Veröffentlicht: (2025)
von: Zhou, Yiqing, et al.
Veröffentlicht: (2025)
RhinoInsight: Improving Deep Research through Control Mechanisms for Model Behavior and Context
von: Lei, Yu, et al.
Veröffentlicht: (2025)
von: Lei, Yu, et al.
Veröffentlicht: (2025)
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
von: Bai, Yuzhuo, et al.
Veröffentlicht: (2026)
von: Bai, Yuzhuo, et al.
Veröffentlicht: (2026)
GATEAU: Selecting Influential Samples for Long Context Alignment
von: Si, Shuzheng, et al.
Veröffentlicht: (2024)
von: Si, Shuzheng, et al.
Veröffentlicht: (2024)
FaithLens: Detecting and Explaining Faithfulness Hallucination
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
From Context to Skills: Can Language Models Learn from Context Skillfully?
von: Si, Shuzheng, et al.
Veröffentlicht: (2026)
von: Si, Shuzheng, et al.
Veröffentlicht: (2026)
Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
LycheeDecode: Accelerating Long-Context LLM Inference via Hybrid-Head Sparse Decoding
von: Lin, Gang, et al.
Veröffentlicht: (2026)
von: Lin, Gang, et al.
Veröffentlicht: (2026)
Speculative Decoding: Performance or Illusion?
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2025)
Online Speculative Decoding
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2023)
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2023)
SpecHub: Provable Acceleration to Multi-Draft Speculative Decoding
von: Sun, Ryan, et al.
Veröffentlicht: (2024)
von: Sun, Ryan, et al.
Veröffentlicht: (2024)
Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
ChartCoder: Advancing Multimodal Large Language Model for Chart-to-Code Generation
von: Zhao, Xuanle, et al.
Veröffentlicht: (2025)
von: Zhao, Xuanle, et al.
Veröffentlicht: (2025)
Beyond Natural Language: LLMs Leveraging Alternative Formats for Enhanced Reasoning and Communication
von: Chen, Weize, et al.
Veröffentlicht: (2024)
von: Chen, Weize, et al.
Veröffentlicht: (2024)
Outdated Issue Aware Decoding for Reasoning Questions on Edited Knowledge
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
EVA-Net: Subject-Independent EEG Motor Decoding with Video-Derived Motor Priors
von: Li, Ziyuan, et al.
Veröffentlicht: (2026)
von: Li, Ziyuan, et al.
Veröffentlicht: (2026)
Efficient Mixture-of-Agents Serving via Tree-Structured Routing, Adaptive Pruning, and Dependency-Aware Prefill-Decode Overlap
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors
von: Bao, Qiming, et al.
Veröffentlicht: (2025)
von: Bao, Qiming, et al.
Veröffentlicht: (2025)
Diversified and Adaptive Negative Sampling on Knowledge Graphs
von: Liu, Ran, et al.
Veröffentlicht: (2024)
von: Liu, Ran, et al.
Veröffentlicht: (2024)
Domain Knowledge is Power: Leveraging Physiological Priors for Self Supervised Representation Learning in Electrocardiography
von: Maghsoodi, Nooshin, et al.
Veröffentlicht: (2025)
von: Maghsoodi, Nooshin, et al.
Veröffentlicht: (2025)
VERITAS: Leveraging Vision Priors and Expert Fusion to Improve Multimodal Data
von: Xu, Tingqiao, et al.
Veröffentlicht: (2025)
von: Xu, Tingqiao, et al.
Veröffentlicht: (2025)
Reason Analogically via Cross-domain Prior Knowledge: An Empirical Study of Cross-domain Knowledge Transfer for In-Context Learning
von: Liu, Le, et al.
Veröffentlicht: (2026)
von: Liu, Le, et al.
Veröffentlicht: (2026)
ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
Gradient Inversion Transcript: Leveraging Robust Generative Priors to Reconstruct Training Data from Gradient Leakage
von: Chen, Xinping, et al.
Veröffentlicht: (2025)
von: Chen, Xinping, et al.
Veröffentlicht: (2025)
Progressive Refinement Regulation for Accelerating Diffusion Language Model Decoding
von: Wan, Lipeng, et al.
Veröffentlicht: (2026)
von: Wan, Lipeng, et al.
Veröffentlicht: (2026)
VistaWise: Building Cost-Effective Agent with Cross-Modal Knowledge Graph for Minecraft
von: Fu, Honghao, et al.
Veröffentlicht: (2025)
von: Fu, Honghao, et al.
Veröffentlicht: (2025)
DeepEdit: Knowledge Editing as Decoding with Constraints
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment based on Anthropic Prior Knowledge
von: Zou, Bo, et al.
Veröffentlicht: (2024)
von: Zou, Bo, et al.
Veröffentlicht: (2024)
CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit
von: Wang, Kangyu, et al.
Veröffentlicht: (2025)
von: Wang, Kangyu, et al.
Veröffentlicht: (2025)
Black-Box Membership Inference Attack for LVLMs via Prior Knowledge-Calibrated Memory Probing
von: Yin, Jinhua, et al.
Veröffentlicht: (2025)
von: Yin, Jinhua, et al.
Veröffentlicht: (2025)
Accelerating Constrained Decoding with Token Space Compression
von: Sullivan, Michael, et al.
Veröffentlicht: (2026)
von: Sullivan, Michael, et al.
Veröffentlicht: (2026)
Structure-R1: Dynamically Leveraging Structural Knowledge in LLM Reasoning through Reinforcement Learning
von: Wu, Junlin, et al.
Veröffentlicht: (2025)
von: Wu, Junlin, et al.
Veröffentlicht: (2025)
LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior
von: Wang, Hanyu, et al.
Veröffentlicht: (2024)
von: Wang, Hanyu, et al.
Veröffentlicht: (2024)
Discerning and Resolving Knowledge Conflicts through Adaptive Decoding with Contextual Information-Entropy Constraint
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2024)
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2024)
Salca: A Sparsity-Aware Hardware Accelerator for Efficient Long-Context Attention Decoding
von: Fan, Wang, et al.
Veröffentlicht: (2026)
von: Fan, Wang, et al.
Veröffentlicht: (2026)
Beyond Gradient and Priors in Privacy Attacks: Leveraging Pooler Layer Inputs of Language Models in Federated Learning
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
Bifurcated Attention: Accelerating Massively Parallel Decoding with Shared Prefixes in LLMs
von: Athiwaratkun, Ben, et al.
Veröffentlicht: (2024)
von: Athiwaratkun, Ben, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
von: Zhou, Yiqing, et al.
Veröffentlicht: (2025) -
RhinoInsight: Improving Deep Research through Control Mechanisms for Model Behavior and Context
von: Lei, Yu, et al.
Veröffentlicht: (2025) -
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
von: Bai, Yuzhuo, et al.
Veröffentlicht: (2026) -
GATEAU: Selecting Influential Samples for Long Context Alignment
von: Si, Shuzheng, et al.
Veröffentlicht: (2024) -
FaithLens: Detecting and Explaining Faithfulness Hallucination
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)