WGRAMMAR: Leverage Prior Knowledge to Accelerate Structured Decoding
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Ran, Liu, Xiaoxuan, Ren, Hao, Chen, Gang, Qi, Fanchao, Sun, Maosong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
di: Zhou, Yiqing, et al.
Pubblicazione: (2025)
di: Zhou, Yiqing, et al.
Pubblicazione: (2025)
RhinoInsight: Improving Deep Research through Control Mechanisms for Model Behavior and Context
di: Lei, Yu, et al.
Pubblicazione: (2025)
di: Lei, Yu, et al.
Pubblicazione: (2025)
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
di: Bai, Yuzhuo, et al.
Pubblicazione: (2026)
di: Bai, Yuzhuo, et al.
Pubblicazione: (2026)
GATEAU: Selecting Influential Samples for Long Context Alignment
di: Si, Shuzheng, et al.
Pubblicazione: (2024)
di: Si, Shuzheng, et al.
Pubblicazione: (2024)
FaithLens: Detecting and Explaining Faithfulness Hallucination
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
From Context to Skills: Can Language Models Learn from Context Skillfully?
di: Si, Shuzheng, et al.
Pubblicazione: (2026)
di: Si, Shuzheng, et al.
Pubblicazione: (2026)
Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
di: Si, Shuzheng, et al.
Pubblicazione: (2025)
LycheeDecode: Accelerating Long-Context LLM Inference via Hybrid-Head Sparse Decoding
di: Lin, Gang, et al.
Pubblicazione: (2026)
di: Lin, Gang, et al.
Pubblicazione: (2026)
Speculative Decoding: Performance or Illusion?
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2025)
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2025)
Online Speculative Decoding
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2023)
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2023)
SpecHub: Provable Acceleration to Multi-Draft Speculative Decoding
di: Sun, Ryan, et al.
Pubblicazione: (2024)
di: Sun, Ryan, et al.
Pubblicazione: (2024)
Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
ChartCoder: Advancing Multimodal Large Language Model for Chart-to-Code Generation
di: Zhao, Xuanle, et al.
Pubblicazione: (2025)
di: Zhao, Xuanle, et al.
Pubblicazione: (2025)
Beyond Natural Language: LLMs Leveraging Alternative Formats for Enhanced Reasoning and Communication
di: Chen, Weize, et al.
Pubblicazione: (2024)
di: Chen, Weize, et al.
Pubblicazione: (2024)
Outdated Issue Aware Decoding for Reasoning Questions on Edited Knowledge
di: Sun, Zengkui, et al.
Pubblicazione: (2024)
di: Sun, Zengkui, et al.
Pubblicazione: (2024)
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases
di: Zeng, Zheni, et al.
Pubblicazione: (2024)
di: Zeng, Zheni, et al.
Pubblicazione: (2024)
EVA-Net: Subject-Independent EEG Motor Decoding with Video-Derived Motor Priors
di: Li, Ziyuan, et al.
Pubblicazione: (2026)
di: Li, Ziyuan, et al.
Pubblicazione: (2026)
Efficient Mixture-of-Agents Serving via Tree-Structured Routing, Adaptive Pruning, and Dependency-Aware Prefill-Decode Overlap
di: Wang, Zijun, et al.
Pubblicazione: (2025)
di: Wang, Zijun, et al.
Pubblicazione: (2025)
Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors
di: Bao, Qiming, et al.
Pubblicazione: (2025)
di: Bao, Qiming, et al.
Pubblicazione: (2025)
Diversified and Adaptive Negative Sampling on Knowledge Graphs
di: Liu, Ran, et al.
Pubblicazione: (2024)
di: Liu, Ran, et al.
Pubblicazione: (2024)
Domain Knowledge is Power: Leveraging Physiological Priors for Self Supervised Representation Learning in Electrocardiography
di: Maghsoodi, Nooshin, et al.
Pubblicazione: (2025)
di: Maghsoodi, Nooshin, et al.
Pubblicazione: (2025)
VERITAS: Leveraging Vision Priors and Expert Fusion to Improve Multimodal Data
di: Xu, Tingqiao, et al.
Pubblicazione: (2025)
di: Xu, Tingqiao, et al.
Pubblicazione: (2025)
Reason Analogically via Cross-domain Prior Knowledge: An Empirical Study of Cross-domain Knowledge Transfer for In-Context Learning
di: Liu, Le, et al.
Pubblicazione: (2026)
di: Liu, Le, et al.
Pubblicazione: (2026)
ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
Gradient Inversion Transcript: Leveraging Robust Generative Priors to Reconstruct Training Data from Gradient Leakage
di: Chen, Xinping, et al.
Pubblicazione: (2025)
di: Chen, Xinping, et al.
Pubblicazione: (2025)
Progressive Refinement Regulation for Accelerating Diffusion Language Model Decoding
di: Wan, Lipeng, et al.
Pubblicazione: (2026)
di: Wan, Lipeng, et al.
Pubblicazione: (2026)
VistaWise: Building Cost-Effective Agent with Cross-Modal Knowledge Graph for Minecraft
di: Fu, Honghao, et al.
Pubblicazione: (2025)
di: Fu, Honghao, et al.
Pubblicazione: (2025)
DeepEdit: Knowledge Editing as Decoding with Constraints
di: Wang, Yiwei, et al.
Pubblicazione: (2024)
di: Wang, Yiwei, et al.
Pubblicazione: (2024)
Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment based on Anthropic Prior Knowledge
di: Zou, Bo, et al.
Pubblicazione: (2024)
di: Zou, Bo, et al.
Pubblicazione: (2024)
CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit
di: Wang, Kangyu, et al.
Pubblicazione: (2025)
di: Wang, Kangyu, et al.
Pubblicazione: (2025)
Black-Box Membership Inference Attack for LVLMs via Prior Knowledge-Calibrated Memory Probing
di: Yin, Jinhua, et al.
Pubblicazione: (2025)
di: Yin, Jinhua, et al.
Pubblicazione: (2025)
Accelerating Constrained Decoding with Token Space Compression
di: Sullivan, Michael, et al.
Pubblicazione: (2026)
di: Sullivan, Michael, et al.
Pubblicazione: (2026)
Structure-R1: Dynamically Leveraging Structural Knowledge in LLM Reasoning through Reinforcement Learning
di: Wu, Junlin, et al.
Pubblicazione: (2025)
di: Wu, Junlin, et al.
Pubblicazione: (2025)
LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior
di: Wang, Hanyu, et al.
Pubblicazione: (2024)
di: Wang, Hanyu, et al.
Pubblicazione: (2024)
Discerning and Resolving Knowledge Conflicts through Adaptive Decoding with Contextual Information-Entropy Constraint
di: Yuan, Xiaowei, et al.
Pubblicazione: (2024)
di: Yuan, Xiaowei, et al.
Pubblicazione: (2024)
Salca: A Sparsity-Aware Hardware Accelerator for Efficient Long-Context Attention Decoding
di: Fan, Wang, et al.
Pubblicazione: (2026)
di: Fan, Wang, et al.
Pubblicazione: (2026)
Beyond Gradient and Priors in Privacy Attacks: Leveraging Pooler Layer Inputs of Language Models in Federated Learning
di: Li, Jianwei, et al.
Pubblicazione: (2023)
di: Li, Jianwei, et al.
Pubblicazione: (2023)
Bifurcated Attention: Accelerating Massively Parallel Decoding with Shared Prefixes in LLMs
di: Athiwaratkun, Ben, et al.
Pubblicazione: (2024)
di: Athiwaratkun, Ben, et al.
Pubblicazione: (2024)
Documenti analoghi
-
From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
di: Zhou, Yiqing, et al.
Pubblicazione: (2025) -
RhinoInsight: Improving Deep Research through Control Mechanisms for Model Behavior and Context
di: Lei, Yu, et al.
Pubblicazione: (2025) -
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
di: Bai, Yuzhuo, et al.
Pubblicazione: (2026) -
GATEAU: Selecting Influential Samples for Long Context Alignment
di: Si, Shuzheng, et al.
Pubblicazione: (2024) -
FaithLens: Detecting and Explaining Faithfulness Hallucination
di: Si, Shuzheng, et al.
Pubblicazione: (2025)