Breaking the Autoregressive Chain: Hyper-Parallel Decoding for Efficient LLM-Based Attribute Value Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Glavas, Theodore, Vedula, Nikhita, Dhyani, Dushyanta, Zhu, Yilun, Malmasi, Shervin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quantile Regression with Large Language Models for Price Prediction
by: Vedula, Nikhita, et al.
Published: (2025)
by: Vedula, Nikhita, et al.
Published: (2025)
From Unstructured to Structured: LLM-Guided Attribute Graphs for Entity Search and Ranking
by: Zhu, Yilun, et al.
Published: (2026)
by: Zhu, Yilun, et al.
Published: (2026)
Hint-Augmented Re-ranking: Efficient Product Search using LLM-Based Query Decomposition
by: Zhu, Yilun, et al.
Published: (2025)
by: Zhu, Yilun, et al.
Published: (2025)
Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context
by: Zhu, Yilun, et al.
Published: (2026)
by: Zhu, Yilun, et al.
Published: (2026)
Question Suggestion for Conversational Shopping Assistants Using Product Metadata
by: Vedula, Nikhita, et al.
Published: (2024)
by: Vedula, Nikhita, et al.
Published: (2024)
Generative Explore-Exploit: Training-free Optimization of Generative Recommender Systems using LLM Optimizers
by: Senel, Lütfi Kerem, et al.
Published: (2024)
by: Senel, Lütfi Kerem, et al.
Published: (2024)
A Modular LLM Framework for Explainable Price Outlier Detection
by: Sartipi, Shadi, et al.
Published: (2026)
by: Sartipi, Shadi, et al.
Published: (2026)
Leveraging Interesting Facts to Enhance User Engagement with Conversational Interfaces
by: Vedula, Nikhita, et al.
Published: (2024)
by: Vedula, Nikhita, et al.
Published: (2024)
Instant Answering in E-Commerce Buyer-Seller Messaging using Message-to-Question Reformulation
by: Fetahu, Besnik, et al.
Published: (2024)
by: Fetahu, Besnik, et al.
Published: (2024)
Identifying High Consideration E-Commerce Search Queries
by: Chen, Zhiyu, et al.
Published: (2024)
by: Chen, Zhiyu, et al.
Published: (2024)
Generative Product Recommendations for Implicit Superlative Queries
by: Dhole, Kaustubh D., et al.
Published: (2025)
by: Dhole, Kaustubh D., et al.
Published: (2025)
Wizard of Shopping: Target-Oriented E-commerce Dialogue Generation with Decision Tree Branching
by: Li, Xiangci, et al.
Published: (2025)
by: Li, Xiangci, et al.
Published: (2025)
Hierarchical Skip Decoding for Efficient Autoregressive Text Generation
by: Zhu, Yunqi, et al.
Published: (2024)
by: Zhu, Yunqi, et al.
Published: (2024)
Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?
by: Li, Pengxiang, et al.
Published: (2026)
by: Li, Pengxiang, et al.
Published: (2026)
EIVEN: Efficient Implicit Attribute Value Extraction using Multimodal LLM
by: Zou, Henry Peng, et al.
Published: (2024)
by: Zou, Henry Peng, et al.
Published: (2024)
MADIAVE: Multi-Agent Debate for Implicit Attribute Value Extraction
by: Huang, Wei-Chieh, et al.
Published: (2025)
by: Huang, Wei-Chieh, et al.
Published: (2025)
ReFusion: A Diffusion Large Language Model with Parallel Autoregressive Decoding
by: Li, Jia-Nan, et al.
Published: (2025)
by: Li, Jia-Nan, et al.
Published: (2025)
Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing
by: Liu, Ziyang
Published: (2026)
by: Liu, Ziyang
Published: (2026)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
by: Islam, Saad Obaid ul, et al.
Published: (2025)
by: Islam, Saad Obaid ul, et al.
Published: (2025)
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction
by: Fang, Chenhao, et al.
Published: (2024)
by: Fang, Chenhao, et al.
Published: (2024)
Parallel Decoder Transformer: Planner-Seeded Latent Coordination for Synchronized Parallel Decoding
by: Robbins, Logan
Published: (2025)
by: Robbins, Logan
Published: (2025)
Bridging the Gap Between Information Seeking and Product Search Systems: Q&A Recommendation for E-commerce
by: Kuzi, Saar, et al.
Published: (2024)
by: Kuzi, Saar, et al.
Published: (2024)
Towards Hyper-Efficient RAG Systems in VecDBs: Distributed Parallel Multi-Resolution Vector Search
by: Liu, Dong, et al.
Published: (2025)
by: Liu, Dong, et al.
Published: (2025)
Falcon: Faster and Parallel Inference of Large Language Models through Enhanced Semi-Autoregressive Drafting and Custom-Designed Decoding Tree
by: Gao, Xiangxiang, et al.
Published: (2024)
by: Gao, Xiangxiang, et al.
Published: (2024)
Think Big, Generate Quick: LLM-to-SLM for Fast Autoregressive Decoding
by: Bergner, Benjamin, et al.
Published: (2024)
by: Bergner, Benjamin, et al.
Published: (2024)
Cerberus: Efficient Inference with Adaptive Parallel Decoding and Sequential Knowledge Enhancement
by: Liu, Yuxuan, et al.
Published: (2024)
by: Liu, Yuxuan, et al.
Published: (2024)
Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph
by: Liu, Ning, et al.
Published: (2026)
by: Liu, Ning, et al.
Published: (2026)
Dialogue Ontology Relation Extraction via Constrained Chain-of-Thought Decoding
by: Vukovic, Renato, et al.
Published: (2024)
by: Vukovic, Renato, et al.
Published: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
by: Paul, Indraneil, et al.
Published: (2024)
by: Paul, Indraneil, et al.
Published: (2024)
Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
by: Shen, Maohao, et al.
Published: (2025)
by: Shen, Maohao, et al.
Published: (2025)
Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs
by: Pramanik, Vishal, et al.
Published: (2026)
by: Pramanik, Vishal, et al.
Published: (2026)
Interaction Techniques that Encourage Longer Prompts Can Improve Psychological Ownership when Writing with AI
by: Joshi, Nikhita, et al.
Published: (2025)
by: Joshi, Nikhita, et al.
Published: (2025)
Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information
by: Huang, Xin, et al.
Published: (2026)
by: Huang, Xin, et al.
Published: (2026)
PaDeLLM-NER: Parallel Decoding in Large Language Models for Named Entity Recognition
by: Lu, Jinghui, et al.
Published: (2024)
by: Lu, Jinghui, et al.
Published: (2024)
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
by: Sun, Chenxi, et al.
Published: (2024)
by: Sun, Chenxi, et al.
Published: (2024)
Distribution-Aligned Decoding for Efficient LLM Task Adaptation
by: Hu, Senkang, et al.
Published: (2025)
by: Hu, Senkang, et al.
Published: (2025)
Reward-Guided Speculative Decoding for Efficient LLM Reasoning
by: Liao, Baohao, et al.
Published: (2025)
by: Liao, Baohao, et al.
Published: (2025)
ASPD: Unlocking Adaptive Serial-Parallel Decoding by Exploring Intrinsic Parallelism in LLMs
by: Chen, Keyu, et al.
Published: (2025)
by: Chen, Keyu, et al.
Published: (2025)
Sentinel: Decoding Context Utilization via Attention Probing for Efficient LLM Context Compression
by: Zhang, Yong, et al.
Published: (2025)
by: Zhang, Yong, et al.
Published: (2025)
CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit
by: Wang, Kangyu, et al.
Published: (2025)
by: Wang, Kangyu, et al.
Published: (2025)
Similar Items
-
Quantile Regression with Large Language Models for Price Prediction
by: Vedula, Nikhita, et al.
Published: (2025) -
From Unstructured to Structured: LLM-Guided Attribute Graphs for Entity Search and Ranking
by: Zhu, Yilun, et al.
Published: (2026) -
Hint-Augmented Re-ranking: Efficient Product Search using LLM-Based Query Decomposition
by: Zhu, Yilun, et al.
Published: (2025) -
Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context
by: Zhu, Yilun, et al.
Published: (2026) -
Question Suggestion for Conversational Shopping Assistants Using Product Metadata
by: Vedula, Nikhita, et al.
Published: (2024)