Efficient Document Parsing via Parallel Token Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Lei, Zhao, Ze, Li, Meng, Lun, Zhongwang, Yuan, Yi, Lu, Xingjing, Wei, Zheng, Bian, Jiang, Li, Zang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning Bias of Next Token Prediction Training
von: Lin, Pengxiao, et al.
Veröffentlicht: (2025)
von: Lin, Pengxiao, et al.
Veröffentlicht: (2025)
Semantic Parsing for Question Answering over Knowledge Graphs
von: Wei, Sijia, et al.
Veröffentlicht: (2023)
von: Wei, Sijia, et al.
Veröffentlicht: (2023)
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
von: Niu, Junbo, et al.
Veröffentlicht: (2025)
von: Niu, Junbo, et al.
Veröffentlicht: (2025)
LoPT: Lossless Parallel Tokenization Acceleration for Long Context Inference of Large Language Model
von: Shao, Wei, et al.
Veröffentlicht: (2025)
von: Shao, Wei, et al.
Veröffentlicht: (2025)
FastOCR: Dynamic Visual Fixation via KV Cache Pruning for Efficient Document Parsing
von: Tang, Zihan, et al.
Veröffentlicht: (2026)
von: Tang, Zihan, et al.
Veröffentlicht: (2026)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
von: Wang, Baode, et al.
Veröffentlicht: (2025)
von: Wang, Baode, et al.
Veröffentlicht: (2025)
Bilingual Rhetorical Structure Parsing with Large Parallel Annotations
von: Chistova, Elena
Veröffentlicht: (2024)
von: Chistova, Elena
Veröffentlicht: (2024)
Parallel Token Prediction for Language Models
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
TokAlign: Efficient Vocabulary Adaptation via Token Alignment
von: Li, Chong, et al.
Veröffentlicht: (2025)
von: Li, Chong, et al.
Veröffentlicht: (2025)
FTP: A Fine-grained Token-wise Pruner for Large Language Models via Token Routing
von: Li, Zekai, et al.
Veröffentlicht: (2024)
von: Li, Zekai, et al.
Veröffentlicht: (2024)
AdaParse: An Adaptive Parallel PDF Parsing and Resource Scaling Engine
von: Siebenschuh, Carlo, et al.
Veröffentlicht: (2025)
von: Siebenschuh, Carlo, et al.
Veröffentlicht: (2025)
MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale
von: Wang, Bin, et al.
Veröffentlicht: (2026)
von: Wang, Bin, et al.
Veröffentlicht: (2026)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
von: Wang, Baode, et al.
Veröffentlicht: (2025)
von: Wang, Baode, et al.
Veröffentlicht: (2025)
Character-Level Chinese Dependency Parsing via Modeling Latent Intra-Word Structure
von: Hou, Yang, et al.
Veröffentlicht: (2024)
von: Hou, Yang, et al.
Veröffentlicht: (2024)
Doc-Researcher: A Unified System for Multimodal Document Parsing and Deep Research
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
Progressive Document-level Text Simplification via Large Language Models
von: Fang, Dengzhao, et al.
Veröffentlicht: (2025)
von: Fang, Dengzhao, et al.
Veröffentlicht: (2025)
SCORE: A Semantic Evaluation Framework for Generative Document Parsing
von: Li, Renyu, et al.
Veröffentlicht: (2025)
von: Li, Renyu, et al.
Veröffentlicht: (2025)
Hierarchical Document Parsing via Large Margin Feature Matching and Heuristics
von: Kiet, Duong Anh
Veröffentlicht: (2025)
von: Kiet, Duong Anh
Veröffentlicht: (2025)
Unsupervised Mutual Learning of Discourse Parsing and Topic Segmentation in Dialogue
von: Xu, Jiahui, et al.
Veröffentlicht: (2024)
von: Xu, Jiahui, et al.
Veröffentlicht: (2024)
QoSBERT: An Uncertainty-Aware Approach based on Pre-trained Language Models for Service Quality Prediction
von: Wang, Ziliang, et al.
Veröffentlicht: (2025)
von: Wang, Ziliang, et al.
Veröffentlicht: (2025)
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
von: Li, Guanghao, et al.
Veröffentlicht: (2025)
von: Li, Guanghao, et al.
Veröffentlicht: (2025)
Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing
von: Zheng, Tong, et al.
Veröffentlicht: (2026)
von: Zheng, Tong, et al.
Veröffentlicht: (2026)
Textless Dependency Parsing by Labeled Sequence Prediction
von: Kando, Shunsuke, et al.
Veröffentlicht: (2024)
von: Kando, Shunsuke, et al.
Veröffentlicht: (2024)
Dependency Parsing is More Parameter-Efficient with Normalization
von: Gajo, Paolo, et al.
Veröffentlicht: (2025)
von: Gajo, Paolo, et al.
Veröffentlicht: (2025)
Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
CTkvr: KV Cache Retrieval for Long-Context LLMs via Centroid then Token Indexing
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
Predicting Rewards Alongside Tokens: Non-disruptive Parameter Insertion for Efficient Inference Intervention in Large Language Model
von: Yuan, Chenhan, et al.
Veröffentlicht: (2024)
von: Yuan, Chenhan, et al.
Veröffentlicht: (2024)
Beyond the Next Token: Towards Prompt-Robust Zero-Shot Classification via Efficient Multi-Token Prediction
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
Mind the Gap No More: Achieving Zero-Gap Multimodal Integration via One Tokenizer
von: Li, Yanan, et al.
Veröffentlicht: (2026)
von: Li, Yanan, et al.
Veröffentlicht: (2026)
Efficient Vision-Language Reasoning via Adaptive Token Pruning
von: Li, Xue, et al.
Veröffentlicht: (2025)
von: Li, Xue, et al.
Veröffentlicht: (2025)
Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context
von: Zhu, Yilun, et al.
Veröffentlicht: (2026)
von: Zhu, Yilun, et al.
Veröffentlicht: (2026)
Conditional Unigram Tokenization with Parallel Data
von: Vico, Gianluca, et al.
Veröffentlicht: (2025)
von: Vico, Gianluca, et al.
Veröffentlicht: (2025)
Semi-Parametric Retrieval via Binary Bag-of-Tokens Index
von: Zhou, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhou, Jiawei, et al.
Veröffentlicht: (2024)
High-order Joint Constituency and Dependency Parsing
von: Gu, Yanggan, et al.
Veröffentlicht: (2023)
von: Gu, Yanggan, et al.
Veröffentlicht: (2023)
ECLIPSE: Semantic Entropy-LCS for Cross-Lingual Industrial Log Parsing
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Efficient Training-Free Multi-Token Prediction via Embedding-Space Probing
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
Efficient Sparse Attention needs Adaptive Token Release
von: Zhang, Chaoran, et al.
Veröffentlicht: (2024)
von: Zhang, Chaoran, et al.
Veröffentlicht: (2024)
Empirical Analysis for Unsupervised Universal Dependency Parse Tree Aggregation
von: Kulkarni, Adithya, et al.
Veröffentlicht: (2024)
von: Kulkarni, Adithya, et al.
Veröffentlicht: (2024)
DocFusion: A Unified Framework for Document Parsing Tasks
von: Chai, Mingxu, et al.
Veröffentlicht: (2024)
von: Chai, Mingxu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reasoning Bias of Next Token Prediction Training
von: Lin, Pengxiao, et al.
Veröffentlicht: (2025) -
Semantic Parsing for Question Answering over Knowledge Graphs
von: Wei, Sijia, et al.
Veröffentlicht: (2023) -
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
von: Niu, Junbo, et al.
Veröffentlicht: (2025) -
LoPT: Lossless Parallel Tokenization Acceleration for Long Context Inference of Large Language Model
von: Shao, Wei, et al.
Veröffentlicht: (2025) -
FastOCR: Dynamic Visual Fixation via KV Cache Pruning for Efficient Document Parsing
von: Tang, Zihan, et al.
Veröffentlicht: (2026)