PARL: Position-Aware Relation Learning Network for Document Layout Analysis
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Fuyuan, Yu, Dianyu, Ren, He, Liu, Nayu, Kang, Xiaomian, Qiu, Delai, Zhang, Fa, Zhen, Genpeng, Liu, Shengping, Liang, Jiaen, Huang, Wei, Wang, Yining, Zhu, Junnan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing
di: Liu, Fuyuan, et al.
Pubblicazione: (2026)
di: Liu, Fuyuan, et al.
Pubblicazione: (2026)
FocalOrder: Focal Preference Optimization for Reading Order Detection
di: Liu, Fuyuan, et al.
Pubblicazione: (2026)
di: Liu, Fuyuan, et al.
Pubblicazione: (2026)
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
di: Mei, Yuxiang, et al.
Pubblicazione: (2026)
di: Mei, Yuxiang, et al.
Pubblicazione: (2026)
LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification
di: Li, Xuan, et al.
Pubblicazione: (2026)
di: Li, Xuan, et al.
Pubblicazione: (2026)
Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning
di: Wei, Shuyu, et al.
Pubblicazione: (2026)
di: Wei, Shuyu, et al.
Pubblicazione: (2026)
VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models
di: Hu, Rui, et al.
Pubblicazione: (2025)
di: Hu, Rui, et al.
Pubblicazione: (2025)
Semantic Pivots Enable Cross-Lingual Transfer in Large Language Models
di: He, Kaiyu, et al.
Pubblicazione: (2025)
di: He, Kaiyu, et al.
Pubblicazione: (2025)
Investigating and Enhancing Vision-Audio Capability in Omnimodal Large Language Models
di: Hu, Rui, et al.
Pubblicazione: (2025)
di: Hu, Rui, et al.
Pubblicazione: (2025)
ASP2LJ : An Adversarial Self-Play Laywer Augmented Legal Judgment Framework
di: Chang, Ao, et al.
Pubblicazione: (2025)
di: Chang, Ao, et al.
Pubblicazione: (2025)
Transparentize the Internal and External Knowledge Utilization in LLMs with Trustworthy Citation
di: Shen, Jiajun, et al.
Pubblicazione: (2025)
di: Shen, Jiajun, et al.
Pubblicazione: (2025)
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations
di: Lin, Jiaen, et al.
Pubblicazione: (2025)
di: Lin, Jiaen, et al.
Pubblicazione: (2025)
PARL-MT: Learning to Call Functions in Multi-Turn Conversation with Progress Awareness
di: Chai, Huacan, et al.
Pubblicazione: (2025)
di: Chai, Huacan, et al.
Pubblicazione: (2025)
BayesRAG: Probabilistic Mutual Evidence Corroboration for Multimodal Retrieval-Augmented Generation
di: Li, Xuan, et al.
Pubblicazione: (2026)
di: Li, Xuan, et al.
Pubblicazione: (2026)
DEL SUJETO AL PARL’ÊTRE
di: Pablo D. Muñoz
Pubblicazione: (2021)
di: Pablo D. Muñoz
Pubblicazione: (2021)
Self-Modifying State Modeling for Simultaneous Machine Translation
di: Yu, Donglei, et al.
Pubblicazione: (2024)
di: Yu, Donglei, et al.
Pubblicazione: (2024)
PARL: Prompt-based Agents for Reinforcement Learning
di: Resendiz, Yarik Menchaca, et al.
Pubblicazione: (2025)
di: Resendiz, Yarik Menchaca, et al.
Pubblicazione: (2025)
How Much Can RAG Help the Reasoning of LLM?
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
Tackling the Inherent Difficulty of Noise Filtering in RAG
di: Liu, Jingyu, et al.
Pubblicazione: (2026)
di: Liu, Jingyu, et al.
Pubblicazione: (2026)
Solution to the 10th ABAW Expression Recognition Challenge: A Robust Multimodal Framework with Safe Cross-Attention and Modality Dropout
di: Yu, Jun, et al.
Pubblicazione: (2026)
di: Yu, Jun, et al.
Pubblicazione: (2026)
ReLayout: Versatile and Structure-Preserving Design Layout Editing via Relation-Aware Design Reconstruction
di: Lin, Jiawei, et al.
Pubblicazione: (2026)
di: Lin, Jiawei, et al.
Pubblicazione: (2026)
Cracking Factual Knowledge: A Comprehensive Analysis of Degenerate Knowledge Neurons in Large Language Models
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
ControlLM: Crafting Diverse Personalities for Language Models
di: Weng, Yixuan, et al.
Pubblicazione: (2024)
di: Weng, Yixuan, et al.
Pubblicazione: (2024)
ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding
di: Liu, Xiao, et al.
Pubblicazione: (2026)
di: Liu, Xiao, et al.
Pubblicazione: (2026)
Relation-Aware Diffusion Model for Controllable Poster Layout Generation
di: Li, Fengheng, et al.
Pubblicazione: (2023)
di: Li, Fengheng, et al.
Pubblicazione: (2023)
LAPDoc: Layout-Aware Prompting for Documents
di: Lamott, Marcel, et al.
Pubblicazione: (2024)
di: Lamott, Marcel, et al.
Pubblicazione: (2024)
Instance-free Text to Point Cloud Localization with Relative Position Awareness
di: Wang, Lichao, et al.
Pubblicazione: (2024)
di: Wang, Lichao, et al.
Pubblicazione: (2024)
From Past To Path: Masked History Learning for Next-Item Prediction in Generative Recommendation
di: Wei, KaiWen, et al.
Pubblicazione: (2025)
di: Wei, KaiWen, et al.
Pubblicazione: (2025)
Hydrogen Leakage Simulation and Optimization of Sensor Layout for Hydrogen Supply Systems in Fuel Cell Truck
di: Yanwei Cui, et al.
Pubblicazione: (2025)
di: Yanwei Cui, et al.
Pubblicazione: (2025)
Sustainability management evolution: literature review and consolidative model
di: Ivete Delai
Pubblicazione: (2016)
di: Ivete Delai
Pubblicazione: (2016)
DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception
di: Zhao, Zhiyuan, et al.
Pubblicazione: (2024)
di: Zhao, Zhiyuan, et al.
Pubblicazione: (2024)
Position IDs Matter: An Enhanced Position Layout for Efficient Context Compression in Large Language Models
di: Zhao, Runsong, et al.
Pubblicazione: (2024)
di: Zhao, Runsong, et al.
Pubblicazione: (2024)
Cluster-Aware Grid Layout
di: Zhou, Yuxing, et al.
Pubblicazione: (2023)
di: Zhou, Yuxing, et al.
Pubblicazione: (2023)
LANS: A Layout-Aware Neural Solver for Plane Geometry Problem
di: Li, Zhong-Zhi, et al.
Pubblicazione: (2023)
di: Li, Zhong-Zhi, et al.
Pubblicazione: (2023)
Optimizing Feature Set for Click-Through Rate Prediction
di: Lyu, Fuyuan, et al.
Pubblicazione: (2023)
di: Lyu, Fuyuan, et al.
Pubblicazione: (2023)
Anchoring Emotions in Text: Robust Multimodal Fusion for Mimicry Intensity Estimation
di: Zhu, Lingsi, et al.
Pubblicazione: (2026)
di: Zhu, Lingsi, et al.
Pubblicazione: (2026)
OmniDocLayout: Towards Diverse Document Layout Generation via Coarse-to-Fine LLM Learning
di: Kang, Hengrui, et al.
Pubblicazione: (2025)
di: Kang, Hengrui, et al.
Pubblicazione: (2025)
Neural-Network-based NLOS Identification in Angular Domain at 60-GHz
di: Lyu, Pengfei, et al.
Pubblicazione: (2021)
di: Lyu, Pengfei, et al.
Pubblicazione: (2021)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
di: Wang, Baode, et al.
Pubblicazione: (2025)
di: Wang, Baode, et al.
Pubblicazione: (2025)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
di: Wang, Baode, et al.
Pubblicazione: (2025)
di: Wang, Baode, et al.
Pubblicazione: (2025)
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
di: Sun, Ting, et al.
Pubblicazione: (2025)
di: Sun, Ting, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing
di: Liu, Fuyuan, et al.
Pubblicazione: (2026) -
FocalOrder: Focal Preference Optimization for Reading Order Detection
di: Liu, Fuyuan, et al.
Pubblicazione: (2026) -
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
di: Mei, Yuxiang, et al.
Pubblicazione: (2026) -
LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification
di: Li, Xuan, et al.
Pubblicazione: (2026) -
Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning
di: Wei, Shuyu, et al.
Pubblicazione: (2026)