PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Ting, Cui, Cheng, Du, Yuning, Liu, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2024)
PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition
von: Liu, Hongen, et al.
Veröffentlicht: (2025)
von: Liu, Hongen, et al.
Veröffentlicht: (2025)
OmniDocLayout: Towards Diverse Document Layout Generation via Coarse-to-Fine LLM Learning
von: Kang, Hengrui, et al.
Veröffentlicht: (2025)
von: Kang, Hengrui, et al.
Veröffentlicht: (2025)
uLayout: Unified Room Layout Estimation for Perspective and Panoramic Images
von: Lee, Jonathan, et al.
Veröffentlicht: (2025)
von: Lee, Jonathan, et al.
Veröffentlicht: (2025)
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
von: Sun, Fan-Yun, et al.
Veröffentlicht: (2024)
von: Sun, Fan-Yun, et al.
Veröffentlicht: (2024)
PP-DocBee2: Improved Baselines with Efficient Data for Multimodal Document Understanding
von: Huang, Kui, et al.
Veröffentlicht: (2025)
von: Huang, Kui, et al.
Veröffentlicht: (2025)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
von: Brioschi, Riccardo, et al.
Veröffentlicht: (2025)
von: Brioschi, Riccardo, et al.
Veröffentlicht: (2025)
DocCogito: Aligning Layout Cognition and Step-Level Grounded Reasoning for Document Understanding
von: Wu, Yuchuan, et al.
Veröffentlicht: (2026)
von: Wu, Yuchuan, et al.
Veröffentlicht: (2026)
LayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding
von: Luo, Chuwei, et al.
Veröffentlicht: (2024)
von: Luo, Chuwei, et al.
Veröffentlicht: (2024)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
Self-training Room Layout Estimation via Geometry-aware Ray-casting
von: Solarte, Bolivar, et al.
Veröffentlicht: (2024)
von: Solarte, Bolivar, et al.
Veröffentlicht: (2024)
PP-DocBee: Improving Multimodal Document Understanding Through a Bag of Tricks
von: Ni, Feng, et al.
Veröffentlicht: (2025)
von: Ni, Feng, et al.
Veröffentlicht: (2025)
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
von: Zhang, Ting, et al.
Veröffentlicht: (2024)
von: Zhang, Ting, et al.
Veröffentlicht: (2024)
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
von: Iwai, Shoma, et al.
Veröffentlicht: (2024)
von: Iwai, Shoma, et al.
Veröffentlicht: (2024)
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Towards Khmer Scene Document Layout Detection
von: Kong, Marry, et al.
Veröffentlicht: (2026)
von: Kong, Marry, et al.
Veröffentlicht: (2026)
DocSAM: Unified Document Image Segmentation via Query Decomposition and Heterogeneous Mixed Learning
von: Li, Xiao-Hui, et al.
Veröffentlicht: (2025)
von: Li, Xiao-Hui, et al.
Veröffentlicht: (2025)
LED Benchmark: Diagnosing Structural Layout Errors for Document Layout Analysis
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
ReLayout: Versatile and Structure-Preserving Design Layout Editing via Relation-Aware Design Reconstruction
von: Lin, Jiawei, et al.
Veröffentlicht: (2026)
von: Lin, Jiawei, et al.
Veröffentlicht: (2026)
Cross-Domain Document Layout Analysis Using Document Style Guide
von: Wu, Xingjiao, et al.
Veröffentlicht: (2022)
von: Wu, Xingjiao, et al.
Veröffentlicht: (2022)
No More Ambiguity in 360° Room Layout via Bi-Layout Estimation
von: Tsai, Yu-Ju, et al.
Veröffentlicht: (2024)
von: Tsai, Yu-Ju, et al.
Veröffentlicht: (2024)
Consistent Image Layout Editing with Diffusion Models
von: Xia, Tao, et al.
Veröffentlicht: (2025)
von: Xia, Tao, et al.
Veröffentlicht: (2025)
LANS: A Layout-Aware Neural Solver for Plane Geometry Problem
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2023)
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2023)
ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis
von: Zhang, Xike, et al.
Veröffentlicht: (2026)
von: Zhang, Xike, et al.
Veröffentlicht: (2026)
LayoutCoT: Unleashing the Deep Reasoning Potential of Large Language Models for Layout Generation
von: Shi, Hengyu, et al.
Veröffentlicht: (2025)
von: Shi, Hengyu, et al.
Veröffentlicht: (2025)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
von: Bieri, Valentin, et al.
Veröffentlicht: (2025)
von: Bieri, Valentin, et al.
Veröffentlicht: (2025)
VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer
von: Yu, Ning, et al.
Veröffentlicht: (2022)
von: Yu, Ning, et al.
Veröffentlicht: (2022)
LayoutLLM: Large Language Model Instruction Tuning for Visually Rich Document Understanding
von: Fujitake, Masato
Veröffentlicht: (2024)
von: Fujitake, Masato
Veröffentlicht: (2024)
KH-FUNSD: A Hierarchical and Fine-Grained Layout Analysis Dataset for Low-Resource Khmer Business Document
von: Thuon, Nimol, et al.
Veröffentlicht: (2025)
von: Thuon, Nimol, et al.
Veröffentlicht: (2025)
Constrained Layout Generation with Factor Graphs
von: Dupty, Mohammed Haroon, et al.
Veröffentlicht: (2024)
von: Dupty, Mohammed Haroon, et al.
Veröffentlicht: (2024)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
Enhancing Object Coherence in Layout-to-Image Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2023)
von: Wang, Yibin, et al.
Veröffentlicht: (2023)
Training-free Composite Scene Generation for Layout-to-Image Synthesis
von: Liu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2024)
LayoutRAG: Retrieval-Augmented Model for Content-agnostic Conditional Layout Generation
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
LayoutFlow: Flow Matching for Layout Generation
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
SciPostLayout: A Dataset for Layout Analysis and Layout Generation of Scientific Posters
von: Tanaka, Shohei, et al.
Veröffentlicht: (2024)
von: Tanaka, Shohei, et al.
Veröffentlicht: (2024)
SemLayoutDiff: Semantic Layout Generation with Diffusion Model for Indoor Scene Synthesis
von: Sun, Xiaohao, et al.
Veröffentlicht: (2025)
von: Sun, Xiaohao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2024) -
PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition
von: Liu, Hongen, et al.
Veröffentlicht: (2025) -
OmniDocLayout: Towards Diverse Document Layout Generation via Coarse-to-Fine LLM Learning
von: Kang, Hengrui, et al.
Veröffentlicht: (2025) -
uLayout: Unified Room Layout Estimation for Perspective and Panoramic Images
von: Lee, Jonathan, et al.
Veröffentlicht: (2025) -
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
von: Sun, Fan-Yun, et al.
Veröffentlicht: (2024)