UnSupDLA: Towards Unsupervised Document Layout Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sheikh, Talha Uddin, Shehzadi, Tahira, Hashmi, Khurram Azeem, Stricker, Didier, Afzal, Muhammad Zeshan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Hybrid Approach for Document Layout Analysis in Document images
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
von: Hashmi, Khurram Azeem, et al.
Veröffentlicht: (2024)
von: Hashmi, Khurram Azeem, et al.
Veröffentlicht: (2024)
VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
von: Lu, Zhijing, et al.
Veröffentlicht: (2026)
von: Lu, Zhijing, et al.
Veröffentlicht: (2026)
End-to-End Semi-Supervised approach with Modulated Object Queries for Table Detection in Documents
von: Ehsan, Iqraa, et al.
Veröffentlicht: (2024)
von: Ehsan, Iqraa, et al.
Veröffentlicht: (2024)
Semi-Supervised Object Detection: A Survey on Progress from CNN to Transformer
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
Towards End-to-End Semi-Supervised Table Detection with Semantic Aligned Matching Transformer
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
CICA: Content-Injected Contrastive Alignment for Zero-Shot Document Image Classification
von: Sinha, Sankalp, et al.
Veröffentlicht: (2024)
von: Sinha, Sankalp, et al.
Veröffentlicht: (2024)
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
Classroom-Inspired Multi-Mentor Distillation with Adaptive Learning Strategies
von: Sarode, Shalini, et al.
Veröffentlicht: (2024)
von: Sarode, Shalini, et al.
Veröffentlicht: (2024)
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
von: Khan, Mohammad Sadil, et al.
Veröffentlicht: (2024)
von: Khan, Mohammad Sadil, et al.
Veröffentlicht: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
von: Nazir, Danish, et al.
Veröffentlicht: (2022)
von: Nazir, Danish, et al.
Veröffentlicht: (2022)
SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
von: Usama, Muhammad, et al.
Veröffentlicht: (2025)
von: Usama, Muhammad, et al.
Veröffentlicht: (2025)
Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
ReConText3D: Replay-based Continual Text-to-3D Generation
von: Khan, Muhammad Ahmed Ullah, et al.
Veröffentlicht: (2026)
von: Khan, Muhammad Ahmed Ullah, et al.
Veröffentlicht: (2026)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
Human Pose Descriptions and Subject-Focused Attention for Improved Zero-Shot Transfer in Human-Centric Classification Tasks
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
DreamCAD: Scaling Multi-modal CAD Generation using Differentiable Parametric Surfaces
von: Khan, Mohammad Sadil, et al.
Veröffentlicht: (2026)
von: Khan, Mohammad Sadil, et al.
Veröffentlicht: (2026)
PromptDLA: A Domain-aware Prompt Document Layout Analysis Framework with Descriptive Knowledge as a Cue
von: Zhang, Zirui, et al.
Veröffentlicht: (2026)
von: Zhang, Zirui, et al.
Veröffentlicht: (2026)
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation
von: Sinha, Sankalp, et al.
Veröffentlicht: (2024)
von: Sinha, Sankalp, et al.
Veröffentlicht: (2024)
Towards Unconstrained 2D Pose Estimation of the Human Spine
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2025)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2025)
SIMSPINE: A Biomechanics-Aware Simulation Framework for 3D Spine Motion Annotation and Benchmarking
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2026)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2026)
PoseAdapt: Sustainable Human Pose Estimation via Continual Learning Benchmarks and Toolkit
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
Beyond Averages: Open-Vocabulary 3D Scene Understanding with Gaussian Splatting and Bag of Embeddings
von: Arafa, Abdalla, et al.
Veröffentlicht: (2025)
von: Arafa, Abdalla, et al.
Veröffentlicht: (2025)
Object-Centric 2D Gaussian Splatting: Background Removal and Occlusion-Aware Pruning for Compact Object Models
von: Rogge, Marcel, et al.
Veröffentlicht: (2025)
von: Rogge, Marcel, et al.
Veröffentlicht: (2025)
Jointly Learning Spatial, Angular, and Temporal Information for Enhanced Lane Detection
von: Alam, Muhammad Zeshan
Veröffentlicht: (2024)
von: Alam, Muhammad Zeshan
Veröffentlicht: (2024)
Towards Minimal Focal Stack in Shape from Focus
von: Ashfaq, Khurram, et al.
Veröffentlicht: (2026)
von: Ashfaq, Khurram, et al.
Veröffentlicht: (2026)
Small Object Detection with YOLO: A Performance Analysis Across Model Versions and Hardware
von: Tariq, Muhammad Fasih, et al.
Veröffentlicht: (2025)
von: Tariq, Muhammad Fasih, et al.
Veröffentlicht: (2025)
RMS-FlowNet++: Efficient and Robust Multi-Scale Scene Flow Estimation for Large-Scale Point Clouds
von: Battrawy, Ramy, et al.
Veröffentlicht: (2024)
von: Battrawy, Ramy, et al.
Veröffentlicht: (2024)
ShapeAug: Occlusion Augmentation for Event Camera Data
von: Bendig, Katharina, et al.
Veröffentlicht: (2024)
von: Bendig, Katharina, et al.
Veröffentlicht: (2024)
ShapeAug++: More Realistic Shape Augmentation for Event Data
von: Bendig, Katharina, et al.
Veröffentlicht: (2024)
von: Bendig, Katharina, et al.
Veröffentlicht: (2024)
MILE: Mixture of Incremental LoRA Experts for Continual Semantic Segmentation across Domains and Modalities
von: Muralidhara, Shishir, et al.
Veröffentlicht: (2026)
von: Muralidhara, Shishir, et al.
Veröffentlicht: (2026)
G3FA: Geometry-guided GAN for Face Animation
von: Javanmardi, Alireza, et al.
Veröffentlicht: (2024)
von: Javanmardi, Alireza, et al.
Veröffentlicht: (2024)
EgoFlowNet: Non-Rigid Scene Flow from Point Clouds with Ego-Motion Support
von: Battrawy, Ramy, et al.
Veröffentlicht: (2024)
von: Battrawy, Ramy, et al.
Veröffentlicht: (2024)
Sensor Generalization for Adaptive Sensing in Event-based Object Detection via Joint Distribution Training
von: Saha, Aheli, et al.
Veröffentlicht: (2026)
von: Saha, Aheli, et al.
Veröffentlicht: (2026)
Domain-Incremental Semantic Segmentation for Autonomous Driving under Adverse Driving Conditions
von: Muralidhara, Shishir, et al.
Veröffentlicht: (2025)
von: Muralidhara, Shishir, et al.
Veröffentlicht: (2025)
PanoSAMic: Panoramic Image Segmentation from SAM Feature Encoding and Dual View Fusion
von: Chamseddine, Mahdi, et al.
Veröffentlicht: (2026)
von: Chamseddine, Mahdi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Hybrid Approach for Document Layout Analysis in Document images
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024) -
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024) -
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
von: Hashmi, Khurram Azeem, et al.
Veröffentlicht: (2024) -
VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
von: Lu, Zhijing, et al.
Veröffentlicht: (2026) -
End-to-End Semi-Supervised approach with Modulated Object Queries for Table Detection in Documents
von: Ehsan, Iqraa, et al.
Veröffentlicht: (2024)