From Engineering Diagrams to Graphs: Digitizing P&IDs with Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Stürmer, Jan Marius, Graumann, Marius, Koch, Tobias |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rotation Invariance in Floor Plan Digitization using Zernike Moments
by: Graumann, Marius, et al.
Published: (2025)
by: Graumann, Marius, et al.
Published: (2025)
DenseBEV: Transforming BEV Grid Cells into 3D Objects
by: Dähling, Marius, et al.
Published: (2025)
by: Dähling, Marius, et al.
Published: (2025)
Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning
by: Cudlenco, Nicolae, et al.
Published: (2026)
by: Cudlenco, Nicolae, et al.
Published: (2026)
With Great Context Comes Great Prediction Power: Classifying Objects via Geo-Semantic Scene Graphs
by: Constantinescu, Ciprian, et al.
Published: (2025)
by: Constantinescu, Ciprian, et al.
Published: (2025)
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
by: Masala, Mihai, et al.
Published: (2025)
by: Masala, Mihai, et al.
Published: (2025)
CrowdQuery: Density-Guided Query Module for Enhanced 2D and 3D Detection in Crowded Scenes
by: Dähling, Marius, et al.
Published: (2025)
by: Dähling, Marius, et al.
Published: (2025)
Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning
by: Mihai-Cristian, Pîrvu, et al.
Published: (2025)
by: Mihai-Cristian, Pîrvu, et al.
Published: (2025)
From One to the Power of Many: Invariance to Multi-LiDAR Perception from Single-Sensor Datasets
by: Uecker, Marc, et al.
Published: (2024)
by: Uecker, Marc, et al.
Published: (2024)
Towards Zero-Shot & Explainable Video Description by Reasoning over Graphs of Events in Space and Time
by: Masala, Mihai, et al.
Published: (2025)
by: Masala, Mihai, et al.
Published: (2025)
Multiple Random Masking Autoencoder Ensembles for Robust Multimodal Semi-supervised Learning
by: Todoran, Alexandru-Raul, et al.
Published: (2024)
by: Todoran, Alexandru-Raul, et al.
Published: (2024)
Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation
by: Airinei, Daniel, et al.
Published: (2025)
by: Airinei, Daniel, et al.
Published: (2025)
Cerberus: Attribute-based person re-identification using semantic IDs
by: Eom, Chanho, et al.
Published: (2024)
by: Eom, Chanho, et al.
Published: (2024)
MCP-MedSAM: A Powerful Lightweight Medical Segment Anything Model Trained with a Single GPU in Just One Day
by: Lyu, Donghang, et al.
Published: (2024)
by: Lyu, Donghang, et al.
Published: (2024)
Swin-LiteMedSAM: A Lightweight Box-Based Segment Anything Model for Large-Scale Medical Image Datasets
by: Gao, Ruochen, et al.
Published: (2024)
by: Gao, Ruochen, et al.
Published: (2024)
VoD: Learning Volume of Differences for Video-Based Deepfake Detection
by: Xu, Ying, et al.
Published: (2025)
by: Xu, Ying, et al.
Published: (2025)
Clinical DVH metrics as a loss function for 3D dose prediction in head and neck radiotherapy
by: Gao, Ruochen, et al.
Published: (2026)
by: Gao, Ruochen, et al.
Published: (2026)
GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models
by: Cudlenco, Nicolae, et al.
Published: (2026)
by: Cudlenco, Nicolae, et al.
Published: (2026)
Efficient Data Representation for Motion Forecasting: A Scene-Specific Trajectory Set Approach
by: Vivekanandan, Abhishek, et al.
Published: (2024)
by: Vivekanandan, Abhishek, et al.
Published: (2024)
Can you see me now? Blind spot estimation for autonomous vehicles using scenario-based simulation with random reference sensors
by: Uecker, Marc, et al.
Published: (2024)
by: Uecker, Marc, et al.
Published: (2024)
A Survey on Intermediate Fusion Methods for Collaborative Perception Categorized by Real World Challenges
by: Yazgan, Melih, et al.
Published: (2024)
by: Yazgan, Melih, et al.
Published: (2024)
3D Reconstruction of the Human Colon from Capsule Endoscope Video
by: Floor, Pål Anders, et al.
Published: (2024)
by: Floor, Pål Anders, et al.
Published: (2024)
Learning from Random Subspace Exploration: Generalized Test-Time Augmentation with Self-supervised Distillation
by: Jelea, Andrei, et al.
Published: (2025)
by: Jelea, Andrei, et al.
Published: (2025)
Calibrating the Full Predictive Class Distribution of 3D Object Detectors for Autonomous Driving
by: Schröder, Cornelius, et al.
Published: (2025)
by: Schröder, Cornelius, et al.
Published: (2025)
Recall to Predict: Grounding Motion Forecasting in Interpretable Motion Bank
by: Vivekanandan, Abhishek, et al.
Published: (2026)
by: Vivekanandan, Abhishek, et al.
Published: (2026)
Contrast & Compress: Learning Lightweight Embeddings for Short Trajectories
by: Vivekanandan, Abhishek, et al.
Published: (2025)
by: Vivekanandan, Abhishek, et al.
Published: (2025)
Closer to Ground Truth: Realistic Shape and Appearance Labeled Data Generation for Unsupervised Underwater Image Segmentation
by: Jelea, Andrei, et al.
Published: (2025)
by: Jelea, Andrei, et al.
Published: (2025)
Paired Competing Neurons Improving STDP Supervised Local Learning In Spiking Neural Networks
by: Goupy, Gaspard, et al.
Published: (2023)
by: Goupy, Gaspard, et al.
Published: (2023)
Enginuity: Building an Open Multi-Domain Dataset of Complex Engineering Diagrams
by: Seefried, Ethan, et al.
Published: (2026)
by: Seefried, Ethan, et al.
Published: (2026)
SPoT: Subpixel Placement of Tokens in Vision Transformers
by: Hjelkrem-Tan, Martine, et al.
Published: (2025)
by: Hjelkrem-Tan, Martine, et al.
Published: (2025)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
Bridging GANs and Bayesian Neural Networks via Partial Stochasticity
by: Filippone, Maurizio, et al.
Published: (2025)
by: Filippone, Maurizio, et al.
Published: (2025)
Generative Recall, Dense Reranking: Learning Multi-View Semantic IDs for Efficient Text-to-Video Retrieval
by: Zhao, Zecheng, et al.
Published: (2026)
by: Zhao, Zecheng, et al.
Published: (2026)
ADA-Track++: End-to-End Multi-Camera 3D Multi-Object Tracking with Alternating Detection and Association
by: Ding, Shuxiao, et al.
Published: (2024)
by: Ding, Shuxiao, et al.
Published: (2024)
Maia: A Real-time Non-Verbal Chat for Human-AI Interaction
by: Costea, Dragos, et al.
Published: (2024)
by: Costea, Dragos, et al.
Published: (2024)
Sample-Specific Output Constraints for Neural Networks
by: Brosowsky, Mathis, et al.
Published: (2020)
by: Brosowsky, Mathis, et al.
Published: (2020)
A 3D mesh convolution-based autoencoder for geometry compression
by: Bregeon, Germain, et al.
Published: (2026)
by: Bregeon, Germain, et al.
Published: (2026)
Hybrid Video Anomaly Detection for Anomalous Scenarios in Autonomous Driving
by: Bogdoll, Daniel, et al.
Published: (2024)
by: Bogdoll, Daniel, et al.
Published: (2024)
Uncovering Grounding IDs: How External Cues Shape Multimodal Binding
by: Hasani, Hosein, et al.
Published: (2025)
by: Hasani, Hosein, et al.
Published: (2025)
eMotion-GAN: A Motion-based GAN for Photorealistic and Facial Expression Preserving Frontal View Synthesis
by: Ikne, Omar, et al.
Published: (2024)
by: Ikne, Omar, et al.
Published: (2024)
Multi-modal video data-pipelines for machine learning with minimal human supervision
by: Pîrvu, Mihai-Cristian, et al.
Published: (2025)
by: Pîrvu, Mihai-Cristian, et al.
Published: (2025)
Similar Items
-
Rotation Invariance in Floor Plan Digitization using Zernike Moments
by: Graumann, Marius, et al.
Published: (2025) -
DenseBEV: Transforming BEV Grid Cells into 3D Objects
by: Dähling, Marius, et al.
Published: (2025) -
Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning
by: Cudlenco, Nicolae, et al.
Published: (2026) -
With Great Context Comes Great Prediction Power: Classifying Objects via Geo-Semantic Scene Graphs
by: Constantinescu, Ciprian, et al.
Published: (2025) -
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
by: Masala, Mihai, et al.
Published: (2025)