FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts
Fuente:
arXiv
Salvato in:
| Autori principali: | de Avalle, Guillermo Gil, Maruster, Laura, Sloot, Eric, Emmanouilidis, Christos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Procedural Knowledge Extraction from Industrial Troubleshooting Guides Using Vision Language Models
di: de Avalle, Guillermo Gil, et al.
Pubblicazione: (2026)
di: de Avalle, Guillermo Gil, et al.
Pubblicazione: (2026)
FlowLearn: Evaluating Large Vision-Language Models on Flowchart Understanding
di: Pan, Huitong, et al.
Pubblicazione: (2024)
di: Pan, Huitong, et al.
Pubblicazione: (2024)
JSynFlow: Japanese Synthesised Flowchart Visual Question Answering Dataset built with Large Language Models
di: Sasaki, Hiroshi
Pubblicazione: (2026)
di: Sasaki, Hiroshi
Pubblicazione: (2026)
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026)
First Multi-Dimensional Evaluation of Flowchart Comprehension for Multimodal Large Language Models
di: Zhang, Enming, et al.
Pubblicazione: (2024)
di: Zhang, Enming, et al.
Pubblicazione: (2024)
EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements Engineering
di: Dou, Zhifei, et al.
Pubblicazione: (2026)
di: Dou, Zhifei, et al.
Pubblicazione: (2026)
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
Arrow-Guided VLM: Enhancing Flowchart Understanding via Arrow Direction Encoding
di: Omasa, Takamitsu, et al.
Pubblicazione: (2025)
di: Omasa, Takamitsu, et al.
Pubblicazione: (2025)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
di: Yuan, Kun, et al.
Pubblicazione: (2024)
di: Yuan, Kun, et al.
Pubblicazione: (2024)
Towards Making Flowchart Images Machine Interpretable
di: Shukla, Shreya, et al.
Pubblicazione: (2025)
di: Shukla, Shreya, et al.
Pubblicazione: (2025)
Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models
di: McLaughlin, Oliver, et al.
Pubblicazione: (2026)
di: McLaughlin, Oliver, et al.
Pubblicazione: (2026)
Guiding Video Prediction with Explicit Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
Optical Flow Matters: an Empirical Comparative Study on Fusing Monocular Extracted Modalities for Better Steering
di: Makiyeh, Fouad, et al.
Pubblicazione: (2024)
di: Makiyeh, Fouad, et al.
Pubblicazione: (2024)
Separating Knowledge and Perception with Procedural Data
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2025)
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2025)
FC-Attack: Jailbreaking Multimodal Large Language Models via Auto-Generated Flowcharts
di: Zhang, Ziyi, et al.
Pubblicazione: (2025)
di: Zhang, Ziyi, et al.
Pubblicazione: (2025)
FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation
di: Bill, Eric Tillmann, et al.
Pubblicazione: (2026)
di: Bill, Eric Tillmann, et al.
Pubblicazione: (2026)
Procedural terrain generation with style transfer
di: Merizzi, Fabio
Pubblicazione: (2024)
di: Merizzi, Fabio
Pubblicazione: (2024)
Learning Robust Intervention Representations with Delta Embeddings
di: Alimisis, Panagiotis, et al.
Pubblicazione: (2025)
di: Alimisis, Panagiotis, et al.
Pubblicazione: (2025)
Spatiotemporal Object Detection for Improved Aerial Vehicle Detection in Traffic Monitoring
di: Telegraph, Kristina, et al.
Pubblicazione: (2024)
di: Telegraph, Kristina, et al.
Pubblicazione: (2024)
Motion-Boundary-Driven Unsupervised Surgical Instrument Segmentation in Low-Quality Optical Flow
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Detection-Fusion for Knowledge Graph Extraction from Videos
di: Das, Taniya, et al.
Pubblicazione: (2024)
di: Das, Taniya, et al.
Pubblicazione: (2024)
CogFlow: Bridging Perception and Reasoning through Knowledge Internalization for Visual Mathematical Problem Solving
di: Chen, Shuhang, et al.
Pubblicazione: (2026)
di: Chen, Shuhang, et al.
Pubblicazione: (2026)
Enhanced Cascade Prostate Cancer Classifier in mp-MRI Utilizing Recall Feedback Adaptive Loss and Prior Knowledge-Based Feature Extraction
di: Luo, Kun, et al.
Pubblicazione: (2024)
di: Luo, Kun, et al.
Pubblicazione: (2024)
PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs
di: Assouel, Rim, et al.
Pubblicazione: (2026)
di: Assouel, Rim, et al.
Pubblicazione: (2026)
Less is More: Label-Guided Summarization of Procedural and Instructional Videos
di: Rajpal, Shreya, et al.
Pubblicazione: (2026)
di: Rajpal, Shreya, et al.
Pubblicazione: (2026)
SceneX: Procedural Controllable Large-scale Scene Generation
di: Zhou, Mengqi, et al.
Pubblicazione: (2024)
di: Zhou, Mengqi, et al.
Pubblicazione: (2024)
ViPro: Enabling and Controlling Video Prediction for Complex Dynamical Scenarios using Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence
di: Chen, Yifan, et al.
Pubblicazione: (2026)
di: Chen, Yifan, et al.
Pubblicazione: (2026)
CityX: Controllable Procedural Content Generation for Unbounded 3D Cities
di: Zhang, Shougao, et al.
Pubblicazione: (2024)
di: Zhang, Shougao, et al.
Pubblicazione: (2024)
Modality Translation for Object Detection Adaptation Without Forgetting Prior Knowledge
di: Medeiros, Heitor Rapela, et al.
Pubblicazione: (2024)
di: Medeiros, Heitor Rapela, et al.
Pubblicazione: (2024)
FACE: Faithful Automatic Concept Extraction
di: Bhusal, Dipkamal, et al.
Pubblicazione: (2025)
di: Bhusal, Dipkamal, et al.
Pubblicazione: (2025)
Sharingan: Extract User Action Sequence from Desktop Recordings
di: Chen, Yanting, et al.
Pubblicazione: (2024)
di: Chen, Yanting, et al.
Pubblicazione: (2024)
IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly
di: Wen, Di, et al.
Pubblicazione: (2026)
di: Wen, Di, et al.
Pubblicazione: (2026)
ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding
di: Wang, Xucheng, et al.
Pubblicazione: (2026)
di: Wang, Xucheng, et al.
Pubblicazione: (2026)
A Stitch in Time: Learning Procedural Workflow via Self-Supervised Plackett-Luce Ranking
di: Che, Chengan, et al.
Pubblicazione: (2025)
di: Che, Chengan, et al.
Pubblicazione: (2025)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
di: Stilz, Florian, et al.
Pubblicazione: (2026)
di: Stilz, Florian, et al.
Pubblicazione: (2026)
Designing and Generating Diverse, Equitable Face Image Datasets for Face Verification Tasks
di: Baltsou, Georgia, et al.
Pubblicazione: (2025)
di: Baltsou, Georgia, et al.
Pubblicazione: (2025)
Personalized Federated Learning for Cross-view Geo-localization
di: Anagnostopoulos, Christos, et al.
Pubblicazione: (2024)
di: Anagnostopoulos, Christos, et al.
Pubblicazione: (2024)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
di: Papadimitriou, Christos, et al.
Pubblicazione: (2024)
di: Papadimitriou, Christos, et al.
Pubblicazione: (2024)
WeatherDG: LLM-assisted Diffusion Model for Procedural Weather Generation in Domain-Generalized Semantic Segmentation
di: Qian, Chenghao, et al.
Pubblicazione: (2024)
di: Qian, Chenghao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Procedural Knowledge Extraction from Industrial Troubleshooting Guides Using Vision Language Models
di: de Avalle, Guillermo Gil, et al.
Pubblicazione: (2026) -
FlowLearn: Evaluating Large Vision-Language Models on Flowchart Understanding
di: Pan, Huitong, et al.
Pubblicazione: (2024) -
JSynFlow: Japanese Synthesised Flowchart Visual Question Answering Dataset built with Large Language Models
di: Sasaki, Hiroshi
Pubblicazione: (2026) -
SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding
di: Radwan, Ahmed Y., et al.
Pubblicazione: (2026) -
First Multi-Dimensional Evaluation of Flowchart Comprehension for Multimodal Large Language Models
di: Zhang, Enming, et al.
Pubblicazione: (2024)