A vision-based framework for human behavior understanding in industrial assembly lines
Fuente:
arXiv
Saved in:
| Main Authors: | Papoutsakis, Konstantinos, Bakalos, Nikolaos, Fragkoulis, Konstantinos, Zacharia, Athena, Kapetadimitri, Georgia, Pateraki, Maria |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IndustryShapes: An RGB-D Benchmark dataset for 6D object pose estimation of industrial assembly components and tools
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
Learning using privileged information for segmenting tumors on digital mammograms
by: Tzortzis, Ioannis N., et al.
Published: (2024)
by: Tzortzis, Ioannis N., et al.
Published: (2024)
Anticipating Object State Changes in Long Procedural Videos
by: Manousaki, Victoria, et al.
Published: (2024)
by: Manousaki, Victoria, et al.
Published: (2024)
Understanding Multimodal Complementarity for Single-Frame Action Anticipation
by: Benavent-Lledo, Manuel, et al.
Published: (2026)
by: Benavent-Lledo, Manuel, et al.
Published: (2026)
Action Anticipation at a Glimpse: To What Extent Can Multimodal Cues Replace Video?
by: Benavent-Lledo, Manuel, et al.
Published: (2025)
by: Benavent-Lledo, Manuel, et al.
Published: (2025)
What really matters for person re-identification? A Mixture-of-Experts Framework for Semantic Attribute Importance
by: Psalta, Athena, et al.
Published: (2025)
by: Psalta, Athena, et al.
Published: (2025)
Recognizing Unseen States of Unknown Objects by Leveraging Knowledge Graphs
by: Gouidis, Filipos, et al.
Published: (2023)
by: Gouidis, Filipos, et al.
Published: (2023)
Transformer-based assignment decision network for multiple object tracking
by: Psalta, Athena, et al.
Published: (2022)
by: Psalta, Athena, et al.
Published: (2022)
Addressing single object tracking in satellite imagery through prompt-engineered solutions
by: Psalta, Athena, et al.
Published: (2024)
by: Psalta, Athena, et al.
Published: (2024)
Multi-task Learning For Joint Action and Gesture Recognition
by: Spathis, Konstantinos, et al.
Published: (2025)
by: Spathis, Konstantinos, et al.
Published: (2025)
SHARC: Reference point driven Spherical Harmonic Representation for Complex Shapes
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
Shape Representation using Gaussian Process mixture models
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026)
Fusing Domain-Specific Content from Large Language Models into Knowledge Graphs for Enhanced Zero Shot Object State Classification
by: Gouidis, Filippos, et al.
Published: (2024)
by: Gouidis, Filippos, et al.
Published: (2024)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
by: Peirone, Simone Alberto, et al.
Published: (2025)
by: Peirone, Simone Alberto, et al.
Published: (2025)
COBRA -- COnfidence score Based on shape Regression Analysis for method-independent quality assessment of object pose estimation from single images
by: Sapoutzoglou, Panagiotis, et al.
Published: (2024)
by: Sapoutzoglou, Panagiotis, et al.
Published: (2024)
Generalization vs. Specialization: Evaluating Segment Anything Model (SAM3) Zero-Shot Segmentation Against Fine-Tuned YOLO Detectors
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Introducing DEFORMISE: A deep learning framework for dementia diagnosis in the elderly using optimized MRI slice selection
by: Ntampakis, Nikolaos, et al.
Published: (2024)
by: Ntampakis, Nikolaos, et al.
Published: (2024)
Naiad: Novel Agentic Intelligent Autonomous System for Inland Water Monitoring
by: Baltzi, Eirini, et al.
Published: (2025)
by: Baltzi, Eirini, et al.
Published: (2025)
Zero-Shot Generative De-identification: Inversion-Free Flow for Privacy-Preserving Skin Image Analysis
by: Moutselos, Konstantinos, et al.
Published: (2026)
by: Moutselos, Konstantinos, et al.
Published: (2026)
Zero-shot Segmentation of Skin Conditions: Erythema with Edit-Friendly Inversion
by: Moutselos, Konstantinos, et al.
Published: (2025)
by: Moutselos, Konstantinos, et al.
Published: (2025)
Towards Practical Single-shot Motion Synthesis
by: Roditakis, Konstantinos, et al.
Published: (2024)
by: Roditakis, Konstantinos, et al.
Published: (2024)
Vision-Based Mistake Analysis in Procedural Activities: A Review of Advances and Challenges
by: Bacharidis, Konstantinos, et al.
Published: (2025)
by: Bacharidis, Konstantinos, et al.
Published: (2025)
Plant Disease Detection through Multimodal Large Language Models and Convolutional Neural Networks
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
AI-driven visual monitoring of industrial assembly tasks
by: Nardon, Mattia, et al.
Published: (2025)
by: Nardon, Mattia, et al.
Published: (2025)
Logios : An open source Greek Polytonic Optical Character Recognition system
by: Konstantinos, Perifanos, et al.
Published: (2025)
by: Konstantinos, Perifanos, et al.
Published: (2025)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
by: Spanos, Nikolaos, et al.
Published: (2025)
by: Spanos, Nikolaos, et al.
Published: (2025)
Evaluating Reliability in Medical DNNs: A Critical Analysis of Feature and Confidence-Based OOD Detection
by: Anthony, Harry, et al.
Published: (2024)
by: Anthony, Harry, et al.
Published: (2024)
A hierarchical semantic segmentation framework for computer vision-based bridge damage detection
by: Liu, Jingxiao, et al.
Published: (2022)
by: Liu, Jingxiao, et al.
Published: (2022)
Do large language vision models understand 3D shapes?
by: Eppel, Sagi
Published: (2024)
by: Eppel, Sagi
Published: (2024)
Intelligent Sampling Consensus for Homography Estimation in Football Videos Using Featureless Unpaired Points
by: Nousias, George, et al.
Published: (2023)
by: Nousias, George, et al.
Published: (2023)
Distilling Vision Transformers for Distortion-Robust Representation Learning
by: Alexis, Konstantinos, et al.
Published: (2026)
by: Alexis, Konstantinos, et al.
Published: (2026)
TropNNC: Structured Neural Network Compression Using Tropical Geometry
by: Fotopoulos, Konstantinos, et al.
Published: (2024)
by: Fotopoulos, Konstantinos, et al.
Published: (2024)
MMFusion: Combining Image Forensic Filters for Visual Manipulation Detection and Localization
by: Triaridis, Kostas, et al.
Published: (2023)
by: Triaridis, Kostas, et al.
Published: (2023)
MEDiC: Multi-objective Exploration of Distillation from CLIP
by: Georgiou, Konstantinos, et al.
Published: (2026)
by: Georgiou, Konstantinos, et al.
Published: (2026)
PReP: Efficient context-based shape retrieval for missing parts
by: Fotis, Vlassis, et al.
Published: (2024)
by: Fotis, Vlassis, et al.
Published: (2024)
Generating metamers of human scene understanding
by: Raina, Ritik, et al.
Published: (2026)
by: Raina, Ritik, et al.
Published: (2026)
IConE: Batch Independent Collapse Prevention for Self-Supervised Representation Learning
by: Almpanakis, Konstantinos, et al.
Published: (2026)
by: Almpanakis, Konstantinos, et al.
Published: (2026)
Just rotate it! Uncertainty estimation in closed-source models via multiple queries
by: Pitas, Konstantinos, et al.
Published: (2024)
by: Pitas, Konstantinos, et al.
Published: (2024)
A comprehensive framework for occluded human pose estimation
by: Xu, Linhao, et al.
Published: (2023)
by: Xu, Linhao, et al.
Published: (2023)
Fine-Grained ImageNet Classification in the Wild
by: Lymperaiou, Maria, et al.
Published: (2023)
by: Lymperaiou, Maria, et al.
Published: (2023)
Similar Items
-
IndustryShapes: An RGB-D Benchmark dataset for 6D object pose estimation of industrial assembly components and tools
by: Sapoutzoglou, Panagiotis, et al.
Published: (2026) -
Learning using privileged information for segmenting tumors on digital mammograms
by: Tzortzis, Ioannis N., et al.
Published: (2024) -
Anticipating Object State Changes in Long Procedural Videos
by: Manousaki, Victoria, et al.
Published: (2024) -
Understanding Multimodal Complementarity for Single-Frame Action Anticipation
by: Benavent-Lledo, Manuel, et al.
Published: (2026) -
Action Anticipation at a Glimpse: To What Extent Can Multimodal Cues Replace Video?
by: Benavent-Lledo, Manuel, et al.
Published: (2025)