Synthetic Industrial Object Detection: GenAI vs. Feature-Based Methods
Fuente:
arXiv
Guardado en:
| Autores principales: | Araya-Martinez, Jose Moises, Reig, Adrián Sanchis, Mohan, Gautham, Sardari, Sarvenaz, Lambrecht, Jens, Krüger, Jörg |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2025)
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2025)
SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2026)
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2026)
Image Reconstruction as a Tool for Feature Analysis
por: Allakhverdov, Eduard, et al.
Publicado: (2025)
por: Allakhverdov, Eduard, et al.
Publicado: (2025)
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
por: Liu, Yichen, et al.
Publicado: (2024)
por: Liu, Yichen, et al.
Publicado: (2024)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
por: Jin, Hang, et al.
Publicado: (2025)
por: Jin, Hang, et al.
Publicado: (2025)
Surg$Σ$: A Spectrum of Large-Scale Multimodal Data and Foundation Models for Surgical Intelligence
por: Zeng, Zhitao, et al.
Publicado: (2026)
por: Zeng, Zhitao, et al.
Publicado: (2026)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
por: Zinnen, Mathias, et al.
Publicado: (2025)
por: Zinnen, Mathias, et al.
Publicado: (2025)
Predicting Pedestrian Crossing Behavior in Germany and Japan: Insights into Model Transferability
por: Zhang, Chi, et al.
Publicado: (2024)
por: Zhang, Chi, et al.
Publicado: (2024)
Force-Aware 3D Contact Modeling for Stable Grasp Generation
por: Chen, Zhuo, et al.
Publicado: (2025)
por: Chen, Zhuo, et al.
Publicado: (2025)
Predicting and Analyzing Pedestrian Crossing Behavior at Unsignalized Crossings
por: Zhang, Chi, et al.
Publicado: (2024)
por: Zhang, Chi, et al.
Publicado: (2024)
When Less is Enough: Adaptive Token Reduction for Efficient Image Representation
por: Allakhverdov, Eduard, et al.
Publicado: (2025)
por: Allakhverdov, Eduard, et al.
Publicado: (2025)
RefineFormer3D: Efficient 3D Medical Image Segmentation via Adaptive Multi-Scale Transformer with Cross Attention Fusion
por: Tyagi, Kavyansh, et al.
Publicado: (2026)
por: Tyagi, Kavyansh, et al.
Publicado: (2026)
On the Inherent Robustness of One-Stage Object Detection against Out-of-Distribution Data
por: Martinez-Seras, Aitor, et al.
Publicado: (2024)
por: Martinez-Seras, Aitor, et al.
Publicado: (2024)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
por: Mohammad, Noor Islam S.
Publicado: (2025)
por: Mohammad, Noor Islam S.
Publicado: (2025)
ClustViT: Clustering-based Token Merging for Semantic Segmentation
por: Montello, Fabio, et al.
Publicado: (2025)
por: Montello, Fabio, et al.
Publicado: (2025)
Deep Learning Approaches for Human Action Recognition in Video Data
por: Xie, Yufei
Publicado: (2024)
por: Xie, Yufei
Publicado: (2024)
LatentForensics: Towards frugal deepfake detection in the StyleGAN latent space
por: Delmas, Matthieu, et al.
Publicado: (2023)
por: Delmas, Matthieu, et al.
Publicado: (2023)
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
por: Li, Yiyue, et al.
Publicado: (2025)
por: Li, Yiyue, et al.
Publicado: (2025)
Multi-scale Temporal Prediction via Incremental Generation and Multi-agent Collaboration
por: Zeng, Zhitao, et al.
Publicado: (2025)
por: Zeng, Zhitao, et al.
Publicado: (2025)
PlaneSAM: Multimodal Plane Instance Segmentation Using the Segment Anything Model
por: Deng, Zhongchen, et al.
Publicado: (2024)
por: Deng, Zhongchen, et al.
Publicado: (2024)
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
por: Zeng, Zhitao, et al.
Publicado: (2025)
por: Zeng, Zhitao, et al.
Publicado: (2025)
UrbanAlign: Post-hoc Semantic Calibration for VLM-Human Preference Alignment
por: Zhang, Yecheng, et al.
Publicado: (2026)
por: Zhang, Yecheng, et al.
Publicado: (2026)
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
por: Tian, Jie, et al.
Publicado: (2025)
por: Tian, Jie, et al.
Publicado: (2025)
Video-CoE: Reinforcing Video Event Prediction via Chain of Events
por: Su, Qile, et al.
Publicado: (2026)
por: Su, Qile, et al.
Publicado: (2026)
Learning Discriminative Spatio-temporal Representations for Semi-supervised Action Recognition
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
Meaning over Motion: A Semantic-First Approach to 360° Viewport Prediction
por: Khah, Arman Nik, et al.
Publicado: (2026)
por: Khah, Arman Nik, et al.
Publicado: (2026)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
por: Sáez, Arnau Igualde, et al.
Publicado: (2025)
por: Sáez, Arnau Igualde, et al.
Publicado: (2025)
AI-Enhanced Precision in Sport Taekwondo: Increasing Fairness, Speed, and Trust in Competition (FST.ai)
por: Shariatmadar, Keivan, et al.
Publicado: (2025)
por: Shariatmadar, Keivan, et al.
Publicado: (2025)
Learning Association via Track-Detection Matching for Multi-Object Tracking
por: Adžemović, Momir
Publicado: (2025)
por: Adžemović, Momir
Publicado: (2025)
MB-DSMIL-CL-PL: Scalable Weakly Supervised Ovarian Cancer Subtype Classification and Localisation Using Contrastive and Prototype Learning with Frozen Patch Features
por: Jenkins, Marcus, et al.
Publicado: (2026)
por: Jenkins, Marcus, et al.
Publicado: (2026)
On Memory: A comparison of memory mechanisms in world models
por: Laird, Eli J., et al.
Publicado: (2025)
por: Laird, Eli J., et al.
Publicado: (2025)
LRVS-Fashion: Extending Visual Search with Referring Instructions
por: Lepage, Simon, et al.
Publicado: (2023)
por: Lepage, Simon, et al.
Publicado: (2023)
Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation
por: Ayanzadeh, Aydin, et al.
Publicado: (2025)
por: Ayanzadeh, Aydin, et al.
Publicado: (2025)
E Pluribus Unum Interpretable Convolutional Neural Networks
por: Dimas, George, et al.
Publicado: (2022)
por: Dimas, George, et al.
Publicado: (2022)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
por: Baharoon, Mohammed, et al.
Publicado: (2024)
por: Baharoon, Mohammed, et al.
Publicado: (2024)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
por: Rashid, Muhammad, et al.
Publicado: (2026)
por: Rashid, Muhammad, et al.
Publicado: (2026)
GSEdit: Efficient Text-Guided Editing of 3D Objects via Gaussian Splatting
por: Palandra, Francesco, et al.
Publicado: (2024)
por: Palandra, Francesco, et al.
Publicado: (2024)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
por: Zhang, Junbin, et al.
Publicado: (2022)
por: Zhang, Junbin, et al.
Publicado: (2022)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
por: Zarei, Mohammad, et al.
Publicado: (2025)
por: Zarei, Mohammad, et al.
Publicado: (2025)
LAESI: Leaf Area Estimation with Synthetic Imagery
por: Kałużny, Jacek, et al.
Publicado: (2024)
por: Kałużny, Jacek, et al.
Publicado: (2024)
Ejemplares similares
-
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2025) -
SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception
por: Araya-Martinez, Jose Moises, et al.
Publicado: (2026) -
Image Reconstruction as a Tool for Feature Analysis
por: Allakhverdov, Eduard, et al.
Publicado: (2025) -
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
por: Liu, Yichen, et al.
Publicado: (2024) -
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
por: Jin, Hang, et al.
Publicado: (2025)