Process Integrated Computer Vision for Real-Time Failure Prediction in Steel Rolling Mill
Fuente:
arXiv
Saved in:
| Main Authors: | Kurrey, Vaibhav, Pujari, Sivakalyan, Gupta, Gagan Raj |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Action Recognition based Industrial Safety Violation Detection
by: Reddy, Surya N, et al.
Published: (2024)
by: Reddy, Surya N, et al.
Published: (2024)
Comparing Computational Pathology Foundation Models using Representational Similarity Analysis
by: Mishra, Vaibhav, et al.
Published: (2025)
by: Mishra, Vaibhav, et al.
Published: (2025)
Wildfire Detection Using Vision Transformer with the Wildfire Dataset
by: Vuppari, Gowtham Raj, et al.
Published: (2025)
by: Vuppari, Gowtham Raj, et al.
Published: (2025)
A Lightweight Group Multiscale Bidirectional Interactive Network for Real-Time Steel Surface Defect Detection
by: Zhang, Yong, et al.
Published: (2025)
by: Zhang, Yong, et al.
Published: (2025)
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
by: Zhao, Shuxian, et al.
Published: (2026)
by: Zhao, Shuxian, et al.
Published: (2026)
HydroVision: Predicting Optically Active Parameters in Surface Water Using Computer Vision
by: Deshmukh, Shubham Laxmikant, et al.
Published: (2025)
by: Deshmukh, Shubham Laxmikant, et al.
Published: (2025)
The Geometry of Representational Failures in Vision Language Models
by: Savietto, Daniele, et al.
Published: (2026)
by: Savietto, Daniele, et al.
Published: (2026)
Real-Time Fusion of Visual and Chart Data for Enhanced Maritime Vision
by: Kreis, Marten, et al.
Published: (2025)
by: Kreis, Marten, et al.
Published: (2025)
Feedback Driven Multi Stereo Vision System for Real-Time Event Analysis
by: Benkedadra, Mohamed, et al.
Published: (2025)
by: Benkedadra, Mohamed, et al.
Published: (2025)
Attention-Based Real-Time Defenses for Physical Adversarial Attacks in Vision Applications
by: Rossolini, Giulio, et al.
Published: (2023)
by: Rossolini, Giulio, et al.
Published: (2023)
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation
by: Shah, Arya, et al.
Published: (2026)
by: Shah, Arya, et al.
Published: (2026)
Discovering Failure Modes in Vision-Language Models using RL
by: Jain, Kanishk, et al.
Published: (2026)
by: Jain, Kanishk, et al.
Published: (2026)
Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics
by: Roy, Subhadeep, et al.
Published: (2026)
by: Roy, Subhadeep, et al.
Published: (2026)
Beyond Vision: How Large Language Models Interpret Facial Expressions from Valence-Arousal Values
by: Mehra, Vaibhav, et al.
Published: (2025)
by: Mehra, Vaibhav, et al.
Published: (2025)
Temporally-Grounded Language Generation: A Benchmark for Real-Time Vision-Language Models
by: Yu, Keunwoo Peter, et al.
Published: (2025)
by: Yu, Keunwoo Peter, et al.
Published: (2025)
Let's Roll a BiFTA: Bi-refinement for Fine-grained Text-visual Alignment in Vision-Language Models
by: Sun, Yuhao, et al.
Published: (2026)
by: Sun, Yuhao, et al.
Published: (2026)
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
by: Agnur, Bharath Kumar
Published: (2024)
by: Agnur, Bharath Kumar
Published: (2024)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
by: Shah, Arya, et al.
Published: (2025)
by: Shah, Arya, et al.
Published: (2025)
Towards Explainable LiDAR Point Cloud Semantic Segmentation via Gradient Based Target Localization
by: Kuriyal, Abhishek, et al.
Published: (2024)
by: Kuriyal, Abhishek, et al.
Published: (2024)
Vehicle detection from GSV imagery: Predicting travel behaviour for cycling and motorcycling using Computer Vision
by: Kyriaki, et al.
Published: (2025)
by: Kyriaki, et al.
Published: (2025)
Vision-Language Models display a strong gender bias
by: Konavoor, Aiswarya, et al.
Published: (2025)
by: Konavoor, Aiswarya, et al.
Published: (2025)
MIRROR: Multimodal Cognitive Reframing Therapy for Rolling with Resistance
by: Kim, Subin, et al.
Published: (2025)
by: Kim, Subin, et al.
Published: (2025)
WebSight: A Vision-First Architecture for Robust Web Agents
by: Bhathal, Tanvir, et al.
Published: (2025)
by: Bhathal, Tanvir, et al.
Published: (2025)
VARS: Vision-based Assessment of Risk in Security Systems
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
Symbolic Rule Extraction from Attention-Guided Sparse Representations in Vision Transformers
by: Padalkar, Parth, et al.
Published: (2025)
by: Padalkar, Parth, et al.
Published: (2025)
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
Deep Learning-Based Computer Vision Models for Early Cancer Detection Using Multimodal Medical Imaging and Radiogenomic Integration Frameworks
by: Oghenekaro, Emmanuella Avwerosuoghene
Published: (2025)
by: Oghenekaro, Emmanuella Avwerosuoghene
Published: (2025)
Accelerating Vision Transformers on Brain Processing Unit
by: Tang, Jinchi, et al.
Published: (2026)
by: Tang, Jinchi, et al.
Published: (2026)
A Deep Learning Framework for Real-Time Image Processing in Medical Diagnostics: Enhancing Accuracy and Speed in Clinical Applications
by: Filvantorkaman, Melika, et al.
Published: (2025)
by: Filvantorkaman, Melika, et al.
Published: (2025)
Training Vision-Language Process Reward Models for Test-Time Scaling in Multimodal Reasoning: Key Insights and Lessons Learned
by: Ong, Brandon, et al.
Published: (2025)
by: Ong, Brandon, et al.
Published: (2025)
Revisiting the Integration of Convolution and Attention for Vision Backbone
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
Your Vision-Language Model Can't Even Count to 20: Exposing the Failures of VLMs in Compositional Counting
by: Guo, Xuyang, et al.
Published: (2025)
by: Guo, Xuyang, et al.
Published: (2025)
Dynamic Weight Adjustment for Knowledge Distillation: Leveraging Vision Transformer for High-Accuracy Lung Cancer Detection and Real-Time Deployment
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
Video Detector: A Dual-Phase Vision-Based System for Real-Time Traffic Intersection Control and Intelligent Transportation Analysis
by: Şen, Mustafa Fatih, et al.
Published: (2026)
by: Şen, Mustafa Fatih, et al.
Published: (2026)
Stabilizing Open-Set Test-Time Adaptation via Primary-Auxiliary Filtering and Knowledge-Integrated Prediction
by: Lee, Byung-Joon, et al.
Published: (2025)
by: Lee, Byung-Joon, et al.
Published: (2025)
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
by: Titiya, Prasham, et al.
Published: (2025)
by: Titiya, Prasham, et al.
Published: (2025)
Vision without Images: End-to-End Computer Vision from Single Compressive Measurements
by: Pan, Fengpu, et al.
Published: (2025)
by: Pan, Fengpu, et al.
Published: (2025)
Edge Reliability Gap in Vision-Language Models: Quantifying Failure Modes of Compressed VLMs Under Visual Corruption
by: Erol, Mehmet Kaan
Published: (2026)
by: Erol, Mehmet Kaan
Published: (2026)
NanoVLMs: How small can we go and still make coherent Vision Language Models?
by: Agarwalla, Mukund, et al.
Published: (2025)
by: Agarwalla, Mukund, et al.
Published: (2025)
A Novel Framework For Text Detection From Natural Scene Images With Complex Background
by: Kaladagi, Basavaraj, et al.
Published: (2024)
by: Kaladagi, Basavaraj, et al.
Published: (2024)
Similar Items
-
Action Recognition based Industrial Safety Violation Detection
by: Reddy, Surya N, et al.
Published: (2024) -
Comparing Computational Pathology Foundation Models using Representational Similarity Analysis
by: Mishra, Vaibhav, et al.
Published: (2025) -
Wildfire Detection Using Vision Transformer with the Wildfire Dataset
by: Vuppari, Gowtham Raj, et al.
Published: (2025) -
A Lightweight Group Multiscale Bidirectional Interactive Network for Real-Time Steel Surface Defect Detection
by: Zhang, Yong, et al.
Published: (2025) -
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
by: Zhao, Shuxian, et al.
Published: (2026)