TSViT: A Time Series Vision Transformer for Fault Diagnosis

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Zhang, Shouhua, Zhou, Jiehan, Ma, Xue, Pirttikangas, Susanna, Yang, Chunsheng
Natura: Preprint
Pubblicazione: 2023
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910769965170688
author Zhang, Shouhua
Zhou, Jiehan
Ma, Xue
Pirttikangas, Susanna
Yang, Chunsheng
author_facet Zhang, Shouhua
Zhou, Jiehan
Ma, Xue
Pirttikangas, Susanna
Yang, Chunsheng
contents Traditional fault diagnosis methods using Convolutional Neural Networks (CNNs) often struggle with capturing the temporal dynamics of vibration signals. To overcome this, the application of Transformer-based Vision Transformer (ViT) methods to fault diagnosis is gaining attraction. Nonetheless, these methods typically require extensive preprocessing, which increases computational complexity, potentially reducing the efficiency of the diagnosis process. Addressing this gap, this paper presents the Time Series Vision Transformer (TSViT), tailored for effective fault diagnosis. TSViT incorporates a convolutional layer to extract local features from vibration signals, alongside a transformer encoder to discern long-term temporal patterns. A thorough experimental comparison on three diverse datasets demonstrates TSViT's effectiveness and adaptability. Moreover, the paper delves into the influence of hyperparameter tuning on the model's performance, computational demand, and parameter count. Remarkably, TSViT achieves an unprecedented 100% average accuracy on two test sets and 99.99% on another, showcasing its exceptional diagnostic capabilities.
format Preprint
id arxiv_https___arxiv_org_abs_2311_06916
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle TSViT: A Time Series Vision Transformer for Fault Diagnosis
Zhang, Shouhua
Zhou, Jiehan
Ma, Xue
Pirttikangas, Susanna
Yang, Chunsheng
Systems and Control
Artificial Intelligence
Traditional fault diagnosis methods using Convolutional Neural Networks (CNNs) often struggle with capturing the temporal dynamics of vibration signals. To overcome this, the application of Transformer-based Vision Transformer (ViT) methods to fault diagnosis is gaining attraction. Nonetheless, these methods typically require extensive preprocessing, which increases computational complexity, potentially reducing the efficiency of the diagnosis process. Addressing this gap, this paper presents the Time Series Vision Transformer (TSViT), tailored for effective fault diagnosis. TSViT incorporates a convolutional layer to extract local features from vibration signals, alongside a transformer encoder to discern long-term temporal patterns. A thorough experimental comparison on three diverse datasets demonstrates TSViT's effectiveness and adaptability. Moreover, the paper delves into the influence of hyperparameter tuning on the model's performance, computational demand, and parameter count. Remarkably, TSViT achieves an unprecedented 100% average accuracy on two test sets and 99.99% on another, showcasing its exceptional diagnostic capabilities.
title TSViT: A Time Series Vision Transformer for Fault Diagnosis
topic Systems and Control
Artificial Intelligence
url https://arxiv.org/abs/2311.06916