paper-with-me

홈 › Papers

Evaluating the Role of Data Enrichment Approaches Towards Rare Event Analysis in Manufacturing

2024-07-01 · Chathurangi Shyalika, Ruwan Wickramarachchi, Fadi El Kalach, Ramy Harik, Amit Sheth

Rare events are occurrences that take place with a significantly lower frequency than more common regular events. In manufacturing, predicting such events is particularly important, as they lead to unplanned downtime, shortening equipment lifespan, and high energy consumption. The occurrence of events is considered frequently-rare if observed in more than 10% of all instances, very-rare if it is 1-5%, moderately-rare if it is 5-10%, and extremely-rare if less than 1%. The rarity of events is inversely correlated with the maturity of a manufacturing industry. Typically, the rarity of events affects the multivariate data generated within a manufacturing process to be highly imbalanced, which leads to bias in predictive models. This paper evaluates the role of data enrichment techniques combined with supervised machine-learning techniques for rare event detection and prediction. To address the data scarcity, we use time series data augmentation and sampling methods to amplify the dataset with more multivariate features and data points while preserving the underlying time series patterns in the combined alterations. Imputation techniques are used in handling null values in datasets. Considering 15 learning models ranging from statistical learning to machine learning to deep learning methods, the best-performing model for the selected datasets is obtained and the efficacy of data enrichment is evaluated. Based on this evaluation, our results find that the enrichment procedure enhances up to 48% of F1 measure in rare failure event detection and prediction of supervised prediction models. We also conduct empirical and ablation experiments on the datasets to derive dataset-specific novel insights. Finally, we investigate the interpretability aspect of models for rare event prediction, considering multiple methods.

📄 PDF Abstract BibTeX arXiv:2407.01644

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationEvent DetectionImputationPredictionTime Series

Similar Papers 제목 키워드 기반

A Study of Variable-Role-based Feature Enrichment in Neural Models of Code

2023-03-08 · Aftab Hussain, Md Rafiqul Islam Rabin, Bowen Xu, David Lo 외

Although deep neural models substantially reduce the overhead of feature engineering, the features readily available in the inputs might significantly impact training cost and the performance of the models. In this paper…

Feature Engineering

Mind the Gap: Data Enrichment in Dependency Parsing of Elliptical Constructions

2018-11-01 · WS 2018 11 · Kira Droganova, Filip Ginter, Jenna Kanerva, Daniel Zeman

In this paper, we focus on parsing rare and non-trivial constructions, in particular ellipsis. We report on several experiments in enrichment of training data for this specific construction, evaluated on five languages: …

Dependency Parsing

Enrichment Score: a better quantitative metric for evaluating the enrichment capacity of molecular docking models

2022-10-19 · Ian Scott Knight, Slava Naprienko, John J. Irwin

The standard quantitative metric for evaluating enrichment capacity known as $\textit{LogAUC}$ depends on a cutoff parameter that controls what the minimum value of the log-scaled x-axis is. Unless this parameter is chos…

Molecular Docking

VCU at Semeval-2016 Task 14: Evaluating definitional-based similarity measure for semantic taxonomy enrichment

2016-06-01 · SEMEVAL 2016 6 · Bridget McInnes
Information RetrievalSemantic Textual SimilarityWord Sense Disambiguation

Enriching the WebNLG corpus

2018-11-01 · WS 2018 11 · Thiago Castro Ferreira, Diego Moussallem, Emiel Krahmer, S Wubben 외

This paper describes the enrichment of WebNLG corpus (Gardent et al., 2017a,b), with the aim to further extend its usefulness as a resource for evaluating common NLG tasks, including Discourse Ordering, Lexicalization an…

Machine TranslationReferring ExpressionReferring expression generationText Generation+1