Leveraging Pre-trained Models for Failure Analysis Triplets Generation
Pre-trained Language Models recently gained traction in the Natural Language Processing (NLP) domain for text summarization, generation and question-answering tasks. This stems from the innovation introduced in Transformer models and their overwhelming performance compared with Recurrent Neural Network Models (Long Short Term Memory (LSTM)). In this paper, we leverage the attention mechanism of pre-trained causal language models such as Transformer model for the downstream task of generating Failure Analysis Triplets (FATs) - a sequence of steps for analyzing defected components in the semiconductor industry. We compare different transformer models for this generative task and observe that Generative Pre-trained Transformer 2 (GPT2) outperformed other transformer model for the failure analysis triplet generation (FATG) task. In particular, we observe that GPT2 (trained on 1.5B parameters) outperforms pre-trained BERT, BART and GPT3 by a large margin on ROUGE. Furthermore, we introduce Levenshstein Sequential Evaluation metric (LESE) for better evaluation of the structured FAT data and show that it compares exactly with human judgment than existing metrics.
Code (0)
등록된 구현이 없습니다.
Tasks
Question AnsweringText SummarizationTripletMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Linguistic Structures as Weak Supervision for Visual Scene Graph Generation
Prior work in scene graph generation requires categorical supervision at the level of triplets - subjects and objects, and predicates that relate them, either with or without bounding box information. However, scene grap…
Graph GenerationScene Graph GenerationMulti-Reward as Condition for Instruction-based Image Editing
High-quality training triplets (instruction, original image, edited image) are essential for instruction-based image editing. Predominant training datasets (e.g., InsPix2Pix) are created using text-to-image generative mo…
DescriptiveInstruction FollowingWeak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection
Vision-language (VL) models often exhibit a limited understanding of complex expressions of visual objects (e.g., attributes, shapes, and their relations), given complex and diverse language queries. Traditional approach…
Contrastive Learningobject-detectionObject DetectionSynthetic Data GenerationReMoT: Reinforcement Learning with Motion Contrast Triplets
We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a critical failure point in navigation, robotics, and autonomous driving. ReM…
Reinforcement LearningAutonomous DrivingEnhanced Data Transfer Cooperating with Artificial Triplets for Scene Graph Generation
This work focuses on training dataset enhancement of informative relational triplets for Scene Graph Generation (SGG). Due to the lack of effective supervision, the current SGG model predictions perform poorly for inform…
Graph GenerationScene Graph GenerationTriplet