paper-with-me

홈 › Papers

What to Prioritize? Natural Language Processing for the Development of a Modern Bug Tracking Solution in Hardware Development

2021-09-28 · Thi Thu Hang Do, Markus Dobler, Niklas Kühl

Managing large numbers of incoming bug reports and finding the most critical issues in hardware development is time consuming, but crucial in order to reduce development costs. In this paper, we present an approach to predict the time to fix, the risk and the complexity of debugging and resolution of a bug report using different supervised machine learning algorithms, namely Random Forest, Naive Bayes, SVM, MLP and XGBoost. Further, we investigate the effect of the application of active learning and we evaluate the impact of different text representation techniques, namely TF-IDF, Word2Vec, Universal Sentence Encoder and XLNet on the model's performance. The evaluation shows that a combination of text embeddings generated through the Universal Sentence Encoder and MLP as classifier outperforms all other methods, and is well suited to predict the risk and complexity of bug tickets.

📄 PDF Abstract BibTeX arXiv:2109.13825

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningSentence

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Residual Connection 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
SentencePiece 설명 없음

Similar Papers 제목 키워드 기반

Phonemic Transcription of Low-Resource Languages: To What Extent can Preprocessing be Automated?

2020-05-01 · LREC 2020 5 · Guillaume Wisniewski, S{\'e}verine Guillaume, Alexis Michaud

Automatic Speech Recognition for low-resource languages has been an active field of research for more than a decade. It holds promise for facilitating the urgent task of documenting the world{'}s dwindling linguistic div…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Representing Affect Information in Word Embeddings

2022-09-21 · Yuhan Zhang, Wenqi Chen, Ruihan Zhang, Xiajie Zhang

A growing body of research in natural language processing (NLP) and natural language understanding (NLU) is investigating human-like knowledge learned or encoded in the word embeddings from large language models. This is…

Natural Language UnderstandingWord Embeddings

Is “good enough” good enough? Ethical and responsible development of sign language technologies

2021-08-01 · MTSummit 2021 8 · Maartje De Meulder

This paper identifies some common and specific pitfalls in the development of sign language technologies targeted at deaf communities, with a specific focus on signing avatars. It makes the call to urgently interrogate s…

Not Every AI Problem is a Data Problem: We Should Be Intentional About Data Scaling

2025-01-23 · Tanya Rodchenko, Natasha Noy, Nino Scherrer, Jennifer Prendki

While Large Language Models require more and more data to train and scale, rather than looking for any data to acquire, we should consider what types of tasks are more likely to benefit from data scaling. We should be in…

Quantum Natural Language Processing

2024-03-28 · Dominic Widdows, Willie Aboumrad, Dohun Kim, Sayonee Ray 외

Language processing is at the heart of current developments in artificial intelligence, and quantum computers are becoming available at the same time. This has led to great interest in quantum natural language processing…

Word Embeddings