paper-with-me

홈 › Papers

A Novel Approach to Part Name Discovery in Noisy Text

2018-06-01 · NAACL 2018 6 · Nobal Bikram Niraula, Daniel Whyatt, Anne Kao

As a specialized example of information extraction, part name extraction is an area that presents unique challenges. Part names are typically multi-word terms longer than two words. There is little consistency in how terms are described in noisy free text, with variations spawned by typos, ad hoc abbreviations, acronyms, and incomplete names. This makes search and analyses of parts in these data extremely challenging. In this paper, we present our algorithm, PANDA (Part Name Discovery Analytics), based on a unique method that exploits statistical, linguistic and machine learning techniques to discover part names in noisy text such as that in manufacturing quality documentation, supply chain management records, service communication logs, and maintenance reports. Experiments show that PANDA is scalable and outperforms existing techniques significantly.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

Discovery of interpretable structural model errors by combining Bayesian sparse regression and data assimilation: A chaotic Kuramoto-Sivashinsky test case

2021-10-01 · Rambod Mojgani, Ashesh Chattopadhyay, Pedram Hassanzadeh

Models of many engineering and natural systems are imperfect. The discrepancy between the mathematical representations of a true physical system and its imperfect model is called the model error. These model errors can l…

Equation Discovery

Frustratingly Easy Truth Discovery

2019-05-02 · Reshef Meir, Ofra Amir, Omer Ben-Porat, Tsviel Ben-Shabat 외

Truth discovery is a general name for a broad range of statistical methods aimed to extract the correct answers to questions, based on multiple answers coming from noisy sources. For example, workers in a crowdsourcing p…

Automated Early Leaderboard Generation From Comparative Tables

2018-02-13 · Mayank Singh, Rajdeep Sarkar, Atharva Vyas, Pawan Goyal 외

A leaderboard is a tabular presentation of performance scores of the best competing techniques that address a specific scientific problem. Manually maintained leaderboards take time to emerge, which induces a latency in …

Bidirectional LSTM for Named Entity Recognition in Twitter Messages

2016-12-01 · WS 2016 12 · Nut Limsopatham, Nigel Collier

In this paper, we present our approach for named entity recognition in Twitter messages that we used in our participation in the Named Entity Recognition in Twitter shared task at the COLING 2016 Workshop on Noisy User-g…

Feature Engineeringnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1

Named Entity Recognition in the Legal Domain using a Pointer Generator Network

2020-12-17 · Stavroula Skylaki, Ali Oskooei, Omar Bari, Nadja Herger 외

Named Entity Recognition (NER) is the task of identifying and classifying named entities in unstructured text. In the legal domain, named entities of interest may include the case parties, judges, names of courts, case n…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1