paper-with-me

홈 › Papers

The Interplay Between Lexical and Syntactic Resources in Incremental Parsebanking

2014-05-01 · LREC 2014 5 · Victoria Ros{\'e}n, Petter Haugereid, Martha Thunes, Gyri S. Losnegaard, Helge Dyvik

Automatic syntactic analysis of a corpus requires detailed lexical and morphological information that cannot always be harvested from traditional dictionaries. In building the INESS Norwegian treebank, it is often the case that necessary lexical information is missing in the morphology or lexicon. The approach used to build the treebank is incremental parsebanking; a corpus is parsed with an existing grammar, and the analyses are efficiently disambiguated by annotators. When the intended analysis is unavailable after parsing, the reason is often that necessary information is not available in the lexicon. INESS has therefore implemented a text preprocessing interface where annotators can enter unrecognized words before parsing. This may concern words that are unknown to the morphology and/or lexicon, and also words that are known, but for which important information is missing. When this information is added, either during text preprocessing or during disambiguation, the result is that after reparsing the intended analysis can be chosen and stored in the treebank. The lexical information added to the lexicon in this way may be of great interest both to lexicographers and to other language technology efforts, and the enriched lexical resource being developed will be made available at the end of the project.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Optical Character Recognition (OCR)

Similar Papers 제목 키워드 기반

The Interplay between Lexical Resources and Natural Language Processing

2018-07-02 · NAACL 2018 6 · Jose Camacho-Collados, Luis Espinosa-Anke, Mohammad Taher Pilehvar

Incorporating linguistic, world and common sense knowledge into AI/NLP systems is currently an important research area, with several open problems and challenges. At the same time, processing and storing this knowledge i…

Common Sense Reasoning

Towards an environment for the production and the validation of lexical semantic resources

2014-05-01 · LREC 2014 5 · Mika{\"e}l Morardo, {\'E}ric Villemonte de la Clergerie

We present the components of a processing chain for the creation, visualization, and validation of lexical resources (formed of terms and relations between terms). The core of the chain is a component for building lexica…

Question Answering

Extracting semantic relations from Portuguese corpora using lexical-syntactic patterns

2014-05-01 · LREC 2014 5 · Raquel Amaro

The growing investment on automatic extraction procedures, together with the need for extensive resources, makes semi-automatic construction a new viable and efficient strategy for developing of language resources, combi…

Information RetrievalMachine TranslationWord Sense Disambiguation

Phrase Pair Mappings for Hindi-English Statistical Machine Translation

2017-10-05 · Sreelekha. S, Pushpak Bhattacharyya

In this paper, we present our work on the creation of lexical resources for the Machine Translation between English and Hindi. We describes the development of phrase pair mappings for our experiments and the comparative …

Machine TranslationTranslation

Acquisition of Syntactic Simplification Rules for French

2012-05-01 · LREC 2012 5 · Violeta Seretan

Text simplification is the process of reducing the lexical and syntactic complexity of a text while attempting to preserve (most of) its information content. It has recently emerged as an important research area, which h…

Information RetrievalMachine TranslationSentence CompressionText Simplification+1