The LIMA Multilingual Analyzer Made Free: FLOSS Resources Adaptation and Correction
At CEA LIST, we have decided to release our multilingual analyzer LIMA as Free software. As we were not proprietary of all the language resources used we had to select and adapt free ones in order to attain results good enough and equivalent to those obtained with our previous ones. For English and French, we found and adapted a full-form dictionary and an annotated corpus for learning part-of-speech tagging models.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalMorphological AnalysisPart-Of-Speech TaggingSimilar Papers 제목 키워드 기반
A Morphological Analyzer for Gulf Arabic Verbs
We present CALIMAGLF, a Gulf Arabic morphological analyzer currently covering over 2,600 verbal lemmas. We describe in detail the process of building the analyzer starting from phonetic dictionary entries to fully inflec…
Morphological TaggingPart-Of-Speech TaggingPOSTAGAn Arabic Morphological Analyzer and Generator with Copious Features
We introduce CALIMA-Star, a very rich Arabic morphological analyzer and generator that provides functional and form-based morphological features as well as built-in tokenization, phonological representation, lexical rati…
From LIMA to DeepLIMA: following a new path of interoperability
In this article, we describe the architecture of the LIMA (Libre Multilingual Analyzer) framework and its recent evolution with the addition of new text analysis modules based on deep neural networks. We extended the fun…
FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation
Recent Open-Vocabulary Semantic Segmentation (OVSS) models extend the CLIP model to segmentation while maintaining the use of multiple templates (e.g., a photo of <class>, a sketch of a <class>, etc.) for constructing cl…
Open Vocabulary Semantic SegmentationOpen-Vocabulary Semantic SegmentationSemantic SegmentationNouveaut\'es de l'analyseur linguistique LIMA (What's New in the LIMA Language Analyzer)
LIMA est un analyseur linguistique libre d{'}envergure industrielle. Nous pr{\'e}sentons ici ses {\'e}volutions depuis la derni{\`e}re publication en 2014.