Web services and data mining: combining linguistic tools for Polish with an analytical platform
In this paper we present a new combination of existing language tools for Polish with a popular data mining platform intended to help researchers from digital humanities perform computational analyses without any programming. The toolset includes RapidMiner Studio, a software solution offering graphical setup of integrated analytical processes and Multiservice, a Web service offering access to several state-of-the-art linguistic tools for Polish. The setting is verified in a simple task of counting frequencies of unknown words in a small corpus.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Web Service integration platform for Polish linguistic resources
This paper presents a robust linguistic Web service framework for Polish, combining several mature offline linguistic tools in a common online platform. The toolset comprise paragraph-, sentence- and token-level segmente…
SentenceComputational Linguistics Applications for Multimedia Services
We present Computational Linguistics Applications for Multimedia Services (CLAMS), a platform that provides access to computational content analysis tools for archival multimedia material that appear in different media, …
Parallel Data, Tools and Interfaces in OPUS
This paper presents the current status of OPUS, a growing language resource of parallel corpora and related tools. The focus in OPUS is to provide freely available data sets in various formats together with basic annotat…
Machine TranslationTranslationWord Sense DisambiguationLanguage Processing Infrastructure in the XLike Project
This paper presents the linguistic analysis tools and its infrastructure developed within the XLike project. The main goal of the implemented tools is to provide a set of functionalities for supporting some of the main o…
Dependency ParsingSemantic Role LabelingWord Sense DisambiguationThe evolution of argumentation mining: From models to social media and emerging tools
Argumentation mining is a rising subject in the computational linguistics domain focusing on extracting structured arguments from natural text, often from unstructured or noisy text. The initial approaches on modeling ar…