T2K\textasciicircum2: a System for Automatically Extracting and Organizing Knowledge from Texts
In this paper, we present T2K{\textasciicircum}2, a suite of tools for automatically extracting domain―specific knowledge from collections of Italian and English texts. T2K{\textasciicircum}2 (Text―To―Knowledge v2) relies on a battery of tools for Natural Language Processing (NLP), statistical text analysis and machine learning which are dynamically integrated to provide an accurate and incremental representation of the content of vast repositories of unstructured documents. Extracted knowledge ranges from domain―specific entities and named entities to the relations connecting them and can be used for indexing document collections with respect to different information types. T2K{\textasciicircum}2 also includes linguistic profiling functionalities aimed at supporting the user in constructing the acquisition corpus, e.g. in selecting texts belonging to the same genre or characterized by the same degree of specialization or in monitoring the added value of newly inserted documents. T2K{\textasciicircum}2 is a web application which can be accessed from any browser through a personal account which has been tested in a wide range of domains.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A New Method for Evaluating Automatically Learned Terminological Taxonomies
Abstract Evaluating a taxonomy learned automatically against an existing gold standard is a very complex problem, because differences stem from the number, label, depth and ordering of the taxonomy nodes. In this paper w…
GS^2: Graph-based Spatial Distribution Optimization for Compact 3D Gaussian Splatting
3D Gaussian Splatting (3DGS) has demonstrated breakthrough performance in novel view synthesis and real-time rendering. Nevertheless, its practicality is constrained by the high memory cost due to a huge number of Gaussi…
Novel View SynthesisAIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models
AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing them remains heavily manual, requiring practitioners to design archit…
Extracting Commonsense Properties from Embeddings with Limited Human Guidance
Intelligent systems require common sense, but automatically extracting this knowledge from text can be difficult. We propose and assess methods for extracting one type of commonsense knowledge, object-property comparison…
Active LearningCommon Sense ReasoningZero-Shot Learning