paper-with-me

KIND

Kessler Italian Named-entities Dataset

홈페이지 · 논문 3편

KIND is an Italian dataset for Named-Entity Recognition. It contains more than one million tokens with the annotation covering three classes: persons, locations, and organizations. Most of the dataset (around 600K tokens) contains manual gold annotations in three different domains: news, literature, and political discourses.

Texts Italian