paper-with-me

Papers

ICE: Idiom and Collocation Extractor for Research and Education

2017-04-01 · EACL 2017 4 · Vasanthi Vuppuluri, Shahryar Baki, An Nguyen, Rakesh Verma

Collocation and idiom extraction are well-known challenges with many potential applications in Natural Language Processing (NLP). Our experimental, open-source software system, called ICE, is a python package for flexibly extracting collocations and idioms, currently in English. It also has a competitive POS tagger that can be used alone or as part of collocation/idiom extraction. ICE is available free of cost for research and educational uses in two user-friendly formats. This paper gives an overview of ICE and its performance, and briefly describes the research underlying the extraction algorithms.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

POSQuestion Answering

Similar Papers 제목 키워드 기반

Idiomify -- Building a Collocation-supplemented Reverse Dictionary of English Idioms with Word2Vec for non-native learners

2022-04-12 · Eu-Bin Kim

The aim of idiomify is to build a collocation-supplemented reverse dictionary of idioms for the non-native learners of English. We aim to do so because the reverse dictionary could help the non-natives explore idioms on …

Reverse Dictionary

Rolling the DICE on Idiomaticity: How LLMs Fail to Grasp Context

2024-10-21 · Maggie Mi, Aline Villavicencio, Nafise Sadat Moosavi

Human processing of idioms relies on understanding the contextual sentences in which idioms occur, as well as language-intrinsic features such as frequency and speaker-intrinsic factors like familiarity. While LLMs have …

Sentence

Dedicated Language Resources for Interdisciplinary Research on Multiword Expressions: Best Thing since Sliced Bread

2020-05-01 · LREC 2020 5 · Ferdy Hubers, Catia Cucchiarini, Helmer Strik

Multiword expressions such as idioms (beat about the bush), collocations (plastic surgery) and lexical bundles (in the middle of) are challenging for disciplines like Natural Language Processing (NLP), psycholinguistics …

Language Acquisition

Data-driven Identification of Idioms in Song Lyrics

2021-08-01 · ACL (MWE) 2021 8 · Miriam Amin, Peter Fankhauser, Marc Kupietz, Roman Schneider

The automatic recognition of idioms poses a challenging problem for NLP applications. Whereas native speakers can intuitively handle multiword expressions whose compositional meanings are hard to trace back to individual…

KoWit-24: A Richly Annotated Dataset of Wordplay in News Headlines

2025-03-03 · Alexander Baranov, Anna Palatkina, Yulia Makovka, Pavel Braslavski

We present KoWit-24, a dataset with fine-grained annotation of wordplay in 2,700 Russian news headlines. KoWit-24 annotations include the presence of wordplay, its type, wordplay anchors, and words/phrases the wordplay r…