paper-with-me

홈 › Papers

Architectures of Meaning, A Systematic Corpus Analysis of NLP Systems

2021-07-16 · Oskar Wysocki, Malina Florea, Donal Landers, Andre Freitas

This paper proposes a novel statistical corpus analysis framework targeted towards the interpretation of Natural Language Processing (NLP) architectural patterns at scale. The proposed approach combines saturation-based lexicon construction, statistical corpus analysis methods and graph collocations to induce a synthesis representation of NLP architectural patterns from corpora. The framework is validated in the full corpus of Semeval tasks and demonstrated coherent architectural patterns which can be used to answer architectural questions on a data-driven fashion, providing a systematic mechanism to interpret a largely dynamic and exponentially growing field.

📄 PDF Abstract BibTeX arXiv:2107.08124

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Comparison by Conversion: Reverse-Engineering UCCA from Syntax and Lexical Semantics

2020-11-02 · COLING 2020 8 · Daniel Hershcovich, Nathan Schneider, Dotan Dvir, Jakob Prange 외

Building robust natural language understanding systems will require a clear characterization of whether and how various linguistic meaning representations complement each other. To perform a systematic comparative analys…

Natural Language UnderstandingSentence

Building a Chinese AMR Bank with Concept and Relation Alignments

2019-07-01 · LILT 2019 7 · Bin Li, Yuan Wen, Li Song, Weiguang Qu 외

Abstract Meaning Representation (AMR) is a meaning representation framework in which the meaning of a full sentence is represented as a single-rooted, acyclic, directed graph. In this article, we describe an on-going pro…

Abstract Meaning RepresentationRelationSentence

Characterizing Variation in Crowd-Sourced Data for Training Neural Language Generators to Produce Stylistically Varied Outputs

2018-09-14 · WS 2018 11 · Juraj Juraska, Marilyn Walker

One of the biggest challenges of end-to-end language generation from meaning representations in dialogue systems is making the outputs more natural and varied. Here we take a large corpus of 50K crowd-sourced utterances …

Text Generation

Spectral Signatures of Large Language Models

2026-07-03 · Zhuoying Zhang, Ishan V. Prasad, Yuanzhe Hu, Zihang Liu 외 arxiv

The rapidly growing repository of publicly available large language models (LLMs) presents significant challenges for systematic management and quantification at scale, such as model lineage tracing, licensing, and evalu…

MIST: a Large-Scale Annotated Resource and Neural Models for Functions of Modal Verbs in English Scientific Text

2022-12-14 · Sophie Henning, Nicole Macher, Stefan Grünewald, Annemarie Friedrich

Modal verbs (e.g., "can", "should", or "must") occur highly frequently in scientific articles. Decoding their function is not straightforward: they are often used for hedging, but they may also denote abilities and restr…

Articles