paper-with-me

홈 › Papers

Does It Capture STEL? A Modular, Similarity-based Linguistic Style Evaluation Framework

2021-09-10 · EMNLP 2021 11 · Anna Wegmann, Dong Nguyen

Style is an integral part of natural language. However, evaluation methods for style measures are rare, often task-specific and usually do not control for content. We propose the modular, fine-grained and content-controlled similarity-based STyle EvaLuation framework (STEL) to test the performance of any model that can compare two sentences on style. We illustrate STEL with two general dimensions of style (formal/informal and simple/complex) as well as two specific characteristics of style (contrac'tion and numb3r substitution). We find that BERT-based methods outperform simple versions of commonly used style measures like 3-grams, punctuation frequency and LIWC-based approaches. We invite the addition of further tasks and task instances to STEL and hope to facilitate the improvement of style-sensitive measures.

📄 PDF Abstract BibTeX arXiv:2109.04817

Code (1)

nlpsoc/stel 공식 구현 tf

Similar Papers 제목 키워드 기반

lingvis.io - A Linguistic Visual Analytics Framework

2019-07-01 · ACL 2019 7 · Mennatallah El-Assady, Wolfgang Jentner, Fabian Sperrle, Rita Sevastjanova 외

We present a modular framework for the rapid-prototyping of linguistic, web-based, visual analytics applications. Our framework gives developers access to a rich set of machine learning and natural language processing st…

BIG-bench Machine Learning

Structural epitome: a way to summarize one’s visual experience

2010-12-01 · NeurIPS 2010 12 · Nebojsa Jojic, Alessandro Perina, Vittorio Murino

In order to study the properties of total visual input in humans, a single subject wore a camera for two weeks capturing, on average, an image every 20 seconds (www.research.microsoft.com/~jojic/aihs). The resulting new …

Clustering

Sentence Modeling via Multiple Word Embeddings and Multi-level Comparison for Semantic Textual Similarity

2018-05-21 · Huy Nguyen Tien, Minh Nguyen Le, Yamasaki Tomohiro, Izuha Tatsuya

Different word embedding models capture different aspects of linguistic properties. This inspired us to propose a model (M-MaxLSTM-CNN) for employing multiple sets of word embeddings for evaluating sentence similarity/re…

Natural Language InferenceRelationSemantic Textual SimilaritySentence+7

A Linguistics-Aware LLM Watermarking via Syntactic Predictability

2025-10-10 · Shinwoo Park, Hyejin Park, Hyeseon An, Yo-Sub Han arxiv

As large language models (LLMs) continue to advance rapidly, reliable governance tools have become critical. Publicly verifiable watermarking is particularly essential for fostering a trustworthy AI ecosystem. A central …

Linguistic Minimal Pairs Elicit Linguistic Similarity in Large Language Models

2024-09-19 · Xinyu Zhou, Delong Chen, Samuel Cahyawijaya, Xufeng Duan 외

We introduce a novel analysis that leverages linguistic minimal pairs to probe the internal linguistic representations of Large Language Models (LLMs). By measuring the similarity between LLM activation differences acros…

Semantic SimilaritySemantic Textual Similarity