paper-with-me

Papers

SLICE: Supersense-based Lightweight Interpretable Contextual Embeddings

2020-12-01 · COLING 2020 8 · Cindy Aloui, Carlos Ramisch, Alexis Nasr, Lucie Barque

Contextualised embeddings such as BERT have become de facto state-of-the-art references in many NLP applications, thanks to their impressive performances. However, their opaqueness makes it hard to interpret their behaviour. SLICE is a hybrid model that combines supersense labels with contextual embeddings. We introduce a weakly supervised method to learn interpretable embeddings from raw corpora and small lists of seed words. Our model is able to represent both a word and its context as embeddings into the same compact space, whose dimensions correspond to interpretable supersenses. We assess the model in a task of supersense tagging for French nouns. The little amount of supervision required makes it particularly well suited for low-resourced scenarios. Thanks to its interpretability, we perform linguistic analyses about the predicted supersenses in terms of input word and context representations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Weight Decay 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Residual Connection 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Multi-Head Attention 설명 없음

Similar Papers 제목 키워드 기반

Changing the Basis of Contextual Representations with Explicit Semantics

2021-08-01 · ACL 2021 5 · Tam{\'a}s Ficsor, G{\'a}bor Berend

The application of transformer-based contextual representations has became a de facto solution for solving complex NLP tasks. Despite their successes, such representations are arguably opaque as their latent dimensions a…

Supersense Embeddings: A Unified Model for Supersense Interpretation, Prediction, and Utilization

2016-08-01 · ACL 2016 8 · Lucie Flekova, Iryna Gurevych
Dependency ParsingDocument ClassificationGeneral ClassificationInformation Retrieval+8

My Case, For an Adposition: Lexical Polysemy of Adpositions and Case Markers in Finnish and Latin

2022-06-01 · LREC 2022 6 · Daniel Chen, Mans Hulden

Adpositions and case markers contain a high degree of polysemy and participate in unique semantic role configurations. We present a novel application of the SNACS supersense hierarchy to Finnish and Latin data by manuall…

Clustering

Car Drag Coefficient Prediction from 3D Point Clouds Using a Slice-Based Surrogate Model

2026-01-05 · Utkarsh Singh, Absaar Ali, Adarsh Roy arxiv

The automotive industry's pursuit of enhanced fuel economy and performance necessitates efficient aerodynamic design. However, traditional evaluation methods such as computational fluid dynamics (CFD) and wind tunnel tes…

Point Clouds

Putting Context in SNACS: A 5-Way Classification of Adpositional Pragmatic Markers

2022-06-01 · LREC (LAW) 2022 6 · Yang Janet Liu, Jena D. Hwang, Nathan Schneider, Vivek Srikumar

The SNACS framework provides a network of semantic labels called supersenses for annotating adpositional semantics in corpora. In this work, we consider English prepositions (and prepositional phrases) that are chiefly p…