paper-with-me

홈 › Papers

Comparing Text Representations: A Theory-Driven Approach

2021-09-15 · EMNLP 2021 11 · Gregory Yauney, David Mimno

Much of the progress in contemporary NLP has come from learning representations, such as masked language model (MLM) contextual embeddings, that turn challenging problems into simple classification tasks. But how do we quantify and explain this effect? We adapt general tools from computational learning theory to fit the specific characteristics of text datasets and present a method to evaluate the compatibility between representations and tasks. Even though many tasks can be easily solved with simple bag-of-words (BOW) representations, BOW does poorly on hard natural language inference tasks. For one such task we find that BOW cannot distinguish between real and randomized labelings, while pre-trained MLM representations show 72x greater distinction between real and random labelings than BOW. This method provides a calibrated, quantitative measure of the difficulty of a classification-based NLP task, enabling comparisons between representations without requiring empirical evaluations that may be sensitive to initializations and hyperparameters. The method provides a fresh perspective on the patterns in a dataset and the alignment of those patterns with specific labels.

📄 PDF Abstract BibTeX arXiv:2109.07458

Code (1)

gyauney/data-label-alignment 공식 구현

Tasks

Language ModelingLanguage ModellingLearning TheoryNatural Language Inference

Similar Papers 제목 키워드 기반

Anna Karenina Strikes Again: Pre-Trained LLM Embeddings May Favor High-Performing Learners

2024-06-06 · Abigail Gurin Schleifer, Beata Beigman Klebanov, Moriah Ariely, Giora Alexandron

Unsupervised clustering of student responses to open-ended questions into behavioral and cognitive profiles using pre-trained LLM embeddings is an emerging technique, but little is known about how well this captures peda…

Clustering

Data-Driven Spectral Prediction for Accelerating Large-Scale Electronic Structure Calculations

2026-05-29 · Abhiram Badrinarayanan, Davor Davidovic, Edoardo Di Napoli, Jurica Novak 외 arxiv

Simulating large molecular systems comprising thousands of atoms requires highly scalable methodologies. While modern Density Functional Theory (DFT) codes exhibit linear scaling, solving the associated large, sparse gen…

Theory of Mind for Multi-Agent Collaboration via Large Language Models

2023-10-16 · Huao Li, Yu Quan Chong, Simon Stepputtis, Joseph Campbell 외

While Large Language Models (LLMs) have demonstrated impressive accomplishments in both reasoning and planning, their abilities in multi-agent collaborations remains largely unexplored. This study evaluates LLM-based age…

HallucinationMulti-agent Reinforcement Learning

Gauge theory and twins paradox of disentangled representations

2019-06-24 · X. Dong, L. Zhou

Achieving disentangled representations of information is one of the key goals of deep network based machine learning system. Recently there are more discussions on this issue. In this paper, by comparing the geometric st…

BIG-bench Machine Learning

Comparing linear structure-based and data-driven latent spatial representations for sequence prediction

2019-08-19 · Myriam Bontonou, Carlos Lassance, Vincent Gripon, Nicolas Farrugia

Predicting the future of Graph-supported Time Series (GTS) is a key challenge in many domains, such as climate monitoring, finance or neuroimaging. Yet it is a highly difficult problem as it requires to account jointly f…

Time SeriesTime Series Analysis