paper-with-me

홈 › Papers

Learning Non-linguistic Skills without Sacrificing Linguistic Proficiency

2023-05-14 · Mandar Sharma, Nikhil Muralidhar, Naren Ramakrishnan

The field of Math-NLP has witnessed significant growth in recent years, motivated by the desire to expand LLM performance to the learning of non-linguistic notions (numerals, and subsequently, arithmetic reasoning). However, non-linguistic skill injection typically comes at a cost for LLMs: it leads to catastrophic forgetting of core linguistic skills, a consequence that often remains unaddressed in the literature. As Math-NLP has been able to create LLMs that can closely approximate the mathematical skills of a grade-schooler or the arithmetic reasoning skills of a calculator, the practicality of these models fail if they concomitantly shed their linguistic capabilities. In this work, we take a closer look into the phenomena of catastrophic forgetting as it pertains to LLMs and subsequently offer a novel framework for non-linguistic skill injection for LLMs based on information theoretic interventions and skill-specific losses that enable the learning of strict arithmetic reasoning. Our model outperforms the state-of-the-art both on injected non-linguistic skills and on linguistic knowledge retention, and does so with a fraction of the non-linguistic training data (1/4) and zero additional synthetic linguistic training data.

📄 PDF Abstract BibTeX arXiv:2305.08246

Code (1)

mandar-sharma/skill-lm 공식 구현 pytorch

Tasks

Arithmetic ReasoningMath

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models

2025-11-18 · Mohammad Zbeeb, Hasan Abed Al Kader Hammoud, Sina Mukalled, Nadine Rizk 외 arxiv

We present AraLingBench: a fully human annotated benchmark for evaluating the Arabic linguistic competence of large language models (LLMs). The benchmark spans five core categories: grammar, morphology, spelling, reading…

Reading Comprehension

SyntaxGym: An Online Platform for Targeted Evaluation of Language Models

2020-07-01 · ACL 2020 6 · Jon Gauthier, Jennifer Hu, Ethan Wilcox, Peng Qian 외

Targeted syntactic evaluations have yielded insights into the generalizations learned by neural network language models. However, this line of research requires an uncommon confluence of skills: both the theoretical know…

Experimental DesignLanguage ModelingLanguage Modelling

Coursebook Texts as a Helping Hand for Classifying Linguistic Complexity in Language Learners' Writings

2016-12-01 · WS 2016 12 · Ildik{\'o} Pil{\'a}n, David Alfter, Elena Volodina

We bring together knowledge from two different types of language learning data, texts learners read and texts they write, to improve linguistic complexity classification in the latter. Linguistic complexity in the foreig…

ClassificationDomain AdaptationGeneral Classification

Feature-based analysis of oral narratives from Afrikaans and isiXhosa children

2025-07-17 · Emma Sharratt, Annelien Smith, Retief Louw, Daleen Klop 외 arxiv

Oral narrative skills are strong predictors of later literacy development. This study examines the features of oral narratives from children who were identified by experts as requiring intervention. Using simple machine …

An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment

2025-08-27 · Tien-Hong Lo, Szu-Yu Chen, Yao-Ting Sung, Berlin Chen arxiv

A recent line of research on automated speaking assessment (ASA) has benefited from self-supervised learning (SSL) representations, which capture rich acoustic and linguistic patterns in non-native speech without underly…

Self-Supervised Learning