paper-with-me

Papers

PhysNLU: A Language Resource for Evaluating Natural Language Understanding and Explanation Coherence in Physics

2022-01-12 · LREC 2022 6 · Jordan Meadows, Zili Zhou, Andre Freitas

In order for language models to aid physics research, they must first encode representations of mathematical and natural language discourse which lead to coherent explanations, with correct ordering and relevance of statements. We present a collection of datasets developed to evaluate the performance of language models in this regard, which measure capabilities with respect to sentence ordering, position, section prediction, and discourse coherence. Analysis of the data reveals equations and sub-disciplines which are most common in physics discourse, as well as the sentence-level frequency of equations and expressions. We present baselines that demonstrate how contemporary language models are challenged by coherence related tasks in physics, even when trained on mathematical natural language objectives.

📄 PDF Abstract BibTeX arXiv:2201.04275

Code (1)

jmeadows17/physnlu 공식 구현

Tasks

PositionSentenceSentence Ordering

Similar Papers 제목 키워드 기반

IndoNLU: Benchmark and Resources for Evaluating Indonesian Natural Language Understanding

2020-09-11 · Asian Chapter of the Association for Computational Linguistics 2020 · Bryan Wilie, Karissa Vincentio, Genta Indra Winata, Samuel Cahyawijaya 외

Although Indonesian is known to be the fourth most frequently used language over the internet, the research progress on this language in the natural language processing (NLP) is slow-moving due to a lack of available res…

BenchmarkingDiversityNatural Language UnderstandingSentence+1

BanglaNLG and BanglaT5: Benchmarks and Resources for Evaluating Low-Resource Natural Language Generation in Bangla

2022-05-23 · Abhik Bhattacharjee, Tahmid Hasan, Wasi Uddin Ahmad, Rifat Shahriyar

This work presents BanglaNLG, a comprehensive benchmark for evaluating natural language generation (NLG) models in Bangla, a widely spoken yet low-resource language. We aggregate six challenging conditional text generati…

Conditional Text GenerationDialogue GenerationLanguage ModelingLanguage Modelling+1

Consolidating and Developing Benchmarking Datasets for the Nepali Natural Language Understanding Tasks

2024-11-28 · Jinu Nyachhyon, Mridul Sharma, Prajwal Thapa, Bal Krishna Bal

The Nepali language has distinct linguistic features, especially its complex script (Devanagari script), morphology, and various dialects, which pose a unique challenge for natural language processing (NLP) evaluation. W…

BenchmarkingNatural Language InferenceNatural Language UnderstandingSentence+1

Bridging Language Gaps: Enhancing Few-Shot Language Adaptation

2025-08-26 · Philipp Borchert, Jochen De Weerdt, Marie-Francine Moens arxiv

The disparity in language resources poses a challenge in multilingual NLP, with high-resource languages benefiting from extensive data, while low-resource languages lack sufficient data for effective training. Our Contra…

Natural Language UnderstandingNatural Language InferenceCross-Lingual TransferContrastive Learning

On Evaluating and Mitigating Gender Biases in Multilingual Settings

2023-07-04 · Aniket Vashishtha, Kabir Ahuja, Sunayana Sitaram

While understanding and removing gender biases in language models has been a long-standing problem in Natural Language Processing, prior research work has primarily been limited to English. In this work, we investigate s…