paper-with-me

홈 › Papers

From Test-Taking to Test-Making: Examining LLM Authoring of Commonsense Assessment Items

2024-10-18 · Melissa Roemmele, Andrew S. Gordon

LLMs can now perform a variety of complex writing tasks. They also excel in answering questions pertaining to natural language inference and commonsense reasoning. Composing these questions is itself a skilled writing task, so in this paper we consider LLMs as authors of commonsense assessment items. We prompt LLMs to generate items in the style of a prominent benchmark for commonsense reasoning, the Choice of Plausible Alternatives (COPA). We examine the outcome according to analyses facilitated by the LLMs and human annotation. We find that LLMs that succeed in answering the original COPA benchmark are also more successful in authoring their own items.

📄 PDF Abstract BibTeX arXiv:2410.14897

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Language Inference

Similar Papers 제목 키워드 기반

More Effective Ontology Authoring with Test-Driven Development

2018-12-14 · C. Maria Keet, Kieren Davies, Agnieszka Lawrynowicz

Ontology authoring is a complex process, where commonly the automated reasoner is invoked for verification of newly introduced changes, therewith amounting to a time-consuming test-last approach. Test-Driven Development …

test driven development

Developing a Corpus of Indirect Speech Act Schemas

2020-05-01 · LREC 2020 5 · Antonio Roque, Alex Tsuetaki, er, Vasanth Sarathy 외

Resolving Indirect Speech Acts (ISAs), in which the intended meaning of an utterance is not identical to its literal meaning, is essential to enabling the participation of intelligent systems in peoples{'} everyday lives…

Test-Driven Development of ontologies (extended version)

2015-12-19 · C. Maria Keet, Agnieszka Lawrynowicz

Emerging ontology authoring methods to add knowledge to an ontology focus on ameliorating the validation bottleneck. The verification of the newly added axiom is still one of trying and seeing what the reasoner says, bec…

test driven development

Querying Knowledge via Multi-Hop English Questions

2019-07-18 · Tiantian Gao, Paul Fodor, Michael Kifer

The inherent difficulty of knowledge specification and the lack of trained specialists are some of the key obstacles on the way to making intelligent systems based on the knowledge representation and reasoning (KRR) para…

BIG-bench Machine LearningQuestion Answering

AiBAT: Artificial Intelligence/Instructions for Build, Assembly, and Test

2024-10-03 · Benjamin Nuernberger, Anny Liu, Heather Stefanini, Richard Otis 외

Instructions for Build, Assembly, and Test (IBAT) refers to the process used whenever any operation is conducted on hardware, including tests, assembly, and maintenance. Currently, the generation of IBAT documents is tim…