Prompt Engineering and Calibration for Zero-Shot Commonsense Reasoning
Prompt engineering and calibration make large language models excel at reasoning tasks, including multiple choice commonsense reasoning. From a practical perspective, we investigate and evaluate these strategies on smaller language models. Through experiments on five commonsense reasoning benchmarks, we find that each strategy favors certain models, but their joint effects are mostly negative.
Code (0)
등록된 구현이 없습니다.
Tasks
Multiple-choicePrompt EngineeringSimilar Papers 제목 키워드 기반
Hint of Thought prompting: an explainable and zero-shot approach to reasoning tasks with LLMs
Prompting becomes an increasingly important research topic for better utilization of LLMs. Although simple prompting performs well on single-step questions, it cannot permanently activate the correct knowledge path for m…
Arithmetic ReasoningGSM8KLogical ReasoningMath+3Batch Calibration: Rethinking Calibration for In-Context Learning and Prompt Engineering
Prompting and in-context learning (ICL) have become efficient learning paradigms for large language models (LLMs). However, LLMs suffer from prompt brittleness and various bias factors in the prompt, including but not li…
image-classificationImage ClassificationIn-Context LearningNatural Language Understanding+1Zero-Shot Verification-guided Chain of Thoughts
Previous works have demonstrated the effectiveness of Chain-of-Thought (COT) prompts and verifiers in guiding Large Language Models (LLMs) through the space of reasoning. However, most such studies either use a fine-tune…
ConstraintChecker: A Plugin for Large Language Models to Reason on Commonsense Knowledge Bases
Reasoning over Commonsense Knowledge Bases (CSKB), i.e. CSKB reasoning, has been explored as a way to acquire new commonsense knowledge based on reference knowledge in the original CSKBs and external prior knowledge. Des…
Prompt EngineeringZero-Shot LearningZero-Shot Prompting for Implicit Intent Prediction and Recommendation with Commonsense Reasoning
Intelligent virtual assistants are currently designed to perform tasks or services explicitly mentioned by users, so multiple related domains or tasks need to be performed one by one through a long conversation with many…
Language ModelingLanguage Modelling