paper-with-me

홈 › Papers

Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4

2023-12-26 · Sondos Mahmoud Bsharat, Aidar Myrzakhan, Zhiqiang Shen

This paper introduces 26 guiding principles designed to streamline the process of querying and prompting large language models. Our goal is to simplify the underlying concepts of formulating questions for various scales of large language models, examining their abilities, and enhancing user comprehension on the behaviors of different scales of large language models when feeding into different prompts. Extensive experiments are conducted on LLaMA-1/2 (7B, 13B and 70B), GPT-3.5/4 to verify the effectiveness of the proposed principles on instructions and prompts design. We hope that this work can provide a better guide for researchers working on the prompting of large language models. Project page is available at https://github.com/VILA-Lab/ATLAS.

📄 PDF Abstract BibTeX arXiv:2312.16171

Code (2)

vila-lab/atlas 공식 구현
lastmile-ai/aiconfig

Tasks

All

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
Attention 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Multi-Head Attention 설명 없음

Similar Papers 제목 키워드 기반

How Do Language Models Process Ethical Instructions? Deliberation, Consistency, and Other-Recognition Across Four Models

2026-03-11 · Hiroki Fukui arxiv

Alignment safety research assumes that ethical instructions improve model behavior, but how language models internally process such instructions remains unknown. We conducted over 600 multi-agent simulations across four …

Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions

2025-08-22 · Akira Oyama, Shoichi Hasegawa, Akira Taniguchi, Yoshinobu Hagiwara 외 arxiv

Daily life support robots must interpret ambiguous verbal instructions involving demonstratives such as ``Bring me that cup,'' even when objects or users are out of the robot's view. Existing approaches to exophora resol…

Sound Source Localization

LLaMA-Omni: Seamless Speech Interaction with Large Language Models

2024-09-10 · Qingkai Fang, Shoutao Guo, Yan Zhou, Zhengrui Ma 외

Models like GPT-4o enable real-time interaction with large language models (LLMs) through speech, significantly enhancing user experience compared to traditional text-based interaction. However, there is still a lack of …

Structured Uncertainty guided Clarification for LLM Agents

2025-11-11 · Manan Suri, Puneet Mathur, Nedim Lipka, Franck Dernoncourt 외 arxiv

LLM agents with tool-calling capabilities often fail when user instructions are ambiguous or incomplete, leading to incorrect invocations and task failures. Existing approaches operate in unstructured language spaces, ge…

Reinforcement LearningQuestion Selection

PlatoLM: Teaching LLMs in Multi-Round Dialogue via a User Simulator

2023-08-21 · Chuyi Kong, Yaxin Fan, Xiang Wan, Feng Jiang 외

The unparalleled performance of closed-sourced ChatGPT has sparked efforts towards its democratization, with notable strides made by leveraging real user and ChatGPT dialogues, as evidenced by Vicuna. However, due to cha…

DiversityLanguage ModellingLarge Language Model