paper-with-me

홈 › Papers

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

2023-05-23 · Ning Ding, Yulin Chen, Bokai Xu, Yujia Qin, Zhi Zheng, Shengding Hu, Zhiyuan Liu, Maosong Sun, BoWen Zhou

Fine-tuning on instruction data has been widely validated as an effective practice for implementing chat language models like ChatGPT. Scaling the diversity and quality of such data, although straightforward, stands a great chance of leading to improved performance. This paper aims to improve the upper bound of open-source models further. We first provide a systematically designed, diverse, informative, large-scale dataset of instructional conversations, UltraChat, which does not involve human queries. Our objective is to capture the breadth of interactions that a human might have with an AI assistant and employs a comprehensive framework to generate multi-turn conversation iteratively. UltraChat contains 1.5 million high-quality multi-turn dialogues and covers a wide range of topics and instructions. Our statistical analysis of UltraChat reveals its superiority in various key metrics, including scale, average length, diversity, coherence, etc., solidifying its position as a leading open-source dataset. Building upon UltraChat, we fine-tune a LLaMA model to create a powerful conversational model, UltraLLaMA. Our evaluations indicate that UltraLLaMA consistently outperforms other open-source models, including Vicuna, the previously recognized state-of-the-art open-source model. The dataset and the model will be publicly released\footnote{\url{https://github.com/thunlp/UltraChat}}.

📄 PDF Abstract BibTeX arXiv:2305.14233

Code (1)

thunlp/ultrachat 공식 구현 pytorch

Tasks

Diversity

Similar Papers 제목 키워드 기반

Farmer.Chat: Scaling AI-Powered Agricultural Services for Smallholder Farmers

2024-09-13 · Namita Singh, Jacqueline Wang'ombe, Nereah Okanga, Tetyana Zelenska 외

Small and medium-sized agricultural holders face challenges like limited access to localized, timely information, impacting productivity and sustainability. Traditional extension services, which rely on in-person agents,…

Chatbot

Scaling Arabic Medical Chatbots Using Synthetic Data: Enhancing Generative AI with Synthetic Patient Records

2025-09-12 · Abdulrahman Allam, Seif Ahmed, Ali Hamdi, Khaled Shaban arxiv

The development of medical chatbots in Arabic is significantly constrained by the scarcity of large-scale, high-quality annotated datasets. While prior efforts compiled a dataset of 20,000 Arabic patient-doctor interacti…

Data Augmentation

Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data

2024-12-02 · Shuaijiang Zhao, Tingwei Guo, Bajian Xiang, Tongtang Wan 외

The GPT-4o represents a significant milestone in enabling real-time interaction with large language models (LLMs) through speech, its remarkable low latency and high fluency not only capture attention but also stimulate …

Language ModelingLanguage Modelling

SeDi-Instruct: Enhancing Alignment of Language Models through Self-Directed Instruction Generation

2025-02-07 · Jungwoo Kim, Minsang Kim, Sungjin Lee

The rapid evolution of Large Language Models (LLMs) has enabled the industry to develop various AI-based services. Instruction tuning is considered essential in adapting foundation models for target domains to provide hi…

Diversity

Enhancing conversational quality in language learning chatbots: An evaluation of GPT4 for ASR error correction

2023-07-19 · Long Mai, Julie Carson-Berndsen

The integration of natural language processing (NLP) technologies into educational applications has shown promising results, particularly in the language learning domain. Recently, many spoken open-domain chatbots have b…

Semantic Textual SimilaritySTS